production systems at scale and who focuses on making infrastructure predictable and stable. You’ll work across Kubernetes, networking, CI/CD, Cloudflare, and observability to create a platform engineers can trust. What You’ll Do Design, deploy, and maintain production Kubernetes clusters. Own cluster reliability, upgrades, security, and performance. … Build and operate monitoring, logging, and alerting pipelines. Ensure full-stack observability across infrastructure and services. Design and maintain CI/CD pipelines that are fast, reproducible, and safe. Improve deployment strategies (rollouts, canaries, rollbacks). Automate infrastructure provisioning and configuration. Investigate and resolve production incidents. Improve system resilience, redundancy ...