Forward Deployed Infrastructure Engineer | Kubernetes | Distributed Sy

Job Description

Forward Deployed Infrastructure Engineer | Kubernetes | Distributed Systems | Python / Go | Must Hold or Be Eligible for SC ClearanceSalary: £90,000 - £105,000 + RSUsStart: ASAPWorking Model: Hybrid 1-2 days p/w in London (in-office for secure work and on-call)Location: London, UKRelocation: Support available for candidates outside LondonEligibility: Must hold active UK SC clearance, OR be eligible to obtain it (British Passport, ILR, or Settled Status) (Non-Negotiable)?? Non-Negotiables - Please only apply if you have ALL of the following Core technologies:Kubernetes and containerisation (Docker, orchestration at scale)Hands-on production infrastructure experienceInfrastructure as Code (Terraform, Ansible, or similar)Proficient in Python, Go, Java, or a comparable backend languageCI/CD pipeline design and deliveryEnvironment:Comfortable operating in a fast-paced, high-autonomy environmentProven track record building and deploying production systems, not just maintaining themAble to own reliability, monitoring, and operations end-to-endHappy to participate in an on-call rotationEligibility to hold UK SC clearance (British Passport, ILR, or Settled Status)Key skills:Production-grade infrastructure design, deployment, and scalingStrong Kubernetes and containerisation depthModern automation and IaC toolingSolid understanding of distributed systemsOn-call and incident response experienceConfident debugging and optimising across the stackAble to work UK hours and hybrid in LondonRole OverviewWe're partnered with a publicly listed technology company building mission-critical software for some of the most important institutions in the world. This is a forward-deployed infrastructure role where you'll operate at startup speed inside a high-security, high-impact environment - owning reliability, deployments, and automation for systems where uptime genuinely matters.This is a hands-on role focused on production infrastructure, system reliability, and scaling operations, working closely with product and delivery teams in a high-ownership setting with minimal bureaucracy.Key ResponsibilitiesBuild, operate, and maintain high-performance, scalable, and reliable production infrastructureOwn reliability end-to-end including monitoring, alerting, config management, and upgradesDeploy new products and run migrations across production environmentsLead automation efforts to reduce manual toil and improve resilienceDebug, harden, and optimise services with a focus on long-term reliabilityParticipate in an on-call rotation (roughly every 5-6 weeks) for production supportPartner with delivery and product teams on sensible, scalable systems designNice to HaveBackground in defence tech or another high-complexity technical organisationExperience at a major cloud or big-tech companyExposure to modern LLM / AI toolingDistributed systems design experiencePrior SRE or on-call rotation ownershipIf you've built and run production infrastructure at scale and want to work on systems that actually matter, let's talk.TPBN1_UKTJ

Job Details

Company
Optimal IT Recruitment Ltd
Location
London, UK
Employment Type
Full-time
Posted