2,326 to 2,350 of 3,805 Kubernetes Jobs

Site Reliability Engineer

Location
Fenny Stratford, England, United Kingdom
identify and prioritise reliability improvements Experience & Skills Required: Hands‐on experience with Azure Monitoring (Application Insights, Alerts, Action Groups) Strong knowledge of OpenTelemetry (including Kubernetes) Experience with Terraform and GitHub Actions Ability to define SLIs/SLOs and manage error budgets Incident response and post‐incident review experience Familiarity with … Docker and Kubernetes Strong communication and documentation skills Working knowledge of .NET/C# and React/NextJS Experience with cloud cost optimisation Knowledge of Azure networking (DNS, VNets, Firewalls) Understanding of security frameworks (e.g. ISO 27002, NIST CSF) About You: You may come from SRE, DevOps, platform engineering ...

AI Platform Engineer

Location
Knutsford, England, United Kingdom
high-performing capabilities that support millions of interactions every day. To be successful as an AI Platform Engineer, you should have experience with: Kubernetes platform engineering Good experience building, deploying and operating secure, scalable Kubernetes platforms. Cloud engineering across AWS and/or Azure Practical experience designing, automating and supporting … across the software delivery lifecycle. Some other highly valued skills may include OpenShift Experience building or operating platforms using OpenShift or a similar enterprise Kubernetes distribution. Agent security, observability, and evaluation: Knowledge of AI guardrails, evaluation frameworks, identity, access control, monitoring and operational assurance. Knowledge of agent evaluation, guardrails, identity ...

Site Reliability Engineer

Hiring Organisation
Connells Limited
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
identify and prioritise reliability improvements Experience & Skills Required: Hands-on experience with Azure Monitoring (Application Insights, Alerts, Action Groups) Strong knowledge of OpenTelemetry (including Kubernetes) Scripting/automation using PowerShell and/or Azure CLI Experience with Terraform and GitHub Actions Ability to define SLIs/SLOs and manage error … budgets Incident response and post-incident review experience Familiarity with Docker and Kubernetes Strong communication and documentation skills Desirable: Working knowledge of .NET/C# and React/NextJS Experience with cloud cost optimisation Knowledge of Azure networking (DNS, VNets, Firewalls) Understanding of security frameworks (e.g. ISO 27002, NIST ...

Observability Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
maintaining a broad range of observability backendsExtending and maintaining OpenTelemetry, including collectors and SDKsBuilding scalable telemetry pipelinesContributing to Golden Path SDKs and auto-instrumentationEnsuring Kubernetes workloads are fully observable and resilientEmbedding observability standards across platform and application teamsImproving incident response with better telemetry coverageProviding external industry observability experience and input … SaaS Observability platforms at scale, such as DataDog, NewRelic, Dynatrace etc. Strong hands-on experience with OpenTelemetryFamiliarity with public cloud infrastructure, ideally AWSProficiency in Kubernetes and DevOps tooling (e.g. Terraform, ArgoCD, Helm, Jenkins)Experience with metrics, logs and tracing backendsScripting in Go, Python, or similarIndustry background in Observability or SREDesirable ...

Senior DevOps Engineer CGEMJP

Hiring Organisation
Experis
Location
United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
Eligibility (Active SC Desired) Role purpose/summary ACP/Core Cloud Keycloak Entra S3 Mongo DB Postgres Github Actions CI/CD Kubernetes Docker Karpenter Data Migration experience (good to have) All profiles will be reviewed against the required skills and experience. Due to the high number of applications ...

Senior DevOps Engineer CGEMJP00350663

Hiring Organisation
Experis
Location
Nationwide, United Kingdom
Employment Type
Contract
Eligibility (Active SC Desired) Role purpose/summary ACP/Core Cloud Keycloak Entra S3 Mongo DB Postgres Github Actions CI/CD Kubernetes Docker Karpenter Data Migration experience (good to have) All profiles will be reviewed against the required skills and experience. Due to the high number of applications ...

Consulting Principal - Enterprise Architect – Hybrid Cloud

Location
Greater London, England, United Kingdom
patterns, with proven experience designing landing zones and core cloud building blocks (compute, storage, network) Hands-on understanding of containerization and orchestration (e.g. Docker, Kubernetes) within cloud-native and hybrid architectures Strong grasp of cloud security concepts, particularly role-based access control (RBAC), identity governance and least-privilege design, plus … equivalent enterprise architecture framework) certification Cloud vendor certifications at architect level (AWS/Azure/GCP) Container platform certifications (e.g. CKA) or hands-on Kubernetes delivery experience Experience with FinOps, sustainability or platform engineering initiatives We're excited to meet people who share our mission and can make an impact ...

Infrastructure Engineer – HPC & Compute

Location
Cambridgeshire and Peterborough, England, United Kingdom
interest in engineers who have supported high-performance or compute-intensive workloads. Key experience Strong Linux infrastructure experience Bare-metal server and compute environments Kubernetes and containerised infrastructure Terraform/Infrastructure-as-Code Public cloud infrastructure – ideally GCP or AWS Infrastructure automation Networking and storage CPU and/ ...

Senior Software Engineer, Full-Stack

Location
Greater London, England, United Kingdom
care more about depth than the specific language); experience with React is a plus Familiarity with cloud infrastructure (GCP preferred) and containerised deployment (Kubernetes) Experience building at scale in a fast-paced startup or high-growth environment Bonus: exposure to machine learning workflows or ML infrastructure Tech stack … looking for experience across all of these — as long as you're open to learning, Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most ...

Director of Software Engineering and ML - Agentic Commerce - Executive Director

Location
Greater London, England, United Kingdom
Life Cycle and an advanced understanding of agile methodologies such as CI/CD, Application Resiliency, and Security Practical cloud native experience Proficiency with Kubernetes and Amazon EKS, micro-VM isolation, and sidecar patterns, with the depth to direct security design across multiple layers of the stack, from network … skills Experience with agent frameworks and protocols (MCP, A2A, Google ADK, LangGraph) and agent evaluation practices Experience with Databricks, MLflow, and model serving on Kubernetes in AWS Experience delivering AI systems through model risk management or equivalent regulatory review Experience running a team within a federated platform model, contributing ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Location
Auchentibber, Scotland, United Kingdom
model serving infrastructure, bringing strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You’ll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. If you enjoy solving hard production problems and making platforms measurably better, you’ll find meaningful … incident management, root-cause analysis, runbooks, and reliability patterns Practical knowledge of observability and instrumentation across metrics, logs, and traces Hands-on experience with Kubernetes and container-based orchestration platforms, including managed cloud variants Experience hosting and serving large language models on cloud-based infrastructure and local GPU environments Knowledge ...

Software Engineer, GPU Infrastructure- ChatGPT Engineering

Location
Greater London, England, United Kingdom
Reliability Engineering (SRE), Infrastructure Engineering, or Platform Engineering. Have built software that automates operational workflows rather than relying on manual processes. Have experience with Kubernetes, Linux systems, container orchestration, or distributed infrastructure. Understand infrastructure observability, monitoring, capacity planning, and incident management. Enjoy identifying cross-team pain points and building reusable … Experience designing and operating highly available distributed systems. Experience with GPU infrastructure, high-performance computing, ML infrastructure, or large-scale compute platforms. Experience with Kubernetes, cloud infrastructure, Linux, networking, and observability tooling. Excellent debugging, systems design, and operational problem-solving skills. Strong communication skills and experience collaborating across engineering organizations. ...

Data Platform Engineer I

Location
Greater London, England, United Kingdom
helpful to have experience/expertise/knowledge in the following (in rough priority order): AWS Terraform/OpenTofu IAC Postgres CDC systems Kubernetes (EKS) Python Docker An interest or experience in data/network security would be a nice-to-have Datadog/Grafana/Prometheus Data related products … proactively to scope problems and solve and deliver pragmatic solutions Our Stack... Python as our main programming language Terraform for our infrastructure definition Kubernetes for data services and task orchestration Argo CD for application deployments Airflow purely for job scheduling and tracking CircleCI for continuous deployment Parquet and Delta file ...

Data Engineer

Location
Greater London, England, United Kingdom
experience Extensive knowledge of AWS/cloud services (ECS, EKS, Lambda, Kinesis, DynamoDB, EMR, Athena) Hands-on experience with orchestration and workflow tools (Docker, Kubernetes, Airflow, Prefect, Dagster) Familiar with CI/CD pipelines, infrastructure as code, and DBT Able to explain complex data concepts clearly to non-technical stakeholders ...

Senior Cloud Developer, Unified Supply Chain

Hiring Organisation
Aveva Group
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
issues, risks, and opportunities for improvement Essential requirements Experience and strong understanding of core Azure services: compute, storage, security, monitoring and diagnostics. Skilled in Kubernetes container orchestration platforms, including Azure Kubernetes Service (AKS) Highly proficient in C# with experience developing microservices, RESTful APIs, etc in a CI/CD environment. ...

Principal Software Engineer - Customer Platforms

Hiring Organisation
Marks & Spencer
Location
United Kingdom, UK
Employment Type
Full-time
them Ability to lead senior engineers and technical customers to a desired outcome, without prescribing it Authoritative skills at cloud computing (network, security, serverless, Kubernetes etc) and automation Experience with implementation of Observability and Reliability using market technologies (e.g.: New Relic) Good experience with Performance Engineering (load testing, derivations, tuning … ownership. Demonstrable entrepreneurship in previous organisation(s) Tech StackM&S uses a variety of technologies including;Java, Spring, SpringBOOT, MicronautReact, Next.js, Typescript, AngularAzure Cloud, Kubernetes, Dynatrace (observability)SQL Server, MongoDBIgnite, RedisWhat's In It For YouWorking at M&S means being part of something bigger - helping to deliver quality, value ...

Staff Platform Engineer

Location
Greater London, England, United Kingdom
implement self‐service capabilities, reusable golden paths, and standardised service templates for efficient software delivery. Provide technical direction across platform domains including compute, Kubernetes, networking, secrets management, observability, CI/CD, and infrastructure as code. Enable AI‐assisted engineering workflows to enhance productivity, automate routine activities, and improve decision‐making. … organisations. Skilled at identifying and prioritising high‐impact technical challenges aligned to platform strategy and business outcomes. Demonstrates strong technical expertise across cloud platforms, Kubernetes, infrastructure as code, CI/CD, observability, networking, and security. Proficient in at least one modern programming language such as Python, Go, or Java, with ...

SDE II, ML Infra Services, Annapurna Labs

Hiring Organisation
Amazon Development Center U.S., Inc
Location
Seattle, Washington, United States
Employment Type
Permanent
Salary
USD Annual
design ownership for SDEs Small, senior team: where every person owns major components and drives architectural decisions AI infrastructure: Work at the intersection of Kubernetes, custom silicon, and large-scale ML workloads Diverse Experiences We value diverse experiences and non-traditional career paths. If your career is just starting … complex software or computing infrastructure that has been successfully delivered to customers - Experience with AWS Services including EC2, Lambda, S3, DynamoDB, SQS - Experience in Kubernetes, Docker or containers ecosystem, or experience managing full application stacks from the OS up through custom applications and experience in any Bigdata architecture - Experience with ...

Aws DevOps Engineer

Location
Greater London, England, United Kingdom
Helm. Develop and optimize infrastructure-as-code using Terraform and related automation tools. Support the containerization and migration of legacy batch components into modern Kubernetes environments. Collaborate with application and infrastructure teams to improve build, deployment, and operational processes. Maintain and enhance version control, artifact registries, and pipeline integrations. Ensure ...

Software Engineer, New Grad - Production Infrastructure

Location
City Of London, England, United Kingdom
intelligence analysts and economic forecasters You’ll join our Production Infrastructure organisation, made up of small teams of engineers working on: Environment Platform: a Kubernetes-based PaaS spanning hundreds of production clusters Apollo: secure, fleet-wide deployment and change-management for complex microservice suites Signals: our full suite of observability … given problem. Right now, we use: A variety of languages, including Java and Go for backend and Typescript for frontend Open-source technologies like Kubernetes, Cilium, Envoy, Grafana, React, and Redux Industry-standard tooling, including Gradle and GitHub, and agentic tools like Windsurf & Cline What We Value Ability to communicate ...

DevSecOps Lead

Location
Greater London, England, United Kingdom
design principles Terraform and Infrastructure as Code CI/CD pipelines and automation AWS security, IAM and networking Security tooling and vulnerability management Kubernetes/EKS/ECS Leading and mentoring engineers Experience within government, financial services or regulated environments Desirable: Government security standards/NCSC principles ForgeRock, Okta, Auth0 ...

Senior DevOps Engineer

Location
Greater London, England, United Kingdom
operating in a Senior capacity in your current role Prior experience working with On‐Prem technologies Strong Containerisation/Orchestration work with Kubernetes Strong hands‐on experience with Terraform for IaC Wealth of work across CI/CD, Monitoring, and Automation Any background will be considered, but a Software/ ...

REST & .NET Developer (ONSITE – Basildon)

Location
United Kingdom
middleware • Ability to work in a hybrid setup (Basildon-based) Nice to have: • Experience with Azure (App Services, API Management) • Knowledge of containerisation (Docker, Kubernetes) • Familiarity with CI/CD pipelines and DevOps practices What we offer: • Competitive salary • Flexible hybrid working model • Opportunity to work on large-scale transformation ...

Software Manager - Remote

Location
United Kingdom
platform and on premise solutions in the Rail sector. The role combines people leadership, technical direction, and strategic ownership, with a strong focus on Kubernetes, C#/.NET, and shaping the adoption of AI-driven capabilities across our products and services. You will play a key role in defining ...

Senior Software Engineer - AI Foundations

Location
Greater London, England, United Kingdom
data-access rules into software you can maintain, working with Security and TechOps. Operate in AWS: Deploy and support services on AWS and Kubernetes, and make sound reliability, performance and cost trade-offs. Drive adoption: Work directly with engineers and non-technical teams, run workshops, write clear docs and turn … LangChain, MCP servers Knowledge and retrieval: RAG, search, embeddings, knowledge graphs Evals and safety: Inspect AI, Ragas, OpenAI Evals, NeMo Guardrails, red teaming Platform: Kubernetes, sandboxed or cloud dev environments, AWS Bedrock Backend: Django Observability: Datadog, OpenTelemetry Are you ready for a career with us? We want to ensure ...