1 to 25 of 78 Remote/Hybrid Prometheus Jobs in London

DevOps Engineer, Studios

Hiring Organisation
iMG world
Location
London, United Kingdom
Salary
£ 60 K
teams.Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar.Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack.Ensure high availability and disaster recovery strategies are in place and tested regularly.Collaborate closely with development, QA, and operations teams to streamline ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, United Kingdom
Salary
£ 80 K
/Azure/GCP (at least one).Containers & Orchestration: Docker, Kubernetes.Infrastructure as Code: Terraform, CloudFormation.Configuration Management: Ansible, Chef, or Puppet.Scripting: Bash, Shell, Python.Monitoring & Logging: Prometheus, Grafana, ELK stack, Splunk.Version Control: Git (GitHub/GitLab/Bitbucket).Security: IAM, secrets management, vulnerability scanning.Experience & Qualifications8+ years of experience in DevOps/Site ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
least one).Containers & Orchestration: Docker, Kubernetes. Infrastructure as Code: Terraform, CloudFormation. Configuration Management: Ansible, Chef, or Puppet. Scripting: Bash, Shell, Python. Monitoring & Logging: Prometheus, Grafana, ELK stack, Splunk. Version Control: Git (GitHub/GitLab/Bitbucket).Security: IAM, secrets management, vulnerability scanning. Experience & Qualifications8+ years of experience in DevOps/ ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
templates, Ansible, or similar. Experience with SQL Server, PostgreSQL, or other relational databases. Experience with monitoring tools such as Azure Monitor, Application Insights, Grafana, Prometheus, Datadog, or similar. Experience supporting change control, incident management, and audit-ready documentation. Technical Skills Preferred experience with: Python, PowerShell, Bash, Go, JavaScript, C# ...

Senior Azure DevOps Engineer

Location
Greater London, England, United Kingdom
tools such as GitLab CI/CD, GitHub Actions, Jenkins or similar You’ll need experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, Open You’ll need experience designing and implementing scalable systems capable of handling high workloads Role details Work model: Hybrid Location: London, hybrid ...

Director, Applied AI & Agentic Platform Engineering

Location
Greater London, England, United Kingdom
Azure exposure optional (not a dependency) Containers: Docker, Kubernetes (GKE/EKS) IaC: Terraform CI/CD: GitHub Actions, Jenkins Observability: Splunk, ELK, Prometheus, Grafana Security & Compliance Secure coding, API security, Zero Trust Data privacy, encryption, access control Regulatory compliance and AI governance (MRM) What We Offer A leadership role ...

Full Stack Engineer - AI Enabled - Senior Vice President

Hiring Organisation
Citigroup
Location
London, United Kingdom
Salary
£ 80 K
strong sense of ownership.Preferred Qualifications:Experience in the financial services industry.Knowledge of domain-driven design and clean architecture principles.Experience with observability tools (e.g., Prometheus, Grafana, ELK stack).Contributions to open-source projects or active participation in developer communities.What we’ll provide you:By joining Citi London, you will not only ...

Full Stack Engineer - AI Enabled - Senior Vice President

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
ownership. Preferred Qualifications: Experience in the financial services industry. Knowledge of domain-driven design and clean architecture principles. Experience with observability tools (e.g., Prometheus, Grafana, ELK stack).Contributions to open-source projects or active participation in developer communities. What we'll provide you: By joining Citi London, you will ...

Lead Dev Ops Engineer

Location
Greater London, England, United Kingdom
ability to architect secure, performant, and highly available cloud solutions. Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence. Demonstrated leadership skills in managing complex, high‐pressure situations and guiding teams through incident ...

Lead Dev Ops Engineer

Hiring Organisation
Easyjet
Location
London, United Kingdom
Salary
£ 80 K
with the ability to architect secure, performant, and highly available cloud solutions.Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence.Demonstrated leadership skills in managing complex, high-pressure situations and guiding teams through incident ...

Senior Devops Engineer

Location
Greater London, England, United Kingdom
industry standard software Preferred/Bonus Experience with MLOps/LLMOps (Softwares such as Sagemaker, Kubeflow or ZenML). Deployment of on‐premise Kubernetes Prometheus (or other stacks) observability Experience with AWS Karpenter & Compute Optimizer Compliance literacy - ISO 27001, NIST SSDF/OWASP SAMM, GDPR basics Why Oxford Dynamics? Join ...

Senior Consultant, DE, TC, FS

Location
Greater London, England, United Kingdom
/DAST/SCA concepts, secrets management, policy-as-code awareness and secure SDLC practices Monitoring & operations Azure Monitor, Application Insights, Log Analytics, Grafana, Prometheus or comparable monitoring/logging tools, incident support and RCA Jira or Azure Boards, sprint ceremonies, backlog support, technical documentation, collaboration with distributed delivery teams ...

Lead Performance Test Engineer

Location
Greater London, England, United Kingdom
Ability to lead, coach and develop technical specialists while maintaining delivery quality and standards.## **Desirable Skills & Experience*** Experience with Dynatrace, AppDynamics, New Relic, Grafana, Prometheus or Azure Monitor.* Knowledge of Azure, AWS or Google Cloud Platform.* Experience with Kubernetes and Infrastructure as Code (IaC).* Understanding of Site Reliability Engineering ...

Intermediate/Senior DevOps

Hiring Organisation
Global Relay
Location
London, United Kingdom
Salary
£ 80 K
Podman, Kubernetes, VMWareOperating Systems: Linux or iOSContinuous Integration/Build and deployment automation: Jenkins, SonarQube, Artifactory, Bitbucket, Maven, Xcode Build, HelmInstrumentation and monitoring: Loki, Prometheus, Grafana, Mimir, Tempo TracingLanguages and frameworks: Bash, Java or Kotlin, Groovy, Python, ReactJS, SwiftWhere you have knowledge gaps, training and mentoring will be provided.About ...

Senior Systems Engineer, Production

Hiring Organisation
Clio
Location
London, United Kingdom
Salary
£ 100 K
infrastructure at scale.Strong expertise with Terraform and infrastructure automation.Solid understanding of AWS services such as ECS/EKS, RDS, LambdaFamiliarity with observability platforms (Datadog, Prometheus, Grafana, etc.).Proficiency in containerization (Docker, Kubernetes).Experience with CI/CD systems such as Buildkite, GitHub Actions, etc.Familiarity with Linux systems administration, networking ...

Senior Systems Engineer, Production

Hiring Organisation
Clio
Location
London, UK
Employment Type
Full-time
scale. Strong expertise with Terraform and infrastructure automation. Solid understanding of AWS services such as ECS/EKS, RDS, LambdaFamiliarity with observability platforms (Datadog, Prometheus, Grafana, etc.).Proficiency in containerization (Docker, Kubernetes).Experience with CI/CD systems such as Buildkite, GitHub Actions, etc. Familiarity with Linux systems administration, networking ...

Director, Applied AI & Agentic Platform Engineering

Hiring Organisation
Citigroup
Location
London, United Kingdom
Salary
£ 100 K
knowledge of AWS; Azure exposure optional (not a dependency)Containers: Docker, Kubernetes (GKE/EKS)IaC: TerraformCI/CD: GitHub Actions, JenkinsObservability: Splunk, ELK, Prometheus, GrafanaSecurity & ComplianceSecure coding, API security, Zero TrustData privacy, encryption, access controlRegulatory compliance and AI governance (MRM)What We OfferA leadership role in a high-priority ...

Site Reliability Engineer - Core

Hiring Organisation
Blockchain
Location
London, United Kingdom
Salary
£ 70 K
allocation, network and/or internals.Experience working with cloud solutions (GCP or AWS).Deep understanding and demonstrable experience with modern monitoring tools such as Prometheus, Datadog, Grafana, TelegrafExperience with infrastructure as code tools. Experience with complex Terraform deployments is a plus.Solid background with configuration management tools. Experience with Saltstack ...

Manager, System and Platform Operations

Hiring Organisation
Publicis Media
Location
London, United Kingdom
Salary
£ 80 K
.Solid understanding of networking, security, and system architecture.Proficient in scripting languages (Java, Golang, Python, Bash, or similar).Experience with monitoring and observability tools (DataDog, Prometheus, Grafana).Knowledge of database management systems (PostgreSQL, Bigtable).Understanding of API and microservices architecture.Strong people leadership skills with at least a year in leading ...

Site Reliability Engineer

Location
City of Westminster, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Senior MLOps Engineer

Hiring Organisation
MFK Recruitment
Location
London, United Kingdom
Salary
£ 80 K
skills, including multi-stage or multi-architecture builds.Experience building CI/CD pipelines for Machine Learning systems.Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog or similar.Linux systems administration and shell-scripting experience.Strong software engineering practices, including Git, testing and code reviews.Experience delivering AI or Machine Learning systems ...

Operations Engineering Lead

Hiring Organisation
Willis Towers Watson
Location
London, United Kingdom
Salary
£ 100 K
with Infrastructure as Code tools such as OpenTofu, Terraform, or Cloud formation (OpenTofu preferred)Expertise in observability - monitoring, logging, alerting, and APM (e.g. Datadog, Prometheus, Grafana, CloudWatch)Solid understanding of Docker and containerisation technologiesDeep CI/CD pipeline ownership and release governance in high-stakes environmentsKnowledge of AWS security best ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Operations Engineering Lead

Location
Greater London, England, United Kingdom
with Infrastructure as Code tools such as OpenTofu, Terraform, or Cloud formation (OpenTofu preferred) Expertise in observability - monitoring, logging, alerting, and APM (e.g. Datadog, Prometheus, Grafana, CloudWatch) Solid understanding of Docker and containerisation technologies Deep CI/CD pipeline ownership and release governance in high-stakes environments Knowledge ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
C++) with a bias toward automation.Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale.Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements.Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...