1 to 25 of 334 Permanent Prometheus Jobs in London

DevOps Engineer (Security Cleared)

Hiring Organisation
Solirius Consulting
Location
London, United Kingdom
Salary
£ 60 K
Python, Bash, PowerShell, or similar languages.Experience with configuration management tools such as Ansible, Puppet, or Chef.Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog.Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles.Experience working in Agile and DevOps delivery environments.BenefitsCompetitive SalaryBonus SchemePrivate ...

DevOps Engineer (Security Cleared)

Location
Greater London, England, United Kingdom
PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile and DevOps delivery environments. ...

DevOps Engineer, Studios

Hiring Organisation
iMG world
Location
London, United Kingdom
Salary
£ 60 K
teams.Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar.Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack.Ensure high availability and disaster recovery strategies are in place and tested regularly.Collaborate closely with development, QA, and operations teams to streamline ...

Senior DevOps Platform Engineer

Hiring Organisation
LEAP29
Location
London, United Kingdom
Salary
£ 80 K
healthcareExperience with GitOps practices and tools such as ArgoCDKnowledge of service mesh technologies including Istio or similarExperience with monitoring and observability platforms such as Prometheus, Grafana, Dynatrace, New Relic, Splunk or ELKExperience supporting hybrid cloud environments across AWS, Azure, GCP or private cloudExperience with security tooling including Vault, SonarQube, OWASP ...

DevOps Solution Architect

Location
London, United Kingdom
align multiple teams and stakeholders. Desirable Skills Experience defining platform strategies or roadmaps at an organisational level. Knowledge of observability and monitoring solutions (Prometheus, Grafana, ELK, Splunk). Experience implementing SRE practices (SLIs, SLOs, error budgets). Familiarity with compliance frameworks and continuous compliance tooling. Experience with large-scale ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, United Kingdom
Salary
£ 80 K
/Azure/GCP (at least one).Containers & Orchestration: Docker, Kubernetes.Infrastructure as Code: Terraform, CloudFormation.Configuration Management: Ansible, Chef, or Puppet.Scripting: Bash, Shell, Python.Monitoring & Logging: Prometheus, Grafana, ELK stack, Splunk.Version Control: Git (GitHub/GitLab/Bitbucket).Security: IAM, secrets management, vulnerability scanning.Experience & Qualifications8+ years of experience in DevOps/Site ...

Senior DevOps Engineer (Azure)

Hiring Organisation
Darktrace
Location
London, United Kingdom
Salary
£ 80 K
Kubernetes, along with supporting technologies like ArgoCD and Helm. You should also be familiar and well-versed in monitoring, logging, and observability tools including: Prometheus; Grafana; Loki; OpenTelemetry; and the ELK stack. Amongst this, you should be able to demonstrate:Proven experience as a DevOps Engineer, with a solid background ...

AWS DevOps Engineer

Hiring Organisation
83zero Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£75000 - £80000/annum benefits, perks, healthcare
PowerShell Good understanding of AWS networking, including VPCs, subnets, routing, load balancing and security groups Experience with monitoring and logging tools such as Prometheus, Grafana, ELK or CloudWatch Understanding of cloud security and Zero Trust principles Strong troubleshooting and problem-solving ability Experience working within Agile delivery environments Security Clearance ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
templates, Ansible, or similar. Experience with SQL Server, PostgreSQL, or other relational databases. Experience with monitoring tools such as Azure Monitor, Application Insights, Grafana, Prometheus, Datadog, or similar. Experience supporting change control, incident management, and audit-ready documentation. Technical Skills Preferred experience with: Python, PowerShell, Bash, Go, JavaScript, C# ...

Senior Azure DevOps Engineer

Location
Greater London, England, United Kingdom
tools such as GitLab CI/CD, GitHub Actions, Jenkins or similar You’ll need experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, Open You’ll need experience designing and implementing scalable systems capable of handling high workloads Role details Work model: Hybrid Location: London, hybrid ...

Senior DevOps Engineer | London, Hybrid | up to 125k

Hiring Organisation
Source Group International
Location
London, United Kingdom
Salary
£ 120 K
environments.Drive containerisation and orchestration using Docker and Kubernetes (including managed services such as EKS/AKS).Improve observability with monitoring, logging and alerting (e.g. Prometheus, Grafana, ELK/EFK, CloudWatch).Embed security best practice: IAM, secrets management, patching, vulnerability remediation and secure configuration.Support incident response and problem management, participating ...

Senior DevOps Engineer | London, Hybrid | up to £125k

Location
Greater London, England, United Kingdom
containerisation and orchestration using Docker and Kubernetes (including managed services such as EKS/AKS). Improve observability with monitoring, logging and alerting (e.g. Prometheus, Grafana, ELK/EFK, CloudWatch). Embed security best practice: IAM, secrets management, patching, vulnerability remediation and secure configuration. Support incident response and problem management ...

devops engineer for customs IT services

Location
Greater London, England, United Kingdom
skills with Docker and orchestration experience with Kubernetes Basic scripting skills in Python, Bash, or PowerShell Experience with monitoring and logging tools such as Prometheus, Grafana, CloudWatch, or ELK Experience using DevOps tools including Jira, Confluence, and Artifactory Nice to have: improving deployment workflows, automation to minimise manual tasks, secure ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
Establish benchmarking and telemetry to track unit economics and throughput for training and serving AI models. Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience ...

ML Ops Engineer

Hiring Organisation
Anaplan
Location
London, United Kingdom
Salary
£ 80 K
infrastructure waste.Establish benchmarking and telemetry to track unit economics and throughput for training and serving AI models.Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow.Your SkillsHands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated ...

AWS DevOps Engineer

Location
Greater London, England, United Kingdom
pipelines using GitLab CI, Jenkins, or ArgoCD Deploy and manage containerised applications with Docker and Kubernetes Implement observability and monitoring solutions using tools like Prometheus, Grafana, ELK, and CloudWatch Improve platform scalability, resilience, automation, and security posture Work within Agile delivery teams on large-scale digital transformation programmes Contribute ...

Platform Engineer

Location
Greater London, England, United Kingdom
tools such as Terraform or CloudFormation. Knowledge of CI/CD platforms and deployment automation. Experience with monitoring and observability tools such as CloudWatch, Prometheus, Grafana, ELK, or similar. Good understanding of Linux systems administration. Experience troubleshooting complex application and infrastructure issues. Knowledge of networking principles including DNS, load balancing ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
/CD, or CircleCI Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch Excellent analytical and troubleshooting abilities, especially within complex distributed systems Proven experience handling incident management and conducting blameless postmortems, including ...

AI Platform Engineer

Location
City Of London, England, United Kingdom
with AWS and/or Azure, Kubernetes, and Docker . Experience with GitOps, ideally Argo CD or Flux. Familiarity with observability tooling such as Prometheus, Grafana, Datadog, Splunk, ELK, or OpenTelemetry. Understanding of platform security, secrets management, policy-as-code, and SRE practices. Experience working in Legal, professional services ...

DevOps Engineer

Location
Greater London, England, United Kingdom
related tooling. Manage and improve Kubernetes platforms, GitOps workflows, Helm charts, and application onboarding. Develop monitoring, logging, alerting, and observability capabilities using Grafana, Prometheus, DataDog, Loki, and CloudWatch. Support cloud networking and connectivity across AWS accounts, regions, on-prem environments, and other cloud platforms we use. Implement and maintain secure ...

Software Engineering Tech Lead (SRE + AI)

Location
Greater London, England, United Kingdom
experience building LLM pipelines, AI Agents, Model Context Protocol (MCP) servers/clients, RAG architectures, and evaluation frameworks. Observability & Telemetry: Experience with OpenTelemetry (OTel), Prometheus, Grafana, Splunk, ThousandEyes, or distributed tracing systems. Cloud & Infrastructure: Expertise in public cloud providers (AWS, GCP, Azure), Terraform/IaC, and GitOps/CI/ ...

Software Engineering Tech Lead (SRE + AI)

Hiring Organisation
CISCO Systems
Location
London, United Kingdom
Salary
£ 80 K
experience building LLM pipelines, AI Agents, Model Context Protocol (MCP) servers/clients, RAG architectures, and evaluation frameworks.Observability & Telemetry: Experience with OpenTelemetry (OTel), Prometheus, Grafana, Splunk, ThousandEyes, or distributed tracing systems.Cloud & Infrastructure: Expertise in public cloud providers (AWS, GCP, Azure), Terraform/IaC, and GitOps/CI/CD pipelines ...

Lead Dev Ops Engineer

Location
Greater London, England, United Kingdom
ability to architect secure, performant, and highly available cloud solutions. Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence. Demonstrated leadership skills in managing complex, high‐pressure situations and guiding teams through incident ...

Director of DevOps & SRE

Location
Greater London, England, United Kingdom
tooling. Networking fundamentals (TCP/IP, DNS, load balancing) and the ability to partner effectively with network and security teams. Observability experience with Grafana, Prometheus, or equivalent, and a track record of building signal rather than noise. Comfort operating in a security and compliance-conscious environment: financial services, private markets ...

Devops Engineer

Location
Greater London, England, United Kingdom
Puppet, or similar) Solid scripting ability in Bash and at least one higher‐level language (Python preferred) Experience with monitoring and observability tooling (e.g. Prometheus, Grafana, Datadog, or similar) Strong incident diagnosis skills—able to work from vague symptoms to root cause using logs, metrics, and reasoning Comfortable working with ...