201 to 225 of 344 Prometheus Jobs in London

Principal DevOps Engineer

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
Salary
£ 120 K
lead the evolution of DevOps tools Kubernetes, Jenkins, Gitlab, Terraform, and more.Optimising automation and performance.Champion containerisation and high performance base images.Elevate monitoring systems Zabbix, Prometheus, Thanos ensuring 24/7 operational excellence.Secure infrastructure access management, balancing innovation with ironclad security. They're offering a career defining role in a company ...

Cloud Operations Engineer (remote – London)

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
Experienced and Certified in cloud computing with AWSExperience in a public cloud such as AWSKnowledge of monitoring and alerting technologies such as Grafana, Prometheus,Expereince of working with of Docker & Kubernetes and Container technology in productionWindows and Linux Operating System Management TechniquesSolid understanding of the OSI ModelExperience in database technology ...

Software Engineer — Observability Instrumentation

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
such as Terraform, ArgoCD, Helm or JenkinsInterest in AI engineering and SRE practices to improve incident response and RCADesirable but not essential experience includes:Prometheus/PromQL, VictoriaMetrics, OpenSearch, Grafana or similar observability backendsAuto-instrumentation, distributed tracing, structured logging or trace/metric correlationKafka or telemetry pipeline architecturesWhy should ...

Senior Site Reliability Engineer

Location
City Of London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, United Kingdom
Salary
£ 70 K
delivery speed and reducing deployment risk.Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Core AI Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
servicesFamiliarity with sandboxing and workload isolation technologiesExperience in quantitative finance or low-latency systemsAWS experience particularly in hybrid environmentsExperience with observability tooling such as Prometheus, Grafana or OpenTelemetryContributions to open-source projects in relevant domainsWhy join us Highly competitive compensation plus annual discretionary bonusLunch provided (via Just Eat for Business ...

Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
speed and reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
speed and reducing deployment risk.* **Adaptable & Problem-Solver**: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.* **Ownership & Quality**: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Site Reliability Engineer - Private Cloud Compute

Location
Greater London, England, United Kingdom
high-level programming language like: Java, Go, Python, or Perl Proclivity towards efficient programming emphasizing improvement via complexity analysis. Experience with Kubernetes, Nginx, Envoy, Prometheus, and/or Docker. Preferred Qualifications Understanding of standard networking protocols and components such as: HTTP, DNS, ECMP, TCP/IP, ICMP, the OSI Model ...

eTrading Platform Product Owner

Hiring Organisation
ING Banking
Location
London, United Kingdom
Salary
£ 80 K
supporting root-cause analysis and measurable improvement actions. Improve observability by strengthening meaningful monitoring, alerting, service-level measures and reporting using tools such as Prometheus and Grafana. Drive automation through Azure DevOps, pipelines, YAML and Ansible, reducing manual activity and improving repeatability, control and delivery confidence. Support robust change, release ...

Software Engineer, AI Libraries

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
Experience working with large GPU clusters or distributed training environments.Familiarity with distributed training techniques such as DDP or FSDP.Experience with observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry.Experience with data pipeline orchestration tools such as Airflow, Flyte, Ray, Metaflow, or Argo Workflows.Experience with containerisation and infrastructure tooling such ...

Neo4j Platform Consultant

Hiring Organisation
NTT DATA
Location
London, United Kingdom
Salary
£ 80 K
scaling)Security (RBAC, authentication, data protection)Preferred SkillsExperience with Graph RAG workloads and traversal optimizationDevOps tools (Docker, Kubernetes, CI/CD pipelines)Monitoring tools (Prometheus, Grafana, etc.)Experience with other graph databases (Neptune, TigerGraph)Job ExpectationsEnsure stable, secure, and high-performing Neo4j platform operationsEnable efficient graph query execution ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
Autonomous delivery and constructive collaboration with application engineers. Experience supporting business-critical services through an on-call rota. Added Bonus: Experience with Honeycomb, OpenTelemetry, Prometheus, Splunk, including SLO-led practices. Pragmatic use of AI-assisted engineering tools to improve quality and productivity. At Zopa we value flexible ways of working. ...

Software Engineer

Location
Greater London, England, United Kingdom
with large GPU clusters or distributed training environments. Familiarity with distributed training techniques such as DDP or FSDP. Experience with observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry. Experience with data pipeline orchestration tools such as Airflow, Flyte, Ray, Metaflow, or Argo Workflows. Experience with containerisation and infrastructure tooling ...

Infrastructure Software Engineering – Platform & Build

Location
Greater London, England, United Kingdom
extensibility, such as Bazel, Buck, Pants, Please, etc. Experience with infrastructure as code (Terraform, OpenTofu, or Pulumi) Experience with monitoring and observability tooling (Prometheus, Grafana, or similar) Working knowledge and practice of DevOps/SRE principles: SLOs, alerting design, incident management, and on‐call practice Strong proficiency in at least ...

Platform Engineer

Hiring Organisation
Lendable
Location
London, United Kingdom
Salary
£ 80 K
good grounding in Linux, networking and security fundamentalsExperience troubleshooting production systems and working through operational issuesUseful experienceTerragrunt or HelmGitHub Actions, ArgoCD or FluxDatadog, Prometheus or GrafanaPostgreSQL or MySQLSupporting developer-experience or internal-platform improvementsWorking with security tooling or policy-as-codeWe don’t expect every box to be ticked. ...

Lead SRE - Chase UK

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Lead SRE - Chase UK

Location
Westminster, West End, United Kingdom
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Platform Engineer

Location
Greater London, England, United Kingdom
hands-on with Kubernetes, and enough Terraform to have opinions about how to structure it. Familiarity with CI/CD pipelines and observability tooling (Prometheus, Grafana, Datadog, or similar). Strong technical foundation: Demonstrated ability to write production-quality code and solve hard technical problems (experience with Python, TypeScript/ ...

AI Engineer

Hiring Organisation
AECOM
Location
London, United Kingdom
Salary
£ 80 K
systems, Prompt engineering, Data processing Preferred Skills Understanding of optimizing both CPU-bound and GPU-bound workloads. Knowledge of monitoring and observability tools (e.g. Prometheus, Grafana, Elastic stack). Experience with Infrastructure as Code (Terraform). Experience with Azure Experience with Machine Learning Experience in building product for the construction ...

Principal Machine Learning Infrastructure Engineer London, United Kingdom

Location
Greater London, England, United Kingdom
consume data Experience building model serving infrastructure with latency and throughput requirements Familiarity with experiment tracking tools (Weights & Biases, MLflow) and observability stacks (Prometheus, Grafana) What we offer Equity options – share in our success and growth. 10% employer pension contribution – invest in your future. Free office lunches – great food ...

Engineer II, Site Reliability (Hybrid, London)

Location
Greater London, England, United Kingdom
drive to make things better Bias towards small development projects and the occasional larger projects Have experience with modern monitoring and telemetry stacks (ELK, Prometheus, Grafana, Zabbix) Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding Ability to lead incident analysis ...

Lead SRE - Chase UK

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 80 K
components, including service discovery, ingress, networking, and load balancing.Experience with Kubernetes.Experience with cloud computing services.Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger.Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Lead SRE - Chase UK

Location
Greater London, England, United Kingdom
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Software Engineer (ML Projects)

Location
Greater London, England, United Kingdom
cloud‐native TeamCity for CI/CD (lots of teams are releasing code 15-20 times per day!) Terraform Prometheus and Grafana If you have built and deployed complex Python applications or have hands‐on experience with generative AI and LLMs, we would be especially keen to talk. ...