1 to 25 of 40 Remote/Hybrid Prometheus Jobs in London

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
templates, Ansible, or similar. Experience with SQL Server, PostgreSQL, or other relational databases. Experience with monitoring tools such as Azure Monitor, Application Insights, Grafana, Prometheus, Datadog, or similar. Experience supporting change control, incident management, and audit-ready documentation. Technical Skills Preferred experience with: Python, PowerShell, Bash, Go, JavaScript, C# ...

Senior Azure DevOps Engineer

Location
Greater London, England, United Kingdom
tools such as GitLab CI/CD, GitHub Actions, Jenkins or similar You’ll need experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, Open You’ll need experience designing and implementing scalable systems capable of handling high workloads Role details Work model: Hybrid Location: London, hybrid ...

Lead Dev Ops Engineer

Location
Greater London, England, United Kingdom
ability to architect secure, performant, and highly available cloud solutions. Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence. Demonstrated leadership skills in managing complex, high‐pressure situations and guiding teams through incident ...

Senior Devops Engineer

Location
Greater London, England, United Kingdom
industry standard software Preferred/Bonus Experience with MLOps/LLMOps (Softwares such as Sagemaker, Kubeflow or ZenML). Deployment of on‐premise Kubernetes Prometheus (or other stacks) observability Experience with AWS Karpenter & Compute Optimizer Compliance literacy - ISO 27001, NIST SSDF/OWASP SAMM, GDPR basics Why Oxford Dynamics? Join ...

DevOps Engineer

Location
Greater London, England, United Kingdom
Networking: Strong understanding of IP networking, physical network topology, and network security. Desirable Skills and Experience Experience with monitoring and observability tooling such as Prometheus, Kibana, Nagios, or Splunk. Familiarity with backup and recovery platforms such as CommVault. Experience working in security-cleared or regulated environments (SC/DV). ...

Senior Consultant, DE, TC, FS

Location
Greater London, England, United Kingdom
/DAST/SCA concepts, secrets management, policy-as-code awareness and secure SDLC practices Monitoring & operations Azure Monitor, Application Insights, Log Analytics, Grafana, Prometheus or comparable monitoring/logging tools, incident support and RCA Jira or Azure Boards, sprint ceremonies, backlog support, technical documentation, collaboration with distributed delivery teams ...

Lead Performance Test Engineer

Location
Greater London, England, United Kingdom
Ability to lead, coach and develop technical specialists while maintaining delivery quality and standards.## **Desirable Skills & Experience*** Experience with Dynatrace, AppDynamics, New Relic, Grafana, Prometheus or Azure Monitor.* Knowledge of Azure, AWS or Google Cloud Platform.* Experience with Kubernetes and Infrastructure as Code (IaC).* Understanding of Site Reliability Engineering ...

Site Reliability Engineer

Location
City of Westminster, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
analyze internet traffic patterns across multiple dimensions using flow-based tools. Experience working with alerting, monitoring and visibility tools (such as Graphite/Grafana, Prometheus, or Splunk). Knowledge across cloud hosting solutions (i.e., GCP, AWS and Azure). Knowledge of DevOps practices and CI/CD pipelines (ie. ...

Software Engineering Manager (Test & Devops)

Location
Greater London, England, United Kingdom
Experience managing complex build environments (CMake, Conan, Ceedling) or cloud-native infrastructure-as-code (Terraform, CDK) Familiarity with observability tooling such as Grafana and Prometheus Benefits: Company equity plan so all employees share in the success of the company Salary-sacrifice pension scheme Private medical, dental and vision insurance (medical ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
performance troubleshooting within Linux environments Hands-on experience operating containerised workloads using Docker and Kubernetes Experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, OpenTelemetry or the ELK Stack Experience with Infrastructure as Code and automation tooling such as Terraform and Ansible Experience building and managing … Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft Azure/Azure/Prometheus/Grafana/Loki/OpenTelemetry/ELK Stack/GitLab CI/CD/GitHub Actions/Jenkins/ArgoCD/Helm/ ...

AWS Architect

Location
London, United Kingdom
CodePipeline, GitLab, Jenkins) with full auditability, rollback mechanisms, and artifact management. Monitoring and Reliability: Implement centralized logging, tracing, and proactive monitoring using AWS CloudWatch, Prometheus, Grafana, and ELK Stack to maintain strict uptime SLAs. Stakeholder Engagement: Act as an AWS container subject matter expert, delivering technical designs and guiding multi ...

Lead DevSecOps Engineer

Location
Greater London, England, United Kingdom
e.g., EC2 to EKS, or cross-cloud) with a focus on data integrity and minimal downtime Ability to implement standardized telemetry pipelines (e.g., OpenTelemetry, Prometheus, or ELK) that provide developers with out-of-the-box visibility into their services Familiarity with automated policy enforcement and compliance-as-code (e.g. ...

DevOps Engineer - Cloud Trading Infra (Hybrid, London)

Location
Greater London, England, United Kingdom
research and trading workflows. The ideal candidate has hands-on AWS, IaC (Terraform/Ansible/Puppet), container tech (Docker/Kubernetes), and monitoring (Prometheus/Grafana) experience. Familiarity with PostgreSQL/MySQL databases is a plus. #J-18808-Ljbffr ...

Software engineering specialist

Hiring Organisation
Randstad Digital
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £700 per day
deployments, and tool setups directly into developer workflows. Tech Stack Delivery: Automate the provisioning and configuration of our operational ecosystem, including tools across observability ( Prometheus, Elastic, Checkmk ), security/compliance ( Tenable, Red Hat Satellite ), service registry/IPAM ( NetBox ), container orchestration ( ArgoCD ), and artifact management ( Artifactory, GitLab ). Engineering Standards ...

Platform Engineer AWS IaC

Hiring Organisation
Client Server
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
seeking to improve operational efficiency through scripting, tooling and automation. You'll monitor and optimise AWS environments for performance, reliability and cost using CloudWatch, Prometheus, Grafana and other observability tools, implement AWS security best practice, including IAM, roles, policies and least privilege access. As a senior member of the team ...

Lead Java Developer

Location
Greater London, England, United Kingdom
gRPC etc Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc. Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches Hands‐on knowledge ...

Lead Java Developer

Location
Greater London, England, United Kingdom
gRPC etc* Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning* Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc.* Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches* Hands-on knowledge ...

Principal AI Quality Engineer

Location
Greater London, England, United Kingdom
practices at the leading edge. Experience with CI/CD tooling, e.g. Jenkins, Azure DevOps, and Octopus Deploy. Experience with observability tooling such as Prometheus, Grafana, and Sumo Logic. Benefits Holidays. We all need to rest so you get 25 basic holidays with the option to grow ...

Observability SME/Architect/Consultant

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
Making Cross-functional Collaboration Technical ExpertiseCandidates should demonstrate experience with one or more of the following technologies and platforms:Observability Platforms Dynatrace Splunk Grafana Prometheus Elastic/ELK Stack AppDynamics New Relic Service Management & IT Operations ServiceNow ServiceNow Event Management ServiceNow ITOM CMDB and Dependency Mapping Solutions Cloud Monitoring Azure ...

Python Backend Developer

Location
Greater London, England, United Kingdom
frontend work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Senior Backend Engineer | AI Platform

Location
Greater London, England, United Kingdom
organization. Tech Stack: Backend Python FastAPI Agent Development Kit (ADK) Datastores PostgreSQL BigQuery Firestore Infrastructure Google Cloud Platform (GCP) RabbitMQ Terraform Monitoring & Observability Grafana Prometheus Langfuse incident.io Sentry What to Expect from Our Hiring Process At Plum, we value a lot the time you devote to the hiring process, this ...

Platform Engineer

Location
Greater London, England, United Kingdom
Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest … supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow ...