1 to 25 of 87 Remote/Hybrid Prometheus Jobs in London

DevOps Engineer, Studios

Location
Greater London, England, United Kingdom
infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar. Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack. Ensure high availability and disaster recovery strategies are in place and tested regularly. Collaborate closely with development, QA, and operations teams ...

Cloud and Platform Engineer-Consultant-AI and Digital Factory

Location
Greater London, England, United Kingdom
operating against SLIs/SLOs/error budgets• Incident response, on-call practices, and post-incident review• Observability and monitoring: Grafana, Dynatrace, CloudWatch, Prometheus, Datadog, OpenTelemetryGeneral• Scripting (bash/shell)• Testing tooling: Selenium, Cucumber, etc.• Ability to flexibly support a variety of technologies (Java Spring Boot, NodeJS, SQL/NoSQL ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
least one).Containers & Orchestration: Docker, Kubernetes. Infrastructure as Code: Terraform, CloudFormation. Configuration Management: Ansible, Chef, or Puppet. Scripting: Bash, Shell, Python. Monitoring & Logging: Prometheus, Grafana, ELK stack, Splunk. Version Control: Git (GitHub/GitLab/Bitbucket).Security: IAM, secrets management, vulnerability scanning. Experience & Qualifications8+ years of experience in DevOps/ ...

DevOps Engineer

Location
Greater London, England, United Kingdom
tools such as Terraform or CloudFormation. Knowledge of CI/CD pipelines and deployment automation. Experience with monitoring and observability tools such as CloudWatch, Prometheus, Grafana, ELK, or similar. Good understanding of Linux systems administration. Experience troubleshooting complex application and infrastructure issues. Knowledge of networking principles including DNS, load balancing ...

Senior Azure DevOps Engineer

Location
Greater London, England, United Kingdom
tools such as GitLab CI/CD, GitHub Actions, Jenkins or similar You’ll need experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, Open You’ll need experience designing and implementing scalable systems capable of handling high workloads Role details Work model: Hybrid Location: London, hybrid ...

Platform Engineer

Location
Greater London, England, United Kingdom
tools such as Terraform or CloudFormation. Knowledge of CI/CD platforms and deployment automation. Experience with monitoring and observability tools such as CloudWatch, Prometheus, Grafana, ELK, or similar. Good understanding of Linux systems administration. Experience troubleshooting complex application and infrastructure issues. Knowledge of networking principles including DNS, load balancing ...

Platform Engineer

Location
Greater London, England, United Kingdom
Google Cloud Platform. Security: Experience with tools for delivering SCA, SAST, DAST capabilities. Monitoring and Logging: Proficiency with tools like Splunk, Dynatrace, Datadog, Prometheus, Grafana. Version Control: Strong understanding of Git and version control practices. Scripting: Skills in scripting languages like Bash, PowerShell, or Perl. Containerization: Familiarity with Docker ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
GitLab CI/CD, or CircleCI. Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch). Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems. Experience of Incident management and blameless ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
/CD, or CircleCI Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch Excellent analytical and troubleshooting abilities, especially within complex distributed systems Proven experience handling incident management and conducting blameless postmortems, including ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
payments, or enterprise SaaS platforms* Exposure to event-driven architecture (Kafka, RabbitMQ)* Familiarity with infrastructure-as-code tools (Terraform, CloudFormation)* Understanding of observability tools (Prometheus, Grafana, ELK stack) #J-18808-Ljbffr ...

Senior Software Engineer

Location
City of Westminster, England, United Kingdom
payments, or enterprise SaaS platforms Exposure to event-driven architecture (Kafka, RabbitMQ) Familiarity with infrastructure-as-code tools (Terraform, CloudFormation) Understanding of observability tools (Prometheus, Grafana, ELK stack) L’état d’esprit Edenred - Nous sommes une entreprise unique. Nous recherchons de nouveaux collaborateurs prêts à prendre part à l’aventure ...

SC / NPPV3 DevOps Engineer - Azure

Location
Greater London, England, United Kingdom
secure cloud landing zones, governance and infrastructure aligned with security and compliance requirements. Implement monitoring and observability using Azure Monitor, Log Analytics, Application Insights, Prometheus and Grafana. Automate infrastructure and operational processes using Python, PowerShell and Bash. Troubleshoot complex production issues and improve platform performance, resilience and availability. Work closely ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Hiring Organisation
FactSet Research Systems
Location
London, UK
Employment Type
Full-time
English both verbal and writtenundefinedAdditional Technical SkillsCloud Platforms:(e.g. AWS, GCP, Azure)CI/CD Tooling:(e.g. GitHub Actions, ArgoCD, Harness)Monitoring & Observability:(e.g. Prometheus, Grafana, Coralogix, OpenTelemetry)Infrastructure as Code:(e.g. Terraform, Pulumi)Config Management: (e.g. Ansible, Puppet, Chef)Programming/Scripting:(e.g. Python, Go, Bash)Soft Skills & General ...

Java Backend Developer

Location
City Of London, England, United Kingdom
Microservices and distributed application architecture. Experience with Docker and cloud-native application development. Familiarity with monitoring and observability tools such as Splunk, Dynatrace, Datadog, Prometheus, or Grafana. Experience with CI/CD tools such as Jenkins, GitLab CI, GitHub Actions, or Azure DevOps. Experience supporting business-critical production applications. Strong ...

Azure CloudOps Engineer

Location
Greater London, England, United Kingdom
supporting Kubernetes, Docker, and microservices-based environments. Experience operating hybrid cloud and on-premise infrastructure. Experience implementing observability platforms such as Splunk, Datadog, Grafana, Prometheus, Dynatrace, or similar technologies. Experience with AI-driven operational tooling and automated incident response solutions. Object-oriented programming experience. Experience working within Product Engineering, Software ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Location
Greater London, England, United Kingdom
written undefined Additional Technical Skills Cloud Platforms: (e.g. AWS, GCP, Azure) CI/CD Tooling: (e.g. GitHub Actions, ArgoCD, Harness) Monitoring & Observability: (e.g. Prometheus, Grafana, Coralogix, OpenTelemetry) Infrastructure as Code: (e.g. Terraform, Pulumi) Config Management: (e.g. Ansible, Puppet, Chef) Programming/Scripting: (e.g. Python, Go, Bash) Soft Skills & General Requirements ...

Test Environment Manager

Location
Greater London, England, United Kingdom
Linux shell scripting Strong understanding of CI/CD pipelines and tools (e.g., Jenkins, GitLab. ADS). Familiarity with monitoring and logging tools (e.g., Prometheus, Grafana. Splunk) Knowledge of UFT, Selenium, Azure Dev Ops Soft Skills: Excellent problem-solving and troubleshooting abilities. Strong communication and collaboration skills, with the ability ...

DevOps Engineer

Location
Greater London, England, United Kingdom
Networking: Strong understanding of IP networking, physical network topology, and network security. Desirable Skills and Experience Experience with monitoring and observability tooling such as Prometheus, Kibana, Nagios, or Splunk. Familiarity with backup and recovery platforms such as CommVault. Experience working in security-cleared or regulated environments (SC/DV). ...

Senior Software Developer (Python)

Location
Greater London, England, United Kingdom
similar frameworks. Strong SQL skills and experience with analytical databases such as BigQuery, PostgreSQL, or ClickHouse. Familiarity with observability and monitoring tools such as Prometheus, Grafana, Splunk, or OpenTelemetry. Experience deploying and supporting large-scale distributed systems in Kubernetes environments. Knowledge of graph analytics, network modelling, optimisation algorithms, or telecommunications ...

Senior System Engineer London

Location
Greater London, England, United Kingdom
Proven experience deploying, managing, and troubleshooting Kubernetes clusters in production environments. Strong understanding of observability and monitoring practices, including experience with tools such as Prometheus, Grafana, or similar platforms. Demonstrated ability to work confidently across both Unix based systems (Ubuntu, RHEL, or similar distributions) and Windows environments. Experience with scripting ...

DevOps Engineer

Location
Greater London, England, United Kingdom
frameworks relevant to cloud environments and experience with policy‐as‐code frameworks for automated compliance and guardrails. Exposure to observability platforms such as Datadog, Prometheus/Grafana, or the OpenTelemetry ecosystem. Experience with container image hardening and scanning (Trivy, Grype, or similar). Experience with using AI tooling as well ...

Infrastructure Engineer (Platform) - London

Location
Greater London, England, United Kingdom
within various deployment scenarios (multi-tenant and on-prem). Setup and maintain observability and monitoring strategies. Requirements: MUST HAVE Observability and monitoring (Datadog, Prometheus, Grafana, ...) Good knowledge of a modern programming language (ideally Python or JS/Typescript) NICE TO HAVE ML Ops or Data Engineering Experience architecting ...

Junior DevOps Engineer

Location
Greater London, England, United Kingdom
ADDS . Strong understanding of IP Networking and physical network setups. Desirable Skills and Experience: Exposure to tools such as CommVault , Nagios , Kibana , Prometheus , or Splunk . Ability to identify and communicate technical issues effectively, including applying root cause analysis. Strong communication and presentation skills, with the ability to influence ...

Java Software Engineer

Location
Greater London, England, United Kingdom
execute unit, integration, and performance tests using appropriate testing frameworks and tools. Monitor, troubleshoot, and optimize applications and infrastructure using New Relic, Grafana, Prometheus, and Bosun. Hands‐on experience with microservices architecture and Kafka. Excellent communication skills and ability to thrive in a fast-paced environment. Collaborate with cross-functional ...

Oracle OCI Multi Cloud Engineer

Location
City Of London, England, United Kingdom
provisioning Build and maintain CI/CD pipelines for infrastructure and database change deployment Implement monitoring and observability solutions (OCI Monitoring, Azure Monitor, CloudWatch, Prometheus/Grafana) Automate routine DBA and EBS administration tasks Required Skills & Experience Oracle & Database Solid experience in Oracle Database administration (11g, 12c, 19c) Strong experience ...