51 to 75 of 175 Remote/Hybrid Prometheus Jobs

Senior Software Engineer

Location
United Kingdom
gaming industry or large‐scale interactive systems. Familiarity with messaging systems such as Kafka, RabbitMQ or similar Proficiency with monitoring tools like Prometheus, Grafana, or ELK Stack Experience with AI/ML integration in backend systems. Why Join Us At Lockwood, you’ll be part of an inclusive, creative ...

Senior Software Engineer

Location
United Kingdom
Desirable: Experience in the gaming industry or large-scale interactive systems. Familiarity with messaging systems such as Kafka,RabbitMQor similar Proficiencywith monitoring tools like Prometheus, Grafana, or ELK Stack Experience with AI/ML integration in backend systems. Why Join Us At Lockwood,you’llbe part of an inclusive, creative ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Embedded DevOps Engineer

Location
Harwell, England, United Kingdom
rigs. Familiarity with embedded Linux, cross-compilation toolchains, RTOS environments or FPGA development and deployment workflows. Experience with monitoring and observability platforms such asOpenTelemetry, Prometheus, Grafana or Loki. Knowledge of secure software supply-chain practices, including dependency scanning, artifact signing, software bills of materials, secrets management and vulnerability management. Experience ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
analyze internet traffic patterns across multiple dimensions using flow-based tools. Experience working with alerting, monitoring and visibility tools (such as Graphite/Grafana, Prometheus, or Splunk). Knowledge across cloud hosting solutions (i.e., GCP, AWS and Azure). Knowledge of DevOps practices and CI/CD pipelines (ie. ...

Embedded DevOps Engineer

Location
Kidlington, England, United Kingdom
rigs. Familiarity with embedded Linux, cross-compilation toolchains, RTOS environments or FPGA development and deployment workflows. Experience with monitoring and observability platforms such asOpenTelemetry, Prometheus, Grafana or Loki. Knowledge of secure software supply-chain practices, including dependency scanning, artifact signing, software bills of materials, secrets management and vulnerability management. Experience ...

NOC Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£50,000
administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with ...

Release Platform Engineer

Location
Stoke-on-Trent, England, United Kingdom
code using Terraform Experience with relational databases like PostgreSQL and AlloyDB Frontend development skills using React and TypeScript Understanding of observability tools such as Prometheus, Grafana, or Splunk Knowledge of security and compliance standards within source-code management Additional Information Operate and improve release and compliance platforms built in Python ...

Release Platform Engineer

Location
United Kingdom
code using Terraform Experience with relational databases like PostgreSQL and AlloyDB Frontend development skills using React and TypeScript Understanding of observability tools such as Prometheus, Grafana, or Splunk Knowledge of security and compliance standards within source-code management What you will be doing Operate and improve release and compliance platforms ...

Release Platform Engineer

Location
Manchester, England, United Kingdom
code using Terraform Experience with relational databases like PostgreSQL and AlloyDB Frontend development skills using React and TypeScript Understanding of observability tools such as Prometheus, Grafana, or Splunk Knowledge of security and compliance standards within source-code management What you will be doing Operate and improve release and compliance platforms ...

Senior Private Cloud Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
debug compute, networking, storage in large production environments."Nice To Have" Skills and Experience: Exposure to modern cloud native principles. Familiarity with observability tools (Prometheus, Grafana)!Exposure to large-scale or multi-tenant environments. Exposure to GitOps driven and CI/CD pipelines (Jenkins, ArgoCD. FluxCD, etc)Contribution to Open ...

KDB+ Developer

Location
Belfast, Northern Ireland, United Kingdom
quantitative finance, or IoT applications. Knowledge of DevOps practices, CI/CD pipelines, and containerization (Docker, Kubernetes). Familiarity with monitoring tools (Splunk, Grafana, Prometheus, etc.). Background in C++, Python, or Java for integration with KDB+. Location & Workplace Type: This position takes on a Hybrid working model based ...

Head of Site Reliability Engineering (SRE)

Hiring Organisation
Computershare
Location
Bristol, UK
Employment Type
Full-time
skills that you'll have: Robust knowledge of observability and monitoring practices, with hands-on experience implementing and managing platforms such as Dynatrace, Prometheus, Grafana, and Splunk. Good understanding of CI/CD tooling and modern software delivery practices, including Jenkins, GitLab CI, and Azure DevOps. A background spanning both ...

Software Engineering Manager (Test & Devops)

Location
Greater London, England, United Kingdom
Experience managing complex build environments (CMake, Conan, Ceedling) or cloud-native infrastructure-as-code (Terraform, CDK) Familiarity with observability tooling such as Grafana and Prometheus Benefits: Company equity plan so all employees share in the success of the company Salary-sacrifice pension scheme Private medical, dental and vision insurance (medical ...

DevOps Engineer

Location
Reigate, England, United Kingdom
Other highly desirable, but not essential skills are: Experience with: GitOps - ArgoCD or GitOps workflows Zero downtime deployments (blue/green/canary patterns) Prometheus/Grafana Microsoft Azure Certification (Az-104, Az-305, Az-400). Kubernetes Certification (CKA). Terraform Associate/OpenTofu equivalent (when available). Experience ...

Java Engineer, Associate

Hiring Organisation
Hackajob Ltd
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Work From Home
architecture, maintainability, and production readiness. Nice to Have - Experience with Kubernetes, Docker, or cloud-native environments (AWS/GCP). - Exposure to observability tools (Prometheus, Grafana, Open Telemetry). - Scripting experience in Python for automation or data analysis. - Exposure to Prompt Engineering and Agentic AI - Interest in financial systems, accounting ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
understanding of algorithms, data structures, software design, and distributed systems fundamentals. Experience with observability platforms, including distributed tracing, logging, metrics, and tools such as Prometheus, Grafana, ELK, or OpenTelemetry. Experience with site reliability engineering practices, relational databases, Hadoop, big data technologies, or cloud-native services. Familiarity with cloud-native architecture ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
performance troubleshooting within Linux environments Hands-on experience operating containerised workloads using Docker and Kubernetes Experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, OpenTelemetry or the ELK Stack Experience with Infrastructure as Code and automation tooling such as Terraform and Ansible Experience building and managing … Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft Azure/Azure/Prometheus/Grafana/Loki/OpenTelemetry/ELK Stack/GitLab CI/CD/GitHub Actions/Jenkins/ArgoCD/Helm/ ...

DevOps Engineer

Hiring Organisation
SF Partners
Location
Nationwide, United Kingdom
Employment Type
Permanent
Salary
£75000 - £110000/annum excellent training & progression
skills: - Deep technical ownership of IDP's for a large Developer user base - Knowledge of cloud native system design - Observability stack exposure - Grafana, Prometheus, Open Telemetry etc - Experience designing AWS and supporting AWS landing zones - IAC experience - Terraform, Ansible, Redhat etc - Strong experience building and owning multiple Kubernetes clusters ...

Remote Senior Site Reliability Engineer Manager (Remote)

Location
Cambourne, England, United Kingdom
infrastructure and services. Expertise in incident management, including incident response, resolution, and post-mortem analysis. Proficiency in monitoring, alerting, and observability tools such as Prometheus, Grafana, ELK stack or Datadog. Experience with cloud platforms such as AWS, Azure, or GCP, including infrastructure as code tools like Terraform or CloudFormation. Strong ...

Dev Ops Systems Administrator

Hiring Organisation
Proactive Appointments
Location
Woking, Surrey, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
configuration using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement and improve monitoring and observability using Grafana, Prometheus and CloudWatch Drive improvements in system reliability, performance, scalability and security Manage IAM, networking, firewalls, VPNs and cloud security Take ownership of complex infrastructure and cloud ...

Web: Full Stack Tech Lead

Location
Greater London, England, United Kingdom
/GKE), containerized with Docker. Own CI/CD (GitHub Actions), IaC (Terraform), logging/metrics/tracing ( OpenTelemetry , CloudWatch/Stackdriver, Grafana/Prometheus), and SLOs . Optimize p95 latency, throughput, and cost ; manage secrets, networking, VPCs, and build resilient retries/backoffs. 15% Collaborate Work closely with design ...

Platform Engineer

Location
Greater London, England, United Kingdom
networking and security fundamentals* Experience troubleshooting production systems and working through operational issues### Useful experience* Terragrunt or Helm* GitHub Actions, ArgoCD or Flux* Datadog, Prometheus or Grafana* PostgreSQL or MySQL* Supporting developer-experience or internal-platform improvements* Working with security tooling or policy-as-codeWe don’t expect every ...