51 to 75 of 124 Remote/Hybrid Prometheus Jobs

Senior Cloud Architect (all genders)

Hiring Organisation
Lam Research
Location
Villach, Kärnten, Austria
Employment Type
Permanent
Salary
EUR Annual
policy-as-code. Proficiency with CI/CD (GitHub Actions, Azure DevOps, GitLab CI), artifact registries, and pipeline security. Experience with observability (OpenTelemetry, Prometheus/Grafana, Azure Monitor, CloudWatch), SRE practices, and production operations. Excellent communication, influence, and executive storytelling skills; ability to set strategy and lead hands-on. Core ...

Principal Infrastructure Engineer

Hiring Organisation
Sidram tech
Location
San Francisco, California, United States
Employment Type
Permanent
Salary
USD Annual
data and ML platforms including Snowflake and Databricks. Experience with cloud networking (AWS VPC, Azure VNet, GCP VPC). Experience with observability tools (Prometheus, ELK Stack, Grafana, or similar). Experience with service mesh technologies such as Istio or Linkerd. Strong understanding of cloud security, scalability, reliability, and distributed systems. ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Southampton, Hampshire, United Kingdom
Employment Type
Permanent
Salary
£60000/annum Bonus, Pension, Healthcare
administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with ...

Head of Site Reliability Engineering (SRE)

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Terraform and Ansible Automation Platform. Other key skills: Robust knowledge of observability and monitoring practices, and experience implementing and managing platforms such as Dynatrace, Prometheus, Grafana, and Splunk. Good understanding of CI/CD tooling and modern software delivery practices, including Jenkins, GitLab CI, and Azure DevOps. Background spanning both ...

Staff Cloud SRE - AI/ML Platform & GPU Compute

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skills Familiarity with infrastructure‐as‐code (e.g. ...

Head of Engineering

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
DevOps AWS (EKS, Lambda, RDS/Postgres, DynamoDB, SQS, Kinesis, S3, Cognito, Route53, VPC, EC2) Terraform, Kubernetes, Helm, Ansible, Puppet GitLab CI Observability & Data Prometheus, Grafana, OpenSearch Snowflake You don’t need experience in every tool, but you should feel confident in modern cloud-native, infrastructure-as-code environments. ...

Cloud Operations Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
processes**: Collaborate with the team to continuously refine operational processes and documentation.* **Maintain observability tools**: Manage and operate monitoring and observability tools like Graphite, Prometheus, Grafana, Elastic, Nagios, and LogScale.* **Support engineering teams**: Provide exceptional support to internal Product and Engineering teams, meeting their requirements for the Mimecast Cloud Platform. ...

Team Lead - Software Engineering

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
React (Frontend). Infrastructure: GCP, Docker, Kubernetes, Terraform, Jenkins, and GitHub Actions. Data & Storage: PostgreSQL and Google Cloud Storage (GCS). Observability: ELK Stack, Prometheus, and Grafana. What will help you thrive as a Team Lead Our engineers do not just simplify problems – they make deep technical complexity navigable, scalable ...

Senior Platform Engineer

Hiring Organisation
Capgemini
Location
City and Borough of Birmingham, United Kingdom
Employment Type
Full Time
self-service automation that boost developer experience. • Deploying and operating hardened Kubernetes platforms (EKS/AKS). • Creating unified observability stacks with OpenTelemetry, Prometheus and Grafana. • Delivering cloud-native modernisation and transformation of legacy systems. • Engineering high-quality CI/CD pipelines with DevSecOps & supply chain integrity. • Implementing event-driven ...

Senior Java Engineer

Hiring Organisation
Jobleads-UK
Location
Haywards Heath, England, United Kingdom
operational experience, not just familiarity. CI/CD pipelines: GitHub Actions, Jenkins, or GitLab CI. Observability: structured logging, distributed tracing (Jaeger/Zipkin), metrics (Prometheus/Grafana/Datadog). Active experimentation with AI/LLM tooling applied to real engineering problems - automated workflows, reduced manual ops, smarter exception handling. ...

NOC Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Basingstoke, Hampshire, United Kingdom
Employment Type
Permanent
Salary
£60000/annum Bonus, Pension, Healthcare
administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with ...

NOC Engineer, AWS

Hiring Organisation
Spectrum It Recruitment Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with ...

Cloud Architect

Hiring Organisation
Talent
Location
City of London, London, United Kingdom
Have): Prior experience executing large legacy-to-cloud migrations within highly regulated environments. Experience with Kubernetes/EKS tooling such as Argo CD, Flux, Prometheus, Grafana, or Istio. A broader multi-cloud perspective (Azure or GCP). ...

Senior DevOps Engineer

Hiring Organisation
SF Partners Admin
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
skills: - Deep technical ownership of IDP's for a large Developer user base - Knowledge of cloud native system design - Observability stack exposure - Grafana, Prometheus, Open Telemetry etc - Experience designing AWS and supporting AWS landing zones - IAC experience - Terraform, Ansible, Redhat etc - Strong experience building and owning multiple Kubernetes clusters ...

DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
pipelines Knowledge of containerisation technologies such as Docker and orchestration platforms such as ECS Experience with monitoring and observability tools (e.g. CloudWatch, Grafana, Prometheus) Strong understanding of security best practices and cloud governance Excellent troubleshooting and problem-solving skills Experience using AI-assisted development tools to improve engineering workflows What ...

Principal SRE Engineer / Grafana Specialist - (Outside IR35)

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £750 per day + Outside IR35
engineering roles, with at least 5 years in SRE, Observability, or DevOps functions. Hands-on proficiency with observability tools such as Datadog, Grafana, Prometheus, OpenTelemetry. Strong knowledge of distributed systems, microservices, and container orchestration (Kubernetes, Docker). Experience with automation and Infrastructure as Code (Terraform, Ansible) and CI/ ...

Senior Software Engineer

Hiring Organisation
Jobleads-UK
Location
United Kingdom
gaming industry or large‐scale interactive systems. Familiarity with messaging systems such as Kafka, RabbitMQ or similar Proficiency with monitoring tools like Prometheus, Grafana, or ELK Stack Experience with AI/ML integration in backend systems. Why Join Us At Lockwood, you’ll be part of an inclusive, creative ...

Lead Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Code expertise using Terraform. Practical experience with DevOps and infrastructure tooling: CI/CD pipelines Docker Kubernetes (K8s) Experience with monitoring and observability tools: Prometheus Grafana Datadog Experience building and monitoring distributed systems (front‐end and back‐end). Hands‐on experience with SQL and NoSQL databases. AI & Advanced Technologies ...

Application Developer - Senior

Hiring Organisation
Pinnacle Technical Resources
Location
Plano, Texas, United States
Employment Type
Permanent
Salary
USD 65 Annual
service mesh solutions. Familiarity with caching mechanisms (e.g., Redis, Memcached). Understanding of event-driven architectures and patterns. Exposure to monitoring tools like Prometheus, Grafana, or Elasticsearch. Soft Skills: Strong communication skills to collaborate effectively across teams. Ability to work independently and manage multiple tasks in a fast-paced environment. ...

Platform Product Manager

Hiring Organisation
Entain
Location
Arkansas, United Kingdom
Employment Type
Full Time
with Ansible, Kubernetes, and GitLab CI/CD. Proficiency with Git version control and branching strategies. Strong AI skills use Knowledge of monitoring tools (Prometheus, Grafana) and logging solutions (ELK/Datadog). Familiarity with cloud platforms such as AWS or GCP. Strong understanding of infrastructure and application security best ...

Systems Engineer, Production

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
systems such as Buildkite, GitHub Actions, etc. Experience with relational or non-relational datastores in a production environment. Familiarity with observability platforms (Datadog, Prometheus, Grafana, etc.). Familiarity with Linux systems administration, networking, and troubleshooting. Excellent communication and documentation abilities, with a focus on knowledge sharing and team collaboration. Growth ...

Lead DevSecOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
e.g., EC2 to EKS, or cross-cloud) with a focus on data integrity and minimal downtime Ability to implement standardized telemetry pipelines (e.g., OpenTelemetry, Prometheus, or ELK) that provide developers with out-of-the-box visibility into their services Familiarity with automated policy enforcement and compliance-as-code (e.g. ...

Senior Platform Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Terraform experience building Infrastructure as Code Good understanding of Site Reliability Engineering (SRE) principles Experience with observability and monitoring tools such as Dynatrace, Grafana, Prometheus or similar Knowledge of CI/CD pipelines and modern cloud-native engineering practices Experience supporting live production platforms Comfortable leading technical discussions and mentoring ...

Azure Platform Engineering Consultant

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
Terragrunt using modular, reusable implementation patterns. DevOps & Observability: Strong experience with CI/CD tools (Azure DevOps/GitHub Actions) and monitoring stacks (Prometheus, Grafana, Azure Monitor, etc.). FinOps: Practical knowledge of cloud cost optimization, tagging strategies, and right-sizing. Consulting Mindset: Excellent stakeholder management and communication skills ...

Senior ML Ops Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
services Implementing infrastructure‐as‐code (Terraform, Bicep or Pulumi) Implementing monitoring and observability for production systems, including metrics, alerting, logging, and dashboarding (e.g. Prometheus, Grafana) Pipeline orchestration using Dagster, Airflow, Prefect, or similar Bonus points if you have: Practical experience with model serving infrastructure – batch and/or real‐time ...