26 to 50 of 175 Remote/Hybrid Prometheus Jobs

Senior System Engineer London

Location
Greater London, England, United Kingdom
Proven experience deploying, managing, and troubleshooting Kubernetes clusters in production environments. Strong understanding of observability and monitoring practices, including experience with tools such as Prometheus, Grafana, or similar platforms. Demonstrated ability to work confidently across both Unix based systems (Ubuntu, RHEL, or similar distributions) and Windows environments. Experience with scripting ...

Site Reliability Engineer

Hiring Organisation
E-Solutions IT Services UK Ltd
Location
Leeds, West Yorkshire, United Kingdom
Employment Type
Full-Time
Salary
£280.00 - £300.00 per day
scripting skills in Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux). ...

Full Stack Developer (Java / Python / Go)

Location
Belfast, Northern Ireland, United Kingdom
Ansible. Experience with CI/CD tools (Jenkins, Harness, Tekton, etc.) Familiarity with serverless technologies (e.g. AWS Lambda) Understanding of monitoring/observability tools (Prometheus, Grafana, ELK) Cloud certifications are beneficial but not required #J-18808-Ljbffr ...

Java Software Engineer

Location
Greater London, England, United Kingdom
execute unit, integration, and performance tests using appropriate testing frameworks and tools. Monitor, troubleshoot, and optimize applications and infrastructure using New Relic, Grafana, Prometheus, and Bosun. Hands‐on experience with microservices architecture and Kafka. Excellent communication skills and ability to thrive in a fast-paced environment. Collaborate with cross-functional ...

Intermediate/Senior DevOps

Hiring Organisation
Global Relay
Location
London, UK
Employment Type
Full-time
Podman, Kubernetes, VMWareOperating Systems: Linux or iOSContinuous Integration/Build and deployment automation: Jenkins, SonarQube, Artifactory, Bitbucket, Maven, Xcode Build, HelmInstrumentation and monitoring: Loki, Prometheus, Grafana, Mimir, Tempo TracingLanguages and frameworks: Bash, Java or Kotlin, Groovy, Python, ReactJS, SwiftWhere you have knowledge gaps, training and mentoring will be provided. About ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Leeds, England, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Glasgow, Scotland, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Birmingham, England, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Manchester, England, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Greater London, England, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Belfast City District, Northern Ireland, United Kingdom
performance. Help ensure the reliability and scalability of our infrastructure, while maintaining platform SLOs. Troubleshoot site issues using industry-leading tools like Splunk, Prometheus and OpenTelemetry. Own and operate critical datastores, including Cassandra and MySQL. Automate everything with Python, Puppet, Git, Jenkins, and Terraform and leveraging an extensive suite ...

Platform Engineer

Location
Greater London, England, United Kingdom
Streaming to support deployment of new infrastructure, permission changes and access Maintaining and supporting infrastructure for Streaming including the Common Platform Infrastructure e.g. Prometheus Managing upgrades of the Kubernetes deployments in AWS Defining and designing the infrastructure for the future - whether this be for a new application or re-design ...

Senior AWS Cloud DevOps Engineer

Location
Greater London, England, United Kingdom
Route 53, Direct Connect, Load Balancers, API Gateway Practical experience with Kubernetes (EKS) and container orchestration in AWS Familiarity with monitoring tools (CloudWatch, DataDog, Prometheus) and alerting systems Hands-on experience with Terraform and/or AWS CloudFormation for IaC Proficiency in implementing CI/CD pipelines using GitHub Actions ...

Senior DevOps Systems Administrator

Hiring Organisation
Rise Technical Recruitment
Location
Guildford, Surrey, United Kingdom
Employment Type
Permanent
Salary
£60000 - £65000/annum Medical Insurance + Holiday + Pensio
within the business. The Role: Design and deploy AWS/Private Cloud solutions. Build reliable pipelines via GitHub Actions and implement monitoring using Grafana, Prometheus, and CloudWatch. Manage IAM, firewalls, and VPNs while collaborating across teams to improve system reliability. Mentor team members, maintain high-quality documentation, and participate ...

Staff DevOps Engineer

Location
Cambridge, England, United Kingdom
understanding of Software Development Lifecycle. Good understanding of microservice architecture. Good experience in Monitoring k8s clusters, APM, services and infrastructure located in AWS (NewRelic,Prometheus). Ability to create automation tools in Go, Python, Bash, Powershell. Familiarity with Continuous Delivery tools used with microservices. Coding Experience. Additional Information Benefits Private ...

Principal Network Engineer

Location
Greater London, England, United Kingdom
certifications (e.g.,Cisco, Juniper, AWS) Experience in the Media or Broadcast technology sector. Familiarity with network monitoring and observability platforms (e.g., Nautobot, Netbox, Batfish, Prometheus, Grafana, ELK stack, Solarwinds). #J-18808-Ljbffr ...

Senior Software Engineer Python

Location
United Kingdom
Reliability Write secure, well-tested code (unit, integration, end-to-end) and uphold coding standards through code reviews. Contribute to logging, metrics, and alerting (Prometheus/Grafana, ELK/OpenSearch) for the services you build. Support compliance readiness (ISO 27001, GDPR) through secure-by-default design. Cross-Functional & Process Collaborate ...

Senior Azure DevOps Engineer

Location
Milton Keynes, England, United Kingdom
Docker (design, scaling, network and security). API Management. App Services and Azure Functions. Observability using Azure Monitor, Log Analytics and Application Insights, Prometheus and Grafana. Implementation of governance and compliance controls using Azure Policy, Management Groups, and Landing Zone principles. Azure SQL and Managed Instance. Event‐driven architecture (Service ...

Senior Software Engineer Python

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Reliability Write secure, well-tested code (unit, integration, end-to-end) and uphold coding standards through code reviews. Contribute to logging, metrics, and alerting (Prometheus/Grafana, ELK/OpenSearch) for the services you build. Support compliance readiness (ISO 27001, GDPR) through secure-by-default design. Cross-Functional & Process Collaborate ...

Forward Deployed Engineer - Infrastructure

Location
Greater London, England, United Kingdom
Anthos, AWS EKS Anywhere, AWS Outposts A strong background in Go, Python or Java Experience with CockroachDB, AlloyDB, Aurora Experience with observability tools, e.g. Prometheus, Grafana Benefits Highly competitive salary Pension plan (match up to 5%) Life insurance - three times annual salary Competitive maternity (six months fully paid) and paternity ...

Site Reliability Engineer - Core

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
network and/or internals. Experience working with cloud solutions (GCP or AWS).Deep understanding and demonstrable experience with modern monitoring tools such as Prometheus, Datadog, Grafana, TelegrafExperience with infrastructure as code tools. Experience with complex Terraform deployments is a plus. Solid background with configuration management tools. Experience with Saltstack ...

Senior Software Engineer

Location
Reading, England, United Kingdom
cloud‐native deployments CI/CD pipeline experience (GitLab, GitHub or similar) Infrastructure as Code (Terraform or similar) Experience with observability tooling (OpenTelemetry, Prometheus, Grafana, etc.) Strong testing mindset (TDD, automated testing, contract testing) Highly Desirable Experience with API Gateway technologies (e.g. AWS API Gateway) Experience building or operating identity ...

Senior Site Reliability Engineer

Hiring Organisation
GCS
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
Salary
£75000 - £95000/annum Bonus
networking, cloud infrastructure and automation. * Experience with Infrastructure-as-Code. * Strong communication and technical leadership skills. Desirable Skills: * SLOs, SLIs, SLAs and error budgets. * Prometheus, Grafana, Elastic/ELK or OpenTelemetry. * Kubernetes, containers and distributed systems. * Performance and resilience engineering. * Experience driving SRE maturity across engineering teams. * Financial services ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Code (Terraform, Ansible, Puppet or similar) Hands-on experience with Kubernetes, Docker, and cloud platforms (AWS preferred) Experience with monitoring/observability tools (Prometheus, Grafana, ELK, APM tools) Solid understanding of system performance, scalability, and resilience Strong collaboration and communication skills within cross-functional product teams Desirable: Experience working ...

Technical Lead - MuleSoft Hybrid Platform

Location
Beeston, England, United Kingdom
connectivity. Ensure high availability and resilience of integration platforms through Kubernetes scaling strategies. Drive monitoring and observability implementation for hybrid platforms using tools like Prometheus, Grafana or ELK stack. Maintain technical intellectual property and documentation for platform and automation tooling. Research and introduce new automation tools and capabilities to improve ...