326 to 350 of 497 Prometheus Jobs in the UK

Site Reliability Engineer III

Hiring Organisation
CME- Group
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.Incident Response & Operations: Engage with urgency in live … teams, coupled with an eagerness to learn independently and collaboratively.Preferred Qualifications/DesirableObservability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Application ...

Site Reliability Engineer III

Location
Belfast City District, Northern Ireland, United Kingdom
Manage cluster lifecycles, data replication, RBAC, and workload placement. Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection. Incident Response & Operations: Engage with urgency … with an eagerness to learn independently and collaboratively. Preferred Qualifications/Desirable Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana. Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles. Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified ...

Senior Devops/Infrastructure Engineer

Hiring Organisation
Intellectual Capital Resources
Location
London, United Kingdom
Salary
£ 80 K
alongside engineers and telco engineering teams. Key skills/experience required: AWS Kubernetes IaC (Terraform) GitOps (Helm, ArgoCD) Monitoring and alerting for production systems (Prometheus/Grafana or similar) Azure (desirable) MLOps in k8s (Kubeflow etc) Running GPU workloads on K8s (drivers, scheduling, utilisation) Familiarity with AI engineering Experience deploying ...

Lead Cloud Infrastructure Engineer

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
Salary
£ 80 K
Administration and Configuration Management experience, as well as networking experience and an understanding of DevOps principles and practices. Experience with observability tools such as Prometheus and Grafana and working with Kubernetes in production environments at scale is a plus. If you're open to hearing further details about this opportunity ...

Principal DevOps Engineer

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
Salary
£ 120 K
lead the evolution of DevOps tools Kubernetes, Jenkins, Gitlab, Terraform, and more.Optimising automation and performance.Champion containerisation and high performance base images.Elevate monitoring systems Zabbix, Prometheus, Thanos ensuring 24/7 operational excellence.Secure infrastructure access management, balancing innovation with ironclad security. They're offering a career defining role in a company ...

Platform Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
Central London, London, United Kingdom
Employment Type
Permanent
Salary
£75,000
narrow slice of the stack and told to stay in it, this isn'tthe one. Core stack: Azure, Kubernetes (AKS), Terraform, GitHub Actions, Prometheus, Grafana, PowerShell, etc. You don't need to be a senior engineer. The team already has senior people. What they need is someone solid who wants ...

Applications Engineer

Location
Cambridge, England, United Kingdom
similar source-of-truth/IPAM platforms. Cisco Nexus and/or Arista Cloud Vision Dashboard APIs. Automated network testing frameworks. Prometheus, Grafana and Open Telemetry or similar Automated compliance and configuration validation. #J-18808-Ljbffr ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU Limited
Location
Milton Keynes, Buckinghamshire, United Kingdom
Salary
£ 80 K
experience with both Azure, and on-premise virtual machines.Experience withInfrastructure as Code/Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor).Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working.Ability ...

Cloud Operations Engineer (remote – London)

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
Experienced and Certified in cloud computing with AWSExperience in a public cloud such as AWSKnowledge of monitoring and alerting technologies such as Grafana, Prometheus,Expereince of working with of Docker & Kubernetes and Container technology in productionWindows and Linux Operating System Management TechniquesSolid understanding of the OSI ModelExperience in database technology ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT Recruitment
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Permanent
Salary
£75,000
with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code/Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT
Location
Wavendon, Bedfordshire, United Kingdom
Employment Type
Permanent
Salary
GBP 65,000 - 75,000 Annual
with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code/Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. ...

AWS Solutions Architect

Hiring Organisation
Tata Technologies
Location
Gaydon, England, United Kingdom
Code: Terraform and/or CloudFormation, including modular design, version control and environment promotion practices. Monitoring and operations: CloudWatch, CloudTrail, OpenSearch/ELK, Grafana, Prometheus or similar tooling. Operating systems and platform administration: Linux and Windows, scripting and automation using Shell, Python or equivalent. DevOps and CI/ ...

Software Engineer — Observability Instrumentation

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
such as Terraform, ArgoCD, Helm or JenkinsInterest in AI engineering and SRE practices to improve incident response and RCADesirable but not essential experience includes:Prometheus/PromQL, VictoriaMetrics, OpenSearch, Grafana or similar observability backendsAuto-instrumentation, distributed tracing, structured logging or trace/metric correlationKafka or telemetry pipeline architecturesWhy should ...

Site Reliability Engineer

Location
City Of London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, United Kingdom
Salary
£ 70 K
delivery speed and reducing deployment risk.Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Core AI Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
servicesFamiliarity with sandboxing and workload isolation technologiesExperience in quantitative finance or low-latency systemsAWS experience particularly in hybrid environmentsExperience with observability tooling such as Prometheus, Grafana or OpenTelemetryContributions to open-source projects in relevant domainsWhy join us Highly competitive compensation plus annual discretionary bonusLunch provided (via Just Eat for Business ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
speed and reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Site Reliability Engineer - Private Cloud Compute

Location
Greater London, England, United Kingdom
high-level programming language like: Java, Go, Python, or Perl Proclivity towards efficient programming emphasizing improvement via complexity analysis. Experience with Kubernetes, Nginx, Envoy, Prometheus, and/or Docker. Preferred Qualifications Understanding of standard networking protocols and components such as: HTTP, DNS, ECMP, TCP/IP, ICMP, the OSI Model ...

Senior Java Software Engineer- Platform Engineering

Hiring Organisation
Wise
Location
London, United Kingdom
Salary
£ 80 K
/CD platform.Practical experience applying SRE principles, such as service-level objectives, error budgets, and automated incident prevention.Familiarity with observability tooling such as Prometheus, Grafana, Elastic, or distributed tracing.Knowledge of cloud networking and security.Experience working in a regulated environment or with standards such as PCI DSS.Experience with capacity planning, performance ...

Non-Functional Test Specialist

Hiring Organisation
Euroclear
Location
United Kingdom
Salary
£ 70 K
experience in planning, preparing & executing Resilience/Disaster Recovery/Continuity/Failover testing/PerformanceExposure to tools supporting resilience & operational observability (Splunk, Prometheus, Kafka, Chaos tooling etc)Understanding of high-availability architectures, infrastructure redundancy & backup/restore strategiesAre familiar with working within large scale & complex Technology Implementation or changeExperience ...

Staff Software Engineer — DevPlatform (Java, AWS)

Hiring Organisation
Duetto
Location
United Kingdom
Salary
£ 50 K
tune MongoDB, Redis (Redisson), and PostgreSQL for performance and reliability in a high-throughput, multi-tenant environment — and strengthen platform observability through OpenTelemetry, Datadog, Prometheus, Grafana, and Sentry instrumentation.You'll pioneer AI-augmented engineering workflows using Claude Code, Claude MPM, and Duetto's in-house MCP server platform (82+ tools ...

Software Engineer

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
Experience working with large GPU clusters or distributed training environments.Familiarity with distributed training techniques such as DDP or FSDP.Experience with observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry.Experience with data pipeline orchestration tools such as Airflow, Flyte, Ray, Metaflow, or Argo Workflows.Experience with containerisation and infrastructure tooling such ...

Neo4j Platform Consultant

Hiring Organisation
NTT DATA
Location
London, United Kingdom
Salary
£ 80 K
scaling)Security (RBAC, authentication, data protection)Preferred SkillsExperience with Graph RAG workloads and traversal optimizationDevOps tools (Docker, Kubernetes, CI/CD pipelines)Monitoring tools (Prometheus, Grafana, etc.)Experience with other graph databases (Neptune, TigerGraph)Job ExpectationsEnsure stable, secure, and high-performing Neo4j platform operationsEnable efficient graph query execution ...

Software Engineer

Location
Greater London, England, United Kingdom
with large GPU clusters or distributed training environments. Familiarity with distributed training techniques such as DDP or FSDP. Experience with observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry. Experience with data pipeline orchestration tools such as Airflow, Flyte, Ray, Metaflow, or Argo Workflows. Experience with containerisation and infrastructure tooling ...