1 to 25 of 2,649 Observability Jobs in England

Senior Software Engineer

Location
Greater London, England, United Kingdom
services, including AWS.Knowledge of containerisation and orchestration technologies, including Docker, Kubernetes, and OpenShift.Experience implementing Infrastructure as Code using Terraform or similar solutions.Experience with observability and monitoring platforms such as Splunk, Dynatrace, Grafana, and Prometheus.Understanding of DevSecOps principles, secure software development practices, and security-focused engineering approaches.Experience working within regulated Financial ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
fintech, payments, or enterprise SaaS platforms* Exposure to event-driven architecture (Kafka, RabbitMQ)* Familiarity with infrastructure-as-code tools (Terraform, CloudFormation)* Understanding of observability tools (Prometheus, Grafana, ELK stack) #J-18808-Ljbffr ...

DevOps Engineer (Security Cleared)

Location
Greater London, England, United Kingdom
automation using Python, Bash, PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile ...

Senior Software Engineer

Location
City of Westminster, England, United Kingdom
fintech, payments, or enterprise SaaS platforms Exposure to event-driven architecture (Kafka, RabbitMQ) Familiarity with infrastructure-as-code tools (Terraform, CloudFormation) Understanding of observability tools (Prometheus, Grafana, ELK stack) L’état d’esprit Edenred - Nous sommes une entreprise unique. Nous recherchons de nouveaux collaborateurs prêts à prendre part ...

DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
Up to £425 per day + Inside IR-35
legacy platforms and drive adoption of cloud-native technologies Manage and optimise production Kubernetes environments, ensuring scalability, performance and resilience Implement monitoring, alerting and observability solutions using Prometheus and Grafana Embed security best practices throughout the software delivery lifecycle Develop automation tooling using Ansible, Python and Bash Work closely with ...

Principal Cloud Engineer (Terraform), London

Location
Greater London, England, United Kingdom
Apply data quality and validation frameworks to ensure accuracy, completeness, and freshness of cost and usage data across all cloud providers; instrument pipelines with observability tooling to surface issues proactively. Build and maintain reusable data assets — curated datasets, aggregations, and data marts — that power FinOps dashboards, showback/chargeback reporting ...

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
produce, distribute and promote the most critically acclaimed and commercially successful music to delight and entertain fans around the world.As a Senior Observability Engineer, you will be a driving force for technical excellence and strategic vision within our global team. You will be instrumental in architecting, building, and leading … comprehensive observability strategy to ensure the reliability, performance, and scalability of our critical IT systems. This senior role demands a passion for data-driven strategy, a commitment to automation, and the ability to mentor and lead. You will not only solve complex technical challenges but also influence the direction ...

Senior DevOps Engineer | London, Hybrid | up to £125k

Location
Greater London, England, United Kingdom
/or CloudFormation), ensuring repeatable environments. Drive containerisation and orchestration using Docker and Kubernetes (including managed services such as EKS/AKS). Improve observability with monitoring, logging and alerting (e.g. Prometheus, Grafana, ELK/EFK, CloudWatch). Embed security best practice: IAM, secrets management, patching, vulnerability remediation and secure ...

Cloud DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £600 per day + Outside IR-35
GitHub Actions Jenkins Linux administration Bash scripting Python automation IAM and cloud security best practices CI/CD pipeline design and implementation Monitoring and observability tooling Agile delivery experience Desirable Skills Government or wider public sector experience GDS-aligned delivery experience Helm Prometheus Grafana ELK/Elastic Stack Azure ...

Platform Engineer

Location
Greater London, England, United Kingdom
Experience with Infrastructure as Code tools such as Terraform or CloudFormation. Knowledge of CI/CD platforms and deployment automation. Experience with monitoring and observability tools such as CloudWatch, Prometheus, Grafana, ELK, or similar. Good understanding of Linux systems administration. Experience troubleshooting complex application and infrastructure issues. Knowledge of networking ...

Senior AI Engineer

Hiring Organisation
INFUSED SOLUTIONS LIMITED
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
Data Scientists to productionise Machine Learning and NLP models. Develop high-performance RESTful APIs and microservices for AI model serving. Drive system reliability, monitoring, observability and performance optimisation. Champion engineering best practices including CI/CD, automated testing and Infrastructure as Code. Improve integration, interoperability and data exchange across enterprise ...

AWS DevOps Platform Engineer - eSC/eDV Clearance

Location
Leicester, England, United Kingdom
Developed Vetting (DV). Preferred technical and professional experience Experience working within multi‐account AWS environments (e.g. Organizations, landing zones) Familiarity with observability and monitoring tooling (CloudWatch, Prometheus, Grafana, ELK stack) Exposure to security best practices in cloud and container environments (IAM, secrets management, Zero Trust concepts) Experience supporting ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
help ensure that systems used by clinical, operational, and administrative teams remain stable, secure, and available. This role is responsible for building automation, improving observability, reducing manual operational work, supporting integrations, and helping maintain reliable systems that directly impact patient care and business operations. This is a hybrid position with ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Location
Greater London, England, United Kingdom
both verbal and written undefined Additional Technical Skills Cloud Platforms: (e.g. AWS, GCP, Azure) CI/CD Tooling: (e.g. GitHub Actions, ArgoCD, Harness) Monitoring & Observability: (e.g. Prometheus, Grafana, Coralogix, OpenTelemetry) Infrastructure as Code: (e.g. Terraform, Pulumi) Config Management: (e.g. Ansible, Puppet, Chef) Programming/Scripting: (e.g. Python, Go, Bash) Soft ...

DevOps Engineer, Studios

Location
Greater London, England, United Kingdom
across multiple teams. Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar. Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack. Ensure high availability and disaster recovery strategies are in place and tested regularly. Collaborate closely with development ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
West Midlands, England, United Kingdom
ingress controllers. Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications ...

Senior Platform Engineer

Location
Reigate, England, United Kingdom
claims‐based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLIs/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms. General knowledge ...

Production Engineer

Location
Greater London, England, United Kingdom
identify patterns, and drive intelligent automation solutions* Hands-on experience with containerization technologies such as Docker and Kubernetes, including cluster management, deployment, scaling, and observability* Deep practical knowledge of algorithmic trading workflows, including the behaviour, lifecycle, and risk controls of execution algos used across the EMEA markets* Experience designing ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
ingress controllers. Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud‐native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications ...

Production Engineer

Location
City Of London, England, United Kingdom
identify patterns, and drive intelligent automation solutions Hands-on experience with containerization technologies such as Docker and Kubernetes, including cluster management, deployment, scaling, and observability Deep practical knowledge of algorithmic trading workflows, including the behaviour, lifecycle, and risk controls of execution algos used across the EMEA markets Experience designing ...

Software Engineer III - AIML Experimentation, Observability and Governance

Hiring Organisation
Hackajob Ltd
Location
Bournemouth, Dorset, South West, United Kingdom
Employment Type
Permanent
JOB DESCRIPTION As a Software Engineer III at JPMorganChase within AMDP, you will play a pivotal role in an agile team, helping to define, design, build, and maintain state-of-the-art technology products. Leveraging ...

Software Engineer – Python/AWS/Terraform (SC or DV Cleared)

Location
Greater London, England, United Kingdom
event-driven architectures Desirable experience Stronger Python backend development using FastAPI or Flask. Kubernetes or Amazon EKS. DevSecOps and secure-by-design engineering. Observability, monitoring and automated alerting. < REST APIs and distributed systems. Ansible, CloudFormation or AWS CDK. AI or LLM integration. Previous delivery within government, defence or another highly ...

Senior Dev/ML Ops Engineer

Location
Greater London, England, United Kingdom
deployment to monitoring and continuous improvement. Build and maintain robust CI/CD pipelines for both software and ML workflows. Ensure reliability, scalability, observability, and security of production systems and ML infrastructure. Automate deployment, orchestration, and environment management using modern DevOps tooling. Collaborate closely with software engineers, data scientists ...

Technical Product Manager

Hiring Organisation
HUKM Development
Location
Buckinghamshire, England, United Kingdom
such as Kafka, RabbitMQ, or SQS • Experience with API monetisation and developer portal management • Containerisation and orchestration experience with Docker and Kubernetes • Familiarity with observability tooling such as Datadog, New Relic, or Grafana • Background in financial services, fintech, or another regulated industry How we work You will be based ...

AWS Cloud Architect - eSC or eDV Clearance Required

Location
Leicester, England, United Kingdom
with relational databases (Postgres, MySQL, SQL Server) and NoSQL platforms (DynamoDB or Cosmos DB), with optional exposure to graph or vector data stores. Monitoring & Observability Experience with cloud‐native monitoring and logging systems such as CloudWatch or Azure Monitor, plus tools like Application Insights, Prometheus, and Grafana. Security & Governance Strong ...