1 to 25 of 1,319 Remote/Hybrid Observability Jobs

Senior Software Engineer

Location
Greater London, England, United Kingdom
services, including AWS.Knowledge of containerisation and orchestration technologies, including Docker, Kubernetes, and OpenShift.Experience implementing Infrastructure as Code using Terraform or similar solutions.Experience with observability and monitoring platforms such as Splunk, Dynatrace, Grafana, and Prometheus.Understanding of DevSecOps principles, secure software development practices, and security-focused engineering approaches.Experience working within regulated Financial ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
fintech, payments, or enterprise SaaS platforms* Exposure to event-driven architecture (Kafka, RabbitMQ)* Familiarity with infrastructure-as-code tools (Terraform, CloudFormation)* Understanding of observability tools (Prometheus, Grafana, ELK stack) #J-18808-Ljbffr ...

Senior Software Engineer

Location
City of Westminster, England, United Kingdom
fintech, payments, or enterprise SaaS platforms Exposure to event-driven architecture (Kafka, RabbitMQ) Familiarity with infrastructure-as-code tools (Terraform, CloudFormation) Understanding of observability tools (Prometheus, Grafana, ELK stack) L’état d’esprit Edenred - Nous sommes une entreprise unique. Nous recherchons de nouveaux collaborateurs prêts à prendre part ...

DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
Up to £425 per day + Inside IR-35
legacy platforms and drive adoption of cloud-native technologies Manage and optimise production Kubernetes environments, ensuring scalability, performance and resilience Implement monitoring, alerting and observability solutions using Prometheus and Grafana Embed security best practices throughout the software delivery lifecycle Develop automation tooling using Ansible, Python and Bash Work closely with ...

Senior DevOps Engineer | London, Hybrid | up to £125k

Location
Greater London, England, United Kingdom
/or CloudFormation), ensuring repeatable environments. Drive containerisation and orchestration using Docker and Kubernetes (including managed services such as EKS/AKS). Improve observability with monitoring, logging and alerting (e.g. Prometheus, Grafana, ELK/EFK, CloudWatch). Embed security best practice: IAM, secrets management, patching, vulnerability remediation and secure ...

Platform Engineer

Location
Greater London, England, United Kingdom
Experience with Infrastructure as Code tools such as Terraform or CloudFormation. Knowledge of CI/CD platforms and deployment automation. Experience with monitoring and observability tools such as CloudWatch, Prometheus, Grafana, ELK, or similar. Good understanding of Linux systems administration. Experience troubleshooting complex application and infrastructure issues. Knowledge of networking ...

Senior AI Engineer

Hiring Organisation
INFUSED SOLUTIONS LIMITED
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
Data Scientists to productionise Machine Learning and NLP models. Develop high-performance RESTful APIs and microservices for AI model serving. Drive system reliability, monitoring, observability and performance optimisation. Champion engineering best practices including CI/CD, automated testing and Infrastructure as Code. Improve integration, interoperability and data exchange across enterprise ...

Platform Engineer

Location
Greater London, England, United Kingdom
maintainability. Prompt Engineering for Engineering Workflows: Create clear prompts for troubleshooting, documentation, and operational runbooks to improve speed and consistency. AIOps and Intelligent Observability: Apply AI to analyze logs, metrics, traces, alerts, and incident history to detect anomalies, identify root causes, reduce noise, and improve mean time to resolution. ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Hiring Organisation
FactSet Research Systems
Location
London, UK
Employment Type
Full-time
fluent in English both verbal and writtenundefinedAdditional Technical SkillsCloud Platforms:(e.g. AWS, GCP, Azure)CI/CD Tooling:(e.g. GitHub Actions, ArgoCD, Harness)Monitoring & Observability:(e.g. Prometheus, Grafana, Coralogix, OpenTelemetry)Infrastructure as Code:(e.g. Terraform, Pulumi)Config Management: (e.g. Ansible, Puppet, Chef)Programming/Scripting:(e.g. Python, Go, Bash)Soft ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
help ensure that systems used by clinical, operational, and administrative teams remain stable, secure, and available. This role is responsible for building automation, improving observability, reducing manual operational work, supporting integrations, and helping maintain reliable systems that directly impact patient care and business operations. This is a hybrid position with ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Location
Greater London, England, United Kingdom
both verbal and written undefined Additional Technical Skills Cloud Platforms: (e.g. AWS, GCP, Azure) CI/CD Tooling: (e.g. GitHub Actions, ArgoCD, Harness) Monitoring & Observability: (e.g. Prometheus, Grafana, Coralogix, OpenTelemetry) Infrastructure as Code: (e.g. Terraform, Pulumi) Config Management: (e.g. Ansible, Puppet, Chef) Programming/Scripting: (e.g. Python, Go, Bash) Soft ...

DevOps Engineer, Studios

Location
Greater London, England, United Kingdom
across multiple teams. Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar. Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack. Ensure high availability and disaster recovery strategies are in place and tested regularly. Collaborate closely with development ...

Senior Platform Engineer

Hiring Organisation
Willis Towers Watson
Location
Reigate, Surrey, United Kingdom
Salary
£ 70 K
OIDC, JWT and claims‐based authorisation.Experience in platform engineering or SRE roles: building internal platforms, defining SLIs/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing).Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms.General knowledge ...

Senior Platform Engineer

Hiring Organisation
WTW
Location
Surrey, United Kingdom
Employment Type
Full Time
claims‐based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLIs/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms. General knowledge ...

AI Engineer

Hiring Organisation
Ten Group
Location
London, United Kingdom
Salary
£ 70 K
platforms (AWS, GCP, Azure) and infrastructure-as-code (Terraform etc). Hands-on with DevOps/Infra tooling (CI/CD, Docker, K8s) and observability (Prometheus, Grafana, Datadog etc) Experience building distributed systemsKnowledge and hands-on experience with multiple datastores (both SQL and NoSQL) Desired experience in building agents ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
ingress controllers. Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud‐native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications ...

Software Engineer – Python/AWS/Terraform (SC or DV Cleared)

Location
Greater London, England, United Kingdom
event-driven architectures Desirable experience Stronger Python backend development using FastAPI or Flask. Kubernetes or Amazon EKS. DevSecOps and secure-by-design engineering. Observability, monitoring and automated alerting. < REST APIs and distributed systems. Ansible, CloudFormation or AWS CDK. AI or LLM integration. Previous delivery within government, defence or another highly ...

Manager - DE - Technology Consulting - FS

Location
Greater London, England, United Kingdom
multi-disciplinary squads, partnering with product owners, business stakeholders, architects, QA, DevOps, security, data and infrastructure teams. Drive production readiness including CI/CD, observability, resilience, security, automated testing, release management, runbooks, incident response and root cause analysis. Mentor engineers and senior consultants, building capability in modern engineering practices ...

Portfolio Software Full Stack Engineer (Contractor)

Location
Cambridge, England, United Kingdom
well-tested and performant code following engineering best practices. Design data models and work with relational and NoSQL databases where appropriate. Improve application reliability, observability and performance through monitoring, logging and optimisation. Software Engineering & Code Quality Write clean, maintainable and well-documented code following established coding standards. Perform thorough code ...

Senior AI Engineer - Hybrid

Hiring Organisation
Genesis10
Location
Columbus, Ohio, United States
Employment Type
Permanent
Salary
USD Hourly
technologies Experience with secure API integrations and OAuth 2.0/OpenID Connect authorization patterns Understanding of the software development lifecycle, CI/CD, observability, reliability, security, privacy, and responsible AI practices Strong problem-solving skills, attention to detail, technical communication, and ability to collaborate effectively across teams AI Platforms ...

Platform Engineer

Location
Greater London, England, United Kingdom
+ Kubernetes for microservice orchestration using Istio service mesh PostgreSQL for relational db, ElasticSearch for indexing, Redis for caching DataDog, Grafana and OpenTelemetry for observability GitHub for our Version Control and CI (with our own runners) CD: Harness and FluxCD Terraform and Terragrunt as IaaC Python and bash for scripting ...

AI Engineer

Location
Greater London, England, United Kingdom
platforms (AWS, GCP, Azure) and infrastructure‐as‐code (Terraform etc). Hands‐on with DevOps/Infra tooling (CI/CD, Docker, K8s) and observability (Prometheus, Grafana, Datadog etc) Experience building distributed systems Knowledge and hands‐on experience with multiple datastores (both SQL and NoSQL) Desired experience in building agents ...

Senior DevOps Engineer

Hiring Organisation
Leonardo DRS
Location
United Kingdom
Salary
£ 70 K
research, publications, or professional forums in DevOps/automation.Knowledge of service mesh, GitOps, and policy-as-code approaches.Experience with monitoring/logging and observability platforms (ELK, Prometheus, Splunk).This is not an exhaustive list, and we are keen to hear from you even if you might not have experience ...

Senior DevOps Engineer

Location
West of England, England, United Kingdom
publications, or professional forums in DevOps/automation. Knowledge of service mesh, GitOps, and policy-as-code approaches. Experience with monitoring/logging and observability platforms (ELK, Prometheus, Splunk). This is not an exhaustive list, and we are keen to hear from you even if you might not have ...

Software Engineer - Investment Growth

Location
West of England, England, United Kingdom
testing tools and practices (e.g. Jest, Cypress, Backstop, Playwright).* Experience with CI/CD and Trunk Based Development.* Experience with observability tools and practices, including monitoring, logging, and tracing to ensure system reliability and performance.* Understanding of Microservices & principles of RESTful API development, including structuring, documenting, versioning, testing ...