1 to 25 of 1,823 Permanent Observability Jobs in London

Senior Observability Engineer

Location
Greater London, England, United Kingdom
## Senior Observability EngineerApplylocations: City of London - United Kingdomtime type: Full timeposted on: Posted Todayjob requisition id: R\_17538**Job Title**Senior Observability Engineer**Job Description****Senior Observability Engineer****Location: London****Employment type:** Permanent, Full Time**Reporting into:** Senior Engineering Manager – SRE and Observability**About IG Group**IG Group … role**IG Group’s systems move billions of dollars every day – and our clients expect them to be fast, reliable, and transparent. As an Observability Engineer, you will own the platforms and practices that give IG’s engineering teams deep, real-time insight into how those systems behave. This ...

Director, Applied AI & Agentic Platform Engineering

Location
Greater London, England, United Kingdom
working knowledge of AWS; Azure exposure optional (not a dependency) Containers: Docker, Kubernetes (GKE/EKS) IaC: Terraform CI/CD: GitHub Actions, Jenkins Observability: Splunk, ELK, Prometheus, Grafana Security & Compliance Secure coding, API security, Zero Trust Data privacy, encryption, access control Regulatory compliance and AI governance (MRM) What ...

Director, Applied AI & Agentic Platform Engineering

Location
Greater London, England, United Kingdom
working knowledge of AWS; Azure exposure optional (not a dependency) Containers: Docker, Kubernetes (GKE/EKS) IaC: Terraform CI/CD: GitHub Actions, Jenkins Observability: Splunk, ELK, Prometheus, Grafana Security & Compliance Secure coding, API security, Zero Trust Data privacy, encryption, access control Regulatory compliance and AI governance (MRM) What ...

Sr. Technology Architect - GCP and AWS Network architecture - UK, Germany, Netherlands

Hiring Organisation
Infosys Technologies
Location
London, United Kingdom
Salary
£ 80 K
EnablementDesign and implement end-to-end automation pipelines for build, test, release, and operational workflows.Partner with Site Reliability Engineering (SRE) teams to facilitate observability, SLO/SLI tracking, error budgeting, and automated operations.Oversee centralized logging and monitoring using CloudWatch, Cloud Monitoring, Prometheus/Grafana, and SIEM tools.Required Skills & ExperienceTechnical SkillsExpert ...

DevOps Engineer (Security Cleared)

Hiring Organisation
Solirius Consulting
Location
London, United Kingdom
Salary
£ 60 K
scripting and automation using Python, Bash, PowerShell, or similar languages.Experience with configuration management tools such as Ansible, Puppet, or Chef.Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog.Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles.Experience working in Agile and DevOps ...

DevOps Engineer (Security Cleared)

Hiring Organisation
Solirius Consulting
Location
London, UK
Employment Type
Full-time
automation using Python, Bash, PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile ...

DevOps Engineer (Security Cleared)

Location
Greater London, England, United Kingdom
automation using Python, Bash, PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile ...

Senior Software Engineer - Backend

Hiring Organisation
Fitch Ratings
Location
London, United Kingdom
Salary
£ 80 K
cross-functional stakeholders to prioritize work, align technical investments, and achieve business outcomes. Ensure high-quality software delivery through automated testing, code reviews, observability, and engineering governance. Lead resolution of complex technical and operational challenges while improving platform performance, resiliency, and operational excellence. Champion DevSecOps, CI/CD, and automation ...

Principal Cloud Engineer (Terraform), London

Location
Greater London, England, United Kingdom
Apply data quality and validation frameworks to ensure accuracy, completeness, and freshness of cost and usage data across all cloud providers; instrument pipelines with observability tooling to surface issues proactively. Build and maintain reusable data assets — curated datasets, aggregations, and data marts — that power FinOps dashboards, showback/chargeback reporting ...

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
produce, distribute and promote the most critically acclaimed and commercially successful music to delight and entertain fans around the world.As a Senior Observability Engineer, you will be a driving force for technical excellence and strategic vision within our global team. You will be instrumental in architecting, building, and leading … comprehensive observability strategy to ensure the reliability, performance, and scalability of our critical IT systems. This senior role demands a passion for data-driven strategy, a commitment to automation, and the ability to mentor and lead. You will not only solve complex technical challenges but also influence the direction ...

Senior AI Engineer - Perm UK

Hiring Organisation
INFUSED SOLUTIONS LIMITED
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£80,000
Data Scientists to productionise Machine Learning and NLP models. Develop high-performance RESTful APIs and microservices for AI model serving. Drive system reliability, monitoring, observability and performance optimisation. Champion engineering best practices including CI/CD, automated testing and Infrastructure as Code. Improve integration, interoperability and data exchange across enterprise ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
help ensure that systems used by clinical, operational, and administrative teams remain stable, secure, and available. This role is responsible for building automation, improving observability, reducing manual operational work, supporting integrations, and helping maintain reliable systems that directly impact patient care and business operations. This is a hybrid position with ...

DevOps Engineer, Studios

Hiring Organisation
iMG world
Location
London, United Kingdom
Salary
£ 60 K
software delivery across multiple teams.Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar.Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack.Ensure high availability and disaster recovery strategies are in place and tested regularly.Collaborate closely with development ...

Senior DevOps Engineer (Azure)

Hiring Organisation
Darktrace
Location
London, United Kingdom
Salary
£ 80 K
Docker and Kubernetes, along with supporting technologies like ArgoCD and Helm. You should also be familiar and well-versed in monitoring, logging, and observability tools including: Prometheus; Grafana; Loki; OpenTelemetry; and the ELK stack. Amongst this, you should be able to demonstrate:Proven experience as a DevOps Engineer, with ...

Site Reliability Engineer III

Hiring Organisation
Appcast
Location
London, UK
Distribution (SFTP/JScape)—to Google Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.Incident ...

Production Engineer

Location
Greater London, England, United Kingdom
identify patterns, and drive intelligent automation solutions* Hands-on experience with containerization technologies such as Docker and Kubernetes, including cluster management, deployment, scaling, and observability* Deep practical knowledge of algorithmic trading workflows, including the behaviour, lifecycle, and risk controls of execution algos used across the EMEA markets* Experience designing ...

Production Engineer

Location
City Of London, England, United Kingdom
identify patterns, and drive intelligent automation solutions Hands-on experience with containerization technologies such as Docker and Kubernetes, including cluster management, deployment, scaling, and observability Deep practical knowledge of algorithmic trading workflows, including the behaviour, lifecycle, and risk controls of execution algos used across the EMEA markets Experience designing ...

AI Engineer

Hiring Organisation
Formula Recruitment Limited
Location
London, United Kingdom
Salary
£ 80 K
with cloud platforms (AWS, GCP or Azure) and infrastructure-as-code such as TerraformHands-on with DevOps practices (CI/CD, Docker, Kubernetes) and observability tools like Prometheus, Grafana or DatadogExperience with distributed systems, scaling, and both SQL and NoSQL datastoresAgentic workflow experience (autonomous or multi-agent systems ...

Senior AI Engineer - Perm - London

Hiring Organisation
INFUSED SOLUTIONS LIMITED
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
productionise Machine Learning, NLP and Generative AI models. Build and optimise RESTful APIs for AI model serving. Monitor system reliability, model performance and observability across production environments. Drive best practices across CI/CD, automated testing, Infrastructure as Code and MLOps. Improve system integration and interoperability across enterprise platforms. Support ...

Infrastructure Security Engineer

Hiring Organisation
X
Location
London, United Kingdom
Salary
£ 70 K
HIPAA, PCI DSS, NIST CSF)Strong proficiency with Python, Terraform, and configuration management (e.g., Puppet)Hands-on experience building GitHub Actions and workflowsExperience with observability and security operations tooling (e.g., Prometheus, Grafana, CloudWatch, Karma; SIEM platforms such as Wazuh)Experience in building custom cloud security tools or integrationsInterest in leveraging ...

Manager, DE , TC, FS

Location
Greater London, England, United Kingdom
multi‐disciplinary squads, partnering with product owners, business stakeholders, architects, QA, DevOps, security, data and infrastructure teams. Drive production readiness including CI/CD, observability, resilience, security, automated testing, release management, runbooks, incident response and root cause analysis. Mentor engineers and senior consultants, building capability in modern engineering practices ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
eliminate infrastructure waste. Establish benchmarking and telemetry to track unit economics and throughput for training and serving AI models. Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering ...

Senior/Staff Software Engineer, Developer Experience

Location
Greater London, England, United Kingdom
environments Familiarity with infrastructure-as-code principles Experience with container orchestration and management Knowledge of performance testing tools and frameworks Experience with monitoring and observability tools Background in test framework development Strong working knowledge of Helm charts and ArgoCD Infrastructure-as-code experience (Terraform, Pulumi, or similar) What ...

ML Ops Engineer

Hiring Organisation
Anaplan
Location
London, United Kingdom
Salary
£ 80 K
resource allocation to eliminate infrastructure waste.Establish benchmarking and telemetry to track unit economics and throughput for training and serving AI models.Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow.Your SkillsHands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with ...

Senior Software Engineer - Risk & FX

Hiring Organisation
Visa
Location
London, United Kingdom
Salary
£ 80 K
next phase of growth, are written to 12-factor principles and fit into our microservices architecture Cloud-related tools, services, and distributed system observability to support these applications, such as Docker, Kubernetes, ElasticSearch, log management systems, and Datadog APM, to name but a few API specifications, conforming to the OpenAPI ...