1 to 25 of 3,187 Observability Jobs

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
produce, distribute and promote the most critically acclaimed and commercially successful music to delight and entertain fans around the world.As a Senior Observability Engineer, you will be a driving force for technical excellence and strategic vision within our global team. You will be instrumental in architecting, building, and leading … comprehensive observability strategy to ensure the reliability, performance, and scalability of our critical IT systems. This senior role demands a passion for data-driven strategy, a commitment to automation, and the ability to mentor and lead. You will not only solve complex technical challenges but also influence the direction ...

Software Engineering Manager

Hiring Organisation
CyberCoders
Location
London, United Kingdom
Salary
£ 80 K
with product, QA, and operations to deliver reliable production services. The role combines technical leadership, people management, and hands-on engineering to ensure performance, observability, security, and cost-effective operation of cloud-native systems.Key ResponsibilitiesLead, mentor, and grow multiple engineering teams; hire, coach, conduct performance reviews, and promote career development.Define … DevOps practices and pipeline orchestration (Jenkins, GitLab CI, GitHub Actions) to accelerate safe delivery and automated releases.Collaborate with SRE and Ops on monitoring, observability, logging and tracing (Prometheus, Grafana, ELK, Jaeger/Zipkin) to ensure SLAs, incident response and release & incident management.Drive performance tuning, capacity planning and cost optimization ...

Director, Applied AI & Agentic Platform Engineering

Location
Greater London, England, United Kingdom
working knowledge of AWS; Azure exposure optional (not a dependency) Containers: Docker, Kubernetes (GKE/EKS) IaC: Terraform CI/CD: GitHub Actions, Jenkins Observability: Splunk, ELK, Prometheus, Grafana Security & Compliance Secure coding, API security, Zero Trust Data privacy, encryption, access control Regulatory compliance and AI governance (MRM) What ...

Director, Applied AI & Agentic Platform Engineering

Location
Greater London, England, United Kingdom
working knowledge of AWS; Azure exposure optional (not a dependency) Containers: Docker, Kubernetes (GKE/EKS) IaC: Terraform CI/CD: GitHub Actions, Jenkins Observability: Splunk, ELK, Prometheus, Grafana Security & Compliance Secure coding, API security, Zero Trust Data privacy, encryption, access control Regulatory compliance and AI governance (MRM) What ...

Senior Software Engineer

Hiring Organisation
Iris Software
Location
United Kingdom
Salary
£ 70 K
highly scalable solutions and internet-facing traffic levelsPerformance & Scalability: Profiling and benchmarking applications.Application Security: Confident vulnerability management, thread modelling and trackingProduction Support: Knowledge of observability and production support practices. Proficient in debugging complex issues, performance optimization, and production troubleshootingExperience Requirements5-7 years of professional software development experienceProven ability in delivering … systems with tools and services for richer, automated workflowsExpertise with advanced monitoring and APM strategies using Datadog, including custom dashboards and alerting (or similar observability platform)Expertise with modern UI architecture patterns (micro-frontends, SSR/SSG)Knowledge of security best practices and compliance requirements (OAuth2, OIDC, RBAC)Experience with ...

Full Stack Engineer - AI Enabled - Senior Vice President

Hiring Organisation
Citigroup
Location
London, United Kingdom
Salary
£ 80 K
results-oriented with a strong sense of ownership.Preferred Qualifications:Experience in the financial services industry.Knowledge of domain-driven design and clean architecture principles.Experience with observability tools (e.g., Prometheus, Grafana, ELK stack).Contributions to open-source projects or active participation in developer communities.What we’ll provide you:By joining Citi London ...

Full Stack Engineer - AI Enabled - Senior Vice President

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
strong sense of ownership. Preferred Qualifications: Experience in the financial services industry. Knowledge of domain-driven design and clean architecture principles. Experience with observability tools (e.g., Prometheus, Grafana, ELK stack).Contributions to open-source projects or active participation in developer communities. What we'll provide you: By joining Citi London ...

Staff Software Engineer

Location
Greater London, England, United Kingdom
Source control and code management (Git, branching strategies) Software design and development (clean code, code reviews) Service operations and resiliency (fault tolerance, circuit breakers, observability) Metrics definition, instrumentation, and monitoring (Prometheus, Grafana, or equivalent) Scripting (Bash, Python — for automation and tooling) Architecture & Solution Design Distributed systems design and architecture Architectural ...

Staff Software Engineer

Hiring Organisation
BP
Location
London, United Kingdom
Salary
£ 80 K
practicesSource control and code management (Git, branching strategies)Software design and development (clean code, code reviews)Service operations and resiliency (fault tolerance, circuit breakers, observability)Metrics definition, instrumentation, and monitoring (Prometheus, Grafana, or equivalent)Scripting (Bash, Python — for automation and tooling)Architecture & Solution DesignDistributed systems design and architectureArchitectural awareness ...

Senior Lead SRE: Reliability, Observability & Resiliency

Location
Auchentibber, Scotland, United Kingdom
integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch reliability and observability for our most critical platforms. As a Senior Lead Site Reliability/DevOps Engineer at JPMorgan Chase within the Commercial & Investment Bank, you are an integral part … significant business impact through your capabilities and contributions, and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of reliability, observability, and performance challenges that span multiple technologies and applications. Job responsibilities Regularly provides technical guidance and direction on site reliability practices to support the business ...

Senior Site Reliability Engineer

Hiring Organisation
NICE Systems
Location
United Kingdom
Salary
£ 60 K
tools such as Jenkins, GitLab CI/CD, or CircleCI.Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture.Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch).Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems.Experience of Incident management and blameless … have an advantage if you also have:Handson experience of working with large Kubernetes Cluster. Certification will be an added plus.Working experience of Grafana Observability Suite (Loki, Mimir, Tempo).Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity with configuration ...

Senior Cloud Site Reliability Engineer

Hiring Organisation
NICE Systems
Location
London, United Kingdom
Salary
£ 80 K
tools such as Jenkins, GitLab CI/CD, or CircleCI.Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture.Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch).Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems.Experience of Incident management and blameless … have an advantage if you also have:Handson experience of working with large Kubernetes Cluster. Certification will be an added plus.Working experience of Grafana Observability Suite (Loki, Mimir, Tempo).Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity with configuration ...

Senior Cloud Site Reliability Engineer

Hiring Organisation
NICE Systems
Location
Southampton, Hampshire, United Kingdom
Salary
£ 70 K
tools such as Jenkins, GitLab CI/CD, or CircleCI.Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture.Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch).Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems.Experience of Incident management and blameless … have an advantage if you also have:Handson experience of working with large Kubernetes Cluster. Certification will be an added plus.Working experience of Grafana Observability Suite (Loki, Mimir, Tempo).Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity with configuration ...

Software Engineer /Tech Lead – Web Platforms - Technology Consulting

Hiring Organisation
Business Integration Partners
Location
London, United Kingdom
Salary
£ 80 K
hands-on through coding, prototyping, code reviews, troubleshooting and resolution of complex technical challenges.Establish and promote engineering standards across code quality, testing, security, performance, observability and documentation.Guide the delivery of cloud-native, microservices-based and API-led solutions.Oversee technical delivery across development, test, release and production environments.Work closely with Product … infrastructure as code using Terraform, CloudFormation or similar.Knowledge of asynchronous and event-driven architectures, messaging platforms and integration patterns.Experience with monitoring, logging, tracing and observability platforms.Understanding of performance engineering, capacity planning, resilience and high-availability patterns.Experience integrating digital platforms with identity, customer, payment, CRM or other enterprise systems.Experience working within ...

Software Engineer /Tech Lead - Web Platforms - Technology Consulting

Hiring Organisation
Business Integration Partners
Location
London, UK
Employment Type
Full-time
through coding, prototyping, code reviews, troubleshooting and resolution of complex technical challenges. Establish and promote engineering standards across code quality, testing, security, performance, observability and documentation. Guide the delivery of cloud-native, microservices-based and API-led solutions. Oversee technical delivery across development, test, release and production environments. Work closely … code using Terraform, CloudFormation or similar. Knowledge of asynchronous and event-driven architectures, messaging platforms and integration patterns. Experience with monitoring, logging, tracing and observability platforms. Understanding of performance engineering, capacity planning, resilience and high-availability patterns. Experience integrating digital platforms with identity, customer, payment, CRM or other enterprise systems. ...

Strategic DevSecOps Consultant

Hiring Organisation
CloudBees
Location
London, United Kingdom
Salary
£ 80 K
DevSecOps, software delivery modernization, platform engineering, or cloud transformation initiatives.Strong understanding of modern software delivery practices, including CI/CD, Infrastructure as Code, GitOps, observability, security, and cloud-native architectures.Experience designing and implementing scalable DevSecOps and Platform Engineering solutions in enterprise environments.Hands-on experience with public cloud platforms such … DevEx) initiatives, or internal developer platforms (IDPs).Familiarity with AI-enabled software development, agentic workflows, large language models (LLMs), or AI governance practices.Experience with observability and telemetry platforms such as OpenTelemetry, Splunk, Dynatrace, Datadog, AppDynamics, Grafana, or similar technologies.Experience working with large-scale enterprise architecture, governance, compliance, and regulated environments.Thought ...

Senior Java Developer

Hiring Organisation
scrumconnect ltd
Location
Swansea, West Glamorgan, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
projects). Experience implementing cloud-native architectures and event-driven systems. Familiarity with Infrastructure as Code tools such as Terraform or CloudFormation. Experience with observability and monitoring tools such as ELK, Grafana, Prometheus, or Splunk. Relevant Java, AWS, Azure, GCP, Kubernetes, or architecture certifications. Experience with Domain-Driven Design ...

DevOps Engineer (Security Cleared)

Hiring Organisation
Solirius Consulting
Location
London, United Kingdom
Salary
£ 60 K
scripting and automation using Python, Bash, PowerShell, or similar languages.Experience with configuration management tools such as Ansible, Puppet, or Chef.Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog.Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles.Experience working in Agile and DevOps ...

DevOps Engineer (Security Cleared)

Hiring Organisation
Solirius Consulting
Location
London, UK
Employment Type
Full-time
automation using Python, Bash, PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile ...

DevOps Engineer (Security Cleared)

Location
Greater London, England, United Kingdom
automation using Python, Bash, PowerShell, or similar languages. Experience with configuration management tools such as Ansible, Puppet, or Chef. Knowledge of monitoring, logging, and observability platforms such as Prometheus, Grafana, ELK Stack, Splunk, or Datadog. Strong understanding of Linux administration, networking, cloud security, and DevSecOps principles. Experience working in Agile ...

Senior Software Engineer - Backend

Hiring Organisation
Fitch Group
Location
Greater London, United Kingdom
Employment Type
Full Time
cross-functional stakeholders to prioritize work, align technical investments, and achieve business outcomes. Ensure high-quality software delivery through automated testing, code reviews, observability, and engineering governance. Lead resolution of complex technical and operational challenges while improving platform performance, resiliency, and operational excellence. Champion DevSecOps, CI/CD, and automation ...

Senior Solutions Engineer

Hiring Organisation
Kroll
Location
United Kingdom
Salary
£ 70 K
expertise with strong execution focus and cross-functional influence. This role will play a critical part in elevating infrastructure reliability, release quality, automation maturity, observability standards, and architectural scalability.Day-to-day responsibilities:Technical LeadershipActively contribute to infrastructure, CI/CD, automation, and reliability initiatives.Lead improvements in:Infrastructure as Code … stability.Standardize DevOps workflows across teams.Infrastructure & Platform Enablement Oversee infrastructure management across cloud and/or hybrid environments.Champion Infrastructure as Code and automation-first approaches.Improve observability through logging, monitoring, tracing, and alerting.Partner with security teams to ensure infrastructure and pipeline best practices.Roadmap & Execution OwnershipCollaborate with Platform Engineers to define:Quarterly roadmapPrioritized ...

Senior DevOps Platform Engineer

Hiring Organisation
LEAP29
Location
London, United Kingdom
Salary
£ 80 K
platform reliability, scalability and security through automation and engineering best practicesSupport cloud migrations from traditional data centre environments into modern cloud-native platformsImplement monitoring, observability and logging solutions across complex distributed systemsWork closely with development, security and architecture teams to improve software delivery processesRequired ExperienceStrong commercial experience working … financial services, government or healthcareExperience with GitOps practices and tools such as ArgoCDKnowledge of service mesh technologies including Istio or similarExperience with monitoring and observability platforms such as Prometheus, Grafana, Dynatrace, New Relic, Splunk or ELKExperience supporting hybrid cloud environments across AWS, Azure, GCP or private cloudExperience with security tooling ...

Senior Engineer

Location
Greater London, England, United Kingdom
evolution initiatives is strongly desirable. Strong engineering discipline across testing, CI/CD, version control, release practices, and production support. Good understanding of monitoring, observability, and building reliable services for business‐critical environments. Stakeholder Engagement: Proven experience working with C‐level stakeholders, translating complex business requirements into effective technology solutions … Linux systems administration and scripting. Testing: Commitment to automated testing and quality assurance, with experience integrating these practices into CI/CD pipelines. Monitoring & Observability: Familiarity with monitoring, logging, and alerting tools (Prometheus, Grafana, Splunk) is advantageous. Analytical & Problem‐Solving: Excellent analytical and problem‐solving skills, focusing on delivering measurable ...

Senior Engineer

Location
Greater London, England, United Kingdom
evolution initiatives is strongly desirable.* Strong engineering discipline across testing, CI/CD, version control, release practices, and production support.* Good understanding of monitoring, observability, and building reliable services for business-critical environments.* Stakeholder Engagement: Proven experience working with C-level stakeholders, translating complex business requirements into effective technology solutions … Linux systems administration and scripting.* Testing: Commitment to automated testing and quality assurance, with experience integrating these practices into CI/CD pipelines.* Monitoring & Observability: Familiarity with monitoring, logging, and alerting tools (e.g., Prometheus, Grafana, Splunk) is advantageous.* Analytical & Problem-Solving: Excellent analytical and problem-solving skills, with a focus ...