101 to 125 of 156 Observability Jobs in Central London

Full Stack Developer

Location
City Of London, England, United Kingdom
where needed and decide when to extend in n8n versus pull logic back into code. Set our internal engineering standards. Branching, review, deploys, environments, observability, security baseline, dependency hygiene. Build the foundations. Build AI native, by default Treat AI as a first‐class building material, not a bolted‐on feature. ...

Fullstack Engineer

Location
City Of London, England, United Kingdom
things You take ownership of your work and treat the company’s success as your own Nice-to-Haves Experience with analytics, monitoring or observability tooling Background in automotive, marketplaces, or high-SKU B2B platforms Exposure to AI/ML-powered products or startups Prior experience leading UI best practices ...

Senior Site Reliability Engineer

Location
City Of London, England, United Kingdom
development, validation, and optimization of configuration-as-code, improving delivery speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality … applications without these: Hands-on with Helm or Kustomize Experience with GitOps (e.g., Argo CD) Knowledge of secrets management (e.g., HashiCorp Vault) Experience with observability (metrics/logs/tracing) Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Senior Software Engineer - Commercial Trading

Location
City of Westminster, England, United Kingdom
business. You’ll play a key role in shaping technical direction, improving engineering standards and ensuring our products and platforms are built with strong observability, operational excellence and best-in-class engineering practices. Due to high interest, this role may close earlier than advertised. We recommend applying as soon … team uses a variety of modern technologies, including: Backend: Java, Spring, Spring Boot, Micronaut Frontend: React, Next.js, TypeScript, Angular Cloud & Infrastructure: Azure Cloud, Kubernetes Observability: Dynatrace Databases: SQL Server, MongoDB Caching & Performance: Ignite, Redis What’s in it for you? Working at M&S means being part of something bigger ...

Lead SRE - Chase UK

Location
Westminster, West End, United Kingdom
possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with engineering teams to ensure … knowledge of microservice infrastructure components, including service discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes ...

IBM Netcool / Observability Technical Lead

Hiring Organisation
Deerfoot Recruitment Solutions
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£780 - £830 per day
Netcool/Observability Technical Lead Inside IR35 Contract -up to £827pd London Hybrid - 4 Days Onsite/1 Day WFH per Week Banking Are you the person who knows exactly why an ObjectServer failover didn't behave as expected, and how to stop a flood of duplicate events before anyone … shape how thousands of infrastructure and application events are detected, correlated and actioned across EMEA, and you'll have genuine scope to modernise observability capability rather than simply keep the lights on. This is a hands-on technical leadership role with no direct reports, so your influence comes from your ...

AI Metrics & Model Evaluation Lead

Location
City Of London, England, United Kingdom
metrics and model evaluation for cutting‐edge AI products used in highly regulated industries. You’ll define metrics, build dashboards, and ensure observability of product performance from user interaction to model outputs. You will partner with AI leadership, Product and Engineering to drive data‐informed decisions, create golden datasets ...

AI Strategist // £110,000 // Remote

Hiring Organisation
Tenth Revolution Group
Location
Central London / West End, London, United Kingdom
best practices. It would be great if you had; Demonstrable understanding of AI models, generative AI technologies, and intelligent automation architectures Experience implementing monitoring, observability, and altering frameworks for automated systems. Ability to translate business requirements into scalable automation solutions Experience developing automation solutions using platforms such as Copilot Studio ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
persistence platforms Provide technical leadership and architectural guidance across multiple engineering teams Define engineering standards, platform roadmaps and best practices Drive automation, resilience, observability and operational excellence initiatives Support and mentor engineers through code reviews, coaching and technical leadership Collaborate with architects and stakeholders to translate business requirements into technical … automation and DevOps practices Experience mentoring engineers and providing technical leadership Key Technologies AWS Terraform Linux Cassandra Couchbase ScyllaDB Kafka CI/CD Pipelines Observability & Monitoring Platforms Distributed Database Technologies Nice to Have Experience with additional distributed persistence technologies Background in large-scale cloud-native environments Experience defining enterprise platform ...

Lead Site Reliability Engineer

Location
Westminster, West End, United Kingdom
undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives in fast paced front office environments, enjoys … Deep knowledge of reliability engineering principles: SLIs/SLOs, real-time telemetry, disaster recovery planning, capacity planning, and performance tuning. Experience designing and implementing observability frameworks for mission critical systems. Proven ability to lead incident response and drive long term remediation. Solid programming skills in Python, Java, or Kotlin, with ...

Contract - Senior CXE Engineer - Amazon Connect

Hiring Organisation
INNOVATIVE TECH PEOPLE LTD
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
hands-on: model tier selection (Haiku vs. Sonnet vs. Opus), prompt caching, and token budgeting against containment-rate targets Diagnose AI agent performance using observability tooling (agent spans, and CloudWatch) that correlates contact flow logs, conversation transcripts, AI agent spans, tool executions, and token usage to isolate latency, cost … Functions, Kinesis) Infrastructure as code proficiency with AWS CDK or Terraform, including multi-account deployment patterns Experience shipping and supporting production systems, testing discipline, observability instrumentation, and incident debugging. ...

Senior Software Engineer / Senior AI Engineer

Location
City Of London, England, United Kingdom
within a defined problem, building and testing tool use, retrieval pipelines, and agent workflows, integrating AI capabilities into enterprise systems, and contributing to evaluation, observability, and guardrails. You will hold a high bar on code quality, flag risks and blockers early, and work alongside host-function stakeholders to make sure … agentic AI solutions to production standards within a defined technical approach. Implement and test tool use, retrieval pipelines, and agent workflows. Contribute to evaluation, observability, and guardrails for agentic systems. Integrate AI capabilities into existing enterprise workflows and systems. Maintain high code quality and documentation so patterns can be reused. ...

ClickHouse Platform Architect - Greenfield, High-Scale

Location
City Of London, England, United Kingdom
Colehouse Group is seeking a ClickHouse Solutions Architect to build a greenfield, enterprise-scale observability platform for a global banking client. You will drive the architectural decisions from first principles to meet performance, multi-tenancy and retention requirements. The role covers discovery, data-modeling, ingestion, storage tiering and security, across ...

AI Metrics & Model Evaluation Analyst

Location
City Of London, England, United Kingdom
products are used, how models perform, and what changes drive business value in regulated environments. In this role you’ll define metrics, build observability, create dashboards, and partner with Product and Engineering to ensure meaningful, data-driven release decisions that advance AI capabilities. #J-18808-Ljbffr ...

Strategic Global Enterprise Account Director

Location
City Of London, England, United Kingdom
will act as a Strategic Hunter and Executive Orchestrator, rebuilding executive relationships and shaping multi-year pipelines across Cisco’s networking, security, cloud, observability, and services portfolio. Responsibilities include reactivating dormant accounts, architecting 12–24 month #J-18808-Ljbffr ...

Principal Engineer I, Prepurchase Platform (Remote)

Location
City Of London, England, United Kingdom
demand on-sales. You will write production code daily, influence technical direction, and collaborate across multiple teams within the Prepurchase domain. You will drive observability, resilience patterns, and AI-assisted enhancements while embedding across services or working horizontally. #J-18808-Ljbffr ...

Core Platform Developer

Location
City Of London, England, United Kingdom
reliability of internal systems. This person should be comfortable working across multiple areas of the stack, from service frameworks and API enablement to observability, governance, and developer workflows. This is a high-ownership role within a global, fast-moving engineering environment. Key Responsibilities Design and build shared backend services, frameworks … developer tooling that support internal application and service development. Develop common platform capabilities such as service templates, authentication and authorization patterns, API standards, observability integrations, error handling, and shared runtime utilities. Improve the developer experience through better tooling, automation, documentation, onboarding patterns, and paved-road workflows for engineering teams. Help ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
Camden, London, Camden Town, United Kingdom
Employment Type
Permanent
Salary
£65000 - £70000/annum + Remote + Progression
automate deployments using Terraform and Ansible, and build CI/CD pipelines using GitHub Actions. You'll also work across cloud security, networking and observability, while having the opportunity to develop your technical expertise through training and professional certifications. This role would suit an experienced DevOps Engineer looking to work … private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join ...

Senior Associate, Full-Stack Engineer

Location
Westminster, West End, United Kingdom
ways: Design, build, and maintain backend services, batches and APIs, contributing to UI components as needed. Own end-to-end delivery: implementation, testing, deployment, observability, and reliability. Write clean, well-tested code participate in code reviews and continuous improvement. Collaborate with product, design, and operations to translate business needs into … microservices Proficiency in Java with Spring. Experience with CI/CD, automated testing (JUnit/Spock), and containers (Docker). Familiarity with microservices, observability/telemetry (e.g., Splunk, AppDynamics), and cloud deployments. Curiosity to understand the business domain and translate product strategy into technical solutions. How we work: Agile (Scrum ...

Technical Leader

Location
City Of London, England, United Kingdom
manage technical debt, prioritizing improvements that provide meaningful value to the team and the product. Promote engineering best practices around testing, CI/CD, observability, documentation, and operational excellence. Stay current with emerging technologies and industry practices, evaluating when new technologies can provide meaningful improvements to our products and engineering … computer science fundamentals, including data structures, algorithms, concurrency, and system design. Strong understanding of modern software development practices, including CI/CD, TDD, DevOps, observability, and automated quality gates. Confident communication and collaboration skills, with the ability to articulate technical decisions, challenge assumptions, and explain complex technical concepts to both ...

Cloud Platforms Engineer

Location
City Of London, England, United Kingdom
runs the internal platform that PEI’s engineering and data teams build on: the cloud accounts, the reusable infrastructure code, the delivery pipelines, the observability, and the security and cost guardrails that wrap around them. We treat that platform as a product with internal customers – the measure of our work … Desirable: experience with data platform infrastructure such as Databricks, or similar – prior Databricks experience is not required. Desirable: experience with Datadog, or another mature observability platform. Desirable: experience of high‐traffic, international, content‐heavy web platforms. Technical Skills Confident with Linux and containers, and able to debug from the command ...

DevOps Lead

Hiring Organisation
TurleyWay Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
manage and optimise containerised environments using Kubernetes. Oversee source control, CI/CD pipelines and development workflows within GitHub and build and maintain monitoring, observability and alerting solutions using Grafana. To be considered you be able to demonstrate proven experience in a DevOps Lead, Senior DevOps Engineer, or Platform Engineering … expertise in Kubernetes and container orchestration technologies, administering and managing GitHub and CI/CD pipelines. Solid knowledge of Grafana and modern monitoring/observability practices and extensive database exposure, including performance, administration, and optimisation of enterprise database environments. In return we offer a competitive basic salary plus bonus scheme ...

DevOps Manager

Hiring Organisation
TurleyWay Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
manage and optimise containerised environments using Kubernetes. Oversee source control, CI/CD pipelines and development workflows within GitHub and build and maintain monitoring, observability and alerting solutions using Grafana. To be considered you be able to demonstrate proven experience in a DevOps Lead, Senior DevOps Engineer, or Platform Engineering … expertise in Kubernetes and container orchestration technologies, administering and managing GitHub and CI/CD pipelines. Solid knowledge of Grafana and modern monitoring/observability practices and extensive database exposure, including performance, administration, and optimisation of enterprise database environments. In return we offer a competitive basic salary plus bonus scheme ...