51 to 75 of 81 Observability Jobs in Central London

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused on AWS, Kubernetes, Terraform, observability, monitoring, and automation, working closely with software engineering teams to improve platform reliability and developer experience. You'll have genuine ownership and the opportunity to influence … infrastructure using Terraform and Infrastructure as Code principles Develop and optimise CI/CD pipelines using GitHub Actions Build and improve comprehensive monitoring and observability across the platform Implement and maintain effective logging, metrics, tracing, alerting, and dashboards Define and improve SLIs, SLOs, and reliability metrics Proactively identify and resolve ...

Platform / DevOps Engineer

Location
City of Westminster, England, United Kingdom
Manager, EventBridge provisioned as code, sized sensibly, and cost‐aware (you'll make the calls on things like NAT vs VPC endpoints). Own observability and reliability. CloudWatch alarms and dashboards, SNS alerting, data freshness and quality signals, and automated recovery for the pipelines that need it. When something breaks … deliberate applies, and easy rollbacks. Everything is code. Infrastructure, pipelines, access, and policy all live in version control and ship through review. Observability first. If we run it, we can see it - and we get told before our users do. You own what you ship. Strong ownership, low ceremony. ...

Remote Site Reliability Engineer - Cloud-Native Infra

Location
City of Westminster, England, United Kingdom
Site Reliability Engineer to help build, scale and operate the platforms powering Curve's products and services. This role focuses on reliability, security, and observability across our cloud-native stack, with a hybrid work model allowing UK-based remote work and partial office attendance. You will collaborate with engineers, product ...

Data Architect

Hiring Organisation
Experis
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£800 - £860 per day
quality requirements. Produce High Level Designs, Low Level Design principles, Architecture Decision Records and design assurance material. Define non-functional requirements for performance, resilience, observability, scalability, security and maintainability. Support assurance, accreditation and security review activities with architecture evidence and design rationale. Provide technical governance and design oversight during build ...

Senior Site Reliability Engineer

Location
City Of London, England, United Kingdom
development, validation, and optimization of configuration-as-code, improving delivery speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality … applications without these: Hands-on with Helm or Kustomize Experience with GitOps (e.g., Argo CD) Knowledge of secrets management (e.g., HashiCorp Vault) Experience with observability (metrics/logs/tracing) Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Platform Engineer

Location
City Of London, England, United Kingdom
complex platform challenges and guiding architectural decisions across teams Build & Evolve Platform Services: Design and continuously improve internal platform services (CI/CD, infrastructure, observability), treating the platform as a product with clear outcomes Drive Engineering Standards & Best Practices: Establish consistent patterns for security, scalability, and delivery, including Infrastructure … background in CI/CD, automation, and developer platform tooling Experience with .NET/C# ecosystems and modern software architecture Solid understanding of security, observability, and reliability engineering Leadership & Behavioural Capabilities: Demonstrated experience leading engineering teams and technical strategy Strong stakeholder management, with ability to influence senior technical ...

Senior Frontend Engineer — Form-Heavy React/TS

Location
City Of London, England, United Kingdom
onboarding clients and accounts with KYC/KYB workflows. You will design, implement, and own features end-to-end, ensuring security, reliability, and observability from day one. Strong collaboration and clear communication with engineers and non-engineers are essential. #J-18808-Ljbffr ...

Enterprise Platform Architect

Hiring Organisation
Opus Recruitment Solutions
Location
City Of London, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £580/day
standards, policies and patterns. Technical Focus Areas Data and analytics platforms including Cloudera and Tableau. Digital Workplace technologies including Microsoft 365, Azure and E5. Observability and monitoring platforms such as Datadog. Enterprise integration, APIs, API Gateways and Apigee. Enterprise application ecosystems and platform strategy. Experience Required Enterprise, Platform or Domain ...

Data Observability Engineer

Hiring Organisation
Ashdown Group
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
successful multinational technology business is looking for a Data Observability Engineer to join its growing data team in Central London. This role is hybrid youll be able to work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents, and building … scalable observability across a modern enterprise data platform. Youll help ensure data across the organisation is accurate, reliable, and trusted for critical business decision-making. Youll take ownership of data reliability end-to-end, designing and implementing frameworks that monitor data health, detect anomalies, and enforce standards across complex data ...

Director of Software Engineering (AIOps) - Executive Director

Location
City Of London, England, United Kingdom
reins and drive impact, we’ve got an opportunity just for you. As a Director of Software Engineeringat JPMorgan Chase within theEngineer's Observability Platforms team, you lead a technical area and drive impact within teams, technologies, and projects firm wide. Utilize your in-depth knowledge of software, applications, technical … troubleshooting root cause analysis finding for application support teams. Job responsibilities Leads technology and process implementations to achieve functional technology objectives in the Observability Platforms space, providing essential services for Site Reliability Engineers, Operations and Engineers across the whole firm Innovates, designs and delivers technical solutions that can be leveraged ...

Vice President, Full-Stack Engineer

Location
Westminster, West End, United Kingdom
engineering teams set clear objectives, coach talent, and foster succession planning. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure secure-by-design with threat … scalable, observable, testable systems strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud. Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows. Excellent ...

Lead Site Reliability Engineer

Location
Westminster, West End, United Kingdom
undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives in fast paced front office environments, enjoys … Deep knowledge of reliability engineering principles: SLIs/SLOs, real-time telemetry, disaster recovery planning, capacity planning, and performance tuning. Experience designing and implementing observability frameworks for mission critical systems. Proven ability to lead incident response and drive long term remediation. Solid programming skills in Python, Java, or Kotlin, with ...

System Architect | Production Domain

Hiring Organisation
Square One Resources
Location
City, London, United Kingdom
Employment Type
Contract
Contract Rate
GBP 600 - 650 Daily
product and system requirements into complete, consistent, and testable system requirements. Define ECU architectures supporting integration, validation, manufacturing, and production readiness. Ensure architecture provides observability, controllability, diagnosability, and scalability. Collaborate with validation teams to enable SIL, HIL, and EOL testing. Define requirements for flashing, provisioning, diagnostics, fault detection, and production ...

Chief Technology, Infrastructure & Product Engineering Officer

Location
City Of London, England, United Kingdom
transparent executive reporting. Drive delivery consistency across all technology portfolios. Enterprise Platforms, Cloud & DevOps Build enterprise-grade platform capabilities (cloud, integration, data, identity, observability). Own DevOps, CI/CD, automation, SRE and embedded AI capabilities to ensure resilient and scalable production operations. Reduce time-to-market while improving quality … modern service management and tooling. Service Resilience, Reliability & Operational Excellence Ensure availability, performance and resilience of critical services through SRE practices. Drive proactive monitoring, observability, capacity management and automated remediation. Improve incident reduction, root cause resolution, and service stability over time. Technology Risk, Security & Compliance Partnership Partner closely with ...

Senior Software Engineer

Location
City Of London, England, United Kingdom
Make the architectural calls on your domain, write them down, and defend them in front of the tribe. Build in security, availability, reliability and observability from the start, and instrument your services so the answer to what's happening is already in a dashboard. Use AI tooling well. … awkward parts that come with them: idempotency, ordering, retries, poison messages, exactly-once as a promise nobody can keep. Security, availability, reliability and observability built in from the start, not bolted on at the end. You think about who can reach what, you know how your service behaves when ...

Senior Lead Software Engineer - Mobile Engineering

Location
Westminster, West End, United Kingdom
drive measurable improvements in stability and release confidence. Own mobile build/release and operational maturity: CI/CD pipelines, distribution, feature flags, observability, crash/performance monitoring, and incident response. Mentor and coach engineers support team growth through feedback, technical guidance, and strong engineering culture. Communicate clearly with senior … accessibility standards, component libraries). Experience with CI/CD for mobile (e.g., build automation, signing, distribution, feature flags, release trains). Experience with observability and production support practices: crash analytics, performance monitoring, logging, alerting, and operational readiness. Experience leading multiple engineers/teams (people leadership or strong matrix leadership ...

Principal Consultant - Cloud & Engineering

Hiring Organisation
Zhlke Engineering Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
The Role As a Principal Consultant for Cloud & Engineering, you will provide senior technical and engineering leadership across complex client engagements. You will help clients shape and deliver practical strategies for cloud adoption, optimisation and ...

Platform Operations Director

Hiring Organisation
ClearCourse
Location
City of London, London, United Kingdom
business continuity across the group. Internal IT & Systems Manages internal IT and business systems administration (M365, NetSuite, SuccessFactors, SharePoint) -infrastructure, integrations, and IAM. Ensures observability and SRE capability is fit for purpose across cloud, hosted, and end-user environments. Vendor & Cost Management Drives cloud and vendor cost discipline - manages …/CD infrastructure requirements. • Head of Infrastructure & Cloud - Direct report. Hosting strategy, cloud platform, and FinOps execution. Head of SRE - Direct report. Observability, on-call, and DR/BCP processes. • Head of Internal Services - Direct report. Internal IT, business systems, and end-user support. Finance - Direct report. Cloud cost visibility ...

Dataiku Solution Architect

Hiring Organisation
Everforth Quinnox
Location
City of London, London, United Kingdom
Employment Type
Permanent
dashboards use controlled, reconciled, and traceable data from the governed platform. Define access, refresh, performance, lineage, and reconciliation standards for reporting solutions. Support curve observability and the monitoring of data quality, source availability, processing status, and workflow completion. Nonfunctional Architecture Define nonfunctional requirements for performance, scalability, security, availability, resiliency, recoverability … observability, maintainability, and supportability. Design monitoring and alerting across Dataiku, Power Automate, PostgreSQL or Amazon RDS, integrations, WebApps, and reporting components. Establish recovery patterns for failed source deliveries, workflow errors, data-quality issues, integration failures, and interrupted processing. Define capacity and performance considerations for regional processing, historical replay, concurrent users ...

SRE Managing Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£90000 - £100000/annum
include: Define and embed SRE engagement models aligned to modern engineering and traditional ITSM/ITIL practices Establish SLIs, SLOs, and Error Budgets Shape observability strategies using metrics, logs, and traces Design incident response models and post-incident learning loops Reduce toil through automation and engineering excellence Deliver SRE capability … Looking For Extensive experience in SRE, cloud operations, or DevOps Proven consulting or advisory background Experience with AWS, Azure, or GCP Strong observability and incident management expertise Ability to obtain UK SC clearance Modis International Ltd acts as an employment agency for permanent recruitment and an employment business ...

Staff Security Engineer

Location
City Of London, England, United Kingdom
engineer who enjoys solving complex security challenges at scale. You’ll work across engineering, data, AI and digital workplace teams to build security observability, automate control assurance and influence how security is embedded into products, platforms and processes. If you're passionate about turning security data into actionable insight, building … raising the security maturity of a fast-moving technology organisation, we'd love to hear from you. About the role Designing and building security observability capabilities that provide meaningful visibility across systems, infrastructure and applications Developing automated control monitoring, evidence collection and continuous testing solutions that strengthen security governance Partnering ...

Senior Security Engineer: Observability, Automation & Risk

Location
City Of London, England, United Kingdom
seeking a Staff Security Engineer to advance security observability, automate controls and strengthen governance across its global tech estate. This senior individual contributor role partners with engineering, data, AI and digital workplace teams to raise security maturity in a fast-moving environment. You will design scalable security observability, automate assurance ...

Principal Data Engineer

Location
City Of London, England, United Kingdom
follow the identical pattern so they are handover-ready by design. Drive data quality as a first-class, firm-wide concern: establish data contracts, observability, SLA/SLO monitoring, and automated alerting and remediation across ingestion and transformation layers, and hold squads to those standards. Act as the senior technical … with the ability to set standards, conduct code and design reviews, and grow engineers’ capabilities Strong grasp of data quality practices: data contracts, pipeline observability, SLA/SLO definition, and automated alerting and remediation Solid understanding of SQL transformation patterns and modern tooling such as dbt, alongside experience managing ingestion ...