126 to 144 of 144 Observability Jobs in Central London

Senior Pre-Sales Consultant, Enterprise Payments

Location
City Of London, England, United Kingdom
workshops, both virtually and on-site. Lead and contribute to RFI and RFP responses, providing detailed technical input across areas including: Integration Security Resiliency Observability Compliance Identify risks, gaps, constraints, and assumptions while clearly articulating solution trade-offs. Plan, design, and support Proof of Concept (POC) activities, enabling customers … identity frameworks, including OpenID Connect and OAuth Cloud platforms such as: AWS Microsoft Azure Google Cloud Platform (GCP) IBM Cloud Oracle Cloud Monitoring and observability tools, including Splunk and SIEM solutions Pre-Sales and Customer-Facing Experience Proven experience in a customer-facing technical role, solution architecture role ...

Dataiku Solution Architect

Hiring Organisation
Everforth Quinnox
Location
City of London, London, United Kingdom
Employment Type
Permanent
dashboards use controlled, reconciled, and traceable data from the governed platform. Define access, refresh, performance, lineage, and reconciliation standards for reporting solutions. Support curve observability and the monitoring of data quality, source availability, processing status, and workflow completion. Nonfunctional Architecture Define nonfunctional requirements for performance, scalability, security, availability, resiliency, recoverability … observability, maintainability, and supportability. Design monitoring and alerting across Dataiku, Power Automate, PostgreSQL or Amazon RDS, integrations, WebApps, and reporting components. Establish recovery patterns for failed source deliveries, workflow errors, data-quality issues, integration failures, and interrupted processing. Define capacity and performance considerations for regional processing, historical replay, concurrent users ...

Vice President, Production Services Application Support

Location
Westminster, West End, United Kingdom
risk while improving platform stability. Drive initiatives through to completion with strong ownership, accountability, urgency, and quality. Champion an automation-first mindset, leveraging AI, observability, and tooling to reduce manual effort and improve service quality. Identify systemic issues and drive sustainable remediation through process simplification, platform improvements, and close partnership … incident, problem, and change management with measurable improvements in stability and service recovery. Demonstrated automation-first and AI-enabled mindset, with experience driving tooling, observability, and process automation. Strong ownership mentality and execution focus, with the ability to take initiatives from concept through delivery and embed sustainable outcomes. Deep technical ...

Senior AI Full Stack Engineer

Location
City of Westminster, England, United Kingdom
wider product capabilities around it. This includes building chat interfaces, agentic workflows, tool integrations, structured outputs, retrieval and grounding mechanisms, evaluation frameworks, observability, and appropriate human-in-the-loop controls. The role requires strong experience with TypeScript, React and modern full-stack web development. Experience with the Vercel … human-in-the-loop approval processes for sensitive or high-impact actions. Ensure transparency through source attribution, execution tracking, and auditability. AI Quality, Evaluation & Observability Develop evaluation frameworks, test datasets, and quality metrics for AI features. Monitor answer quality, groundedness, tool usage, latency, output accuracy, and operational cost. Implement logging ...

SRE Managing Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£90000 - £100000/annum
include: Define and embed SRE engagement models aligned to modern engineering and traditional ITSM/ITIL practices Establish SLIs, SLOs, and Error Budgets Shape observability strategies using metrics, logs, and traces Design incident response models and post-incident learning loops Reduce toil through automation and engineering excellence Deliver SRE capability … Looking For Extensive experience in SRE, cloud operations, or DevOps Proven consulting or advisory background Experience with AWS, Azure, or GCP Strong observability and incident management expertise Ability to obtain UK SC clearance Modis International Ltd acts as an employment agency for permanent recruitment and an employment business ...

Staff Security Engineer

Location
City Of London, England, United Kingdom
engineer who enjoys solving complex security challenges at scale. You’ll work across engineering, data, AI and digital workplace teams to build security observability, automate control assurance and influence how security is embedded into products, platforms and processes. If you're passionate about turning security data into actionable insight, building … raising the security maturity of a fast-moving technology organisation, we'd love to hear from you. About the role Designing and building security observability capabilities that provide meaningful visibility across systems, infrastructure and applications Developing automated control monitoring, evidence collection and continuous testing solutions that strengthen security governance Partnering ...

Senior Security Engineer: Observability, Automation & Risk

Location
City Of London, England, United Kingdom
seeking a Staff Security Engineer to advance security observability, automate controls and strengthen governance across its global tech estate. This senior individual contributor role partners with engineering, data, AI and digital workplace teams to raise security maturity in a fast-moving environment. You will design scalable security observability, automate assurance ...

Principal Data Engineer

Location
City Of London, England, United Kingdom
follow the identical pattern so they are handover-ready by design. Drive data quality as a first-class, firm-wide concern: establish data contracts, observability, SLA/SLO monitoring, and automated alerting and remediation across ingestion and transformation layers, and hold squads to those standards. Act as the senior technical … with the ability to set standards, conduct code and design reviews, and grow engineers’ capabilities Strong grasp of data quality practices: data contracts, pipeline observability, SLA/SLO definition, and automated alerting and remediation Solid understanding of SQL transformation patterns and modern tooling such as dbt, alongside experience managing ingestion ...

Senior Software Engineer

Location
City Of London, England, United Kingdom
Impact and Responsibilities We are seeking a Senior Software Engineer who thrives on untangling complex systems and modernising core infrastructure without breaking production. This is an exciting opportunity to modernise core C#/SQL systems ...

Director, AIOps & Observability Engineering

Location
City Of London, England, United Kingdom
J.P. Morgan in London and across locations seeks a Director of Software Engineering (AIOps) to lead a large Observability Platforms program. You will drive architectural decisions, mentor engineers, and deliver AI-powered self-healing and root-cause capabilities across the firm. As an Executive Director, you will set strategy ...

ML Platform Engineer: Production ML & Observability

Location
City of Westminster, England, United Kingdom
Engineer to build and operate platform capabilities that move ML models from experimentation to reliable production services. You will own automation, deployment, observability, and controls around the ML lifecycle, collaborating with research, software, platform and product teams. You will design repeatable workflows for training, validation, deployment and retraining, productionise models ...

Staff ML Engineer | Agentic AI & Applied ML | London (Hybrid) | Contract | Inside IR35

Location
City Of London, England, United Kingdom
implementation patterns Designing and evolving production RAG and retrieval architectures Establishing effective LangGraph/LangChain patterns for agentic applications Improving AI evaluation, testing, observability and production monitoring Developing guardrails, controls and approaches to hallucination and model risk Supporting the move towards increasingly high-risk and high-complexity AI/… based applications Retrieval Augmented Generation (RAG) LangChain and/or LangGraph Vector databases and retrieval MLOps and production deployment AI evaluation, testing and observability AI governance, model risk and engineering controls ML frameworks such as PyTorch, TensorFlow or Scikit-learn Experience operating in complex, regulated or high-risk environments would ...

Senior Software Development Engineer

Location
City Of London, England, United Kingdom
expectations and can be reused across multiple brands and platforms. Drive engineering excellence for the services you own by championing code quality, automated testing, observability, performance optimization, and simplification, taking technical responsibility for service health, scalability, resilience, and the ongoing reduction of technical debt and operational overhead. Provide technical mentorship … integrations, and communicating trade-offs to both technical and non-technical stakeholders. Track record of improving operational excellence at the team level through enhanced observability, automation, performance tuning, and data-driven analysis of incidents and customer impact. Hands-on experience integrating or consuming AI/ML-enabled services or platforms ...

Principal AI Engineer

Hiring Organisation
Intellias
Location
City of London, London, United Kingdom
Our client is a leading global investment management firm headquartered in London, managing over $228B in assets. Technology, data science, machine learning, and AI are at the heart of its investment and research ecosystem. The ...

Databricks Champion Architect

Location
City Of London, England, United Kingdom
We’re hiring a Databricks Champion Architect to define and govern our modern lakehouse architecture. You’ll blend solution design, hands‐on technical leadership, and platform enablement; setting patterns, accelerating delivery teams, and ensuring production ...

Principal Splunk Architect

Location
City Of London, England, United Kingdom
onsite gym facilities and medical centre. Role Description Are you a recognised Splunk expert with a passion for operating and continuously improving large-scale observability platforms in a complex global banking environment? We're seeking a Principal Splunk Engineer to lead the day-to-day operation, stability, performance, and evolution … enterprise monitoring and observability platforms. This role focuses on ensuring our Splunk and Cribl environments remain resilient, scalable, and effective in supporting mission-critical banking services. Working within the Operational Intelligence team, you'll act as a senior technical leader responsible for platform reliability, operational excellence, service improvement, and production ...

Cloud FinOps Analyst

Hiring Organisation
Manufacturing Recruitment Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60,000
across Azure and Snowflake environments. A key focus of the role is leading the FinOps optimisation activities, embedding governance frameworks, and overseeing AKS cost observability using tooling such as Power BI, Kubecost etc. The FinOps Analyst partners closely with Engineering, Data, Cloud Operations, and Finance teams to enable a cost … optimisation, and waste elimination. Develop, maintain, and enforce cloud and data platform cost governance frameworks including tagging, budgeting, guardrails, and accountability processes. Oversee cost observability tooling (Kubecost, Snowflake dashboards, cloud cost portals) to ensure visibility of usage, forecasts, and budget performance. Manage budgeting, forecasting, cost allocation, and financial reporting ...

Sr Director, Platform Engineering Data Platform & Agentic Platform

Hiring Organisation
Hackajob Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
operate agent workflow platform capabilities aligned to product-defined standards and interfaces, including traceability, state handling, and convergence patterns Implement production-grade evaluation, observability, auditability, and guardrail mechanisms required for safe AI workflows Implement security controls, access governance, encryption, and audit requirements in partnership with InfoSec while ensuring enterprise SDLC … large-scale SaaS systems with production operations accountability Demonstrated success building and operating platforms adopted by multiple product teams, including reliability discipline (SLOs), observability, and incident management Strong hands-on technical leadership background in distributed systems and platform engineering Deep experience with data platform engineering at scale, including ingestion ...

Technology Support Director, Risk Engagement

Location
City Of London, England, United Kingdom
cross-domain incident, problem and change management challenges, identifying process weaknesses, recurring failure patterns and opportunities to improve operational resilience across Infrastructure Platforms. Uses observability, monitoring and diagnostic information to analyse significant, recurring or cross‐domain issues, test root‐cause hypotheses and help accountable teams define sustainable remediation. Leads focused … analysing complex application or infrastructure issues in large‐scale on‐premises and public‐cloud environments, including dependencies across multiple technology domains. Proficient in using observability, monitoring and diagnostic information to identify patterns, test hypotheses and support root‐cause analysis across distributed technology environments. Experience working in a client‐facing technology ...