3,401 to 3,425 of 3,966 Observability Jobs

Operations Team Lead (Production & Reliability)

Location
Cambridge, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
Wollaton, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
Bath, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
High Wycombe, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
Norwich, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Application Support Analyst

Location
Bradford, England, United Kingdom
analysis capability. Experience using monitoring, logging and defect management tools. Useful, but not essential MuleSoft integration platforms and GoAnywhere MFT Sumo Logic or similar observability tooling Salesforce Apex log inspection and SOQL query analysis As Vanquis continues to strengthen and modernise its technology landscape, you'll have opportunities to deepen ...

Operations Team Lead (Production & Reliability)

Location
City of Edinburgh, Scotland, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Software Engineer, Trading – Cumberland Systematic

Location
City Of London, England, United Kingdom
trading operation with high availability requirements.You will be expected to design and develop trading systems, market data connectivity, execution algorithms, research infrastructure, monitoring and observability tooling, and integrations with DRW’s core services.The team’s existing systems are written in C++ and Python. Candidates should have strong initiative and proven ...

Operations Team Lead (Production & Reliability)

Location
Newcastle upon Tyne, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
Hull and East Yorkshire, England, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

Operations Team Lead (Production & Reliability)

Location
Belfast City District, Northern Ireland, United Kingdom
Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under pressure, clear in communication Systems thinker, fixes root causes, not symptoms How We Think Production is sacred. Clear ownership beats ambiguity. ...

AI Product Analyst (AI Metrics & Model Evaluation)

Hiring Organisation
The Portfolio Group
Location
City of London, London, Castle Baynard, United Kingdom
Employment Type
Permanent
Salary
£80000 - £85000/annum
metrics that matter, retrieval quality, correctness of output, and what users go on to do with what they're given Owning production quality observability, from thumbs-down and regeneration rates through to abandoned tasks and drift in how the product is being used Building the analysis and dashboards the team ...

Microsoft AI & Automation Platform Engineer

Hiring Organisation
Real Technical Solutions
Location
Brighton, East Sussex, United Kingdom
Employment Type
Full-Time
Salary
£67,500 - £72,500 per annum, Negotiable
with Copilot Studio or conversational AI Knowledge of Azure OpenAI/RAG solutions Familiarity with Dataverse, Graph, or enterprise systems Exposure to monitoring and observability tools Interest in emerging technology and innovation Familiarity of AI/Copilot capabilities and governance We are seeking an experienced Microsoft AI & Automation Platform Engineer ...

Developer Experience Engineer New London

Location
Greater London, England, United Kingdom
Design and build the tools customers use to run, optimise, and monitor large language models on Fractile hardware Build interfaces to diagnostic, profiling, and observability tools that help developers understand performance, debug issues, and optimise deployments Write documentation, quickstarts, reference examples, and tutorials that define a developer's first ...

Senior Backend Engineer

Location
Greater London, England, United Kingdom
across the UK and internationally as we expand. Build production AI systems - personalisation engines that act on longitudinal health data at scale. Improve reliability, observability, and the foundations that everything else depends on as the product grows. Work directly with product, design, and clinical teams. No tickets thrown over walls. ...

Senior Technical Product Marketing Manager, Agent Identity & Auth (EMEA)

Location
Greater London, England, United Kingdom
market motions is strongly preferred. Familiarity with the AI connectivity landscape -including awareness of competing and complementary vendors across API gateways, AI observability, vector databases, iPaaS, and agentic orchestration frameworks - is a strong plus. #LI-NS1 About Kong: Kong Inc., the AI Connectivity Company, is building the connectivity layer ...

Member of Technical Staff (EMEA)

Location
Greater London, England, United Kingdom
from ambiguous need to running in production. Build on Inference: Create capabilities on top of the serving stack: routing, model controls, observability, and whatever a deployment turns out to need. Build on Training: Ship tooling around fine-tuning and post-training of LLMs, from data pipelines to evaluation, working alongside ...

AI for Technology Operations Lead

Location
Greater London, England, United Kingdom
technology risk Establish AI governance, assurance, control and model‐monitoring frameworks that support safe adoption and audit scrutiny Bring familiarity with cloud platforms, observability tooling, automation frameworks and platform engineering Build and run multidisciplinary teams across engineering, architecture, data and operations, with a demonstrable commitment to inclusive talent development Govern ...

Principal AI/ML Engineer

Hiring Organisation
King Digital Entertainment
Location
London, UK
Employment Type
Full-time
through meaningful production coding, technical prototyping, and implementation work on critical pathsSet a high bar for production AI/ML engineering, including reliability, observability, maintainability, and quality of executionMentor senior engineers and raise the standard of technical judgment, design quality, and cross-team collaborationLead the shared architecture for AI/ ...

Senior Manager - Core Banking Architect, TC, FS

Hiring Organisation
EY UK
Location
Tower Hamlets, London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
microservices, containers, cloud). Stakeholder leadership with CIO/Chief Architect/COO teams; crisp written architecture artefacts. Fluency in NFR engineering (resilience, security, observability) and release governance in regulated environments. Ideally, you'll also have Experience across multiple core vendors and integration to channels, GL, risk and data platforms. ...

Cloud Business Architect

Location
Reading, England, United Kingdom
data residency for network intelligence platforms. Migration and coexistence strategies for legacy tools and platforms. Standard platform capabilities (data pipelines, AI/ML, automation, observability, security) Architectural principles and target architecture for MSN tools and platforms Optimizes MSN operating models for cloud-based delivery, ensuring cloud investments drive clear financial ...

Principal Software Developer (Web)

Location
City Of London, England, United Kingdom
compatible framework (likely React Native), including architecture, incremental migration strategy, and performance/UX parity Quality & Reliability: Champion code quality, testing, observability (metrics, logs, tracing), and performance across the stack Team Growth: Mentor engineers, assist with hiring, onboarding, and capability development as the team scales. Collaboration: Partner closely with ...

Senior AI Engineer, AI Lab

Hiring Organisation
167 Solutions Ltd
Location
London, United Kingdom
Employment Type
Permanent
Strong candidates might also have: Experience fine-tuning models for style or persona (e.g. chatbots with specific character voices) Exposure to LangSmith or similar observability tools for prompt/LLM testing Knowledge of voice synthesis, cloning, or emotion conditioning in audio pipelines Previous work in media, journalism, podcasting, or content ...

PRINCIPAL / STAFF BACKEND ENGINEER - DISTRIBUTED SYSTEMS

Hiring Organisation
Widenet Consulting
Location
Seattle, Washington, United States
Employment Type
Permanent
Salary
USD 200,000 Annual
without a dedicated SRE or platform team to hand things off to. Engineers in these roles own their systems fully: architecture decisions, implementation, deployment, observability, and on-call. If you prefer writing specs over writing code, or handing finished designs to an ops team, this is not the right fit. ...

Director Software Engineering

Location
Greater London, England, United Kingdom
resource allocation.* Drive planning, prioritisation and execution across multiple engineering teams and programmes.* Maintain strong engineering standards for code quality, testing, release management, observability, security and operational resilience.* Establish clear operating models for technical debt, production incidents, operational risk and partner delivery.* Lead platform modernisation, consolidation and adoption of shared ...