1,151 to 1,175 of 1,973 Observability Jobs

Software Engineer - AI Platform & Agents

Hiring Organisation
Moody's
Location
Greater London, United Kingdom
Employment Type
Full Time
distributed systems, and event-driven architectures in modern programming languages such as Python, TypeScript, Java, Go, or similar Familiarity with MLOps practices, model monitoring, observability, versioning, and automated deployment pipelines preferred Strong problem-solving skills with the ability to navigate ambiguity, experiment rapidly, and deliver impactful solutions that create measurable … business challenges autonomously Optimize applications for performance, scalability, reliability, and cost efficiency while supporting increasing AI adoption and usage Implement engineering best practices around observability, monitoring, testing, security, and operational excellence Establish and champion MLOps practices including model lifecycle management, prompt versioning, automated evaluation, and continuous improvement Build reusable frameworks ...

Sr. Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
while architecting systems that handle high-throughput data pipelines. Participate in building robust export pipelines, streaming architectures, webhook integrations and MCP servers. Maintain high observability and reliability standards using tools like Coralogix, CloudWatch, and Grafana. Participate in on-call rotation and incident response for owned services. What You'll Bring … static site generators). Familiarity with authentication, API gateways, and rate limiting strategies. Experience in compliance standards for APIs and data handling. Experience with observability tools and practices. Languages: Golang (primary) with some TypeScript Monitoring: Coralogix, Grafana, CloudWatch CI/CD & IaC: GitHub Actions, Terraform What We Offer Generous paid ...

Staff Software Engineer - AI

Hiring Organisation
Moody's
Location
Greater London, United Kingdom
Employment Type
Full Time
technologies in production environments • Strong experience designing and implementing application programming interfaces, distributed systems, event-driven architectures, data pipelines, PostgreSQL, MongoDB, Redis, vector databases, observability, and automated deployment pipelines • Demonstrated ability to influence technical direction while remaining close to the codebase, mentoring engineers through design reviews, code reviews, pairing, debugging … maintainability, system performance, reliability, security, scalability, and cost efficiency • Establish engineering best practices through hands-on contribution, code reviews, technical design reviews, automated testing, observability, monitoring, and operational excellence • Champion machine learning operations practices including model lifecycle management, prompt versioning, automated evaluation, deployment pipelines, monitoring, and continuous improvement • Partner with ...

AI Engineer

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
context control, and guardrails. Develop retrieval‐augmented workflows to enhance context, reliability, and performance. Perform quality assurance on AI outputs by implementing robust AI observability practices, including monitoring model behaviour, detecting anomalies, and ensuring visibility into AI performance and reliability. Contribute to ongoing research and development, staying current with emerging … chosen when they are safer, simpler, or more cost effective. Ensure AI-enabled solutions consider full total cost of ownership, including token consumption, performance, observability, and ongoing maintenance, with awareness of cost‐efficiency and model‐selection trade‐ Knowledge Sharing, Mentoring and Governance: Mentor and support both technical and non‐technical ...

AI Engineering Enablement Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
/ML, software, or platform engineering, with exposure to automated testing and infrastructure‐as‐code or policy‐as‐code.* Working knowledge of AI observability (logs, metrics, traces, behavioural signals) and practical methods to evaluate or improve AI system behaviour.* Familiarity with AI risk and governance frameworks (e.g., NIST … FinOps, such as cost‐aware model selection, unit economics, or prompt‐efficiency practices.* Experience with MLOps or AI delivery tooling, or with AI‐specific observability systems.* Participation in industry communities or standards bodies, with the ability to translate external practice into internal adoption.* Experience facilitating workshops or engineering enablement events. ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
quickly while remaining safe, resilient and compliant.* Partnering closely with feature teams, providing hands‐on support, coaching and mentoring on CI/CD, automation, observability and DevOps best practices, and supporting the growth of junior engineers.* Continuously improving platform standards and engineering quality, through design reviews, code reviews, knowledge sharing … approach.* Working knowledge of **infrastructure‐as‐code** (e.g. Terraform) and how it supports automated delivery, rather than being the primary focus.* Experience with **observability and monitoring** tooling such as Dynatrace, Prometheus, Splunk or similar.**And any experience of these would be really useful*** Application development and testing ecosystems (e.g. Java ...

Engineering Manager, DevOps

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
/CD pipelines, and deployment practices across GCP (primary), AWS, and Azure Set and enforce engineering standards for infrastructure‐as‐code, GitOps, DevSecOps, and observability across the team and the wider engineering organisation Lead the design and improvement of containerised deployment workflows using Docker, Kubernetes, and Helm … Azure Key Vault), and security integration into the delivery pipeline as a first‐class concern Identify and address tooling gaps across monitoring, alerting, observability, and incident response; own the on‐call process, runbooks, escalation paths, and post‐incident reviews Directly manage 4/5 DevOps engineers: run consistent ...

Solution Architect (Contract Role)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
## Solution Architect (Contract Role)Applylocations: Londontime type: Full timeposted on: Posted Todayjob requisition id: JR216**Solution Architect -** **(6 Month contract Role)**(Inside IR35, £800 - £950 per day)Willis Re is building its global technology ...

Solution Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Solution Architect Willis Re is building its global technology estate from the ground up, unencumbered by legacy and designed around data, analytics and modern cloud platforms. We are looking for a Solution Architect to shape ...

Solution Architect (Contract Role)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Solution Architect – 6-month Contract Willis Re is building its global technology estate from the ground up, centered on data, analytics, and modern cloud platforms. We are looking for a Solution Architect to shape and ...

Python Technical Lead FinTech

Hiring Organisation
Run-Time Group Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Python Technical Lead Salary: £90-120K work model: Hybrid Were looking for a Python Technical Lead to drive the architecture, development, and delivery of high-performance financial systems. Youll lead a team of engineers ...

Developer Enablement, Technical Architect – Release on Demand (SVP)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
graceful degradation, and zero-downtime deployments. Define SLOs, own the SRE practice for the platform, and be accountable when things need fixing. Define the Observability Strategy. Establish a comprehensive observability framework — distributed tracing, structured logging, metrics, dashboards, and alerting. The platform must be fully understood at all times. You will … NoSQL databases: PostgreSQL, MongoDB or Couchbase Demonstrated SRE or platform engineering experience — SLOs, incident management, reliability engineering at scale Experience defining and implementing observability strategies: distributed tracing, structured logging, metrics and alerting Proven experience leading technical projects and mentoring engineers Highly Desirable skills Experience with SDLC tooling, release management platforms ...

Network Analytics & Automation Leader with AI Platforms

Hiring Organisation
Jobleads-UK
Location
Chester, England, United Kingdom
Overview Automation Technologies and AI/ML-Driven Platforms and Analytics Tools; in the realm of automation technologies and AI/ML-driven observability platforms and analytics tools, the following are essential: Terraform Itential NetDevOps Splunk Python React JS Django Database Technologies Proficiency with database technologies is crucial, including: MySQL ...

Senior Director, Data & AI Platform Engineering

Hiring Organisation
Jobleads-UK
Location
United Kingdom
workflows. You will partner with Platform Product Management to ensure scalability, reliability, security, and broad adoption across product domains. You will drive production-grade observability, governance, and AI safeguards while aligning with #J-18808-Ljbffr ...

Senior Backend Engineer, Pricing Platform (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will design and build pricing features end-to-end, partnering with product teams to ship impactfull updates. You will contribute to architecture, improve reliability, observability, and scale across distributed services in a fintech environment, while mentoring teammates and owning the roadmap for pricing infrastructure. #J-18808-Ljbffr ...

Lead Backend Engineer: Scale Microservices & AI-Driven

Hiring Organisation
Jobleads-UK
Location
United Kingdom
lead engineers, drive Python microservices, and influence architecture for a scalable insurance platform. As a technical leader, you’ll mentor peers, champion reliability and observability, and collaborate with product, design and data teams to deliver impact at scale. #J-18808-Ljbffr ...

Remote ML Platform Engineer - Scale AI Infrastructure

Hiring Organisation
Jobleads-UK
Location
United Kingdom
operate complex models, merging software engineering, cloud infrastructure, and ML to advance platform capabilities. Collaborate with researchers and product engineers, focusing on automation, observability, and developer experience. #J-18808-Ljbffr ...

Senior SRE: Lead Reliability for Scalable Platforms

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
will own SLO/SLI definitions, incident leadership, and a roadmap for platform maturity, working with ECS/EKS, Terraform-based automation, and observability tooling. #J-18808-Ljbffr ...

Staff Software Engineer, Payments Platform & Cloud

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
Software Engineer - Payments Platform will lead the design and delivery of scalable, secure payment capabilities, mentor engineers and drive DevOps, CI/CD and observability across high-volume environments. #J-18808-Ljbffr ...

Platform Engineer: Analytics Platform & CI/CD

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
high-performance, distributed computing products at scale, collaborating with software engineers, data scientists and product owners. You will own infrastructure automation, CI/CD, observability and security in a regulated environment, delivering reliable, auditable platform services across the organisation. #J-18808-Ljbffr ...

Remote Senior Full-Stack Engineer — AI‐Powered Platform

Hiring Organisation
Jobleads-UK
Location
United Kingdom
development in a fully remote environment. You will build scalable services in Java/Node.js, craft frontend with React and TypeScript, and drive reliability, observability, and CI/CD improvements. #J-18808-Ljbffr ...

Senior Backend Engineer — Scalable AI Media Platforms

Hiring Organisation
Jobleads-UK
Location
United Kingdom
work with DevOps on platform infrastructure. Occasional frontend touchpoints help expose backend capabilities. Responsibilities include API and system design, data processing, model integration, observability, and performance optimization to ensure low latency and high #J-18808-Ljbffr ...

Senior Software Engineer - Full-Stack & Distributed Systems

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
This role spans backend, APIs, and data workflows, with cross‐team collaboration. You will design reliable services, work with cloud platforms, implement tests and observability, and take responsibility for incidents and production quality. #J-18808-Ljbffr ...

Senior Enterprise Data Integration Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
workloads in a regulated environment. You will drive API-led and event-driven integration, CI/CD and IaC practices, while ensuring data quality, observability and governance in a fast-growing global financial services group. #J-18808-Ljbffr ...

Senior AI/ML Solutions Architect (GenAI & MLOps)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Field Engineering team to design production‐grade AI solutions on the Databricks platform. You will drive GenAI initiatives, RAG architectures, agentic systems, AI observability, and NLQ of structured data, while mentoring peers and influencing the platform roadmap. Some travel may be required. #J-18808-Ljbffr ...