1,426 to 1,450 of 1,896 Observability Jobs

Senior Software Engineer

Hiring Organisation
Fruition Group
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£90,000
multiple products Integrate external identity, KYC and payments platforms (e.g. Senzing, Auth0, Stripe) Own services end-to-end: API design, MongoDB data modelling, testing, observability, deployment Build secure systems (JWT/OIDC, fine-grained authorisation, IDOR protection, audit logging) Write automated tests (Vitest, Playwright) as part of everyday development Mentor …/AML, fintech or other regulated-industry experience Payment provider integration (e.g. Stripe) Monorepo tooling (pnpm, TurboRepo) and CI/CD (GitHub Actions) Observability tooling (Prometheus, Sentry) We are an equal opportunities employer and welcome applications from all suitably qualified persons regardless of their race, sex, disability, religion/belief ...

DevSecOps Engineering Lead CGEMJP00346044

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
management, including allowlist processes and risk acceptance where required Secrets management and identity/access management Policy enforcement for workloads, container images and infrastructure Observability, monitoring, logging and audit controls Partner with developers to embed secure-by-design engineering and ensure compliance with CLIENT security standards. Enable and govern Infrastructure … compliance tooling (e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager) Containers and orchestration (e.g. Docker, AWS EKS) Infrastructure as Code (e.g. Terraform) Observability (e.g. Grafana, Loki) Scripting and automation (e.g. Python, Bash) Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies) Experience delivering within the UK Government ...

Staff / Senior Software DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
baselines, sound metric aggregation, determinism), and turn CI and test signal into CI‐health and product‐readiness dashboards that drive real decisions. Own Software Observability: Choose the metrics store that scales to many series on daily runs with long‐lived history, making dashboards for observable software. Set Standards: Define … Performance‐analysis support: trustworthy regression baselines, determinism and noise handling, sound metric aggregation (geometric vs arithmetic mean vs median), and fast attribution and bisection. Observability and metrics platforms: CI‐health and readiness dashboards, a metrics store that scales to long‐lived, high‐cardinality time series, and self‐serve access ...

Software Engineer III - Python

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
infrastructure‐as‐code using Terraform within established team patterns across modules, environments, and state management Improve operability of services by adding and using observability tooling including logs, metrics, traces, dashboards, and alerts, and participate in incident response and root‐cause analysis Leverage enterprise‐authorized AI coding assist tools within … implementing application logic and APIs on top of relational data Experience building APIs and microservices using REST or gRPC, including contracts, security basics, and observability Practical experience delivering LLM‐based features as part of software systems, with familiarity with agentic patterns Working knowledge of delivery and operations including CI/ ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
development, validation, and optimization of configuration-as-code, improving delivery speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality … applications without these: Hands-on with Helm or Kustomize Experience with GitOps (e.g., Argo CD) Knowledge of secrets management (e.g., HashiCorp Vault) Experience with observability (metrics/logs/tracing) Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations ...

Staff Software Engineer UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Operations to influence technical roadmaps and delivery. Mentor senior engineers and strengthen technical leadership across the organisation. Lead initiatives that improve engineering quality, testing, observability, and operational excellence. Contribute hands-on to critical systems where your expertise delivers the greatest impact. What it takes to succeed: Extensive experience designing … Experience with microservices, APIs, event-driven architectures, relational databases, and cloud platforms (preferably AWS). Knowledge of CI/CD, Infrastructure as Code, containerisation, observability, and automated testing. Strong communication skills with the ability to influence technical decisions and mentor experienced engineers. Professional fluency in English. Java experience is highly ...

AI Native Software Engineer

Hiring Organisation
BC Forward
Location
Dallas, Texas, United States
Employment Type
Permanent
Salary
USD 8,888 Annual
agents including retrieval (RAG), orchestration workflows, tool/function invocation, and policy-based routing. Build evaluation frameworks for accuracy, latency, and reliability, with observability and monitoring for the agent lifecycle. Integrate with AI providers such as OpenAI, Anthropic, Google Vertex, and open-source models, and build abstraction layers for multi … agents, RAG, and orchestration. Proficiency in Python, Java, or similar backend languages. Experience with CI/CD pipelines, infrastructure as code, and monitoring and observability tools. Hands-on experience with AI platforms such as OpenAI, Claude, or Vertex AI. Preferred Skills: Experience with agent frameworks such as LangGraph, AutoGen ...

Europe | Hybrid Staff Software Engineer UK

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Operations to influence technical roadmaps and delivery. Mentor senior engineers and strengthen technical leadership across the organisation. Lead initiatives that improve engineering quality, testing, observability, and operational excellence. Contribute hands‐on to critical systems where your expertise delivers the greatest impact. What it takes to succeed: Extensive experience designing … Experience with microservices, APIs, event‐driven architectures, relational databases, and cloud platforms (preferably AWS). Knowledge of CI/CD, Infrastructure as Code, containerisation, observability, and automated testing. Strong communication skills with the ability to influence technical decisions and mentor experienced engineers. Professional fluency in English. Java experience is highly ...

Core AI Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
running centralised MCP servers, ensuring secure, reliable access to enterprise tools and data Owning platform reliability, performance and scalability across Kubernetes‐based infrastructure, including observability, capacity planning and incident response Building self‐service tooling and APIs to enable teams to provision and consume AI infrastructure independently Integrating platform services with … services Familiarity with sandboxing and workload isolation technologies Experience in quantitative finance or low‐latency systems AWS experience particularly in hybrid environments Experience with observability tooling such as Prometheus, Grafana or OpenTelemetry Contributions to open‐source projects in relevant domains Why join us? Highly competitive compensation plus annual discretionary bonus ...

Site Reliability Engineer

Hiring Organisation
CGI
Location
Greater London, United Kingdom
Employment Type
Full Time
Site Reliability Engineer (SRE) to join a team supporting multiple data product and platform groups. This role is focused on improving the reliability, scalability, observability, and operational performance of critical data-driven platforms and services across complex production environments. The successful candidate will work closely with engineering, platform, and support … across cloud and containerised environments. - Manage and support Kubernetes clusters and Helm-based deployments across multiple environments. - Implement and enhance monitoring, alerting, logging, and observability solutions to improve platform reliability and operational visibility. - Investigate incidents, analyse logs, identify root causes, and drive timely resolution of production issues. - Participate in incident ...

VP of Software Engineering – Full-Stack & AI

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
engineering teams; set clear objectives, coach talent, and foster succession planning. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure; secure-by-design with threat … scalable, observable, testable systems; strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud. Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows. Excellent ...

Senior Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Build comprehensive automated testing and continuously improve test coverage. Optimize application performance, responsiveness and Core Web Vitals. Own frontend services in production, including monitoring, observability and continuous improvement. Build and optimize CI/CD pipelines for frontend applications. Support incident response, root cause analysis and ongoing service improvements. Champion accessibility … Desirable Skills React Native Next.js Google Cloud Platform (GCP) Kubernetes Design Systems Storybook GraphQL Micro Frontends Performance optimisation and Core Web Vitals Accessibility (WCAG) Observability platforms and frontend monitoring Everyone is Welcome At Nando’s, everyone is welcome. Inspired by our Southern African heritage, we know and value the richness ...

Context Plane Python Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
data sources and services across the firm, including enterprise AI and large language model gateways Own quality across your components: automated testing, code reviews, observability, and resilient, secure service design Partner with Corporate Technology AI, product, and data science colleagues to translate concrete use cases into working, measurable capabilities Contribute … working with cloud infrastructure (AWS) and containerized services (Docker/ECS) Ability to own technical components end-to-end - from design through deployment and observability Strong collaboration skills with the ability to work across engineering, product, and data science disciplines Hands‐on experience using enterprise-authorized AI‐assisted software development ...

Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
undergoing a multi‐year convergence and modernization journey. You will play a pivotal role in shaping our next‐generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives in fast‐paced front‐office environments, enjoys … Deep knowledge of reliability engineering principles: SLIs/SLOs, real-time telemetry, disaster recovery planning, capacity planning, and performance tuning. Experience designing and implementing observability frameworks for mission critical systems. Proven ability to lead incident response and drive long term remediation. Solid programming skills in Python, Java, or Kotlin, with ...

Lead DevSecOps Engineer

Hiring Organisation
Capgemini
Location
City of Bristol, United Kingdom
Employment Type
Full Time
including allowlist processes and risk acceptance where required and secrets management and identity/access management;Policy enforcement for workloads, container images and infrastructure •Observability, monitoring, logging and audit controls; Also tracking and remediation of technical debt and improvig the quality of developed code •Control and governance of change across … compliance tooling (e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager) •Containers and orchestration (e.g. Docker, AWS EKS) ; Infrastructure as Code (e.g. Terraform) •Observability (e.g. Grafana, Loki) ;Scripting and automation (e.g. Python, Bash) •Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies) •Experience delivering within the UK Government ...

Staff / Senior Staff Engineer, AI Agent Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Strongproficiencyin Python or TypeScript, plus solid API, microservices, and event-driven architecture skills. Fluency with modern engineering practice: Git, automated testing, CI/CD, observability, and cloud platforms. Sound judgment about when to trust automation and when to demand human review, and the communication skills to explain that reasoning. Preferred … context engineering; tool use and orchestration; multi-agent design. Evals and Trust Eval design and automation; guardrails and human-in-the-loop gates; AI observability; responsible AI governance. Platform Craft API and event-drivendesign;CI/CD andautomation;cloud-nativeengineering;enterprise integration. Judgment and Impact Systems thinking; pragmatic risk-taking ...

Senior Software Engineer II - Data Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ensure technical consistency.* Design, develop, and maintain generative AI services and reusable components using Python.* Define and promote best practices in engineering, including scalability, observability, testing, and CI/CD.* Contribute to system designs spanning multiple services and modules, aligning with architectural best practices.* Collaborate with product, platform, and research … work collaboratively across functions in an Agile or Kanban environment.**Nice to have:*** Experience operationalizing LLMs or building an internal AI platform.* Familiarity with observability practices (metrics, logging, alerts).* Exposure to knowledge graphs or semantic search systems.Join our team and contribute to a culture of innovation, collaboration, and excellence. ...

Platform Engineer

Hiring Organisation
Hireful
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£80,000
We are recruiting founding Platform Engineers on behalf of a fast-growing enterprise level (global, 500+ staff) software business with a strong engineering culture and a genuine commitment to doing things the right way. They ...

AI & ML Engineer — Production-Ready AI on GCP

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About Charlotte Tilbury Beauty Founded by British makeup artist and beauty entrepreneur Charlotte Tilbury MBE in 2013, Charlotte Tilbury Beauty has revolutionised the face of the global beauty industry by de-coding makeup applications for ...

Lead Site Reliability Engineer – Operations Excellence

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
Join a team where your engineering expertise directly protects the reliability and resilience of systems that matter at scale. At JPMorganChase, we invest in engineers who think beyond the code — who own outcomes, drive operational ...

Technical Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Job Description - Technical Architect Company Details – BritBox International BritBox Office Hub – London Your Location - London Reporting To – Principle Architect Direct Reports – NA Contract Type: Permanent/full time About Us Welcome to BritBox, the go ...

Lead Software Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Our Team: How We Enrich Everyday Life You’ll be joining Bauer Media Audio (BMA), Europe’s biggest commercial audio broadcaster, connecting audiences across nine markets to the music, stories and experiences they love. Bauer ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
POSITION SNAPSHOTRole: Platform EngineerLocation: ManchesterContract: PermanentWhy Us And Why This RoleSeize the opportunity to be part of a unique business with an exciting vision for the future - where we dare to be bold, discover what ...

C++ Observability Engineer

Hiring Organisation
Hays
Location
London, United Kingdom
Employment Type
Contract
Senior C++ Observability EngineerLondon | £850/day Inside IR35 | 6-Month ContractBanking Client | 1 day per month in the office A leading banking organisation is seeking a Senior C++ Observability Engineer to drive observability across critical real-time market data and trading platforms. This is a hands-on technical leadership … insight across low-latency systems.Essential Requirements Strong C++ engineering background Experience in financial services, trading, market data, or low-latency environments Deep expertise in observability engineering (metrics, tracing, logging, telemetry, profiling) Experience with OpenTelemetry, eBPF, and modern monitoring platforms Strong knowledge of Kubernetes and distributed systems Experience driving observability standards ...

Observability Engineer

Hiring Organisation
Capgemini Government Solutions LLC
Location
Fort George G Meade, Maryland, United States
Employment Type
Permanent
Salary
USD 150,000 Annual
Capgemini Government Solutions (CGS) LLC is seeking an Observability Engineer to design, implement, and support enterprise monitoring and observability solutions for mission-critical systems. This role will focus on improving system visibility, reliability, and performance through logging, telemetry, alerting, and dashboard development. The ideal candidate has experience supporting large-scale … working closely with DevOps, SRE, engineering, and cybersecurity teams to ensure systems are observable, resilient, and operationally efficient. Key Responsibilities Design, implement, and maintain observability and monitoring solutions. Manage log collection, ingestion, correlation, and analysis across multiple data sources. Develop dashboards, alerts, and reporting to monitor system health, performance ...