351 to 375 of 4,013 Observability Jobs

Platform Engineer

Location
Greater London, England, United Kingdom
container orchestration platforms; Experience in using a scripting language (such as Bash or Python), for CI/CD and general problem‐solving; Knowledge of observability concepts and ability to utilise monitoring tools (logs, metrics, dashboards, alerts – we use Grafana) to identify problems and find solutions. Nice‐to‐haves Hands ...

Principal Engineer (Gen AI and MACH Architecture)

Location
Greater London, England, United Kingdom
existing models for production use cases; including data preparation, evaluation, versioning and deployment within enterprise governance, cost and reliability constraints. AI cost‐value analysis, observability, governance, testing and evaluation frameworks for production systems. Practical application of Generative AI to marketing and experience challenges; e.g. personalisation, content generation, campaign optimisation; with ...

Senior ETF Engineer, Investment Technology

Hiring Organisation
Invesco
Location
London, UK
Employment Type
Full-time
Step Functions, Aurora).Design and operate CI/CD pipelines and deployment processes to ensure reliable and repeatable releases. Ensure data quality, auditability, and observability through logging, monitoring, lineage tracking, and validation frameworksCollaborate closely with investment teams to translate portfolio construction, risk, and analytics requirements into scalable technical solutions. Continuously ...

Senior Engineering Manager Software engineering London

Location
Greater London, England, United Kingdom
production. Data & platform: SQL, plus experience designing batch and stream data-processing systems. Cloud & tooling: AWS and/or GCP, Docker, Terraform, and observability tooling such as Datadog. Knowledge of Kubernetes, Kafka, and CI/CD pipelines is highly beneficial. Additional Information Bring all of you to work We create ...

Cloud Platform Product Manager - UK Security Clearance eligibility required

Location
Greater London, England, United Kingdom
Support the definition and rollout of outcome focussed SOWs, ensuring clear lines of responsibility between different platform focussed product teams. (eg CI/CD, Observability etc) Ensure delivery aligns with DevSecOps best practices (automation, IaC, continuous assurance). Governance & Reporting Define and monitor KPIs/OKRs for product success, including ...

Principal Engineer (Gen AI and MACH Architecture)

Hiring Organisation
Akqa
Location
London, United Kingdom
Salary
£ 70 K
tuning existing models for production use cases; including data preparation, evaluation, versioning and deployment within enterprise governance, cost and reliability constraints.AI cost-value analysis, observability, governance, testing and evaluation frameworks for production systems.Practical application of Generative AI to marketing and experience challenges; e.g. personalisation, content generation, campaign optimisation; with measurable ...

Principal Engineer (Gen AI and MACH Architecture)

Hiring Organisation
Akqa
Location
London, UK
Employment Type
Full-time
existing models for production use cases; including data preparation, evaluation, versioning and deployment within enterprise governance, cost and reliability constraints. AI cost-value analysis, observability, governance, testing and evaluation frameworks for production systems. Practical application of Generative AI to marketing and experience challenges; e.g. personalisation, content generation, campaign optimisation; with ...

Senior Architect, ADC/Quant

Location
Greater London, England, United Kingdom
compliance requirements typical of the London financial services sector.* Resilience & SRE: Advanced knowledge of building fault-tolerant architectures leveraging modern SRE principles, robust observability, and cloud-native resilience patterns.* Matrix Leadership: Exceptional stakeholder and client communication skills, with a track record of influencing cross-functional engineering pods, product managers ...

Lead Solution Architect

Hiring Organisation
BP
Location
London, United Kingdom
Salary
£ 70 K
intelligent scheduling, anomaly detection, automated P&L attribution, predictive maintenance and agentic workflow orchestration (e.g., AWS AgentCore).Define and govern DevOps, platform engineering and observability standards, including CI/CD pipelines, infrastructure-as-code, containerisation (Docker, Kubernetes), monitoring, alerting and incident response architecture.People, Community & GovernanceMentor and develop the architecture community ...

Lead Solution Architect

Hiring Organisation
BP
Location
London, UK
Employment Type
Full-time
intelligent scheduling, anomaly detection, automated P&L attribution, predictive maintenance and agentic workflow orchestration (e.g., AWS AgentCore).Define and govern DevOps, platform engineering and observability standards, including CI/CD pipelines, infrastructure-as-code, containerisation (Docker, Kubernetes), monitoring, alerting and incident response architecture. People, Community & GovernanceMentor and develop the architecture ...

Senior Software Engineer Ref. 3839

Location
Cheltenham, England, United Kingdom
particularly interested in candidates with experience in technologies and practices such as C++, Golang, Python, Rust, CMake, HELM, YAML, JSON, containerisation, VCPkg, observability, DevOps, Kubernetes, systems administration and Linux. However, we also welcome applications from candidates whose experience aligns with any of the technologies and disciplines outlined above. ...

Senior Software Engineer Ref. 3839

Location
Manchester, England, United Kingdom
particularly interested in candidates with experience in technologies and practices such as C++, Golang, Python, Rust, CMake, HELM, YAML, JSON, containerisation, VCPkg, observability, DevOps, Kubernetes, systems administration and Linux. However, we also welcome applications from candidates whose experience aligns with any of the technologies and disciplines outlined above. ...

Software Engineer — Observability Instrumentation

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
high-impact research - designing systems that scale, accelerate discovery and support innovation across the firm. Take the next step in your career. The roleThe Observability Engineering Team manages access to G-Research's telemetry platforms, ensuring our engineering teams can effectively produce and consume telemetry for their services. … looking for a technically strong, customer-focused Software Engineer to help make observability easier to adopt across the organisation. This role focuses on the producer side: instrumentation patterns, OpenTelemetry SDKs and the collector configurations that help teams emit consistent, high-quality telemetry. This role is suited to someone who enjoys ...

Platform Engineer – Monitoring, Observability & SIEM (MONSO)

Location
Greater London, England, United Kingdom
Platform Engineer – Monitoring, Observability & SIEM (MONSO) For our SPEAR Technology (Security, Platform Engineering, Automation and Runtime) division in London we are looking to hire a: Platform Engineer – Monitoring, Observability & SIEM (MONSO) Like solving puzzles with an inquisitive mind? Think outside the box and challenge the status quo? Prefer simplicity over … proactive ownership? Then consider joining Berenberg’s SPEAR Technology programme. SPEAR consists of our CyberSecurity team and several platform engineering teams responsible for Monitoring, Observability, Kubernetes, Developer Platform, Network, and Datacentre Infrastructure. Due to each team’s compact size, all team members are subject matter experts offering an excellent environment ...

Java Full stack Developer

Hiring Organisation
NEEV LIMITED
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Contract
Contract Rate
£400 per day
enterprise applications using modern Java technologies, cloud-native architectures, and microservices. The ideal candidate will have hands-on experience with Kafka, Kubernetes, API Security, Observability tools, SQL databases, and Spring-based microservices development . Key Responsibilities Design, develop, and maintain scalable Java-based applications using Java 17+, Spring Boot … mechanisms. Build event-driven solutions using Apache Kafka for real-time data processing and messaging. Deploy, manage, and troubleshoot applications on Kubernetes environments. Implement observability solutions using tools such as Splunk, ELK, Grafana, Prometheus, Dynatrace, or AppDynamics. Optimize application performance, scalability, and reliability. Work closely with business stakeholders, architects ...

Platform Engineer

Location
Greater London, England, United Kingdom
efficiently and securely. Working closely with software developers, architects, and delivery teams, you will help establish best practices around cloud infrastructure, CI/CD, observability, security, and application reliability. The role combines hands-on engineering with the opportunity to influence platform standards and development practices across multiple projects. Key Responsibilities … with modern front-end technologies, including React and Vite. Develop and improve CI/CD pipelines, automation, and developer tooling. Implement monitoring, logging, and observability solutions to improve system reliability and performance. Manage and optimise cloud infrastructure, containerised environments, and platform configurations. Ensure security and operational best practices are embedded ...

AWS DevOps Engineer

Location
Greater London, England, United Kingdom
Amazon EKS, automating infrastructure with Terraform, orchestrating containers via EKS, building robust CI/CD with GitOps, and implementing strong monitoring/observability + AWS EKS IAM concepts to support secure, high-performance, cost-optimised systems. You will work closely with developers and SREs in an automation-first, collaborative culture. … Support live production AWS/EKS environments, including on-call rotation, incident response, root cause analysis, and blameless post-mortems. Automate provisioning, configuration, scaling, observability, and secure workload access across AWS services. Collaborate on improving system reliability, observability, security (DevSecOps/EKS IAM), and AWS cost optimization. Qualifications: Strong Linux ...

Senior DevOps Engineer - AVP

Location
Belfast City District, Northern Ireland, United Kingdom
operations on a global scale. In this role, you will apply deep technical expertise across CI/CD pipelines, container orchestration, cloud infrastructure, and observability to deliver resilient, high-quality software systems. Your work will directly shape how Citi's engineering teams build, ship, and monitor production services across … engineering teams. Build and manage containerized workloads on Kubernetes and OpenShift using Helm, ensuring systems are scalable, reliable, and production ready. Architect and maintain observability solutions — including log aggregation with Splunk and Elastic/Kibana, and metrics monitoring with Prometheus and Grafana — to give engineering teams real-time visibility into ...

Director of Platform Engineering

Location
Greater London, England, United Kingdom
ITRS, we make society’s critical technology work. Our mission is to deliver automated and holistic IT observability solutions that safeguard critical applications and enable innovation. We are the only monitoring and observability platform designed for the most demanding and regulated industries — trusted by 90% of Tier 1 capital markets … this team is closely aligned with platform engineering practices, customer delivery priorities and operational standards. The ITRS engineering teams are building a next-generation observability platform with the capability to collect, store and analyse the vast amount of data generated by banks and financial institutions. Requirements As Director of Platform ...

Platform Engineer Graduate Considered

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£55,000
hands-on role with Kubernetes at its core, offering the opportunity to work across container orchestration, CI/CD, Infrastructure as Code, observability and cloud-native technologies. Location: Cambridge, 2 days per week in the office Salary: £40,000 - £60,000 per annum + benefits Requirements for Platform Engineer Graduate … efficiently at scale Work closely with software engineering teams to understand infrastructure requirements and encourage scalable, cloud-native development practices Improve the reliability, observability and performance of Kubernetes clusters Build and maintain CI/CD pipelines to support efficient software delivery Develop and improve Infrastructure as Code using technologies such ...

Member of technical staff (Infrastructure) - Paris

Location
Greater London, England, United Kingdom
Company’s agent platform including client-facing APIs and agent runtimes within various deployment scenarios (multi-tenant and on-prem). Setup and maintain observability and monitoring strategies. Requirements: MUST HAVE Observability and monitoring (Datadog, Prometheus, Grafana, ...) Good knowledge of a modern programming language (ideally Python or JS/ ...

Sr Lead AI Platform Engineer

Location
Auchentibber, Scotland, United Kingdom
Owns the design and build of the team's platform: deployment pipelines, model serving, containerisation, orchestration, and environment management Sets the standard for reliability, observability, and operational excellence across the team's production AI/ML services Builds the tooling and paved paths that let AI engineers ship agentic … record of building deployment and release automation Experience serving, scaling, and monitoring ML models or data-intensive services in production (MLOps) Practical experience with observability tooling (metrics, logging, tracing) and production incident response Experience operating ML/LLM workloads in production (LLMOps, inference reliability, cost/performance management) Strong communication ...

Senior ML Ops Engineer

Hiring Organisation
Harnham - Data & Analytics Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
production-grade orchestration pipelines for model training, inference, monitoring, and retraining. Developing CI/CD pipelines to automate testing, deployment, and operational processes. Implementing observability, monitoring, and alerting frameworks to ensure platform reliability and performance. Driving best practice around Responsible AI, governance, transparency, and auditability. Collaborating with engineering, delivery … practices using Azure DevOps or similar tooling. Experience with infrastructure-as-code technologies such as Terraform, Bicep, or Pulumi. Knowledge of monitoring and observability tooling such as Prometheus and Grafana. Experience with orchestration platforms including Dagster, Airflow, Prefect, or similar. BENEFITS The successful Senior MLOps Engineer will receive the following ...

Senior Platform Engineer (Python)

Location
Slough, England, United Kingdom
workload scheduling, configuration, and runtime resource access. Experience building CI/CD pipelines with GitLab CI, GitHub Actions, Jenkins, or similar tools. Experience with observability tools such as Prometheus, Grafana, and OpenTelemetry. Good understanding of platform security: secrets management, IAM, network isolation, and dependency/supply-chain risks. Strong troubleshooting … teams to identify recurring pain points and turn them into scalable platform features. Automate manual infrastructure operations and build reliable self-service workflows. Implement observability through metrics, structured logging, and alerting. Own platform production issues from troubleshooting and root cause analysis to permanent resolution. Manage and scale on-prem compute ...