251 to 275 of 741 Observability Jobs in the UK excluding London

Forward Deployed Infrastructure Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
APIs, or LLM infrastructure in production environments. Past experience at an early-stage company where you built deployment operations, playbooks, and tooling. Familiarity with observability and incident management in distributed systems (Prometheus, Grafana, Datadog, or similar). Experience with CI/CD pipeline design, deployment automation, or infrastructure self-service ...

Forward Deployed Infrastructure Engineer - Spanish speaking

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
APIs, or LLM infrastructure in production environments. Past experience at an early-stage company where you built deployment operations, playbooks, and tooling. Familiarity with observability and incident management in distributed systems (Prometheus, Grafana, Datadog, or similar). Experience with CI/CD pipeline design, deployment automation, or infrastructure self-service ...

Senior Software Engineer - Pay Sustainable Engineering

Hiring Organisation
Hackajob Ltd
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent
e.g., E2E/Cypress). Backend Excellence: Engineers sophisticated backend solutions involving API versioning, caching strategies, and complex data migration plans. Operational Maturity: Leads observability and SRE practices; defines SLOs, manages incident responses, and conducts blameless post-mortems. Security & Risk: Oversees operational security, including secrets hygiene and dependency risk management ...

Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
drive reliability, scalability and performance across critical banking systems. This role combines hands‐on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation. Responsibilities: * Build and maintain reliable, scalable and secure infrastructure platforms and solutions. * Apply SRE and software engineering practices to improve … lead complex troubleshooting and root cause analysis. * Develop automation using programming and scripting to reduce manual intervention and improve efficiency. * Develop and improve observability, monitoring, instrumentation and performance capabilities. * Use data and reliability metrics to drive continuous improvement and optimisation. * Lead technical discussions, blameless retrospectives and problem‐solving activities. * Work ...

Senior Site Reliability Engineer

Hiring Organisation
GCS
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
Salary
£75000 - £95000/annum Bonus
drive reliability, scalability and performance across critical banking systems. This role combines hands-on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation. Responsibilities: * Build and maintain reliable, scalable and secure infrastructure platforms and solutions. * Apply SRE and software engineering practices to improve … lead complex troubleshooting and root cause analysis. * Develop automation using programming and scripting to reduce manual intervention and improve efficiency. * Develop and improve observability, monitoring, instrumentation and performance capabilities. * Use data and reliability metrics to drive continuous improvement and optimisation. * Lead technical discussions, blameless retrospectives and problem-solving activities. * Work ...

Senior Cloud Engineer, AI Platform SRE

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
tools and prompt changes. It means the SLOs and on-call practice that make reliability an asset commitment rather than a hope, and the observability that makes AI-specific failure modes visible, including drift, silent quality regression, cost blowouts and agent loops. You'll be the operational conscience … Take part in the programme's on-call rotation, and build runbooks clear enough that someone else can use them at 3am. AI-specific observability: Instrument latency, token usage, cost per request, model and agent error rates, retrieval quality and drift, with dashboards and alerting that surface problems before users ...

Senior Cloud Engineer, AI Platform SRE

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
tools and prompt changes. It means the SLOs and on-call practice that make reliability an asset commitment rather than a hope, and the observability that makes AI-specific failure modes visible, including drift, silent quality regression, cost blowouts and agent loops. You'll be the operational conscience … Take part in the programme's on-call rotation, and build runbooks clear enough that someone else can use them at 3am. AI-specific observability: Instrument latency, token usage, cost per request, model and agent error rates, retrieval quality and drift, with dashboards and alerting that surface problems before users ...

Senior Cloud Engineer, AI Platform SRE

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
tools and prompt changes. It means the SLOs and on-call practice that make reliability an asset commitment rather than a hope, and the observability that makes AI-specific failure modes visible, including drift, silent quality regression, cost blowouts and agent loops. You'll be the operational conscience … Take part in the programme's on-call rotation, and build runbooks clear enough that someone else can use them at 3am. AI-specific observability: Instrument latency, token usage, cost per request, model and agent error rates, retrieval quality and drift, with dashboards and alerting that surface problems before users ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
Engineer containerized environments with advanced orchestration, networking, and security. Deliver internal platforms and self-service capabilities that improve developer experience and reduce friction. Implement observability stacks—metrics, logs, traces, and proactive alerting for reliability. Champion security and compliance across infrastructure and delivery pipelines. Architect secure, scalable networking solutions for hybrid … management (e.g., Ansible, Chef). Strong AWS architecture skills and cost optimisation strategies Advanced containerization and orchestration experience (Docker, Kubernetes, etc.). Proficiency in observability tools (Prometheus, Grafana, ELK, OpenTelemetry). Security-first mindset with hands‐on experience in access control, encryption, and incident response. Solid networking knowledge - protocols, routing ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch reliability and observability for our most critical platforms. As a Senior Lead Site Reliability/DevOps Engineer at JPMorgan Chase within the Commercial & Investment Bank, you are an integral part … significant business impact through your capabilities and contributions, and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of reliability, observability, and performance challenges that span multiple technologies and applications. Job responsibilitiesRegularly provides technical guidance and direction on site reliability practices to support the business ...

Azure Platform Engineer - Outside IR35

Hiring Organisation
Be-IT Resourcing
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Full-Time
Salary
£500.00 - £650.00 per day
/CD pipelines, ideally on Azure DevOps, that hold up under real load. Confident across cloud networking, identity and access governance. Treat observability as core engineering work Comfortable directing AI coding tools (Claude Code, Cursor, Copilot) as a daily working habit. Bonus: hybrid on-premise exposure, or a SaaS/ ...

Managing Engineer – Database, Platform

Hiring Organisation
Jobleads-UK
Location
Belfast, Northern Ireland, United Kingdom
product engineering, SRE, security, and cloud teams to define SLAs/SLOs, incident response, root cause analysis, and risk mitigation. Establish monitoring, alerting, and observability for database health and performance. Drive proactive incident prevention. Define and enforce standards, best practices, and governance for database access, security, backups, and compliance. Participate ...

Technical Engineer, Full Stack Java

Hiring Organisation
Clarify Consultancy Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
Kafka. You will demonstrate the ability to build robust, scalable and highly available solutions, with hands-on experience across Docker, Kubernetes, MongoDB and modern observability tooling. A solid grounding in design patterns, Domain-Driven Design and front-end developmentideally Reactis important, along with the ability to produce production-quality code ...

Salesforce Service Cloud & Agentforce Technical Architect

Hiring Organisation
Vallum
Location
Leeds, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
Review and approve technical designs, integration specifications, code quality standards, test automation approach and release plans. Drive non-functional design for performance, resilience, security, observability, compliance, maintainability and platform limits. Mentor technical leads and developers, resolve complex design issues and support production release readiness, cutover and post-deployment assurance. ...

Salesforce Service Cloud & Agentforce Technical Architect

Hiring Organisation
Vallum
Location
Leeds, UK
Employment Type
Full-time
Review and approve technical designs, integration specifications, code quality standards, test automation approach and release plans. Drive non-functional design for performance, resilience, security, observability, compliance, maintainability and platform limits. Mentor technical leads and developers, resolve complex design issues and support production release readiness, cutover and post-deployment assurance. ...

Lead Data Engineer

Hiring Organisation
SRG
Location
Manchester, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £90,000 per annum
across data warehousing and modern lakehouse platforms. Supporting and mentoring Data Engineers, helping to develop capability across the team. Driving improvements in reliability, performance, observability, and data quality. Working closely with Data Scientists, Analysts, and Engineering teams to deliver high-value data solutions. Contributing to long-term platform strategy ...

Software Engineer- Backend

Hiring Organisation
Randstad Technologies
Location
Manchester, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
£64.00 - £65.00 per hour
Software Engineer building and scaling high-performance backend systems. Experience working with modern cloud architectures (AWS, GCP, or Azure). Strong understanding of observability, reliability engineering (SLIs/SLOs), and production monitoring. Experience dealing with highly concurrent systems and A/B testing or experimentation platforms. Excellent communication skills with ...

Software Engineer- Backend

Hiring Organisation
Randstad Digital
Location
Manchester, North West, United Kingdom
Employment Type
Contract
Contract Rate
£64 - £65 per hour
Software Engineer building and scaling high-performance backend systems. Experience working with modern cloud architectures (AWS, GCP, or Azure). Strong understanding of observability, reliability engineering (SLIs/SLOs), and production monitoring. Experience dealing with highly concurrent systems and A/B testing or experimentation platforms. Excellent communication skills with ...

Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
cataloguing AI/ML platforms: Amazon Kiro, SageMaker, or similar AI integration experience Event-driven architecture: EventBridge, SQS, SNS, and Kafka/MSK DevOps & observability: CloudWatch, X-Ray, Dynatrace, CI/CD pipelines Database technologies: DynamoDB, PostgreSQL, DocumentDB, Neptune Please note that this is NOT a remote role, you will ...

Senior Cloudflare Developer- Remote

Hiring Organisation
FDM Group
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£65,000 - £70,000 per annum
across global edge locations. Integrate AI capabilities using Workers AI, Vectorize, and related services where appropriate. Monitor applications and troubleshoot production issues using Cloudflare observability tools About You Delivered one or more production applications primarily hosted on Cloudflare. Deep expertise with Cloudflare Workers and edge computing concepts. Experience implementing stateful ...

Senior Data Engineer

Hiring Organisation
Harnham - Data & Analytics Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£90,000 - £100,000 per annum
focused on scaling the company's data platform to support continued customer growth. Current projects include building a full staging environment, improving monitoring and observability, enhancing engineering tooling, and enabling lower-latency analytics across the business. Working end-to-end across the data platform, you'll design, build and optimise ...

Director of Applications

Hiring Organisation
Oscar Technology
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£130,000 - £150,000 per annum
hallucination and misuse pattern detection, with the Director of Security & Compliance Provide the operational foundation for Agentic AI and Generative AI capability pillars Capacity & Observability Own scaling, rightsizing and optimisation of cloud and SaaS resources Own application monitoring, alerting and on-call response for in-scope platforms Lead cross-platform ...

Data Platform Engineer

Hiring Organisation
Norton Rose Fulbright LLP
Location
Newcastle Upon Tyne, Tyne and Wear, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
service acceptance and transition activity, assessing readiness against availability, recoverability, supportability, security and control requirements. The role will implement and maintain monitoring, alerting and observability across Fabric, Databricks and Azure services. It will develop service health reporting, operational dashboards and exception monitoring, and maintain runbooks, known error records, support documentation ...

Product Owner

Hiring Organisation
Synechron
Location
Sheffield, UK
Employment Type
Full-time
identity, workload identity, agent identity, privileged access, secrets management, certificate lifecycle management, and API security. Understanding of cloud-native security, event-driven identity integration, observability, monitoring, audit, SSDLC, DevSecOps, and automation. Experience working in a regulated financial-services or banking environment. Ability to work effectively across multiple countries, cultures ...

Lead Identity and Security Engineer - 12 Month FTC

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Salary
£90,000
across identity and security. Work collaboratively with architecture, platform, application engineering and security teams to solve complex identity challenges. Contribute to the reliability, scalability, observability and operational maturity of identity services. Produce, maintain and review technical documentation, architecture decisions, standards and security guidance. Identify and address security risks and continuously ...