26 to 50 of 126 Service-Level Objective Jobs in England

Site Reliability Engineer

Hiring Organisation
Trainline
Location
London, UK
Employment Type
Full-time
business goalsOur Tech Stack AWSNew RelicELK stackGrafanaIncident.ioDocker, ECSTerraformGithub ActionsWe'd love to hear from you if you have...Experience of SRE concepts such as SLI, SLO and error budgets. Hands-on experience with observability tooling such as New Relic, Elastic (ELK Stack), Influx, Grafana or similarExperience working with cloud providers (preferably ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
Desirable*** Experience building **internal developer platforms or tooling*** Contributions to **open-source, technical blogs, or public speaking*** Experience working in **regulated environments*** Familiarity with **SLO frameworks and error budget management*** Relevant certifications in your specialist domain## ## **Success Measures*** Improved reliability and performance within your domain of specialism* Adoption ...

Lead SRE - Chase UK

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
building reliable infrastructure and tooling that expedites feature development. Develop meaningful service metrics, user journeys, service-level indicators, service-level objectives, error budgets, dashboards, and actionable alerts. Engage with ...

Lead SRE - Chase UK

Location
Westminster, West End, United Kingdom
building reliable infrastructure and tooling that expedites feature development. Develop meaningful service metrics, user journeys, service-level indicators, service-level objectives, error budgets, dashboards, and actionable alerts. Engage with ...

Lead SRE - Chase UK

Location
Greater London, England, United Kingdom
building reliable infrastructure and tooling that expedites feature development. Develop meaningful service metrics, user journeys, service-level indicators, service-level objectives, error budgets, dashboards, and actionable alerts. Engage with ...

Principal Site Reliability Engineer (Remote)

Hiring Organisation
Raytheon
Location
Wokingham, Berkshire, UK
Employment Type
Full-time
team who value new ideas. YOU will learn how to balance feature development speed and reliability while meeting well-defined service-level objectives. YOU will learn to work collaboratively with your colleagues across the world to deliver great systems rather than just ...

Site Reliability Engineer

Hiring Organisation
Infinity Quest
Location
City of London, London, United Kingdom
leading blameless post-incident reviews Define and implement monitoring, logging, and distributed tracing strategies; build and maintain dashboards; set meaningful alerts; and drive SLO/SLI/SLA and error budget adoption across services Scope technical projects and break them down into user stories and tasks, driving them to completion ...

DevOps Team Manager

Hiring Organisation
Bromcom Computers Plc
Location
Bromley, London, United Kingdom
Employment Type
Permanent
workloads, with the ability to govern standards and technical gates. A track record of operating production, multi-tenant SaaS at scale, including SLA/SLO ownership, major-incident command, root-cause analysis, disaster recovery and service improvement. Willingness and ability to participate personally ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
Proficiency with database development (SQL Server, PostgreSQL, MySQL, etc) Proficiency with defining, right-sizing, tracking, and reporting on Service Level Objectives (SLOs), Service Level Indicators (SLIs), system availability, and the progress and outcomes ...

Cloud Native Specialist

Location
Greater London, England, United Kingdom
Dynatrace into CI/CD pipelines (Jenkins, GitLab, GitHub Actions) and "GitOps" workflows (ArgoCD, Flux).* Help customers implement Service Level Objectives (SLOs) and Error Budgets within their cloud-native stacks to drive SRE maturity.4. Ecosystem Advocacy:* Stay at the forefront ...

Site Reliability Engineer / Senior Engineer

Location
Greater London, England, United Kingdom
support a wide ranging CSR programme + 2 days’ volunteering leave per year Your key responsibilities Defining and implementing Service Level Objectives (SLOs) and embed Site Reliability Engineering (SRE) principles to improve service reliability and operational excellence ...

Principal DevOps Engineer

Location
Greater London, England, United Kingdom
/CD pipelines (artifact versioning, approvals, promotion strategy, policy-as-code where applicable) Establish observability standards using VictoriaMetrics/Prometheus (metrics strategy, alerting, SLO/SLA monitoring, dashboards) Provide production leadership: incident response, RCA/postmortems, reliability improvements, capacity planning Mentor engineers, review designs/code, and raise overall engineering ...

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
standards, providing deep insights into system health and performance. GCO Feature Implementation: Aid with the design of GCO dashboards, log‐based metrics, alerts, and SLO/SLI tracking to provide comprehensive visibility. Essential Skills Strong understanding of SRE concepts, including SLOs, SLIs, error budgets, and toil reduction. Observability Instrumentation: Hands ...

Infrastructure Engineer

Location
Greater London, England, United Kingdom
DevOps Practices AWS, GCP, Azure Python or Go Automation Helm and Terraform Hard Skills Infrastructure Engineering DevOps Automation Containerization Incident Response Root‐Cause Analysis SLO Definition Monitoring and Logging System Design High‐Availability Systems Soft Skills Excellent Communication Collaboration Problem-Solving Ownership Accountability Industry Keywords Generative AI High‐Traffic Platforms ...

Senior Full Stack Engineer - Managed Service

Location
Leeds, England, United Kingdom
PagerDuty, OpsGenie or ServiceNow. Experience defining service level indicators (SLIs), service level objectives (SLOs) and operational metrics. Experience building self-healing or automated operational processes. 25 days Annual Leave (plus bank holidays ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service ...

Site Reliability Engineer

Hiring Organisation
Bristow Holland Ltd
Location
Manchester, United Kingdom
Employment Type
Permanent
Salary
£55000 - £60000/annum - Offering 100% Work from home
with Development and DevOps teams throughout the application release process Balancing the delivery of new functionality with platform reliability and service-level objectives Identifying bottlenecks and proposing improvements across infrastructure and applications Improving the reliability, quality and time-to-market ...

Site Reliability Engineer

Hiring Organisation
Bristow Holland Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£55000 - £60000/annum - Offering 100% Work from home
with Development and DevOps teams throughout the application release process Balancing the delivery of new functionality with platform reliability and service-level objectives Identifying bottlenecks and proposing improvements across infrastructure and applications Improving the reliability, quality and time-to-market ...

sre engineer in fintech

Location
Greater London, England, United Kingdom
resources and Helm for Kubernetes application deployments; Implement and manage monitoring, alerting, and logging solutions; Define, measure, and enforce Service Level Objectives and Service Level Indicators; Participate in the on-call rotation, where ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
easy for downstream teams to use.Raise reliability and data quality• Define data contracts, validation rules, freshness expectations, lineage, and service-level objectives for critical datasets.• Implement automated testing, anomaly detection, alerting, and observability across the data lifecycle.• Own production issues through ...

Observability SRE

Location
Greater London, England, United Kingdom
outside, if need be, towards building and maintaining robust, scalable, highly available production systems in accordance with our service level objectives Preventing production incidents but when they do occur, performing effective incident and problem management and RCA to minimize downtime ...

Jobshare - Sr Lead Software Engineer - Site Reliability Engineer, Python & Infrastructure management - Part time/Jobshare

Location
Greater London, England, United Kingdom
virtual SRE community of practice) that scales reliability improvements across many application flows. Defines and implements standards for: Service cataloging, SLO/SLI frameworks and error budgets, incident response maturity, blameless post-incident reviews, resiliency patterns, capacity, performance, and scalability engineering. Drives service ...

Lead DevOps Engineer

Hiring Organisation
Collinson Group
Location
London, UK
Employment Type
Full-time
automation so routine operational tasks are eliminated, not managed. Observability - Own the observability strategy across the platform using Datadog. Define what good looks like: SLO/SLA dashboards, alerting thresholds, runbooks, and the feedback loops that let teams act on signals before users feel them. Team Leadership & Mentoring - Lead ...