1 to 25 of 36 Site Reliability Engineering Jobs in the West Midlands

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
Core Engineering - Site Reliability Engineering - Associate - Birmingham location_on Birmingham, West Midlands, England, United Kingdom What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build and operate platforms and applications that produce metrics, analyze risk, curate financial reports, enable people processes, support budgeting and financial … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
West Midlands, England, United Kingdom
This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish … understand how individual components interact under load. Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority. Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders. Highly motivated, pro-active ...

Site Reliability Engineer / Platform Lead Datadog

Hiring Organisation
MYO Talent
Location
Solihull, West Midlands, United Kingdom
Employment Type
Contract
Contract Rate
From £450 to £600 per day
Site Reliability Engineer/Platform Engineer/Platform Engineering Lead/Observability Engineer/Observability Lead/Manager/SRE/Datadog/Azure/6-month contract/Hybrid West Midlands/Remote/£500 650 per day. One of our leading clients is seeking … performance, and root-cause analysis across production systems. Desirable: Datadog certifications. Azure certifications. Cloudflare administration experience. Background in Site Reliability Engineering (SRE) or Platform Engineering leadership roles. ...

Site Reliability Engineer Datadog / Azure

Hiring Organisation
MYO Talent
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Contract
Contract Rate
From £450 to £600 per day Inside IR35
Site Reliability & Observability Engineer/Platform Engineer/Datadog Synthetic Monitoring, APM, RUM, Log Management, SLOs, Alerting/Azure/Azure DevOps/Cloudflare/6-month contract/Hybrid West Midlands/Remote/£450 600 per day Inside IR35. One of our leading clients is seeking … with distributed systems, microservices, and cloud-native architectures. Desirable: Datadog certifications. Azure certifications. Cloudflare administration experience. Background in Site Reliability Engineering (SRE) or Platform Engineering leadership roles. ...

Senior Network Engineer- IP

Location
Birmingham, England, United Kingdom
will take the lead on complex, high-impact fault resolution spanning multiple platforms and services, acting as a senior technical escalation point. Applying SRE principles and deep technical knowledge, you will drive improvements in service availability and reliability through end-to-end business ownership – implementing flawless network change, championing … automation and IaC tools (e.g. Ansible, Terraform, Netconf/YANG) to manage network infrastructure at scale and reduce operational toil. Proven ability to apply SRE principles – automation, observability and toil reduction – to improve service availability, with proficiency in a programming or scripting language such as Python. Strong proficiency in building ...

Software Engineer Lead - Site Reliability

Location
Telford, England, United Kingdom
proactive, self-starting engineer who enjoys getting things done and improving the reliability of live digital services, Standard Life could be the place for you. We’re looking for a Lead DevOps Engineer to join our Digital Engineering team. This role is focused on making immediate, practical improvements … GitHub/GitHub Actions, Azure DevOps, Terraform and automated testing. Improve deployment safety, release readiness and operational readiness for customer-facing digital services. Apply SRE principles pragmatically to improve availability, recoverability, monitoring and incident learning. Strengthen monitoring, logging, tracing, alerting and service-health dashboards across digitally connected workloads. Reduce single ...

Senior Backend Engineer

Location
Birmingham, England, United Kingdom
platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: Design, build and operate reliable, secure and scalable cloud platform services supporting critical digital products. Develop and maintain platform tooling … automation, observability, monitoring and CI/CD capabilities. Build software solutions using Python and modern engineering practices. Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets and continual service improvement. Support live ...

Infrastructure / DevOps Engineer

Location
Birmingham, England, United Kingdom
HIPAA compliance — encryption, access controls, audit logging, and network security Partner with engineering teams to optimize system performance, cost, and scalability Grow the SRE practice: runbooks, incident response playbooks, chaos engineering, and reliability reviews Location Remote (US). If you're in the Birmingham, AL area, this … infrastructure requirements (encryption at rest/in transit, audit trails, access controls) Nice to have Site Reliability Engineering background or formal SRE experience Experience supporting real-time or high-throughput systems (voice, streaming, or similar) AWS certifications (Solutions Architect, DevOps Engineer, or SysOps) Experience with multi-region ...

Lead Site Reliability & Observability Engineer

Hiring Organisation
Amtis Professional Ltd
Location
Solihull, West Midlands, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£500 - 575 per day
Lead Site Reliability & Observability Engineer Location: Birmingham Rate: Up to £575/day – Inside IR35 Contract opportunity – Initially 6 months Hybrid working: 2 days on-site per week We're recruiting a Lead Site Reliability & Observability Engineer to lead the implementation of Datadog across Azure … DevOps and GitHub. API, integration and browser-based testing expertise. Terraform experience and a strong understanding of distributed systems and microservices. A background in SRE or Platform Engineering leadership. Datadog or Azure certifications and Cloudflare administration experience would be advantageous. If you are interested in this role and would ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Stoke-On-Trent, Staffordshire, West Midlands, United Kingdom
Employment Type
Permanent, Work From Home
role also includes using AI tools, LLM platforms and coding assistants to boost productivity, support autonomous operations and improve system insight. Working across SRE, development and IT Operations, you will help embed reliability throughout the software development lifecycle, lead technical work and share knowledge that lifts standards across … background with Python, Golang, JavaScript or similar language. Knowledge of modern development practices, including testing, source control and delivery lifecycles. An understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. ...

VP, SRE: Architect Resilient, Scalable Systems

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Vice President in Site Reliability Engineering (SRE) within Core Engineering at the Birmingham location. You will lead reliability efforts across distributed systems, driving SLOs, observability, and incident response while shaping scalable, automated platforms. This role emphasizes design reviews, reliability ...

VP, Site Reliability Engineering: Scale & Resilience

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical … services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services. The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context. ...

SRE Engineer: Core Systems & Automation

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Site Reliability Engineer to join the Compliance Engineering SRE team. You will ensure production services remain healthy, automated, and scalable across cloud-native platforms. Collaborate with engineering to improve reliability, implement monitoring, and reduce downtime while balancing feature velocity with … stability. A strong background in SRE, programming, and ownership is essential. #J-18808-Ljbffr ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
Coventry, West Midlands, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work on realistic DevOps projects that may include: Building CI/CD pipelines Deploying applications to cloud environments … completion, learners may pursue roles such as: Junior DevOps Engineer DevOps Engineer Cloud Support Engineer Platform Engineer Infrastructure Engineer Site Reliability Engineer (SRE) Build and Release Engineer Cloud Operations Engineer Systems Administrator Cloud Infrastructure Engineer Apply Today If you are looking to start a career in DevOps ...

Cloud-Scale Site Reliability Engineer: Automation & Resilience

Location
Birmingham, England, United Kingdom
prominent technology firm in Birmingham is seeking a Site Reliability Engineer tasked with ensuring the reliability, performance, and scalability of critical enterprise systems. This role combines software and systems engineering to enhance automation and support cloud transformations. The ideal candidate will have expertise in Unix, Windows ...

Site Reliability Engineer

Location
Birmingham, England, United Kingdom
Overview Our client is seeking a high-impact Site Reliability Engineer to join a team responsible for ensuring the reliability, performance, and scalability of critical enterprise systems. This role blends software and systems engineering to drive automation, prevent service-impacting incidents, and support transformative cloud initiatives. … work for any US Employer without sponsorship. Benefits & Extras Work on cutting-edge distributed systems and cloud transformations Solve challenging performance, scalability, and reliability problems Collaborate with teams driving automation and monitoring initiatives Exposure to enterprise-scale network and fault-tolerant architectures High-impact role with visibility across technical ...

VP, SRE - Build Resilient, Scalable Systems

Location
West Midlands, England, United Kingdom
Goldman Sachs is seeking a Vice President of Site Reliability Engineering to lead the design and operation of highly available, observable, and resilient platforms at scale. You will partner with engineering leaders to define SLOs/SLIs, drive automation, and reduce toil across multiple teams. Ideal ...

Client Service Delivery, Sr Manager

Location
Birmingham, England, United Kingdom
primary point of contact for service delivery, building strong, trusted client relationships. Translate technical insights into clear business value, highlighting outcomes such as improved reliability and reduced Mean TimeToRecover (MTTR). Communicate the impact of AI-driven service management anddemonstratethe value of platforms such as ServiceNow AIOps, Dynatrace … alignment with ISO20k, Experience with AI Ops tools, frameworks, and implementation strategies. Knowledge of AI-enabled automation and monitoring solutions. Awareness of Site Reliability Engineering principles and practices. Locations Birmingham Additional Information Equal Employment Opportunity Statement All employment decisions shall be made without regard to age, race ...

Lead Platform Engineer (AWS)

Hiring Organisation
Kainos
Location
Birmingham, UK
Employment Type
Full-time
share knowledge and mentor those around you. Your key responsibilities will include: Working as part of a team - You'll work alongside colleagues in engineering, testing, consulting, product management and security capabilities to build, test and deploy software of the highest quality; including the operation and continuous improvement … supporting complex projects. The role requires SC Clearance. Significant expertise in: Designing, building, testing, automating, monitoring and/or operating, enhancing and improving reliability of modern digital service platform in production environments. Using the latest Continuous Delivery and automation techniques for releasing operationally ready software to production, including platform ...

IT Problem Lead

Hiring Organisation
Compass UK & Ireland
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent
improving service reliability. You will turn incident trends, major incident findings and known errors into structured, evidence-based corrective action. Working across Incident Management, SRE, Engineering, Infrastructure & Cloud, Change, Product and Service Ownership, you will establish a proactive problem management approach that identifies underlying causes, strengthens accountability and delivers … forums and driving accountability for corrective action completion across resolver teams. Maintaining accurate known error records, workarounds and links to knowledge management. Partnering with SRE and Engineering teams to convert reliability issues into prioritised engineering improvements. Producing executive-ready management information covering problem themes, risks, actions ...

SRE Associate: Build Reliable Cloud Platforms

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Site Reliability Engineer to join Core Engineering. You will help build, run and maintain high-performing, distributed systems, focusing on reliability, observability and automation across critical services. Responsibilities … include monitoring production services, capacity planning, incident response and driving improvements to SLIs/SLOs. The role requires 3–4 years in software or SRE, a CS/engineering degree, and experience with Java, Python or Perl and modern #J-18808-Ljbffr ...

Senior Applied AI Engineer (Manager) TC

Location
Birmingham, England, United Kingdom
working world. You will work across a diverse portfolio of clients spanning Financial Services, the Public Sector, and the Private Sector. Our Applied AI Engineering teams deliver production-grade AI systems in regulated financial institutions as well as government, health, infrastructure, consumer, industrial and energy organisations. This cross-sector … serverless, IAM and network security. Data engineering depth (Spark/Databricks; ETL/ELT); cloud‐native data + AI architectures. Enterprise integration and SRE principles (SLIs/SLOs, runbooks, rollback). Consulting leadership: stakeholder, budget and risk management; team leadership. Nice to have Graph/big‐data stacks; streaming ...

GCP Cloud Engineer

Location
Warwick, England, United Kingdom
make a real impact - look no further. We're looking for a talented GCP (Google Cloud Platform) Cloud Engineer to join our Systems Engineering team and help us deliver secure, scalable, and efficient cloud solutions that keep the nation's energy flowing. The GCP Cloud Engineer is responsible … Helm). Implement blue/green and canary deployment strategies, automated rollbacks, and robust monitoring/alerting for production workloads. Collaborate with development and SRE teams to embed security and quality gates into the SDLC. Cloud Security & Compliance: Enforce cloud security best practices, including least privilege, encryption at rest/ ...