1 to 25 of 70 Permanent Reliability Engineer Jobs in the UK

Reliability Engineer

Hiring Organisation
E3 Recruitment
Location
Bradford, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
Salary
£55,000
This permanent Reliability Engineer role is working on a site that is investing in expansion and upgrade projects across the operation. Additional benefits include up to 10% employer pension contribution, annual bonus, private healthcare, life assurance, 28 day's holiday plus bank holidays and more! As the Reliability Engineer, you will play a critical role in improving asset reliability and plant performance. Working closely with engineering, maintenance and operations teams, you will identify potential equipment failures, lead root cause analysis and implement effective reliability improvements across the site. Responsibilities of the Reliability Engineer ...

Reliability Engineer

Hiring Organisation
Vantage Consulting
Location
Aberdeenshire, United Kingdom
Employment Type
Permanent
Salary
£55000/annum
Reliability Engineer The Role We are seeking a motivated Reliability Engineer to join our Electronics R&D team. The role focuses on improving product reliability through failure investigation, component qualification, reliability testing, and collaboration with engineering, production, and quality teams. Key Responsibilities Investigate product … component reliability issues and recommend improvements. Support reliability testing, failure analysis, and component qualification. Assess component obsolescence and supply risks. Work with internal teams and external suppliers to improve product and manufacturing reliability. Maintain technical documentation and contribute to high-reliability design practices. Essential Requirements Degree ...

Reliability Engineer

Hiring Organisation
Vantage Consulting
Location
Aberdeenshire, United Kingdom
Employment Type
Permanent
Salary
£55000/annum
Reliability Engineer The Role: The Electronic Engineering department are looking for an enthusiastic and driven Reliability Engineer to join our team. The successful applicant will work primarily within the Electronics R&D team to address product reliability challenges and qualify materials and electronic components … high reliability applications. Main Duties: Track, review and investigate product and component reliability issues. Report on investigation findings and provide suggestions for improvement. Work with engineering, production and quality assurance to assess internal and external reliability issues. Carry out critical component reviews to identify obsolescence or lead ...

Senior Network Reliability Engineer

Location
Wimbledon, England, United Kingdom
Senior Network Reliability Engineer Permanent Position Hybrid Role from our Wimbledon Office with On Call Work Domestic & General is growing fast across 12 countries, protecting the appliances that keep households running and cutting waste by extending product life. Recognised by Great Place to Work and the Inclusive Employers … Standard, we're building a modern, technology lead business and we're looking for a Senior Network Reliability Engineer to drive the performance, reliability and security of our hybrid, enterprise network estate. About the Role This is a hands-on senior engineering position with responsibility ...

Product Reliability Engineer – APM

Location
Greater London, England, United Kingdom
## Product Reliability Engineer - APMLondon, UK · Full-time#### About The PositionCoralogix is a modern, full-stack observability platform transforming how businesses process and understand their data. Our unique architecture powers in-stream analytics without reliance on expensive indexing or hot storage. We specialize in comprehensive monitoring of logs … such as APM, RUM, SIEM, Kubernetes monitoring, and more, enhancing operational efficiency and reducing observability spending by up to 70%.We seek a **Product Reliability Engineer** who ensures that the Coralogix APM Product and Process exceed the quality and reliability standards, establish a competitive edge, and prevent ...

Cloud Operations - Service Reliability Engineer

Location
United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. The role involves: Supporting service reliability, observability and cloud engineering across the following areas: Monitoring, observability and alerting for cloud-hosted services, including infrastructure health, service availability, performance signals ...

Cloud Operations - Service Reliability Engineer

Location
Crumlin, Glamorgan, United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. The role involves: Supporting service reliability, observability and cloud engineering across the following areas: Monitoring, observability and alerting for cloud-hosted services, including infrastructure health, service availability, performance signals ...

Cloud Operations - Service Reliability Engineer

Location
Carrickfergus, County Antrim, United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. The role involves: Supporting service reliability, observability and cloud engineering across the following areas: Monitoring, observability and alerting for cloud-hosted services, including infrastructure health, service availability, performance signals ...

Cloud Operations - Service Reliability Engineer

Location
Hillsborough, West Yorkshire, United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. The role involves: Supporting service reliability, observability and cloud engineering across the following areas: Monitoring, observability and alerting for cloud-hosted services, including infrastructure health, service availability, performance signals ...

Cloud Operations - Service Reliability Engineer

Location
United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … deployment pipelines, configuration management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. Improve the reliability, resilience and performance of cloud-hosted services through monitoring, observability and automation. Design and enhance monitoring, alerting and operational dashboards to provide real-time ...

Cloud Operations - Service Reliability Engineer

Location
Bangor, Caernarfonshire, United Kingdom
What you will do The Service Reliability Engineer is accountable for improving the reliability, observability and operational resilience of cloud-hosted services. The role focuses on monitoring, early issue identification, cloud engineering and automation, using tools and practices such as Bicep, Azure DevOps, GitHub and Ansible … deployment pipelines, configuration management and standards-led delivery; and Promoting service resiliency through proactive issue identification, operational insight, automation and continuous improvement. Improve the reliability, resilience and performance of cloud-hosted services through monitoring, observability and automation. Design and enhance monitoring, alerting and operational dashboards to provide real-time ...

Senior Hybrid Azure Network Reliability Engineer

Location
Wimbledon, England, United Kingdom
Domestic & General Group is seeking a Senior Network Reliability Engineer for a permanent hybrid role based around our Wimbledon office. You will own the reliability, performance and security of our hybrid Azure network, lead major incident resolution, and drive automation and observability across our estate. ...

Service Reliability Engineer - Manchester

Hiring Organisation
Fitch Group
Location
Manchester, United Kingdom
Employment Type
Full Time
myGwork – the largest global platform for the LGBTQ+ business community. Please do not contact the recruiter directly. Fitch Group is currently seeking a Service Reliability Engineer to embed with Fitch Ratings development squads. The role is based out of our Manchester office. As a leading, global financial information … data at Fitch? Visit: https://careers.fitch.group/content/Technology-and-Data/About the Team Fitch Group SRE provides Service Reliability Engineering expertise to Fitch’s development organizations. This squad joins Core Engineering and other SRE groups as part of Cloud Infrastructure & Platform Engineering ...

Senior Cloud Reliability Engineer - AWS & Kubernetes Remote

Location
Manchester, England, United Kingdom
Salve.Inno Consulting is hiring a Senior Cloud Reliability Engineer to own highly available, cloud-native production environments. You will work across AWS and Kubernetes with SRE practices, automation, observability, incident management, and direct engagement with enterprise customers. You will drive improvements, challenge current approaches, and help scale reliability ...

Remote Senior Cloud Reliability Engineer - AWS & Kubernetes

Location
Greater London, England, United Kingdom
Salve.Inno Consulting is seeking a Senior Cloud Reliability Engineer to own highly available, cloud-native production environments on AWS and Kubernetes. You will handle on-call incidents, lead root-cause analyses, and drive permanent improvements, using Terraform and GitOps to automate infrastructure. The role emphasizes production readiness, security ...

APM Platform Reliability Engineer - London

Location
Greater London, England, United Kingdom
Coralogix is hiring a Product Reliability Engineer in London to ensure the APM product and processes meet high reliability standards. You will help reduce engineering interruptions, improve customer satisfaction, and drive product quality through benchmarking and knowledge sharing. The role requires PromQL experience, SaaS metrics background ...

Principal Cloud Reliability Engineer Multi-Cloud

Location
Greater London, England, United Kingdom
Veson Nautical is seeking a Principal Site Reliability Engineer to design, build, and operate scalable cloud infrastructure across AWS and GCP. The role focuses on growing the GCP footprint, with cross-region and cross-cloud alignment, in a hybrid London-based environment. You will lead greenfield projects, improve … existing estates, and influence architectural direction while mentoring peers and collaborating with development teams to ensure reliability and cost efficiency. #J-18808-Ljbffr ...

Senior M&E Reliability Engineer – EMEA Travel

Location
Greater London, England, United Kingdom
Nscale, the GPU cloud for AI, seeks a Reliability Engineer (Mechanical & Electrical) to support its EMEA data centre estate. You’ll be the on-site technical escalation point for M&E systems and help set maintenance baselines. This UK role requires regular travel across Europe/Nordics, with … focus on high-density AI cooling and power. You’ll perform design reviews, audits, RCA/CAPA, and coach site engineers to improve reliability and energy efficiency. #J-18808-Ljbffr ...

Reliability Engineer: Automation & Incident Response

Location
United Kingdom
Barclays Services Corp. seeks a Reliability Engineer – Markets RTB (multiple positions) in Whippany, NJ. Provide production support for business analytics activities involving equity financing and ensure application support for customer-facing systems including trade capture, security lending/borrowing and billing. Collaborate with teams in Glasgow, India ...

Reliability Engineer SME (M&E)

Location
Greater London, England, United Kingdom
work. If you join our team, you’ll be contributing to building the technology that powers the future. About the Role (Job Purpose) The Reliability Engineer (Mechanical & Electrical) provides cross-discipline M&E engineering expertise and hands‐on technical support across Nscale's EMEA data centre estate — owned … supporting high-density AI compute. As GPU rack densities and power draw increase, cooling performance and power resilience become two of the most critical reliability factors across the estate — and neither can sensibly be engineered in isolation from the other. What You’ll be Doing (Responsibilities) M&E Technical ...

SRE Engineer – FinTech Reliability, Observability & Cloud

Location
Greater London, England, United Kingdom
Hamilton Barnes Associates Limited is seeking a Site Reliability Engineer to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies to improve reliability and operational efficiency in a global production environment. ...

AI Reliability Engineer - Scalable Infra | Visa Sponsorship

Location
United Kingdom
Anthropic, a leading AI company, seeks a Software Engineer (AI Reliability Engineering) to design, build, and maintain reliable AI infrastructure supporting large language models like Claude. The role includes defining SLOs, incident response, and cross-team collaboration to improve scalability and safety. Successful candidates will have ...

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

Senior AI/ML Platform Reliability Engineer

Location
Auchentibber, Scotland, United Kingdom
JPMorganChase is seeking a Software Engineer III within the AI/ML Data Platforms group to design and deliver trusted, scalable AI/ML platforms, with a focus on reliability and security. As part of the Reliability Engineering team, you will own non-functional requirements, build tooling ...

Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
reliably in production at scale. In this role, you'll build and operate large language model serving infrastructure, bringing strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You'll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. … enjoy solving hard production problems and making platforms measurably better, you'll find meaningful impact and growth here. As a Lead Software Engineer at JPMorgan Chase in the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability ...