476 to 500 of 589 Site Reliability Engineering Jobs in London

Network Engineer

Location
Greater London, England, United Kingdom
ideas take time to evolve. Together we’re building a world-class platform to amplify our teams’ most powerful ideas. As part of our engineering team, you’ll shape the platforms and tools that drive high-impact research - designing systems that scale, accelerate discovery and support innovation across … looking for? The ideal candidate will have the following experience: Strong background as a Network Engineer in enterprise or large-scale environments Experience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and Jenkins Experience with Cisco and Arista switching and routing ...

Network Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
ideas take time to evolve. Together we're building a world-class platform to amplify our teams' most powerful ideas. As part of our engineering team, you'll shape the platforms and tools that drive high-impact research - designing systems that scale, accelerate discovery and support innovation across … looking for? The ideal candidate will have the following experience: Strong background as a Network Engineer in enterprise or large-scale environmentsExperience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and JenkinsExperience with Cisco and Arista switching and routing, alongside network ...

Tivoli Netcool OMNIbus SME - 11850CF

Hiring Organisation
Proactive Appointments
Location
London, UK
Employment Type
Full-time
diagrams and support procedures. Contribute to the monitoring and observability roadmap and identify opportunities for platform modernisation. Provide technical mentoring and knowledge sharing across engineering and operational teams. Significant hands-on experience administering and supporting IBM Tivoli Netcool Omnibus (NOI) in large-scale enterprise environments. Expert knowledge of Omnibus … integration technologies. Understanding of monitoring across AWS, Azure and Google Cloud. Experience with container and Kubernetes monitoring. Knowledge of enterprise observability frameworks and modern SRE practices. Tivoli Netcool OMNIbus SMEDue to the volume of applications received for positions, it will not be possible to respond to all applications and only ...

Tech Accelerate Grad Program-Software Engineer

Hiring Organisation
LexisNexis Risk Solutions Group
Location
London, UK
Employment Type
Full-time
Solutions Tech Accelerate Graduate Program-Software Engineer IAre you a new or upcoming graduate seeking an entry-level opportunity in the realm of Software Engineering? Look no further than the prestigious Risk Solutions Tech Accelerate Graduate Program. About the BusinessLexisNexis Risk Solutions provides customers with solutions and decision tools … Risk Solutions Technology Graduate Program is a 12-month experience with two six-month rotations. Participants gain hands-on experience in big data, AI, SRE, DevOps, and security engineering. They also get exposure to leading cloud platforms like AWS and Azure. The program focuses on technical and personal development, offering ...

Managing Consultant - Cloud Programme Manager

Location
Greater London, England, United Kingdom
facing environment.* Delivering cloud migration and modernisation with at least one of AWS, Azure or GCP; familiarity with DevOps, FinOps, platform engineering and SRE is desirable.* Strong programme governance capability – business case, roadmap, RAID, dependency management, executive reporting, and scope/commercial management.* Stakeholder management across senior executives, finance … procurement, engineering, security and operations, with the ability to enlist support and commitment from peers in a matrixed organisation.* Experience of programmes involving AI or data workloads, and/or using AI tools to accelerate programme delivery, is highly desirable.* Certifications in one or more of MSP, PMP, SAFe ...

Technical Recruiter

Location
City of Westminster, England, United Kingdom
Role Overview We are seeking a Technical Recruiter on a 6 month fixed term contract to accelerate hiring across AI/ML, engineering, data and cloud/DevOps. This is a delivery-led position with strong stakeholder exposure, requiring advanced sourcing, technical fluency … pace. Key Responsibilities Deliver full-cycle hiring across: AI/ML Engineering & Applied Science Full Stack & Backend Data Engineering, Analytics & Platform DevOps, SRE, Cloud Infrastructure Product, Architecture & Security Build targeted technical talent pipelines using: Ashby ATS & LinkedIn Recruiter GitHub, HuggingFace, Stack Overflow, Kaggle, Discord, ArXiv Modern sourcing automations ...

Data Reliability Engineer

Hiring Organisation
Ashdown Group
Location
London, UK
Employment Type
Full-time
successful multinational technology business is looking for a Data Reliability Engineer to join its growing data team in Central London. This role is hybrid – you'll be able to work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents … data health, detect anomalies, and enforce standards across complex data pipelines and platforms. You'll have experience in Data Engineering, Data Platform, or SRE-style roles, with strong SQL and Python skills and experience working in modern cloud-based data environments. Hands-on experience with data observability tools such ...

AI-Driven Principal Platform Engineer - Multi-Cloud SRE

Location
Greater London, England, United Kingdom
engineering across AWS and Azure, own networking design, and lead IaC, CI/CD, and Kubernetes platforms as products. A strong focus on SRE, AI-enabled workflows, and reliability is required. #J-18808-Ljbffr ...

Senior Software Engineer - London

Location
Greater London, England, United Kingdom
decisions move the needle for our customers daily. We prioritize autonomy and pragmatism, giving you the space to solve complex problems without unnecessary friction. Engineering excellence here is measured by the reliability and simplicity of the systems you build to power a global platform. How we work …/Cypress). Backend Excellence: Engineers sophisticated backend solutions involving API versioning, caching strategies, and complex data migration plans. Operational Maturity: Leads observability and SRE practices; defines SLOs, manages incident responses, and conducts blameless post-mortems. Security & Risk: Oversees operational security, including secrets hygiene and dependency risk management, to ensure ...

Senior Software Engineer - London

Location
Greater London, England, United Kingdom
decisions move the needle for our customers daily. We prioritize autonomy and pragmatism, giving you the space to solve complex problems without unnecessary friction. Engineering excellence here is measured by the reliability and simplicity of the systems you build to power a global platform. How we work …/Cypress). Backend Excellence: Engineers sophisticated backend solutions involving API versioning, caching strategies, and complex data migration plans. Operational Maturity: Leads observability and SRE practices; defines SLOs, manages incident responses, and conducts blameless post-mortems. Security & Risk: Oversees operational security, including secrets hygiene and dependency risk management, to ensure ...

London CloudOps Lead — SRE & Automation Champion

Location
Greater London, England, United Kingdom
will drive reliability, automation, and on-call readiness while guiding architectural decisions. The role requires 10+ years in CloudOps/DevOps/SRE, strong AWS/Linux/Python skills, and hands-on IaC/CI/CD proficiency. DBS UK security clearance eligibility and cross-functional collaboration ...

Operations and SRE Manager

Location
Greater London, England, United Kingdom
positive mark on our business. But that's not all - we're creative problem solvers with an entrepreneurial spirit. About the Role As SRE Manager, you will lead the operational reliability function for ICIS, ensuring stable, resilient, and well-supported production services for both internal and external customers. … RCAs, post-mortems and improvement actions are owned, tracked and completed. Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective. Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support. Protect service levels by ensuring engineering effort ...

Agile Project Manager

Location
Greater London, England, United Kingdom
Scrum Master with broader project management accountability within the DevSecOps DB function. You will enable effective Agile delivery, oversee key project milestones and support engineering teams in driving operational excellence and strategic programs. Responsibilities Facilitate Agile ceremonies including daily stand-ups, sprint planning, reviews and retrospectives Drive continuous improvement … maintain project timelines, tracking delivery against milestones Manage and resolve project risks and issues, ensuring visibility for stakeholders Collaborate closely with DevOps and SRE teams to enhance delivery metrics and reporting via dashboards Support the integration and optimization of CI/CD workflows and pipelines Promote engineering best practices ...

Agile Project Manager

Hiring Organisation
EPAM Systems
Location
London, UK
Employment Type
Full-time
Scrum Master with broader project management accountability within the DevSecOps DB function. You will enable effective Agile delivery, oversee key project milestones and support engineering teams in driving operational excellence and strategic programs. Facilitate Agile ceremonies including daily stand-ups, sprint planning, reviews and retrospectivesDrive continuous improvement through structured … when necessaryCreate and maintain project timelines, tracking delivery against milestonesManage and resolve project risks and issues, ensuring visibility for stakeholdersCollaborate closely with DevOps and SRE teams to enhance delivery metrics and reporting via dashboardsSupport the integration and optimization of CI/CD workflows and pipelinesPromote engineering best practices ...

Release Manager London, United Kingdom Sunnyvale, California USA

Location
Greater London, England, United Kingdom
structure, and disciplined release practices. We're looking for someone pragmatic, methodical, and system-minded - someone who understands how releases fit into the broader engineering and product ecosystem, and who introduces processes and systems where they are needed to achieve stability, clarity, and business goals. Key responsibilities: Release Planning … engineering, operations, and model development, and build processes and frameworks that support clarity, stability, and business goals. Tooling & Automation Influence: Partner with SRE/DevOps to shape CI/CD, testing and release tooling. Defining requirements that improve developer experience and release efficiency. Cross-Functional Collaboration: Work closely with ...

Platform Engineer

Location
Greater London, England, United Kingdom
incident response platform, built to help teams dramatically reduce incident response time and improve reliability. We bring together on‐call, incident response, AI SRE, and status pages in a single platform, giving teams everything they need to respond quickly, reduce downtime, and keep customers in the loop. Since launching … helped over 1,500 companies, including Netflix, Airbnb, and Block, run more than 500,000 incidents. Every month, tens of thousands of responders across Engineering, Product, and Support use incident.io to restore services faster, stay aligned under pressure, and focus on building what matters. We’re a fast‐growing ...

Principal Security Engineer, Product & Infrastructure

Hiring Organisation
Pigment
Location
London, UK
Employment Type
Full-time
reproduce and triage vulnerabilities, dig through infrastructure configuration and build the automation that closes gaps. You'll also set direction and bring product and engineering with you, but that influence comes from technical credibility and not process: the engineers you work with will take you seriously because … triage them, design or validate the mitigation, and confirm it actually worked. Improve the KPIs that tell you whether the process is holding. Detection Engineering - Build and improve our detection capability alongside the infrastructure and engineering teams: identify the signals worth collecting, write rules that catch real attacks ...

Senior/Principal Product Manager - Storage

Location
Greater London, England, United Kingdom
scale. You’ll work across object, block, and file storage, data lifecycle, performance, resilience, APIs, and platform integrations, partnering closely with engineering, infrastructure, SRE, networking, security, commercial teams, and customers. This is a highly technical product role focused on building storage services that are high-performance, reliable, scalable … well as the integration with the existing Radiant services. Translate customer needs into clear product requirements, APIs, workflows, service levels, and priorities. Partner with engineering on storage architecture, performance, durability, availability, networking, and platform integration. Work with customers and internal teams to understand workload requirements around capacity, throughput, latency ...

Senior Cloud Infrastructure Engineer - Remote & High-Impact

Location
Greater London, England, United Kingdom
, deploying workloads with Kubernetes, Fargate, or other container tech, and automating infrastructure with Terraform and CloudFormation. We expect 6+ years in Cloud/SRE/DevOps roles, strong AWS expertise (EC2, S3, IAM, Lambda), Linux proficiency, and hands-on CI/CD experience. #J-18808-Ljbffr ...

IBM Netcool / Observability Technical Lead

Hiring Organisation
Deerfoot Recruitment Solutions
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£780 - £830 per day
Monitoring (ITM), Dynatrace, IBM Instana or OpenTelemetry, REST APIs/JSON/XML, cloud monitoring across AWS, Azure or GCP, container and Kubernetes monitoring, SRE practices, financial services or regulated environment experience, and leading technical improvement initiatives If you've held any of these roles or used these technologies/… Observability Engineer, Systems Monitoring Lead, Infrastructure Monitoring Analyst, ITM Administrator, NOI, ObjectServer, Netcool Impact, WebGUI, Probes and Gateways, ITIL, ServiceNow integration, Dynatrace, Instana, OpenTelemetry, SRE, Observability & Monitoring TPM, Systems Monitoring Technical Lead. Deerfoot Recruitment Solutions Ltd is a leading independent tech recruitment consultancy in the UK. For every CV sent ...

Application Security Engineer

Location
Greater London, England, United Kingdom
delegated credentials and tenant-scoped access, and proving the controls hold under adversarial testing Working as an embedded, trusted partner to product teams, SRE, detection and response, the data platform team, and the MystraAI/ML team, unblocking rather than gating Who You Are Experience within application security in production. … code review, and proof-of-concept exploitation across web services, APIs, and data-serving interfaces. Triaging scanner output doesn't count Production-quality software engineering in at least one mainstream language (such as Python, Go, or TypeScript), and you can read several. You've built and operated security ...

Data Scientist

Location
Greater London, England, United Kingdom
emerging AI technologies and recommend practical applications across the organisation. Data Science & Analytics Perform advanced statistical analysis and modelling across large‐scale operational and engineering datasets. Develop feature engineering, experimentation and model evaluation frameworks. Create insights and recommendations that drive decision making for senior stakeholders. Build and maintain … tools. Desirable Skills Experience within financial services, regulated environments or market infrastructure. Familiarity with Azure AI services and Microsoft AI ecosystem. Experience with observability, SRE or operational analytics use cases. Experience developing AI‐powered automation solutions. Knowledge of Responsible AI and model governance frameworks. Behaviours & Competencies Strong problem‐solving capability ...

Lead DevSecOps Engineer

Location
Greater London, England, United Kingdom
systems that underpin a growing technology platform. This is an opportunity to join at an exciting stage of growth, working with a highly capable engineering team and taking significant ownership across infrastructure, automation and security. The Role: Lead the development and evolution of DevSecOps and cloud infrastructure Design … standards across infrastructure and security Provide technical leadership and support the development of other engineers About You: Strong background in DevOps, Platform, Infrastructure or SRE Excellent understanding of security and secure software development Hands-on experience with cloud platforms, CI/CD, containers and infrastructure-as-code Strong experience automating ...

Senior IP Network Engineer & SRE Lead

Location
Greater London, England, United Kingdom
seeking a Senior Network Engineer- IP to join Professional Services. You will act as a subject matter expert across network engineering and SRE, leading complex fault resolution and end-to-end changes across BT’s fixed network infrastructure. You will champion reliability, automation and observability, delivering high-impact ...

SRE Engineer: Build Reliable, Scalable Systems

Location
Greater London, England, United Kingdom
leading recruitment agency is seeking an SRE Engineer to combine software and IT engineering principles for building reliable systems. Responsibilities include scoping projects, designing software, and automating infrastructure management. Ideal candidates will have proficiency in programming languages like Python or Go, experience with Terraform, and familiarity with CI/ ...