151 to 175 of 224 Incident Management Jobs in London

Technical Lead

Hiring Organisation
17918
Location
London, United Kingdom
Reliability & Operations Own technical operations and platform stability Act as the first point of contact for production incidents and outages Improve monitoring, alerting and incident management processes Reduce operational risk through better documentation and processes Work with external technology partners where specialist support is required Platform & DevOps Improve … deployment processes and release management Introduce Infrastructure as Code and modern DevOps practices Build CI/CD pipelines and automation capabilities Strengthen security, governance and auditability Drive improvements in resilience, scalability and performance Software Development Contribute directly to platform development Support and maintain existing Node.js services Work alongside product ...

Senior Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Core Technical Skills Advanced expertise in Unix/Linux systems administration, including performance tuning, troubleshooting, and scripting. Strong proficiency in Git and source control management, including branching strategies and integration with CI/CD pipelines. Proven experience with infrastructure automation, particularly Ansible, in large enterprise environments. Splunk Platform: hands … with numerous Worker Nodes and extensive Edge node infrastructures. Desirable Skills Experience in development and/or support of Banking applications. Exposure to Major Incident Management processes. Experience working within large, multi‐data‐centre enterprise environments. Business awareness with adaptable communication and flexible approach to evolving priorities. Strong ...

Integration Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
copilots, automation platforms, and intelligent workflows to securely discover, access, and act on enterprise systems and data. MuleSoft AI Gateway and secure AI traffic management Model Context Protocol (MCP) enablement for enterprise APIs and tools MuleSoft Agent Fabric adoption, including tool discovery, agent governance, and agent-to-system orchestration … Anypoint Monitoring, Exchange, and CI/CD practices for AI-enabled integration delivery. Drive best practices for error handling, retries, idempotency, throttling, resiliency, traffic management, and graceful failure modes. Partner with SRE and DevOps teams on observability, capacity planning, incident management, and operational runbooks for agent-driven ...

Senior Software Engineer I

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Product Vision: Partner with Product Managers to prioritize features, translate client requirements into technical specifications, and ensure solutions align with our overarching product vision. Incident Management: Confidently own the resolution of incidents, coordinating mitigation, fixes, and post-mortem artifacts. Collaboration & Mentorship Team Growth: Mentor and support less experienced … engineers, helping them level up technically and adapt to Thought Machine’s engineering philosophy. Stakeholder Management: Translate highly technical problems and architectural decisions into clear concepts for non-technical stakeholders. What We’re Looking For Technical Expertise: Strong proficiency in modern backend languages (ideally Go or Python) and experience ...

Mandarin speaking IT Support

Hiring Organisation
People First
Location
Central London, London, England, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
What You'll be Doing: Carry out the Help Desk first line support duty for IT Centre (Europe) including Help Desk cases, user ID management, production access and Head Office’s service requests and manage the workflow of these requests Carry out machine room environment and devices first line … maintenance, support second line engineer to perform system changes and maintenance Monitor the production systems and applications and follow the incident management procedure to deal with critical messages. Perform first line handling of system and network alerts Carry out daily operational duty and run batch jobs Assist with ...

IT Manager

Hiring Organisation
FundApps
Location
Greater London, United Kingdom
Employment Type
Full Time
Salary
65000 to 75000 GBP Annually
will own FundApps’ global corporate IT strategy roadmap in service of the company goals. The global IT support service, including service levels, escalations, incident management, employee satisfaction and continuous improvement. The Joiner, mover and leaver processes and IAM processes for over 270 users. A fleet of MacOS, Windows … room systems and associated workplace technology across 5 offices. Operation of the corporate IT security and compliance controls in partnership with the security team. Management of IT vendors, contracts, licences and expenditure, balancing cost, employee experience, security and long-term maintainability. Our Tech Stack MacOS and Windows, Okta, Google ...

Senior Software Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
time, to quality, and at scale Lead global engineering and DevOps teams, translating business priorities into effective technical execution Drive operational excellence, including incident management, service performance, and resilience during critical exam periods Partner with regulators and senior stakeholders, acting as a trusted technology leader across internal … education, finance, or healthcare) Deep understanding of cloud, DevSecOps, APIs, and secure system design Experience driving agile transformation and modern engineering practices Strong stakeholder management, including senior leadership and external or regulatory bodies Experience building and scaling global, distributed teams Commercial awareness, including budget ownership and supplier management ...

2nd Line Engineer / NOC Engineer - URGENT

Hiring Organisation
eTech Partners
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£40,000 - £45,000 per annum
/7 operational environment where service availability is paramount. Skills & Experience Previous experience in a 2nd Line Infrastructure Support, Service Desk Strong troubleshooting and incident management skills. Good knowledge of Microsoft technologies, including Windows Server and Microsoft 365. Experience supporting cloud-based infrastructure and backup solutions. Comfortable working ...

Senior Software Engineer, Platform

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
work on it. Make promises, and keep them. Do Sensible Things : Be directly involved in determining how our platform works. Participate in incident management and determine sensible practices as the platform evolves. Garage Door Open : Create and maintain comprehensive internal documentation for systems and processes, ensuring clarity ...

Senior Engineering Manager - Enterprise Trust & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
clean, well‐tested APIs and integrations that hold up under enterprise security reviews. Stability & Operations (S&O) keeps the platform dependable as it scales: incident management and the severity model, service‐level objectives (SLOs) and error budgets, observability and DORA (DevOps Research and Assessment) delivery metrics … tighter discovery‐to‐build‐to‐review loops, and you modelling it in how you run the teams. Own the operational backbone. On‐call health, incident practice, and clear, proactive communication when things go wrong, so reliability is a habit rather than a fire drill. What we’re looking ...

Critical Incident Response Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Palantir is seeking an Incident Management Engineer (IME) to join a centralized team responsible for managing the most critical outages. You will triage, troubleshoot, and coordinate resolutions to restore services as quickly as possible, while maintaining clear, proactive communication across stakeholders. The role requires excellent collaboration ...

Lead Software Engineer - Backend Engineer - Chase UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you've built through their entire lifecycle, including in production and during incident management Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes … e.g., AI-assisted code review/refactoring, test strategy acceleration, incident/root-cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team. Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise ...

Engineering Manager, GenAI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About Pleo Messy spend management is tricky business. And tedious processes are a lose-lose situation for all involved, not just finance. At Pleo, we're changing that. We build spend solutions that make managing money seamless, empowering, and surprisingly effective for finance teams and employees alike - with … gateway and tool registry, agentic runtime, evaluation tooling, and the observability layer that makes AI outputs trustworthy at scale. This is a technical engineering management role for a builder-leader. You will lead a team of strong staff and senior engineers through a critical transition: evolving from a Kotlin ...

T-ISAC Platform Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
delivering secure, scalable workflows, automation capabilities and collaboration tools for threat intelligence sharing and enrichment. Service Reliability & Operations: Lead operational reliability through proactive monitoring, incident management, performance optimisation and secure‐by‐design engineering practices. Integration & Interoperability: Deliver and support integrations, APIs and data‐exchange capabilities that enable secure … defining future‐state architecture, driving technical improvements and evaluating emerging technologies. Security & Compliance: Maintain a secure and compliant platform using automated controls, vulnerability management and adherence to regulatory and governance standards. Collaboration & Community: Foster collaboration across members, analysts, security teams and industry partners to improve adoption and community engagement. ...

Audio Visual Engineer - onsite

Hiring Organisation
Unified Support Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£45,000
easily be picked up by colleagues during an absence. Admin 10% Follow appropriate departmental and company procedures and policies (i.e., change control, problem and incident management) Monitor performance through the scorecard Monthly meetings with on-site primary contact Reporting ticket management Essential: Previous AV support experience … concierge service Skilled AV Engineer possessing good interpersonal skills, and should be comfortable with Senior Management Must be smart and confident in their appearance. Should have proven abilities within the AV industry and/or corporate environment for over 5 years Excellent communication and customer service skills Enthusiastic ...

Senior DevOps Engineer — Cloud IaC & MLOps Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Azure. Ideal candidates have deep Terraform expertise, experience with Kubernetes-based ML tooling, and a proven track record in high‐availability environments with strong incident management skills. #J-18808-Ljbffr ...

Low Latency Trading Systems Engineer C++ - Selby Jennings

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
modern DevOps/SRE practices. Proven track record designing and maintaining scalable, resilient and highly available production platforms. Experience with monitoring, observability and incident management within mission-critical environments. Knowledge of low-latency systems, electronic trading platforms or high-performance financial technology environments. #J-18808-Ljbffr ...

Engineering Manager

Hiring Organisation
Randstad Construction and Property
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 65,000 - 70,000 Annual
smart BMS controls. Showroom Standards: Drive an "engineering-first" culture, maintaining all plant rooms, workshops, and critical infrastructure spaces to immaculate, showcase-ready standards. Incident Management: Direct emergency responses to plant failures or critical system alarms, leading rapid recovery procedures and presenting clear root-cause analyses to client ...

Engineering Manager

Hiring Organisation
Randstad Construction & Property
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£65,000 - £70,000 per annum
smart BMS controls. Showroom Standards: Drive an "engineering-first" culture, maintaining all plant rooms, workshops, and critical infrastructure spaces to immaculate, showcase-ready standards. Incident Management: Direct emergency responses to plant failures or critical system alarms, leading rapid recovery procedures and presenting clear root-cause analyses to client ...

Senior Java Developer

Hiring Organisation
Global
Location
Greater London, United Kingdom
Employment Type
Full Time
code tools (e.g. Terraform), CI/CD tooling (e.g. Jenkins, GitHub Actions), modern Java language features, and good practices for production support and incident management. ...

Principal Platform Engineer (SRE/Cloud)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
supporting a foundation of tools and common services — including our managed Kubernetes clusters, observability tooling, and much more — while setting best practices around incident management and cost controls and empowering teams to build it, run it. The Principal Engineers work together as a unified team with a common … will deliver on Beamery's strategy Engage with other Principal Engineers in setting and advocating company-wide standards for operational excellence, observability, reliability and incident response Take a whole-company view of major incidents, identifying recurring themes and turning them into company-level investments Own and evolve the platform ...

Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Manager to lead a squad responsible for delivering critical healthcare data capabilities at scale. This is a unique opportunity to combine technical leadership, people management and delivery oversight while helping shape the future of health data infrastructure used by researchers across the UK and beyond. You’ll join … continuous improvement and effective Agile ways of working. Overseeing technical decision‐making, balancing delivery pace, quality, resilience and technical debt. Supporting operational excellence through incident management, governance, compliance and service reliability. Communicating progress, risks and strategic updates to senior stakeholders and leadership teams. We don’t expect ...

Senior Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 60,000 - 65,000 Annual
monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience … Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including ...

HR Business Partner

Hiring Organisation
Humanoid
Location
City of London, London, United Kingdom
with leaders to evolve career frameworks/levels for technical roles (e.g., SWE, robotics, embedded, firmware, electrical, mechanical). Clarify dual tracks (IC vs management), promotion readiness, and growth expectations. Spot retention risks in critical technical roles and implement targeted plans. Support leaders through difficult moments with empathy, directness … proven HRBP experience supporting technical teams (Engineering, Hardware, Software, AI/ML). Strong capability in org design, workforce planning, manager coaching, performance management, and employee relations. Comfort operating in ambiguity and driving clarity in fast-moving environments. High judgment with an ability to balance empathy with high standards. ...

Open Banking Governance Analyst

Hiring Organisation
IT Talent Solutions
Location
London, United Kingdom
Employment Type
Contract
PSD2 (EU), UK Open Banking, and Australia's Consumer Data Right (CDR) . In this contract role, you'll be responsible for regulatory reporting, incident management, API performance oversight, and governance activities across several international markets. You'll work closely with engineering, product, compliance, and risk teams … fintech - Strong understanding of EU or UK regulatory frameworks (PSD2, FCA, Open Banking) - Advanced SQL skills and experience working with large datasets - Excellent stakeholder management and cross-functional collaboration skills - Ability to work independently in a complex, multi-jurisdictional environment Nice to Have Open Banking API experience Knowledge ...