376 to 400 of 481 Incident Management Jobs in London

Production Reliability Lead

Location
Greater London, England, United Kingdom
Complexio is seeking an Operations Team Lead to own production reliability and lead a high-performance team. You will drive observability, incident management, and proactive reliability improvements across all live customer-facing systems. You will shape incident response, lead runbooks, and ensure sustainable on-call rotations while ...

Principal Tech Team Lead (Front End)

Location
Greater London, England, United Kingdom
driving best practices for frontend architecture, accessibility, reliability, and performance. This is a hybrid role that combines technical leadership, organisational strategy, and people management . You’ll work closely with the VP of Engineering and Staff Engineers to shape our technical vision, enforce architectural standards, and develop future leaders … Frontend Technical Strategy & Alignment Define and maintain frontend architecture standards (React + TypeScript), ensuring consistency across squads. Set conventions for component architecture, routing, state management, and data‐fetching patterns. Support the design system and shared component library direction, including governance and adoption. Drive front end performance strategy and ensure ...

Head of Cloud Platform Engineering

Location
Greater London, England, United Kingdom
proportional operational effort. Modernising our infrastructure Lead the infrastructure migration towards our target operating model through incremental, low-risk delivery. Develop repeatable, automated lifecycle management for our dedicated customer deployments, with clear isolation and resilience built in. Cloud architecture and cost Own the architecture, security, resilience, and operational effectiveness … early delivery is underway. Executive reporting includes meaningful visibility of service reliability and cloud unit economics. Within your first year Automated provisioning and lifecycle management for our target platform is operating successfully in production. Operational effort grows significantly more slowly than our customer base. The platform engineering organisation ...

Senior Backend Engineer | AI Platform

Location
Greater London, England, United Kingdom
similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. Key responsibilities: Design … methodologies. Build platform capabilities that enable teams to safely deploy, monitor, and iterate on AI-powered applications. Configure and maintain tracing, monitoring, observability, and incident management solutions to ensure platform reliability. Partner closely with product and engineering teams to understand their needs and provide scalable platform solutions. Continuously ...

AMBG Senior Cyber Defense Architect

Hiring Organisation
Avanade
Location
London, UK
Employment Type
Full-time
strategy and business case through to design, mobilisation and delivery assurance. Shape modern SOC and SecOps transformation programmes, including operating model design, tooling consolidation, incident management, detection engineering, automation, threat hunting and continuous improvement. Define and evolve AMBG cyber defence offerings, accelerators and go-to-market assets that … services engagements. Build trusted relationships with client security leaders, Microsoft, Accenture and internal Avanade stakeholders to drive growth and delivery excellence. Contribute to demand management, capability planning, certification pathways and skills development for the AMBG Security team. Represent AMBG Security in the market as a credible cyber defence thought ...

Network SRE - DC – Network WAN

Location
Greater London, England, United Kingdom
/Nautobot Salt Networking(either one of the following) EVPNSegment routing (although I would accept someone with significant MPLSdepth on their resume) KeyResponsibilities Lead Incident Management: Own and resolvecritical network incidents, manage outages, andprovideexpertguidance during high-pressure situations. Advanced Troubleshooting: Diagnose andresolve complex issues across routing, switching, firewalling … andautomation. Innovation Projects: Collaborate on wireless design and AI clusterdeployments to supportcutting-edgeinitiatives. PreferredSkills Experiencewith InfiniBand and AI cluster deployments . Familiaritywith network asset management systems (e.g.,Nautobot). Wirelessdesign experience with Cisco, Mist, Aruba . #J-18808-Ljbffr ...

Software Engineer III - Backend Engineering - Chase UK

Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you've built through their entire lifecycle, including in production and during incident management Required qualifications, capabilities & skills Formal training or certification on software engineering concepts and applied experience Recent hands‐on professional experience as a back … functions across our network. Your efforts will touch lives all over the financial spectrum and across all our divisions: Global Finance, Corporate Treasury, Risk Management, Human Resources, Compliance, Legal, and within the Corporate Administrative Office. You’ll be part of a team specifically built to meet and exceed ...

Lead Data Engineer â International Consumer Bank (Chase UK)

Location
Greater London, England, United Kingdom
focused on specific banking functions and products, providing opportunities to build data pipelines and reporting capabilities for functional areas such as finance and business management, treasury operations, financial crime prevention, regulatory reporting and analytics. We collaborate with product teams such as card payments, electronic payments, lending, customer onboarding, core … date by continuously updating our technologies and patterns Support the products you've built through their entire lifecycle, including in production and during incident management Required Qualifications, Capabilities & Skills Formal training or certification on data engineering concepts and applied experience Recent hands‐on professional experience as a data ...

Software Engineer III - Back-end Engineer - Chase UK

Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you've built through their entire lifecycle, including in production and during incident management Required qualifications, capabilities & skills Formal training or certification on software engineering concepts and applied experience Recent hands-on professional experience as a back … functions across our network. Your efforts will touch lives all over the financial spectrum and across all our divisions: Global Finance, Corporate Treasury, Risk Management, Human Resources, Compliance, Legal, and within the Corporate Administrative Office. You’ll be part of a team specifically built to meet and exceed ...

Senior Manager - Technology and Operational Risk Reporting

Hiring Organisation
SWIFT
Location
London, UK
Employment Type
Full-time
financial institutions. The Senior Manager – Technology and Operational Risk Reporting is responsible for leading the first technology and operational risk reporting, ensuring senior management and governance committees receive timely, accurate and insightful reporting on the organisation's operational and ICT risk profile. The role supports the effective oversight … operational and technology-related risks through high-quality management information, trend analysis and meaningful risk insights. Working across Technology, Risk Management, Cyber Security, Operational Resilience and business functions, the Senior Manager ensures reporting reflects the organisation's current risk landscape and supports informed strategic decision-making. The role ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
logging tools, including Prometheus, AWS CloudWatch, Grafana, OpenTelemetry, Honeycomb, and ELK. Basic knowledge of Site Reliability Engineering (SRE) and experience with alerting and incident management systems like Opsgenie and PagerDuty. Demonstrated capability to develop and maintain robust and scalable Continuous Integration/Continuous Deployment (CI/CD) pipelines. ...

Connect Direct/Managed File Transfer Engineer

Location
Greater London, England, United Kingdom
environments. Knowledge of SFTP, FTPS, SCP, HTTPS, SSH, and secure data transfer protocols. Experience administering Linux and/or Unix environments. Strong troubleshooting and incident management skills. Understanding of network concepts including Firewalls, DNS, load balancing, and TCP/IP. Experience working within ITIL-based support environments. Excellent ...

Trading Applications Specialist

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, UK
Employment Type
Full-time
with strong technical skills to support systematic trading applications and other critical software. You will step into a high responsibility role, taking ownership of incident management and root cause elimination through process improvements. Trading Applications Specialists drive continuous innovation via automation and optimisation of our daily workflows. ...

Software Engineer III - Back-end Engineer - Chase UK

Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you’ve built through their entire lifecycle, including in production and during incident management Required qualifications, capabilities & skill Formal training or certification on software engineering concepts and applied experience Recent hands‐on professional experience as a back ...

Senior Lead Software Engineer- Backend Engineer - Chase UK

Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you've built through their entire lifecycle, including in production and during incident management Required qualifications, capabilities & skills Formal training or certification on software engineering concepts and applied experience Recent hands‐on professional experience as a back ...

Senior Software Engineer - Pricing

Location
Greater London, England, United Kingdom
concepts quickly Able to be a friendly and approachable team member, capable of mentoring more junior colleagues Desirable Attributes Mastersdegree orhigherinaSTEM subject. Experience with incident management best practices Experience with writing andanalysingperformance tests Experience with automation testing Experience writing and deploying software within cloud-nativeenvironments Experience working with ...

Senior Software Engineer - Pricing

Location
Greater London, England, United Kingdom
friendly and approachable team member, capable of mentoring more junior colleagues Desirable Attributes Masters degree or higher in a STEM subject. Experience with incident management best practices Experience with writing and analysing performance tests Experience with automation testing Experience writing and deploying software within cloud‐native environments Experience ...

Senior Lead Software Engineer- Java (Tech Lead/Staff Level) - Chase UK

Location
Greater London, England, United Kingdom
date by continuously updating our technologies and patterns. Support the products you've built through their entire lifecycle, including in production and during incident management Required qualifications, capabilities & skills Formal training or certification on software engineering concepts and applied experience Recent hands‐on professional experience as a back ...

Chief Technology Officer

Location
Greater London, England, United Kingdom
Confidential - Anonymised Chief Technology Officer Search 1 4. Improve velocity and quality Strengthen architecture, development practices, CI/CD, testing, automation, observability, reliability, security, incident management and technical debt management. Use practical measures such as cycle time, deployment frequency, uptime, defect rates, infrastructure efficiency and developer productivity ...

Salesforce Portal Owner / Developer

Location
Greater London, England, United Kingdom
order - GDPR/UK GDPR and FCA obligations (including Consumer Duty), data‐subject requests, retention and access control - and lead GDPR, security and regulatory incident management (containment, remediation and reporting) with the DPO and Compliance. Build and maintain the portal hands‐on in Experience Cloud - Lightning Web Components ...

Principal Data Engineer

Hiring Organisation
ComplyAdvantage
Location
London, UK
Employment Type
Full-time
treating data quality, observability and data contracts as first-class engineering concerns rather than afterthoughts. Strong working understanding of logging, monitoring, alerting and incident management tooling for data systems. Excellent written and verbal communication. You can produce technical documentation that senior leaders and engineers can act on. Ownership ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
highly autonomous, cross-functional squad with genuine ownership across the full software lifecycle - from architecture and development through to production support and incident management. Required Skills: Strong commercial Node.js and TypeScript experience. Solid AWS and Serverless experience. Experience building distributed, resilient APIs. Experience with MongoDB or another document database. ...

Global Support Engineering Manager — AI-Driven Automation

Location
Greater London, England, United Kingdom
Amazon Web Services (AWS) Operations Management (AWSOM) seeks an experienced Support Engineering Manager to lead a global team transforming from reactive troubleshooting to proactive automation and engineering excellence. You will drive incident management at scale, champion automation-driven solutions, and mentor engineers toward strategic problem-solving over ...

Lead, Production AI Solutions & Incident Response

Location
Greater London, England, United Kingdom
Data Engineering, BTS, and governance to ensure stability, compliance, and measurable business value. The role requires leadership of internal teams and partner resources, strong incident management, and a focus on continuous improvement, governance, and operational excellence within AXIS’ AI capability. #J-18808-Ljbffr ...