101 to 125 of 143 Problem Management Jobs in London

Defect Manager

Location
Croydon, England, United Kingdom
team are responsible to scope and develop efficiencies and improvements throughout their own processes and aid development in wider teams. Main responsibilities The management and control of all defects within the in-house development environments. To own and chair daily defect calls with lead Development, Test and Operational Teams … Customer and third parties and to provide clear minutes and actions. Definition, ownership and management of the Defect Process. Liaising with the Problem Management Team to ensure a clear and efficient interaction within overlapping processes, and to promote this with all customers. Ensuring all defects are entered ...

HALO ITSM Configuration Specialist

Hiring Organisation
SSA Digital Recruitment
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£425 - £500/day
Specialist Rate: £500 per day Contract Length: 6 Months IR35 Status: Outside IR35 Location: Fully Remote A leading organisation undergoing a significant IT Service Management transformation is seeking an experienced HALO ITSM Configuration Specialist to join their programme team on an initial 6 month contract. This is a fully … hands on experience configuring and administering HALO ITSM • Proven experience delivering ITSM implementations or service transformation projects • Strong understanding of Incident, Request, Change and Problem Management processes • Experience configuring workflows, service catalogues, forms, approvals and SLAs • Excellent stakeholder management and communication skills • Ability to translate business requirements ...

Senior DevOps Engineer | London, Hybrid | up to £125k

Location
Greater London, England, United Kingdom
/AKS). Improve observability with monitoring, logging and alerting (e.g. Prometheus, Grafana, ELK/EFK, CloudWatch). Embed security best practice: IAM, secrets management, patching, vulnerability remediation and secure configuration. Support incident response and problem management, participating in on‐call rotations and post-incident reviews. Collaborate … . Solid understanding of SDLC, Git workflows, and automated testing practices. Experience implementing and maintaining compliance‐aligned controls (e.g. audit trails, least privilege, change management). Ability to troubleshoot complex production issues and improve system reliability through engineering. Nice to Have Experience with service mesh, API gateways, or platform ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
/CD pipelines to enable safe, frequent and low-risk delivery of changes. Oversee live service reliability, supporting teams through incident and problem management while encouraging a learning-focused, blameless culture. Ensure security, resilience and compliance considerations are understood and embedded into engineering practices. Essential skills … Senior SRE Squad Lead include: Experience leading, supporting and developing engineers, including line management or strong mentoring experience. Strong communication skills, with the ability to explain technical concepts clearly and build effective relationships with a range of stakeholders. Experience of working with cloud platforms such as AWS, Azure ...

Platform Engineer

Location
Greater London, England, United Kingdom
model access, retrieval tooling, and evaluation workflows, so AI capabilities can move from prototype to production using standard platform patterns. CI/CD, Release Management & Test Automation Implement and maintain pipelines-as-code (e.g., GitHub Actions/Azure DevOps YAML) to established patterns, keeping builds, tests, and deployments reliable … fast. Execute and support release activities, following defined GitOps and change‐management processes. Write and maintain automated tests and quality gates within pipelines. Implement evaluation‐based quality gates for AI systems, including evals‐as‐code, regression suites against golden datasets, and human‐review thresholds where required. Infrastructure as Code ...

Cloud Engineer – Fintech

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 70 K
security is designed and implemented into all core IT technologies • Comply with Company polices and procedure processes such as change control and release management • Support the major incident process and problem management to ensure root causes of issues are identified and that effective mitigations are applied • Provide … expertise in system administration of Linux and Microsoft technologies, including the patching, securing and monitoring of such systems. • Understanding of CI/CD pipeline management • Detailed knowledge of networking infrastructure, communication protocols and security is desirable • Demonstrable expertise in web delivered application topologies, including IIS and related technologiesThe environment ...

2nd Line Support

Hiring Organisation
INTEC SELECT LIMITED
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£250.00 - £300.00 per day
Exchange Online, Intune, Teams Execute SOX controls and maintain audit-ready evidence Support audit cycles, walkthroughs and control testing Root cause analysis, incident/problem management, onboarding/offboarding Skills & Experience 3+ years in Service Desk/Technical Support Hands-on M365 admin experience (Entra ID, AD, Intune … Exposure to audit/compliance work - SOX-regulated environments a plus ITIL, CompTIA, or Microsoft AZ-900/AZ-104 certs desirable MacOS endpoint management (ABM) , end device configuration Get in touch or apply below ...

Client Operations - Technical Assistance Associate I

Location
Greater London, England, United Kingdom
teams and clients to ensure a prompt and effective resolution to bugs and issues. The ideal candidate will be analytical with an affinity for problem‐solving and troubleshooting technical and software issues, able to recognize, investigate, and escalat[e] client‐reported issues related to our platforms. Responsibilities Providing support … Interactive Brokers’ platforms Desktop applications (Windows, macOS, and Linux) Mobile applications (Android and iOS) Troubleshooting and support for Interactive Brokers’ web‐based offerings Problem management with a focus on wide‐scale technical issues Requirements Languages: fluency in English and German or French is a must. Any other European ...

End User Support Specialist

Location
Greater London, England, United Kingdom
continuous improvement and service stability. Operational standards are defined by the Client Support Operations Lead. Endpoint platform design and standards remain with EUC. Line management accountability sits with the Client Support Team Lead. Core Responsibilities Area Description Support Execution and Case Ownership Deliver incidents and service requests escalated from … ticket discipline. Understanding of joiner, mover, leaver processes, access governance, and controlled execution of user changes. Good understanding of ITIL-aligned incident, request, and problem management practices. Strong communication, stakeholder handling, and troubleshooting skills in business-facing environments. Ability to work in a process-led operating model with ...

Azure Platform Engineer

Location
City Of London, England, United Kingdom
explain technical issues and trade‐offs to both technical and non‐technical audiences. Nice to have Experience supporting production platforms (monitoring, incident response, problem management, change control). Exposure to hybrid connectivity and on‐premises environments (Active Directory, DNS, firewalls, proxies). Exposure to migration projects (server, database ...

Trading Application Support

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
looking for a Trading Application Engineerto join our high profile client. Our client is and independent research business and a well-known Algo Order Management and Execution Business. They operate as an independent strategic partner to investment businesses by offering insight and guidance through their powerful network of intellectual … links, messaging middleware etc.).Identification and mitigation of risks to service delivery and proactive resolution of issues before they impact service levels. Incident and problem management, including root cause analysis and follow up. Coordination of all relevant parties during client facing outages. Timely escalation of complex issues ...

Service Systems Engineer

Hiring Organisation
ARM
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
Provide advanced troubleshooting and support for live CAD systems in mission-critical public safety environments. Investigate and resolve complex CAD application issues affecting incident management, mobilisation, mapping, geospatial services and external interfaces. Lead technical investigations during major incidents and act as an escalation point for operational issues. Analyse application … messaging and third-party systems. Work directly with customers during incidents, service reviews and technical workshops. Perform root cause analysis (RCA) and contribute to problem management activities. Develop and implement changes while ensuring technical governance and service continuity. Provide technical leadership and recommendations during service-impacting incidents ...

Service Systems Engineer

Hiring Organisation
ARM
Location
Twickenham, London, St. Margarets and North Twickenham, United Kingdom
Employment Type
Permanent
Provide advanced troubleshooting and support for live CAD systems in mission-critical public safety environments. Investigate and resolve complex CAD application issues affecting incident management, mobilisation, mapping, geospatial services and external interfaces. Lead technical investigations during major incidents and act as an escalation point for operational issues. Analyse application … messaging and third-party systems. Work directly with customers during incidents, service reviews and technical workshops. Perform root cause analysis (RCA) and contribute to problem management activities. Develop and implement changes while ensuring technical governance and service continuity. Provide technical leadership and recommendations during service-impacting incidents ...

Lead Data Product Manager

Hiring Organisation
Shell
Location
London, United Kingdom
Salary
£ 70 K
resilience, recoverability and cyber hygiene; meet controls for data privacy, confidentiality and regulatory obligations in trading (e.g., REMIT/EMIR, privacy).Drive incident, change & problem management; measure and report service health and user satisfaction.What you bring Deep commodities data knowledge, ideally within energy trading, including ETRM/Endur … backlogs and adoption outcomes that engineering teams can execute.Strong product and technical judgement across modern data platforms, ingestion patterns, APIs, streaming/batch, schema management, CDC, metadata, lineage and data quality, with the ability to assess what good looks like.Experience leading cross-functional product delivery through Product Owners, architects ...

IT Support Manager

Location
Greater London, England, United Kingdom
complex hardware, network, and operational incidents Monitor and drive SLA, KPI, and incident resolution performance across IT support services Oversee proactive maintenance, troubleshooting, and problem management activities Collaborate with infrastructure engineers, vendors, and colocation partners to resolve issues and deliver improvements Maintain and improve support processes, documentation … issues, including GPU‐based systems, fiber networks. Solid understanding of data center operations, server platforms, and enterprise networking principles. Practical experience with IT service management processes (ITIL/ITSM). Basic proficiency with Linux/Unix operating systems and command‐line tools. Proven ability to manage operational KPIs, incidents ...

Lead Software Engineer - Proxy/SSE Network Security

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
operational stability of software applications and systems Own the US perimeter, proxy, and SSE/SASE engineering roadmap and execution, including intake, prioritization, dependency management, delivery governance, and stakeholder alignment across cybersecurity, network services, operations, and application and platform teams. Define and operationalize standards, reference architectures, and reusable engineering … risk reduction and resilience outcomes. Drive operational excellence at scale for perimeter, proxy, and SSE services in the US, including incident, change, and problem management rigor, observability and resiliency validation practices, automation to improve repeatability and evidence quality, reduction of client and partner impact, and execution of Technology ...

Lead Software Engineer - Proxy/SSE Network Security

Location
London, United Kingdom
operational stability of software applications and systems Own the US perimeter, proxy, and SSE/SASE engineering roadmap and execution, including intake, prioritization, dependency management, delivery governance, and stakeholder alignment across cybersecurity, network services, operations, and application and platform teams. Define and operationalize standards, reference architectures, and reusable engineering … risk reduction and resilience outcomes. Drive operational excellence at scale for perimeter, proxy, and SSE services in the US, including incident, change, and problem management rigor, observability and resiliency validation practices, automation to improve repeatability and evidence quality, reduction of client and partner impact, and execution of Technology ...

Data Engineer

Location
Greater London, England, United Kingdom
backfills, late-arriving data and rerun-safe pipelines Experience with data observability, monitoring and troubleshooting Strong system and data-flow discovery skills Excellent stakeholder management and communication skills Data governance, ownership and stewardship frameworks Compute infrastructure supporting modern data platforms Scheduling and orchestration concepts Data platform security, access controls … auditability Incident, change and problem management/ITIL environments Experience working within large, matrixed enterprise organisations #J-18808-Ljbffr ...

Infrastructure Engineer (BMS) - DV Cleared

Location
Greater London, England, United Kingdom
operation of critical infrastructure within the telecommunications sector. This is a 12-month contract opportunity based in Ipswich or London, working across complex Building Management Systems (BMS), Central Monitoring Systems (CMS), CCTV and wider infrastructure platforms. Key Responsibilities: Architecting and delivering next-generation BMS, CMS and CCTV systems Driving … operations teams to deliver end-to-end services Enhancing platform security, resilience, availability and performance at scale Supporting best practice in incident, change and problem management Creating clear, reusable runbooks, engineering standards and operational documentation Identifying opportunities to improve infrastructure delivery and introduce new technologies Job Requirements: Experience ...

Lead Infrastructure Engineer - Proxy/SSE Network Security

Location
City Of London, England, United Kingdom
operational stability of software applications and systems Own the US perimeter, proxy, and SSE/SASE engineering roadmap and execution, including intake, prioritization, dependency management, delivery governance, and stakeholder alignment across cybersecurity, network services, operations, and application and platform teams. Define and operationalize standards, reference architectures, and reusable engineering … risk reduction and resilience outcomes. Drive operational excellence at scale for perimeter, proxy, and SSE services in the US, including incident, change, and problem management rigor, observability and resiliency validation practices, automation to improve repeatability and evidence quality, reduction of client and partner impact, and execution of Technology ...

Lead Software Engineer - FIXED INCOME UK

Location
Greater London, England, United Kingdom
services industry and their IT systems Practical cloud native experience Experience with observability tooling and practices (structured logging, metrics, distributed tracing) and incident/problem management. Work Style/Ways of Working Strong communication skills, ownership mindset, and ability to collaborate across product, engineering, and operations. Commitment to inclusive ...

Lead Software Engineer - FIXED INCOME UK

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
financial services industry and their IT systemsPractical cloud native experienceExperience with observability tooling and practices (structured logging, metrics, distributed tracing) and incident/problem management. Work Style/Ways of WorkingStrong communication skills, ownership mindset, and ability to collaborate across product, engineering, and operations. Commitment to inclusive teamwork ...

Lead Software Engineer - FIXED INCOME UK

Location
Greater London, England, United Kingdom
services industry and their IT systems Practical cloud native experience Experience with observability tooling and practices (structured logging, metrics, distributed tracing) and incident/problem management. Work Style/Ways of Workin g Strong communication skills, ownership mindset, and ability to collaborate across product, engineering, and operations. Commitment ...

Onsite Senior Data Engineer - London | AWS & Azure, SAP

Location
Greater London, England, United Kingdom
driven thinking across IT and business stakeholders and mentor junior engineers. You will lead the rollout of Data Foundation initiatives, coordinate change, incident and problem management, and present insights to stakeholders to improve decision making. #J-18808-Ljbffr ...

Senior Azure L3: Incident Resolution & Automation

Location
Greater London, England, United Kingdom
will own escalations, coordinate with stakeholders, and drive platform stability across Azure IaaS, PaaS, and DevOps environments. The role emphasizes automation, governance, and proactive problem management to minimize outages and improve service quality for global clients. #J-18808-Ljbffr ...