1 to 25 of 287 Proactive Monitoring Jobs in the UK

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Fleet: Ensure the rock-solid reliability of our existing RDS footprint. You will architect automated strategies for seamless, multi-version upgrades and proactive performance tuning to minimize downtime across hundreds of instances Architect Cross-Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider … agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming ...

Observability Manager

Location
Greater London, England, United Kingdom
roadmap across infrastructure, applications, and services. Deliver scalable observability solutions using metrics, logs, traces, and events to improve service visibility and performance. Enable proactive monitoring, predictive alerting, and faster incident detection, diagnosis, and resolution. Support Major Incident and Problem Management through real-time insights and evidence-based root … Operations, and third-party partners to maximise operational value. Ensure observability practices align with security, compliance, and regulatory requirements. Qualifications & Experience Deep understanding of monitoring frameworks, telemetry, and observability concepts. Hands-on experience with enterprise monitoring tools (e.g., LogicMonitor, ManageEngine, ServiceNow). Proven ability to deliver operational excellence ...

Senior Technical Account Manager, Strategic Industries - Global Financial Services

Location
Greater London, England, United Kingdom
technical guidance to help plan and build solutions using best practices, and proactively keep your customers' AWS environments operationally healthy through application-specific reviews, proactive monitoring, and custom runbooks. You will establish and evolve frameworks governing customer operations on AWS, including DevSecOps practices, Operational Excellence standards, AI strategy … Target Operating Models. You will use generative AI tools fluently to accelerate your own work and, more importantly, to unlock deeper insights and proactive recommendations for your customers. You will guide customers on adopting AWS generative AI services to modernise their operations and drive measurable business outcomes. You will ...

3rd Line Build & Support Engineer

Location
Burton Latimer, England, United Kingdom
deployment of secure, reliable Information Technology (IT) systems. Acting as the escalation point for Second Line Technicians, you will provide in‐depth troubleshooting, proactive system monitoring, performance optimisation, and ensure environments are fully prepared and resilient. You will play a key role in maintaining stability while continuously improving … cloud platforms, applications and end-user environments Monitor, maintain and optimise systems to ensure stability, performance and security, including patching, backup management, antivirus and proactive health checks Troubleshoot and manage a broad technology estate, including physical and virtual servers, cloud platforms (such as Microsoft Azure and Amazon Web Services ...

Data Platform Engineer

Hiring Organisation
Connells Limited
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Permanent
Code Work within the Cloud Platform design pattern to implement technical and financial observability Undertake platform capacity management Support projects where appropriate Undertake proactive monitoring and react to escalations from other IT teams Collaborating with other members of the team. Experience and Skills Required: Essential: Demonstrable experience … strong attention to detail? Desirable: Experience of GitHub, Git Actions, Terraform, Platform as Code and Zero Trust architectures Experience of advanced tools for operational monitoring Ability to operate and influence at all levels within the organisation Experience of tools for Alerting and Monitoring Cloud Cost Monitoring ...

Data Platform Engineer

Hiring Organisation
Connells Group HQ
Location
Milton Keynes, Buckinghamshire, United Kingdom
Employment Type
Full-Time
Salary
£50,000 - £65,000 per annum
Code Work within the Cloud Platform design pattern to implement technical and financial observability Undertake platform capacity management Support projects where appropriate Undertake proactive monitoring and react to escalations from other IT teams Collaborating with other members of the team. Experience and Skills Required: Essential: Demonstrable experience … strong attention to detail Desirable: Experience of GitHub, Git Actions, Terraform, Platform as Code and Zero Trust architectures Experience of advanced tools for operational monitoring Ability to operate and influence at all levels within the organisation Experience of tools for Alerting and Monitoring Cloud Cost Monitoring ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
performance across hundreds of instances. Architect Cross‐Cloud Portability: use CNPG and cloud‐native patterns to keep our database layer provider‐agnostic. Evolve Observability & Monitoring: build proactive monitoring and alerting to detect regressions before they affect customers. Support Replication & Mobility: enable data streaming and zero‐downtime migration ...

Subject Matter Expert (Support&Ops)

Location
Greater London, England, United Kingdom
Monthly/weekly OS and platform patching of Azure VMs and PaaS servicesPatch validation, compliance tracking, and post-patch verification2. Backup & Restore Operations Daily monitoring of Azure Backup jobs (VMs, PaaS-supported workloads)Backup validation, restore testing, and failure remediation3. Monitoring & Alert Management Proactive monitoring using … incidents and escalationsDeep troubleshooting of Azure IaaS, PaaS and DevopsPatch strategy planning and compliance reportingBackup architecture review, restore validation, and DR drillsAdvanced automation and monitoring improvementsProblem management and preventive action implementationLead MSR, QBR, and capacity planning discussionsMentor L2 teams and improve SOPs/KB articles Skills & Experience :7–10+ ...

Director, Credit Risk

Hiring Organisation
Airwallex
Location
London, UK
Employment Type
Full-time
management: Set and oversee counterparty risk standards for banking partners, payment partners, merchants, customers, and other material counterparties. Establish appropriate exposure limits, concentration thresholds, monitoring requirements, and escalation processes. Portfolio monitoring and early warning indicators: Lead proactive monitoring of credit, FX, and counterparty portfolios, including exposure … other relevant risk indicators. Identify emerging trends and escalate material issues before they become systemic. Loss mitigation and remediation: Review and challenge first-line monitoring, collections, collateral, prefunding, reserves, chargeback, and other mitigation strategies. Ensure deteriorating counterparties, customers, merchants, or exposures are identified and addressed promptly. New products ...

SAP Basis / BTP Lead (Hybrid)

Hiring Organisation
Agilent Technologies
Location
Polegate, East Sussex, UK
Employment Type
Full-time
patching, and system refresh activities across ECC, S/4HANA, BW, GTS E4H, MES, and BTP environments. Drive system reliability, availability, and performance through proactive monitoring and governance. Administer SAP BTP landscape including subaccount setup, entitlements, user management, security configuration, and service provisioning. Lead SAP technical architecture decisions … hybrid and cloud-native environments (RISE, hyperscalers).Drive adoption of AIOps for SAP operations (predictive monitoring, anomaly detection, auto-remediation).Implement automation for repetitive BASIS operations (system health checks, patching, job management).Leverage AI/ML tools for incident prediction, root cause analysis (RCA), and performance optimization. Introduce intelligent ...

SysOps Team Lead

Location
Greater London, England, United Kingdom
platforms. Optimise toolset processes (e.g. n‐Able, Acronis, Mimecast, Intune, Microsoft MDE, BitDefender, Qualys, Meraki, etc.) for efficiency and scalability. Good scripting ability Drive proactive monitoring and self‐healing systems. Systems Administration: Manage servers, Active Directory (Entra), M365 and cloud platforms (Azure/AWS). Ensure system reliability ...

Senior Database Specialist

Location
Skipton, England, United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Database Specialist

Location
United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Azure Data Support Engineer

Hiring Organisation
Avanade
Location
London, UK
Employment Type
Full-time
Services environment. This role is responsible for ensuring the stability, performance, and reliability of enterprise-scale data and machine learning workloads on Azure through proactive monitoring, issue investigation, automation, and continuous improvement. The ideal candidate will also play a key role in supporting data warehouse systems by managing ...

Senior Reliability Engineer

Hiring Organisation
Fitch Ratings
Location
London, UK
Employment Type
Full-time
application deployments for reliability, security, and efficiencyIdentify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applicationsImplement proactive monitoring and observability practices to detect and prevent issues before they impact usersDevelop and maintain automation and tooling solutions, including AI-assisted approaches ...

Senior Performance Support Specialist

Hiring Organisation
Unit4
Location
Greater London, United Kingdom
Employment Type
Full Time
Troubleshoot performance issues across new implementations (Project) and existing implementations (BAU) of Unit4 Financials by Coda, this includes Technical and Application aspects. Develop enhanced, proactive monitoring, detection and remediation of performance issues. Develop proactive, continuous performance tuning and processes to maintain optimal performance levels. Comply with … Team Context One of three in the U4F by Coda team, alongside a Senior Performance & Optimisation Consultant and a Database Administrator. Together you deliver proactive performance and technical support before and after go live. Location - this role is fully remote. Occasional customer visits may be required, so candidates should ...

Dynamics 365 F&O Platform Engineer

Location
Greater London, England, United Kingdom
technical bridge between the D365 functional/business teams and the underlying Azure infrastructure. You will be responsible for environment management, deployment governance, platform monitoring, release management, and integration health across F&O and connected systems, ensuring the platform is stable, secure, resilient, and ready to support the business. … versioning, and release processes Monitor and troubleshoot F&O performance issues, including batch jobs, SQL/Azure SQL performance, and overall environment health Develop proactive monitoring, alerting, and operational documentation, driving continuous improvement through automation Manage data integrations between F&O and downstream systems (e.g. Shopify Plus, Patchworks ...

Senior Infrastructure Operations Engineer

Location
Bracknell, England, United Kingdom
solutions using automation and infrastructure-as-code. Continuously improve platform reliability and streamline operational workflows through automation and tooling. Security & Observability: Develop and maintain proactive monitoring, logging, and alerting solutions. Contribute to both offensive and defensive security strategies, including threat simulations and incident playbooks. Operational Excellence: Provide ...

Senior Infrastructure Engineer

Hiring Organisation
Hackajob Ltd
Location
Leicester, Leicestershire, East Midlands, United Kingdom
Employment Type
Permanent
Salary
£80,000
load balancing, and backup/recovery solutions such as Rubrik. Strong Automation & Scripting Skills: Hands-on experience with Terraform, Azure DevOps, PowerShell, and enterprise monitoring tools such as Dynatrace. Exceptional Problem-Solving Abilities: An analytical mindset with a talent for troubleshooting complex technical issues and identifying root causes efficiently. ...

Network Engineer

Hiring Organisation
Third Nexus Group Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Contract
Contract Rate
£375 - £400/annum
Reliability Engineering (SRE) practices. Key Responsibilities Review the existing network data landscape, including current data sources, data quality, and data flows between network management, monitoring, automation, and IT service management platforms. Assess the current NetBox implementation and identify opportunities to improve data modelling, governance, discovery, assurance, and integrations. Develop … manual activity. Define and develop a roadmap towards Network SRE practices. Propose service reliability measures, including relevant KPIs, SLIs and SLOs. Identify opportunities for proactive monitoring, fault prevention and self-healing automation. Work with Operations, Security, Architecture, Platform and ITSM teams to improve resilience and reliability. Required Technical ...

Data Engineer

Location
Nottingham, England, United Kingdom
automated deployment pipelines, CI/CD processes and engineering standards using Azure DevOps and Git-based development practices. Drive Data Quality & Governance - Implement controls, monitoring, lineage and governance frameworks that ensure data remains reliable, secure and compliant. Monitor & Optimise Performance - Continuously improve platform performance, observability and operational resilience through … proactive monitoring and alerting. Collaborate Across The Business - Work closely with Data Architects, Analysts and business stakeholders to deliver solutions that create meaningful value for members and colleagues. About You: Microsoft Fabric Expertise - You have strong hands-on experience with Microsoft Fabric technologies including OneLake, Lakehouse, Warehouse, Data ...

Principal Service Desk Analyst

Location
Leeds, England, United Kingdom
Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve the resolution (both reactive and proactive) of problems throughout the information system lifecycle, including classification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future … incidents.They also process and coordinate appropriate and timely responses to incident reports, including the channelling of requests for help to appropriate functions for resolution, monitoring resolution activity, and keeping customers appraised of progress towards service restoration. ThePrincipal Service Desk Analystis an escalation point for the Service Desk Teamand ...

Principal Service Desk Analyst

Location
Manchester, England, United Kingdom
Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve the resolution (both reactive and proactive) of problems throughout the information system lifecycle, including classification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future … incidents.They also process and coordinate appropriate and timely responses to incident reports, including the channelling of requests for help to appropriate functions for resolution, monitoring resolution activity, and keeping customers appraised of progress towards service restoration. ThePrincipal Service Desk Analystis an escalation point for the Service Desk Teamand ...

Senior Network Engineer

Hiring Organisation
Infoplus Technologies UK Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Full-Time
Salary
£450.00 - £500.00 per day
Ansible, REST APIs and, where applicable, Terraform. Develop reusable automation workflows, operational runbooks and documentation. Perform network device baselining and configuration validation. Improve network monitoring, data quality and operational reliability. Define KPIs, SLIs and SLOs to support Network SRE practices. Identify opportunities for proactive monitoring, fault prevention … MPLS VLANs WAN/LAN SD-WAN Network Security Fundamentals Network Automation Python Ansible REST APIs Terraform – desirable PowerShell – desirable Git/GitHub – desirable Monitoring & Observability Experience with platforms such as: SolarWinds Auvik IP Fabric LogicMonitor Dynatrace Azure Monitor Grafana Preferred Certifications Cisco CCNP/CCIE Juniper JNCIS/ ...

Infrastructure Engineer

Location
Bracknell, England, United Kingdom
customers across the finance, blue light and government sectors. Responsibilities include CI/CD pipeline management, infrastructure as code, containerisation and orchestration, database technologies, monitoring and observability, and security and compliance. What we are looking for We're seeking a motivated infrastructure professional with at least 2 years … support for operational issues and platform reliability Automate workflows and processes to improve efficiency Contribute to infrastructure-as-code and DevOps pipelines Develop proactive monitoring strategies and support security best practices Participate in incident response, threat simulation and operational runbooks Collaborate with development teams to resolve technical issues ...