1 to 25 of 110 Remote/Hybrid Proactive Monitoring Jobs in the UK

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Fleet: Ensure the rock-solid reliability of our existing RDS footprint. You will architect automated strategies for seamless, multi-version upgrades and proactive performance tuning to minimize downtime across hundreds of instances Architect Cross-Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider … agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming ...

Observability Manager

Location
Greater London, England, United Kingdom
roadmap across infrastructure, applications, and services. Deliver scalable observability solutions using metrics, logs, traces, and events to improve service visibility and performance. Enable proactive monitoring, predictive alerting, and faster incident detection, diagnosis, and resolution. Support Major Incident and Problem Management through real-time insights and evidence-based root … Operations, and third-party partners to maximise operational value. Ensure observability practices align with security, compliance, and regulatory requirements. Qualifications & Experience Deep understanding of monitoring frameworks, telemetry, and observability concepts. Hands-on experience with enterprise monitoring tools (e.g., LogicMonitor, ManageEngine, ServiceNow). Proven ability to deliver operational excellence ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
performance across hundreds of instances. Architect Cross‐Cloud Portability: use CNPG and cloud‐native patterns to keep our database layer provider‐agnostic. Evolve Observability & Monitoring: build proactive monitoring and alerting to detect regressions before they affect customers. Support Replication & Mobility: enable data streaming and zero‐downtime migration ...

SAP Basis / BTP Lead (Hybrid)

Hiring Organisation
Agilent Technologies
Location
Polegate, East Sussex, UK
Employment Type
Full-time
patching, and system refresh activities across ECC, S/4HANA, BW, GTS E4H, MES, and BTP environments. Drive system reliability, availability, and performance through proactive monitoring and governance. Administer SAP BTP landscape including subaccount setup, entitlements, user management, security configuration, and service provisioning. Lead SAP technical architecture decisions … hybrid and cloud-native environments (RISE, hyperscalers).Drive adoption of AIOps for SAP operations (predictive monitoring, anomaly detection, auto-remediation).Implement automation for repetitive BASIS operations (system health checks, patching, job management).Leverage AI/ML tools for incident prediction, root cause analysis (RCA), and performance optimization. Introduce intelligent ...

Senior Database Specialist

Location
Skipton, England, United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Database Specialist

Location
United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Reliability Engineer

Hiring Organisation
Fitch Ratings
Location
London, UK
Employment Type
Full-time
application deployments for reliability, security, and efficiencyIdentify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applicationsImplement proactive monitoring and observability practices to detect and prevent issues before they impact usersDevelop and maintain automation and tooling solutions, including AI-assisted approaches ...

Senior Performance Support Specialist

Hiring Organisation
Unit4
Location
Greater London, United Kingdom
Employment Type
Full Time
Troubleshoot performance issues across new implementations (Project) and existing implementations (BAU) of Unit4 Financials by Coda, this includes Technical and Application aspects. Develop enhanced, proactive monitoring, detection and remediation of performance issues. Develop proactive, continuous performance tuning and processes to maintain optimal performance levels. Comply with … Team Context One of three in the U4F by Coda team, alongside a Senior Performance & Optimisation Consultant and a Database Administrator. Together you deliver proactive performance and technical support before and after go live. Location - this role is fully remote. Occasional customer visits may be required, so candidates should ...

Data Engineer

Location
Nottingham, England, United Kingdom
automated deployment pipelines, CI/CD processes and engineering standards using Azure DevOps and Git-based development practices. Drive Data Quality & Governance - Implement controls, monitoring, lineage and governance frameworks that ensure data remains reliable, secure and compliant. Monitor & Optimise Performance - Continuously improve platform performance, observability and operational resilience through … proactive monitoring and alerting. Collaborate Across The Business - Work closely with Data Architects, Analysts and business stakeholders to deliver solutions that create meaningful value for members and colleagues. About You: Microsoft Fabric Expertise - You have strong hands-on experience with Microsoft Fabric technologies including OneLake, Lakehouse, Warehouse, Data ...

Principal Service Desk Analyst

Location
Leeds, England, United Kingdom
Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve the resolution (both reactive and proactive) of problems throughout the information system lifecycle, including classification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future … incidents.They also process and coordinate appropriate and timely responses to incident reports, including the channelling of requests for help to appropriate functions for resolution, monitoring resolution activity, and keeping customers appraised of progress towards service restoration. ThePrincipal Service Desk Analystis an escalation point for the Service Desk Teamand ...

Senior Workday Integrations Product Engineer

Location
Hook, England, United Kingdom
strong software engineering principles including version control, peer reviews, automated testing, CI/CD, and structured release management. Build integrations with observability, operational resilience, proactive monitoring, alerting, logging, automated recovery, and self‐healing error handling by design. Troubleshoot complex production issues, conduct root‐cause analysis, and continuously improve … Practices: Experience with cloud platforms (e.g., GCP or Microsoft Azure), CI/CD pipelines, and automated deployment practices. Operational Tooling & Observability: Strong background in monitoring, observability, automated operational tooling, and IT service management platforms such as ServiceNow. Modernization & Analytics: Proven experience modernizing legacy integrations, reducing technical debt, and leveraging ...

IT Manager

Location
Lincoln, England, United Kingdom
Title: IT Manager Location: Hybrid (Lincoln) Contract: Full-time, Permanent Role Overview We areseekingan experienced and proactive IT Manager to take ownership of our internal IT infrastructure, security, and operational systems. This roleis responsible forensuring the stability, security, and efficiency of core IT services, while supporting business growth … proactively mitigate risk. Business Continuity,Operational Reliability Disaster Recovery:Monitorbackup systems, conduct regular restore testing, and support annual external audits and DR testing initiatives. Proactive Monitoring: Manage system uptime, antivirus/EDR, and SIEM tools to quicklyidentifyand resolve performance anomalies. Network Admin: Support Azure VPN connectivity and applied ...

Principal Service Desk Analyst

Location
Manchester, England, United Kingdom
Alten Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve theresolution (both reactive and proactive) of problems throughout the information system lifecycle, includingclassification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future incidents.They … alsoprocess and coordinateappropriate and timely responses to incident reports, including the channelling ofrequests for help to appropriate functions for resolution, monitoring resolution activity, and keepingcustomersappraised of progress towards service restoration. ThePrincipal Service Desk Analystis anescalation pointfor the Service Desk Teamand isrequired todeliveradvanced problem resolution,working with internaland/ ...

Platform Engineering Manager (Cloud Foundations)

Location
Greater London, England, United Kingdom
psychologically safeenvironment;with leadership tailored to individual strengths and motivations. Platform Engineering, Operations&Reliability Cloudplatform and keyservices arereliable,incident detection and resolutionaresmooth due to proactive monitoring, well‐maintainedalerts/logs, and complete observability coverage. Platform resilience and DR planning/testing strategyisdefined and operational,working across the business … network and key components aremaintainedfor audit, security, knowledgesharing purposes. Whatyou’llbring Proven experience in cloud operations and platform engineeringmanagement, overseeing cloud infrastructureand platforms, monitoring,reliability, and service delivery in production environments. Previoushands‐on experience running large cloud‐based website environmentswith GKE, service mesh, load balancers, CDN/WAF, Kafka ...

Platform Engineering Manager (Cloud Foundations)

Hiring Organisation
Rightmove
Location
London, UK
Employment Type
Full-time
individual strengths and motivations. Platform Engineering, Operations & Reliability Cloud platform and key services are reliable, incident detection and resolution are smooth due to proactive monitoring, well‐maintained alerts/logs, and complete observability coverage. Platform resilience and DR planning/testing strategy is defined and operational, working across … audit, security, knowledge sharing purposes. What you'll bring Proven experience in cloud operations and platform engineering management, overseeing cloud infrastructure and platforms, monitoring, reliability, and service delivery in production environments. Previous hands‐on experience running large cloud-based website environments with GKE, service mesh, load balancers, CDN/ ...

Platform Engineering Manager (Cloud Foundations) London, UK

Location
Greater London, England, United Kingdom
resilience mechanisms are regularly tested, ensuring the platform can recover predictably. The platform is reliable, incident detection and resolution are smooth due to proactive monitoring, well‐maintained alerts/logs, and complete observability coverage. Platform resilience and DR planning/testing strategy is defined and operational, working across … leadership to individual strengths and motivations. What you’ll bring Proven experience in cloud operations and platform engineering management, overseeing cloud infrastructure and platforms, monitoring, reliability, and service delivery in production environments. Hands‐on experience running large cloud‐based website environments with GKE, service mesh, load balancers, CDN/ ...

Virtual Technology Subject Matter Expert (SME)

Location
Sheffield, England, United Kingdom
Cloud PC administration, including planning, provisioning, migration, and users support. Support and maintain OpenShift/VMWare based virtual machine workloads, ensuring scalability and resilience. Proactive Monitoring & End User Experience Proactively monitor end-user experience (DEX), session latency, and logon performance using analytics tools and automated remediation scripts. Manage ...

Senior DevOps Engineer

Location
United Kingdom
infrastructure and application security controls. Support the organisation's ongoing compliance and certification requirements. Reliability & SRE Establish and maintain observability across distributed systems. Develop proactive monitoring, alerting and performance-tuning strategies. Help maintain service-level objectives and platform availability. Investigate and resolve infrastructure and application incidents. Coordinate emergency ...

Senior DevOps Engineer

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Permanent
infrastructure and application security controls. Support the organisation's ongoing compliance and certification requirements. Reliability & SRE Establish and maintain observability across distributed systems. Develop proactive monitoring, alerting and performance-tuning strategies. Help maintain service-level objectives and platform availability. Investigate and resolve infrastructure and application incidents. Coordinate emergency ...

Infrastructure Engineer

Location
Birmingham, England, United Kingdom
support and improve a business-critical UK infrastructure spanning on-premises servers, cloud platforms and networking. You will resolve complex third-line issues, strengthen monitoring and automation, and help migrate services to the cloud while maintaining high security, compliance and governance standards. Key Responsibilities Administer Windows Server, Active Directory … Investigate and resolve complex third-line incidents across servers, cloud and networking. Support AWS and Azure platforms and contribute to cloud migration projects. Implement proactive monitoring, alerting and automation to improve resilience and efficiency. Work with internal teams and technology partners to deliver secure, compliant infrastructure improvements. Maintain ...

Infrastructure Engineer

Location
Tyseley, England, United Kingdom
support and improve a business-critical UK infrastructure spanning on-premises servers, cloud platforms and networking. You will resolve complex third-line issues, strengthen monitoring and automation, and help migrate services to the cloud while maintaining high security, compliance and governance standards. Key Responsibilities Administer Windows Server, Active Directory … Investigate and resolve complex third-line incidents across servers, cloud and networking. Support AWS and Azure platforms and contribute to cloud migration projects. Implement proactive monitoring, alerting and automation to improve resilience and efficiency. Work with internal teams and technology partners to deliver secure, compliant infrastructure improvements. Maintain ...

Focus Principal Engineer

Location
Manchester, England, United Kingdom
flourish. Resilient and comfortable prioritising in demanding situations. Highly trustworthy and able to operate with integrity and discretion at all times. Energetic and proactive, will enjoy motivating others with your “can do, will do” attitude. Able to operate with minimal brief, and a fast‐moving set of changing priorities. … taking high level solution and enterprise architecture artefacts and translating them into workable designs and work packages. You’ll drive forward and own the proactive monitoring of the products and services within your Value Stream. You’ll participate in Enterprise planning events to support the delivery and planning ...

Integrations Engineer (Fixed Term) based in Derby

Location
East Midlands, England, United Kingdom
tools). Develop and optimise ETL/ELT processes (extraction, transformation, loading) to meet business, reporting, and analytics requirements. Implement data quality, validation, and monitoring controls to ensure accurate, reliable, and timely data flows. Troubleshoot, resolve, and prevent integration incidents; continuously improve performance, scalability, and resiliency. Collaborate with stakeholders … cloud services while managing connectivity, security and operational constraints. Significant experience of owning the support, troubleshooting and optimisation of live data pipelines , including proactive monitoring, root‐cause analysis, incident resolution, performance tuning and the prevention of recurring issues. Strong experience of leading technical engagement with stakeholders and third ...

Server & Cloud Engineer

Hiring Organisation
Data Careers
Location
Lincolnshire, East Midlands, United Kingdom
Employment Type
Contract, Work From Home
Virtualisation technologies including VMware and VxRail. Microsoft Server technologies, storage and related infrastructure services. Third-line infrastructure support within a complex technical environment. Infrastructure monitoring, performance management and service availability. Desirable Experience Experience working within a secure public sector, policing or government environment. Knowledge of ITIL service management principles. ...

BI Engineer, SRE (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
primarily building features — and that's not sustainable. We're looking for an SRE to own that space: service health, incident response, infrastructure monitoring, and making sure we're not blindly burning cloud budget. The BI Engineer, SRE will ensure the availability, performance, and security of the Business Intelligence … team's GCP-hosted APIs and data infrastructure. This role is responsible for proactive monitoring, incident response, and continuous improvement of platform reliability across a cloud-native stack. The engineer will work closely with backend and data engineers to maintain service health and drive operational excellence. This position ...