1 to 25 of 154 Remote/Hybrid Permanent Proactive Monitoring Jobs

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Fleet: Ensure the rock-solid reliability of our existing RDS footprint. You will architect automated strategies for seamless, multi-version upgrades and proactive performance tuning to minimize downtime across hundreds of instances Architect Cross-Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider … agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming ...

Observability Manager

Location
Greater London, England, United Kingdom
roadmap across infrastructure, applications, and services. Deliver scalable observability solutions using metrics, logs, traces, and events to improve service visibility and performance. Enable proactive monitoring, predictive alerting, and faster incident detection, diagnosis, and resolution. Support Major Incident and Problem Management through real-time insights and evidence-based root … Operations, and third-party partners to maximise operational value. Ensure observability practices align with security, compliance, and regulatory requirements. Qualifications & Experience Deep understanding of monitoring frameworks, telemetry, and observability concepts. Hands-on experience with enterprise monitoring tools (e.g., LogicMonitor, ManageEngine, ServiceNow). Proven ability to deliver operational excellence ...

Infrastructure Engineer

Location
Bradford, England, United Kingdom
apps and services to provide endpoint management and secure access Providing an SME level escalation point for our Service Desk and App Support teams Proactive monitoring, patching and troubleshooting of endpoints and systems Documentation, knowledge sharing and continuous improvement. Some of the keytechnologies we use: Primarily Microsoft technologies … experience with Microsoft Azure across PaaS and IaaS domains Proficiency in LAN/WAN technologies, switches, firewalls, routing protocols, and network security Experience using monitoring tools for proactive issue resolution. Skilled in troubleshooting complex technical issues and performing root cause analysis. You’ll also be someone who enjoys ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
performance across hundreds of instances. Architect Cross‐Cloud Portability: use CNPG and cloud‐native patterns to keep our database layer provider‐agnostic. Evolve Observability & Monitoring: build proactive monitoring and alerting to detect regressions before they affect customers. Support Replication & Mobility: enable data streaming and zero‐downtime migration ...

Infrastructure Engineer

Hiring Organisation
Signet Jewelers
Location
Watford, Hertfordshire, South East, United Kingdom
Employment Type
Permanent
Supporting and administering our UK infrastructure across servers, cloud platforms and networking Investigating and resolving complex 3rd line technical issues Designing and implementing proactive monitoring and alerting to prevent future incidents Managing Windows Server environments, Active Directory, Group Policy and core infrastructure services Supporting our AWS and Azure … Active Directory & Group Policy Office 365 AWS (ECS, Amazon MQ, RDS, S3, Workspaces, VPCs) Microsoft Azure PowerShell Meraki networking SQL Server SolarWinds (or similar monitoring platforms) Remote Access technologies With around 150 Windows servers, multiple AWS environments and Azure services, there's plenty of opportunity to broaden your technical ...

Senior Performance Support Specialist

Hiring Organisation
Unit4
Location
Greater London, United Kingdom
Employment Type
Full Time
Troubleshoot performance issues across new implementations (Project) and existing implementations (BAU) of Unit4 Financials by Coda, this includes Technical and Application aspects. Develop enhanced, proactive monitoring, detection and remediation of performance issues. Develop proactive, continuous performance tuning and processes to maintain optimal performance levels. Comply with … Team Context One of three in the U4F by Coda team, alongside a Senior Performance & Optimisation Consultant and a Database Administrator. Together you deliver proactive performance and technical support before and after go live. Location - this role is fully remote. Occasional customer visits may be required, so candidates should ...

Senior Database Specialist

Location
Skipton, England, United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Database Specialist

Location
Skipton, England, United Kingdom
engineering practices to support current and future business needs.Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling.Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing.Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures.Establishing and maintaining … Managed Instance, Cosmos DB or Azure PostgreSQL, or other cloud environmentsFamiliarity with DevOps practices, CI/CD pipelines and modern engineering methodologies.Experience with monitoring and observability platforms.Knowledge of data platform technologies, analytics services or data engineering practices.What’s In It For YouYour work matters.And the way we reward ...

Principal Service Desk Analyst

Location
Leeds, England, United Kingdom
Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve the resolution (both reactive and proactive) of problems throughout the information system lifecycle, including classification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future … incidents.They also process and coordinate appropriate and timely responses to incident reports, including the channelling of requests for help to appropriate functions for resolution, monitoring resolution activity, and keeping customers appraised of progress towards service restoration. ThePrincipal Service Desk Analystis an escalation point for the Service Desk Teamand ...

Platform Engineering Manager (Cloud Foundations)

Location
Greater London, England, United Kingdom
psychologically safeenvironment;with leadership tailored to individual strengths and motivations. Platform Engineering, Operations&Reliability Cloudplatform and keyservices arereliable,incident detection and resolutionaresmooth due to proactive monitoring, well‐maintainedalerts/logs, and complete observability coverage. Platform resilience and DR planning/testing strategyisdefined and operational,working across the business … network and key components aremaintainedfor audit, security, knowledgesharing purposes. Whatyou’llbring Proven experience in cloud operations and platform engineeringmanagement, overseeing cloud infrastructureand platforms, monitoring,reliability, and service delivery in production environments. Previoushands‐on experience running large cloud‐based website environmentswith GKE, service mesh, load balancers, CDN/WAF, Kafka ...

Platform Engineering Manager (Cloud Foundations) London, UK

Location
Greater London, England, United Kingdom
resilience mechanisms are regularly tested, ensuring the platform can recover predictably. The platform is reliable, incident detection and resolution are smooth due to proactive monitoring, well‐maintained alerts/logs, and complete observability coverage. Platform resilience and DR planning/testing strategy is defined and operational, working across … leadership to individual strengths and motivations. What you’ll bring Proven experience in cloud operations and platform engineering management, overseeing cloud infrastructure and platforms, monitoring, reliability, and service delivery in production environments. Hands‐on experience running large cloud‐based website environments with GKE, service mesh, load balancers, CDN/ ...

DATA ENGINEER

Hiring Organisation
Talent Sure Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£40,000 - £80,000 per annum
analytics platform. Data Integration: Integrate large, diverse datasets from various sources while maintaining exceptionally high quality and governance standards. Platform Orchestration: Maintain the orchestration, proactive monitoring, and general performance of data platform components. Engineering Excellence: Champion high coding and data practice standards, helping to improve engineering processes across … experience with Docker and Kubernetes, or exposure to production-level Generative AI, is highly beneficial. Collaboration: Strong communication and collaboration skills, coupled with a proactive, inquisitive attitude. Salary, Benefits & Culture This organisation has been officially recognised as a leading UK employer, offering an inclusive, supportive, and modern environment where ...

Senior Workday Integrations Product Engineer

Location
Hook, England, United Kingdom
strong software engineering principles including version control, peer reviews, automated testing, CI/CD, and structured release management. Build integrations with observability, operational resilience, proactive monitoring, alerting, logging, automated recovery, and self‐healing error handling by design. Troubleshoot complex production issues, conduct root‐cause analysis, and continuously improve … Practices: Experience with cloud platforms (e.g., GCP or Microsoft Azure), CI/CD pipelines, and automated deployment practices. Operational Tooling & Observability: Strong background in monitoring, observability, automated operational tooling, and IT service management platforms such as ServiceNow. Modernization & Analytics: Proven experience modernizing legacy integrations, reducing technical debt, and leveraging ...

Senior DevOps Engineer

Location
United Kingdom
infrastructure and application security controls. Support the organisation's ongoing compliance and certification requirements. Reliability & SRE Establish and maintain observability across distributed systems. Develop proactive monitoring, alerting and performance-tuning strategies. Help maintain service-level objectives and platform availability. Investigate and resolve infrastructure and application incidents. Coordinate emergency ...

Senior DevOps Engineer

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Permanent
infrastructure and application security controls. Support the organisation's ongoing compliance and certification requirements. Reliability & SRE Establish and maintain observability across distributed systems. Develop proactive monitoring, alerting and performance-tuning strategies. Help maintain service-level objectives and platform availability. Investigate and resolve infrastructure and application incidents. Coordinate emergency ...

IT Manager

Location
Lincoln, England, United Kingdom
Title: IT Manager Location: Hybrid (Lincoln) Contract: Full-time, Permanent Role Overview We areseekingan experienced and proactive IT Manager to take ownership of our internal IT infrastructure, security, and operational systems. This roleis responsible forensuring the stability, security, and efficiency of core IT services, while supporting business growth … proactively mitigate risk. Business Continuity,Operational Reliability Disaster Recovery:Monitorbackup systems, conduct regular restore testing, and support annual external audits and DR testing initiatives. Proactive Monitoring: Manage system uptime, antivirus/EDR, and SIEM tools to quicklyidentifyand resolve performance anomalies. Network Admin: Support Azure VPN connectivity and applied ...

Principal Service Desk Analyst

Location
Manchester, England, United Kingdom
Alten Group in early 2022. The Principal Service Desk Analyst, like theSenior and Service Desk Analyst, is there to achieve theresolution (both reactive and proactive) of problems throughout the information system lifecycle, includingclassification, prioritisation and initiation of action, documentation of root causes and implementation of remediesto prevent future incidents.They … alsoprocess and coordinateappropriate and timely responses to incident reports, including the channelling ofrequests for help to appropriate functions for resolution, monitoring resolution activity, and keepingcustomersappraised of progress towards service restoration. ThePrincipal Service Desk Analystis anescalation pointfor the Service Desk Teamand isrequired todeliveradvanced problem resolution,working with internaland/ ...

Virtual Technology Subject Matter Expert (SME)

Location
Sheffield, England, United Kingdom
Cloud PC administration, including planning, provisioning, migration, and users support. Support and maintain OpenShift/VMWare based virtual machine workloads, ensuring scalability and resilience. Proactive Monitoring & End User Experience Proactively monitor end-user experience (DEX), session latency, and logon performance using analytics tools and automated remediation scripts. Manage ...

Data Engineer

Hiring Organisation
Connells Group HQ
Location
Milton Keynes, Buckinghamshire, United Kingdom
Employment Type
Full-Time
Salary
£45,000 - £55,000 per annum
ensure organisation-wide consistency. Apply Agile principles for iterative and collaborative development. Ensure data pipeline quality, reliability and performance. Develop, test and implement monitoring to ensure effective operation. Support cross-functional data projects, providing expertise as required. Deliver data solutions that support project goals and business outcomes. Proactively monitor ...

Senior DevOps Engineer

Hiring Organisation
Centrica - CHP
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent, Work From Home
Build and enhance CI/CD pipelines, working closely with engineering teams to enable faster, safer and more efficient software delivery. Automate infrastructure, deployments, monitoring and operational processes to improve reliability and reduce manual effort. Leverage AWS services such as EC2, S3 and Lambda, alongside Kubernetes, to deliver modern … cloud-native solutions. Ensure application availability, performance and resilience through proactive monitoring, incident management and disaster recovery planning. Create clear technical documentation and provide guidance to development teams, helping drive best practices and continuous improvement. What's in it for you? Enjoy a generous market salary, along with ...

Senior DevOps Engineer

Location
Leeds, England, United Kingdom
Build and enhance CI/CD pipelines, working closely with engineering teams to enable faster, safer and more efficient software delivery. Automate infrastructure, deployments, monitoring and operational processes to improve reliability and reduce manual effort. Leverage AWS services such as EC2, S3 and Lambda, alongside Kubernetes, to deliver modern … cloud-native solutions. Ensure application availability, performance and resilience through proactive monitoring, incident management and disaster recovery planning. Create clear technical documentation and provide guidance to development teams, helping drive best practices and continuous improvement. What's in it for you? Enjoy a generous market salary, along with ...

Site Reliability Engineer (SRE) - Glasgow, UK

Location
Glasgow, Scotland, United Kingdom
DevOps expertise and moderate development experience will also be considered* This position sits at the intersection of Operations Engineering and Cloud Infrastructure requiring a proactive individual who can automate processes improve system reliability troubleshoot production issues and contribute to application enhancements when required**Your Skills:**Site Reliability Operational Support … businesscritical applications and cloud infrastructure* Ensure high system availability performance scalability and reliability* Participate in incident management root cause analysis and problem resolution* Implement proactive monitoring alerting and observability solutions* Reduce operational overhead through automation and selfhealing mechanisms* Support production releases and deployment activitiesAWS Cloud Engineering* Design deploy ...

Security Analyst: 2nd Line

Location
Newcastle upon Tyne, England, United Kingdom
work. Within the role you will be responsible for performing the day-to-day maintenance of the Security Operations Centre. These responsibilities will include proactive monitoring of customer’s security posture as well as reactive actions to control a breach should this occur. Typical tasks will include triage … Duties And Responsibilities Taking the lead on investigating and responding to security incidents, carrying out forensic analysis and driving issues through to resolution. Proactively monitoring customer environments, threat hunting, and identifying risks before they become problems. Conducting vulnerability assessments and helping customers strengthen their security posture through effective remediation. ...

Integrations Engineer (Fixed Term) based in Derby

Location
East Midlands, England, United Kingdom
tools). Develop and optimise ETL/ELT processes (extraction, transformation, loading) to meet business, reporting, and analytics requirements. Implement data quality, validation, and monitoring controls to ensure accurate, reliable, and timely data flows. Troubleshoot, resolve, and prevent integration incidents; continuously improve performance, scalability, and resiliency. Collaborate with stakeholders … cloud services while managing connectivity, security and operational constraints. Significant experience of owning the support, troubleshooting and optimisation of live data pipelines , including proactive monitoring, root‐cause analysis, incident resolution, performance tuning and the prevention of recurring issues. Strong experience of leading technical engagement with stakeholders and third ...

z/OS Mainframe Storage Engineer

Location
Greater London, England, United Kingdom
/OS Mainframe Storage Engineer to manage, support, and optimize enterprise mainframe storage infrastructure. This role is responsible for the installation, configuration, administration, maintenance, monitoring, and modernization of storage platforms supporting IBM z/OS environments. The successful candidate will oversee all aspects of storage management, including disk … recovery, disaster recovery, storage capacity planning, replication technologies, and storage software lifecycle management. The role also involves infrastructure design, hardware upgrades, automation initiatives, and proactive monitoring to ensure maximum performance, availability, and resilience of business-critical systems. The ideal candidate will possess deep expertise in IBM mainframe storage ...