1 to 25 of 313 Proactive Monitoring Jobs in the UK

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Fleet: Ensure the rock-solid reliability of our existing RDS footprint. You will architect automated strategies for seamless, multi-version upgrades and proactive performance tuning to minimize downtime across hundreds of instances Architect Cross-Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider … agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming ...

Observability Manager

Hiring Organisation
Appcast
Location
Remote, UK
strategy, standards, and roadmap across infrastructure, applications, and services.Deliver scalable observability solutions using metrics, logs, traces, and events to improve service visibility and performance.Enable proactive monitoring, predictive alerting, and faster incident detection, diagnosis, and resolution.Support Major Incident and Problem Management through real-time insights and evidence-based root … Service Delivery, Engineering, Operations, and third-party partners to maximise operational value.Ensure observability practices align with security, compliance, and regulatory requirements.Qualifications & ExperienceDeep understanding of monitoring frameworks, telemetry, and observability concepts.Hands-on experience with enterprise monitoring tools (e.g., LogicMonitor, ManageEngine, ServiceNow).Proven ability to deliver operational excellence through proactive ...

Observability Manager

Hiring Organisation
Merlin Entertainments
Location
United Kingdom
Salary
£ 55 K
strategy, standards, and roadmap across infrastructure, applications, and services.Deliver scalable observability solutions using metrics, logs, traces, and events to improve service visibility and performance.Enable proactive monitoring, predictive alerting, and faster incident detection, diagnosis, and resolution.Support Major Incident and Problem Management through real-time insights and evidence-based root … Service Delivery, Engineering, Operations, and third-party partners to maximise operational value.Ensure observability practices align with security, compliance, and regulatory requirements.Qualifications & ExperienceDeep understanding of monitoring frameworks, telemetry, and observability concepts.Hands-on experience with enterprise monitoring tools (e.g., LogicMonitor, ManageEngine, ServiceNow).Proven ability to deliver operational excellence through proactive ...

Senior Technical Account Manager, Strategic Industries - Global Financial Services

Hiring Organisation
AmazonWebServices
Location
London, United Kingdom
Salary
£ 70 K
technical guidance to help plan and build solutions using best practices, and proactively keep your customers' AWS environments operationally healthy through application-specific reviews, proactive monitoring, and custom runbooks. You will establish and evolve frameworks governing customer operations on AWS, including DevSecOps practices, Operational Excellence standards, AI strategy … Target Operating Models. You will use generative AI tools fluently to accelerate your own work and, more importantly, to unlock deeper insights and proactive recommendations for your customers. You will guide customers on adopting AWS generative AI services to modernise their operations and drive measurable business outcomes. You will ...

Senior Technical Account Manager, Strategic Industries - Global Financial Services

Location
City of Westminster, England, United Kingdom
technical guidance to help plan and build solutions using best practices, and proactively keep your customers' AWS environments operationally healthy through application-specific reviews, proactive monitoring, and custom runbooks. You will establish and evolve frameworks governing customer operations on AWS, including DevSecOps practices, Operational Excellence standards, AI strategy … Target Operating Models. You will use generative AI tools fluently to accelerate your own work and, more importantly, to unlock deeper insights and proactive recommendations for your customers. You will guide customers on adopting AWS generative AI services to modernise their operations and drive measurable business outcomes. You will ...

Senior Technical Account Manager, Strategic Industries - Global Financial Services

Location
Greater London, England, United Kingdom
technical guidance to help plan and build solutions using best practices, and proactively keep your customers' AWS environments operationally healthy through application-specific reviews, proactive monitoring, and custom runbooks. You will establish and evolve frameworks governing customer operations on AWS, including DevSecOps practices, Operational Excellence standards, AI strategy … Target Operating Models. You will use generative AI tools fluently to accelerate your own work and, more importantly, to unlock deeper insights and proactive recommendations for your customers. You will guide customers on adopting AWS generative AI services to modernise their operations and drive measurable business outcomes. You will ...

3rd Line Build & Support Engineer

Location
Burton Latimer, England, United Kingdom
deployment of secure, reliable Information Technology (IT) systems. Acting as the escalation point for Second Line Technicians, you will provide in‐depth troubleshooting, proactive system monitoring, performance optimisation, and ensure environments are fully prepared and resilient. You will play a key role in maintaining stability while continuously improving … cloud platforms, applications and end-user environments Monitor, maintain and optimise systems to ensure stability, performance and security, including patching, backup management, antivirus and proactive health checks Troubleshoot and manage a broad technology estate, including physical and virtual servers, cloud platforms (such as Microsoft Azure and Amazon Web Services ...

Storage System Administrator - Belfast, UK

Location
Belfast City District, Northern Ireland, United Kingdom
System Administration & Operations Administer and manage Storage/Unix servers with a strong focus on regular operation work. Perform routine system health checks, capacity monitoring, and performance tuning Manage user accounts, permissions, and security configurations Patching & Upgrades Plan and execute OS patching, kernel updates, and security remediation activities Perform … implement permanent fixesAutomation & Optimization Develop and maintain scripts (Shell/Python) using Ansible automation platform of operational tasks Improve system reliability and efficiency through proactive monitoring and tuning Monitoring & Compliance Work with enterprise monitoring tools (e.g., Datadog, Dynatrace or equivalent) Ensure systems adhere to security, compliance ...

Director, Credit Risk

Hiring Organisation
Airwallex
Location
London, United Kingdom
Salary
£ 80 K
management: Set and oversee counterparty risk standards for banking partners, payment partners, merchants, customers, and other material counterparties. Establish appropriate exposure limits, concentration thresholds, monitoring requirements, and escalation processes.Portfolio monitoring and early warning indicators: Lead proactive monitoring of credit, FX, and counterparty portfolios, including exposure, concentration … other relevant risk indicators. Identify emerging trends and escalate material issues before they become systemic.Loss mitigation and remediation: Review and challenge first-line monitoring, collections, collateral, prefunding, reserves, chargeback, and other mitigation strategies. Ensure deteriorating counterparties, customers, merchants, or exposures are identified and addressed promptly.New products and market expansion ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
performance across hundreds of instances. Architect Cross‐Cloud Portability: use CNPG and cloud‐native patterns to keep our database layer provider‐agnostic. Evolve Observability & Monitoring: build proactive monitoring and alerting to detect regressions before they affect customers. Support Replication & Mobility: enable data streaming and zero‐downtime migration ...

Subject Matter Expert (Support&Ops)

Location
Greater London, England, United Kingdom
Monthly/weekly OS and platform patching of Azure VMs and PaaS servicesPatch validation, compliance tracking, and post-patch verification2. Backup & Restore Operations Daily monitoring of Azure Backup jobs (VMs, PaaS-supported workloads)Backup validation, restore testing, and failure remediation3. Monitoring & Alert Management Proactive monitoring using … incidents and escalationsDeep troubleshooting of Azure IaaS, PaaS and DevopsPatch strategy planning and compliance reportingBackup architecture review, restore validation, and DR drillsAdvanced automation and monitoring improvementsProblem management and preventive action implementationLead MSR, QBR, and capacity planning discussionsMentor L2 teams and improve SOPs/KB articles Skills & Experience :7–10+ ...

Director, Credit Risk

Location
Greater London, England, United Kingdom
management: Set and oversee counterparty risk standards for banking partners, payment partners, merchants, customers, and other material counterparties. Establish appropriate exposure limits, concentration thresholds, monitoring requirements, and escalation processes. Portfolio monitoring and early warning indicators: Lead proactive monitoring of credit, FX, and counterparty portfolios, including exposure … other relevant risk indicators. Identify emerging trends and elevate material issues before they become systemic. Loss mitigation and remediation: Review and challenge first‐line monitoring, collections, collateral, prefunding, reserves, chargeback, and other mitigation strategies. Ensure deteriorating counterparties, customers, merchants, or exposures are identified and addressed promptly. New products ...

Senior DevOps Engineer (Fixed Term)

Location
Nottingham, England, United Kingdom
validation, deployment readiness and release governance processes. Creating and maintaining Infrastructure as Code, deployment scripts, reusable templates and configuration standards. Enhancing observability through logging, monitoring, alerting, dashboards and operational telemetry. Supporting service reliability, resilience and operational readiness through proactive monitoring and continuous improvement initiatives. Investigating deployment failures … supporting enterprise applications, APIs, databases and service-oriented architectures. Knowledge of containerisation and orchestration technologies such as Docker, Kubernetes, OpenShift and Helm. Experience implementing monitoring, logging, alerting and observability solutions. A solid understanding of operational resilience, deployment risk management, incident investigation and continuous improvement practices. Experience working across both ...

SAP Basis / BTP Lead (Hybrid)

Hiring Organisation
Agilent Technologies
Location
Polegate, East Sussex, United Kingdom
Salary
£ 100 K
upgrades, patching, and system refresh activities across ECC, S/4HANA, BW, GTS E4H, MES, and BTP environments.Drive system reliability, availability, and performance through proactive monitoring and governance.Administer SAP BTP landscape including subaccount setup, entitlements, user management, security configuration, and service provisioning.Lead SAP technical architecture decisions for hybrid … cloud-native environments (RISE, hyperscalers).Drive adoption of AIOps for SAP operations (predictive monitoring, anomaly detection, auto-remediation).Implement automation for repetitive BASIS operations (system health checks, patching, job management).Leverage AI/ML tools for incident prediction, root cause analysis (RCA), and performance optimization.Introduce intelligent monitoring dashboards ...

Senior DevOps Engineer

Location
Greater London, England, United Kingdom
performing technology team supporting a large-scale eCommerce and digital estate . You’ll take ownership of production and non-production environments, deployment pipelines, monitoring and operational excellence, while working closely with development, IT and third-party partners to deliver secure, resilient and scalable digital services . Key Responsibilities … improve CI/CD pipelines across Azure DevOps and GitHub. Support and optimise production environments, resolving complex incidents and problems. Drive monitoring, observability, performance and reliability improvements. Develop and maintain Infrastructure as Code using Terraform, Pulumi or ARM. Work closely with development teams to ensure applications are designed ...

Senior Reliability Engineer

Hiring Organisation
Fitch Group
Location
Greater London, United Kingdom
Employment Type
Full Time
reliability, security, and efficiency Identify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applications Implement proactive monitoring and observability practices to detect and prevent issues before they impact users Develop and maintain automation and tooling solutions, including AI-assisted approaches ...

DevOps Engineer

Hiring Organisation
Microlise
Location
Nottingham, Nottinghamshire, East Midlands, United Kingdom
Employment Type
Temporary
Salary
£45,000
validation, deployment readiness and release governance processes. Creating and maintaining Infrastructure as Code, deployment scripts, reusable templates and configuration standards. Enhancing observability through logging, monitoring, alerting, dashboards and operational telemetry. Supporting service reliability, resilience and operational readiness through proactive monitoring and continuous improvement initiatives. Investigating deployment failures … supporting enterprise applications, APIs, databases and service-oriented architectures. Knowledge of containerisation and orchestration technologies such as Docker, Kubernetes, OpenShift and Helm. Experience implementing monitoring, logging, alerting and observability solutions. A solid understanding of operational resilience, deployment risk management, incident investigation and continuous improvement practices. Experience working across both ...

Azure Data Support Engineer

Hiring Organisation
Avanade
Location
London, United Kingdom
Salary
£ 80 K
Services environment. This role is responsible for ensuring the stability, performance, and reliability of enterprise-scale data and machine learning workloads on Azure through proactive monitoring, issue investigation, automation, and continuous improvement.The ideal candidate will also play a key role in supporting data warehouse systems by managing ...

SysOps Team Lead

Location
Greater London, England, United Kingdom
platforms. Optimise toolset processes (e.g. n‐Able, Acronis, Mimecast, Intune, Microsoft MDE, BitDefender, Qualys, Meraki, etc.) for efficiency and scalability. Good scripting ability Drive proactive monitoring and self‐healing systems. Systems Administration: Manage servers, Active Directory (Entra), M365 and cloud platforms (Azure/AWS). Ensure system reliability ...

Senior Performance Support Specialist

Hiring Organisation
UNIT4 Group
Location
London, United Kingdom
Salary
£ 80 K
Troubleshoot performance issues across new implementations (Project) and existing implementations (BAU) of Unit4 Financials by Coda, this includes Technical and Application aspects. Develop enhanced, proactive monitoring, detection and remediation of performance issues. Develop proactive, continuous performance tuning and processes to maintain optimal performance levels. Comply with … policies. Team ContextOne of three in the U4F by Coda team, alongside a Senior Performance & Optimisation Consultant and a Database Administrator.Together you deliver proactive performance and technical support before and after go live.Location - this role is fully remote. Occasional customer visits may be required, so candidates should be comfortable ...

Senior Performance Support Specialist

Hiring Organisation
Unit4
Location
Greater London, United Kingdom
Employment Type
Full Time
Troubleshoot performance issues across new implementations (Project) and existing implementations (BAU) of Unit4 Financials by Coda, this includes Technical and Application aspects. Develop enhanced, proactive monitoring, detection and remediation of performance issues. Develop proactive, continuous performance tuning and processes to maintain optimal performance levels. Comply with … Team Context One of three in the U4F by Coda team, alongside a Senior Performance & Optimisation Consultant and a Database Administrator. Together you deliver proactive performance and technical support before and after go live. Location - this role is fully remote. Occasional customer visits may be required, so candidates should ...

Senior Database Specialist

Location
United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Database Specialist

Location
Skipton, England, United Kingdom
support current and future business needs. Leading improvements in database automation, infrastructure as code, deployment pipelines and operational tooling. Driving platform reliability through proactive monitoring, performance optimisation, capacity management and resilience testing. Investigating and resolving complex database incidents, conducting root cause analysis and implementing preventative measures. Establishing … Instance, Cosmos DB or Azure PostgreSQL, or other cloud environments. Familiarity with DevOps practices, CI/CD pipelines and modern engineering methodologies. Experience with monitoring and observability platforms. Knowledge of data platform technologies, analytics services or data engineering practices. What’s In It For You Your work matters. ...

Senior Reliability Engineer

Hiring Organisation
Fitch Ratings
Location
London, United Kingdom
Salary
£ 80 K
application deployments for reliability, security, and efficiencyIdentify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applicationsImplement proactive monitoring and observability practices to detect and prevent issues before they impact usersDevelop and maintain automation and tooling solutions, including AI-assisted approaches ...

Data Engineer

Location
Nottingham, England, United Kingdom
automated deployment pipelines, CI/CD processes and engineering standards using Azure DevOps and Git-based development practices. Drive Data Quality & Governance - Implement controls, monitoring, lineage and governance frameworks that ensure data remains reliable, secure and compliant. Monitor & Optimise Performance - Continuously improve platform performance, observability and operational resilience through … proactive monitoring and alerting. Collaborate Across The Business - Work closely with Data Architects, Analysts and business stakeholders to deliver solutions that create meaningful value for members and colleagues. About You: Microsoft Fabric Expertise - You have strong hands-on experience with Microsoft Fabric technologies including OneLake, Lakehouse, Warehouse, Data ...