51 to 75 of 290 Performance Monitoring Jobs

Senior AI Engineer| London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
deployment strategy. Experience crafting architectures that encompass data preprocessing, RAG pipelines, agent orchestration, MCP‐based tool and system integration, model integration, guardrails, and performance, cost, and latency optimization. LLMOps, Evaluation & Optimization : Experience operationalizing LLM and agentic applications—building evaluation harnesses and offline/online metrics for quality, groundedness … safety; implementing observability, tracing, and monitoring; continuously optimizing accuracy, cost, and latency. Familiarity with guardrails, red‐team, and responsible deployment of AI systems in production. Communication Skills : Excellent verbal and written communication to engage with clients, articulate technical concepts to non‐technical stakeholders, and collaborate with cross‐functional teams. ...

Azure Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Service Plans, scaling models, host processes and environment lifecycles. Oversee Azure SQL Databases and Managed Instances, including availability, firewalls, private endpoints, Entra ID integration, performance monitoring, backups and recovery readiness. Manage certificates and secrets across Azure workloads using Azure Key Vault, App Service certificates and SSL/… Terraform, Bicep or ARM templates, with structured Resource Groups aligned to Microsoft Cloud Adoption Framework governance. Perform regular health checks, administrative oversight and capacity monitoring for Power Apps environments, Power Automate flows and custom connections integrated with Azure platform services. Establish curated Azure Deployment Environments and self-service infrastructure ...

Lead Platform Operations Manager

Hiring Organisation
United Kingdom Government
Location
Reading, Berkshire, United Kingdom
Salary
£ 70 K
operations. Guide teams in maintaining and enhancing infrastructure and platforms. Risk & Compliance Management: Implement strategies to identify and mitigate potential Data Platform issues proactively. Performance Monitoring & Improvement: Monitor performance metrics and implement data-driven strategies to improve efficiency, scalability and capacity. Drive continuous improvement initiatives that enhance … deliver the Operational part of infrastructure, maintenance and operational needs of the data platform service. Support contract management, service-level agreements, vendor and supplier performance reviews. You will: Lead the Data Operations function in DPS which will provide all the operations governance. Engage with CDIO functions to ensure service ...

Observability SME | 1 year | London, UK (Hybrid - 3 days/week in office)

Hiring Organisation
Hamilton Barnes
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
GBP 450 - 475 Daily
Microsoft Azure, for a leading organisation. The role requires establishing observability standards, ensuring end-to-end visibility across business-critical platforms, and enabling proactive monitoring, faster incident resolution, and improved platform reliability through modern observability practices - with deep expertise in Grafana, OpenTelemetry, distributed tracing, SRE, event-driven architecture … Responsibilities Define and implement enterprise observability strategies, standards, and governance frameworks Design and manage observability solutions covering metrics, logs, traces, and application telemetry Establish monitoring, alerting, and diagnostics best practices across cloud-native platforms and microservices Design and implement distributed tracing solutions using OpenTelemetry and modern observability tools Develop ...

Technical Operations Manager

Hiring Organisation
Equitix Management Services
Location
Bellshill, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent, Work From Home
Principal Responsibilities: The Technical Operations Manager will support the day-to-day management of assets within our portfolio along and provide technical support and performance analysis. The role location is flexible as EMS have offices throughout the UK, however the preferred location for this role is Glasgow. Primary responsibilities … management of the portfolio of solar projects that we manage on behalf of our clients, however the role will also involve data driven performance analysis and technical investigations and discussions with project service providers. The successful candidate will report to the Solar Subsector Lead but will be expected ...

Senior Manager R&D Quality Analytics & Portfolio Insights

Hiring Organisation
BeiGene
Location
United Kingdom
Salary
£ 80 K
develops, implements, and maintains fit-for-purpose analytical methods and independent oversight-support tools tailored to regulated R&D activities. These may include statistical monitoring, predictive risk models, anomaly-detection methods, patient-safety and adverse-event reporting analytics, data-informed audit packages, risk indicators, dashboards, automated monitoring, natural … Based Quality Management, audit and inspection readiness, vendor, process, and system oversight, and targeted risk mitigation.The position integrates and interprets complex clinical, safety, laboratory, monitoring, operational, and quality data to identify study-, site-, subject-, vendor-, process-, product-, and portfolio-level signals that may affect participant protection, patient safety, data ...

Enterprise Architect - AI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
scale. This is a technical hardware-and-software architect role, not a strategy-only position. The successful candidate operates comfortably across GPU infrastructure, high-performance networking, model training and inference pipelines, and the AI risk/governance disciplines increasingly demanded by regulators and enterprise boards. The Enterprise Architect will … Authority (TDA) sessions for AI engagements, owning governance gates, decision records, and design sign-off. Architect AI infrastructure spanning GPU/accelerator compute, high-performance interconnects ,parallel/high-throughput storage, and orchestration Design the AI software stack: training and fine-tuning pipelines, distributed training frameworks, inference/serving ...

Dynatrace/Observability Engineer £528 per day INSIDE

Hiring Organisation
Hays
Location
Telford, Shropshire, West Midlands, United Kingdom
Employment Type
Contract
Contract Rate
Up to £528 per day + £528 per day INSIDE
days min per month Mandatory Certifications: Dynatrace Associate Certification As a Dynatrace/Observability Engineer, you will be responsible for designing, implementing, and supporting monitoring solutions across a range of technologies and platforms, ensuring service stability, performance insight, and proactive incident management. Key Responsibilities: Translate high-level monitoring and non-functional requirements (NFRs) into actionable configurations in Dynatrace. Deliver full-stack observability solutions, including application-aware network performance monitoring (NPM), synthetics, log analytics, and infrastructure metrics. Collaborate with architects and project teams to integrate monitoring into solution designs and test strategies. Maintain and enhance ...

Senior Business Intelligence Officer

Hiring Organisation
LANCASHIRE COUNTY COUNCIL
Location
Preston, Lancashire, United Kingdom
Salary
£ 55 K
develop our understanding of the people who live in Lancashire and insight into Lancashire's economic and cultural landscape. We develop business critical reports, performance analyses and support statutory and regulatory requirements. The following role present an exciting opportunity to join a dynamic team at the heart … datasets. The production of reports and the development of statistical tools within the analysis.Development of new data requirements and reporting arrangements to facilitate improved performance monitoring, management and evaluation (including the use of data dashboards and scorecards).Querying of data and implementing data quality assurances and procedures.Support colleagues ...

Senior Business Intelligence Officer

Hiring Organisation
LANCASHIRE COUNTY COUNCIL
Location
Brighton, East Sussex, United Kingdom
Salary
£ 60 K
develop our understanding of the people who live in Lancashire and insight into Lancashire's economic and cultural landscape. We develop business critical reports, performance analyses and support statutory and regulatory requirements. The following role present an exciting opportunity to join a dynamic team at the heart … datasets. The production of reports and the development of statistical tools within the analysis.Development of new data requirements and reporting arrangements to facilitate improved performance monitoring, management and evaluation (including the use of data dashboards and scorecards).Querying of data and implementing data quality assurances and procedures.Support colleagues ...

Senior AI Engineer

Hiring Organisation
MarkIT Placements
Location
West London, London, United Kingdom
Employment Type
Permanent, Work From Home
gapped environments Build production-ready pipelines from data ingestion through to inference Experience with observability for AI systems, including agent behaviour, model performance, and failure modes Collaborate with engineers, product leads, and customers to translate requirements into working systems Contribute to evaluation frameworks, system integration, and performance tracking … premises, or sovereign cloud Preferred Experience with multimodal reasoning Experience with edge or offline AI deployments Familiarity with Kubernetes (EKS/OpenShift) for monitoring and managing deployed applications MLOps experience - model evaluation, monitoring, reproducibility Observability tooling for agentic systems (model drift, agent behaviour, performance monitoring) Experience ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, UK
Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog, sprint … relevant governance forums. Core Competencies, Knowledge, and Experience Azure Expertise: Strong hands-on experience with Microsoft Azure services, including compute, networking, storage, identity, security, monitoring, and platform services. IaC Proficiency: Proven experience delivering cloud infrastructure using Infrastructure as Code (preferably Terraform and Terragrunt) with a strong understanding of modular ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, England, United Kingdom
Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog, sprint … relevant governance forums. Core Competencies, Knowledge, and Experience Azure Expertise: Strong hands-on experience with Microsoft Azure services, including compute, networking, storage, identity, security, monitoring, and platform services. IaC Proficiency: Proven experience delivering cloud infrastructure using Infrastructure as Code (preferably Terraform and Terragrunt) with a strong understanding of modular ...

Senior Dev Ops Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
This role will focus on the delivery, operation, and continuous optimisation of hybrid infrastructure platforms across both on‐premises and cloud environments. Responsibilities Maintain monitoring and alerting solutions Develop and support automation frameworks Manage CI/CD pipelines Provide technical support to development teams to enable the efficient delivery … maintain Kubernetes clusters on‐prem and in cloud Access and IAM management in Google Cloud, Looker, BigQuery, etc. Database management (MySQL, PostgreSQL, MongoDB) Performance monitoring and fine‐tuning Log collection and analysis Monitoring and alerting setup and management Establish best practices and standards and maintain documentation Build ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
this is the right place. About the Role We are seeking highly skilled and experienced Platform Product Engineers to join our Security, Infrastructure and Performance team. This is a crucial, dual‐faceted role that combines high‐level engineering strategy with hands‐on operational excellence. The successful candidates will … Engineering (SRE) mindset. The successful candidates will be instrumental in defining and upholding Service Level Objectives (SLOs) and Service Level Indicators (SLIs), implementing effective monitoring and alerting strategies, and leading operational incident response processes. What you'll do The core responsibility is to implement, maintain, and continuously improve ...

Enterprise Architect - AI

Hiring Organisation
World Wide Technology
Location
London, United Kingdom
Salary
£ 70 K
scale.This is a technical hardware-and-software architect role, not a strategy-only position. The successful candidate operates comfortably across GPU infrastructure, high-performance networking, model training and inference pipelines, and the AI risk/governance disciplines increasingly demanded by regulators and enterprise boards.The Enterprise Architect will lead technical … Design Authority (TDA) sessions for AI engagements, owning governance gates, decision records, and design sign-off.Architect AI infrastructure spanning GPU/accelerator compute, high-performance interconnects ,parallel/high-throughput storage, and orchestrationDesign the AI software stack: training and fine-tuning pipelines, distributed training frameworks, inference/serving platforms ...

Senior DevOps Engineer

Hiring Organisation
Northrop Grumman
Location
Cheltenham, Gloucestershire, United Kingdom
Salary
£ 60 K
Google CloudPlatforms Experience of designing, deploying and administering Linux or Unix based solutions would be advantageous e.g. using virtualisation, containerisation, Infrastructure as CodeExperience with performance monitoring tools e.g. Elastic Stack, Grafana, CheckMK.Your Benefits:Flexible working schedules - we offer flexible and hybrid working arrangements. Talk … Healthcare, Dental, Life Assurance and Pension. Benefits you can flex include Critical Illness Cover, Health Cash Plan, and Health Assessments.Employee Incentive Programme – exceptional performance is recognised through our annual incentive programme which is awarded to top performers who excelCareer Development – opportunity for ongoing professional development and career growth opportunitiesYour ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
London, United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
Edinburgh, Midlothian, United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
Darlington, County Durham, United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
Birmingham, West Midlands (County), United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

Senior Site Reliability Engineer

Hiring Organisation
United Kingdom Government
Location
Milton Keynes, Buckinghamshire, United Kingdom
Salary
£ 60 K
work as users expect.Main responsibilitiesAs a Senior Site Reliability Engineer you will work to give development teams the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines.You’ll evangelise product teams about … rota for which you will receive an additional allowance.Specific projects the team are working on include rolling out an observability tool to enhance system monitoring and incident response and streamlining deployment processes to reduce downtime and speed up feature delivery.You will be using:Amazon Web Services AzureAWS CodePipelines ...

AI Engineer Defence - 3 months Contract

Hiring Organisation
Jobleads-UK
Location
Oxford, England, United Kingdom
gapped environments Build production-ready pipelines from data ingestion through to inference Experience with observability for AI systems, including agent behaviour, model performance, and failure modes Collaborate with engineers, product leads, and customers to translate requirements into working systems Contribute to evaluation frameworks, system integration, and performance tracking … premises, or sovereign cloud Preferred Experience with multimodal reasoning Experience with edge or offline AI deployments Familiarity with Kubernetes (EKS/OpenShift) for monitoring and managing deployed applications MLOps experience - model evaluation, monitoring, reproducibility Observability tooling for agentic systems (model drift, agent behaviour, performance monitoring) Experience ...