76 to 100 of 103 Observability Jobs in the South East

NOC Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Basingstoke, Hampshire, United Kingdom
Employment Type
Permanent
Salary
£60000/annum Bonus, Pension, Healthcare
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

NOC Engineer, AWS

Hiring Organisation
Spectrum It Recruitment Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

Azure Platform Engineer

Hiring Organisation
EMBS Engineering
Location
Newbury, West Berkshire, Berkshire, United Kingdom
Employment Type
Permanent
Salary
£65000 - £75000/annum + Benefits
Azure platform templates and engineering patterns to support consistent platform adoption. Build and improve CI/CD pipelines, deployment automation and release processes. Implement observability, monitoring, logging, resilience and operational readiness across the platform. Embed FinOps principles, improving cloud cost visibility and optimisation. Work closely with engineering teams throughout sprint … engineering patterns. Strong understanding of DevOps and DevSecOps practices including CI/CD, source control, automated testing and release management. Experience with Azure observability, monitoring, alerting, logging and platform reliability. Practical knowledge of cloud security, governance and enterprise engineering standards. Experience applying FinOps principles including tagging strategies, right-sizing ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Stansted Mountfitchet, Essex, UK
Employment Type
Full-time
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, ... LFWQ1_UKTJ ...

Performance and Monitoring Engineer

Hiring Organisation
17918
Location
Stansted, Essex, United Kingdom
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, ... WKCL1_UKTJ ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Stansted, Essex, United Kingdom
Employment Type
Permanent
Salary
GBP 50,000 Annual
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, click apply for full ...

Data Services Sales Leader – Remote-First, EMEA/APJ

Hiring Organisation
Jobleads-UK
Location
Windsor, England, United Kingdom
Manager for our Data Services portfolio to lead a high-performing team of sales specialists across EMEA and APJ, driving the ARR target for Observability and Cyber Resilience solutions in large enterprise accounts. You will develop and execute a comprehensive sales strategy, oversee pipeline reviews and revenue forecasting, coach ...

Junior Azure Engineer

Hiring Organisation
COMPUTACENTER (UK) LIMITED
Location
South East London, London, United Kingdom
Employment Type
Permanent
containers Implement security and governance controls (RBAC, Azure Policy, Management Groups) Build and support landing zones and foundational cloud environments Manage monitoring and observability using Azure Monitor, Log Analytics, and alerts Support CI/CD pipelines using Azure DevOps or GitHub Actions Automate operational tasks using PowerShell, Azure … experience with Infrastructure-as-Code (Bicep, ARM, or Terraform) Solid understanding of Azure networking and cloud architecture principles Experience with monitoring, logging, and observability tools Ability to troubleshoot and resolve complex cloud issues Experience with automation and scripting (PowerShell, Azure CLI) Strong collaboration and communication skills Desirable Experience with Azure ...

Azure Platform Engineering Consultant

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
platform templates and landing zone patterns. CI/CD & Automation: Build and refine automated deployment pipelines, environment management, and release practices. Platform Quality: Embed observability (monitoring, logging, alerting), resilience, security, and FinOps principles directly into platform assets. Co-Delivery & Knowledge Transfer: Work closely alongside client engineering teams to pair, document … Core compute, networking, storage, identity, security, and platform services. Infrastructure as Code: Strong proficiency with Terraform AND Terragrunt using modular, reusable implementation patterns. DevOps & Observability: Strong experience with CI/CD tools (Azure DevOps/GitHub Actions) and monitoring stacks (Prometheus, Grafana, Azure Monitor, etc.). FinOps: Practical knowledge ...

Principal Cloud Architect

Hiring Organisation
TXP
Location
Southampton, Hampshire, South East, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £600 per day
delivery teams. The successful candidate will provide manager-level technical leadership across DevOps, cloud platforms, Infrastructure as Code, CI/CD, networking, security, observability and reliability engineering. They will help shape enterprise-scale transformation, hybrid cloud strategy and platform services aligned to the Azure Well-Architected Framework, ensuring solutions … compute/storage design. Evaluate platform changes including major provider upgrades (AzureRM/Cloudflare), DR and high availability improvements, cost optimisation strategies, and observability frameworks. Lead technical designs for large-scale refactoring and provider upgrades, environment creation pipelines, secure container registry access, identity integration and Zero Trust patterns, and event ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT Recruitment
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Permanent
Salary
£75,000
experience with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code/Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways … Engineer Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure servers and networks, and automate application lifecycles. Regularly use Datadog and other observability tools for application performance monitoring. Implement new ways of working, helping to shape how the organisation responds and recovers to incidents. Take ownership of incident ...

SRE Technical Lead

Hiring Organisation
Adecco
Location
Reading, Berkshire, United Kingdom
Employment Type
Permanent
Salary
GBP 70,000 - 90,000 Annual
remediation Act as the technical escalation point for major incidents and high-risk releases Lead blameless post-incident reviews and ensure continuous improvement Establish observability and capacity management practices using modern tooling Identify and eliminate systemic reliability risks and operational inefficiencies Collaborate with engineering, platform, security, and operations teams across … Experience working in multi-cloud or hybrid cloud environments Strong understanding of SRE principles (SLOs, SLAs, error budgets, reliability engineering) Hands-on experience with observability tooling (eg, Prometheus, Grafana, OpenTelemetry, Loki, Tempo) Strong knowledge of Infrastructure as Code and GitOps (eg, Helm, Kustomize, ArgoCD, Tekton) Experience with CI/ ...

Site Reliability Engineer

Hiring Organisation
Connells Group HQ
Location
Milton Keynes, Buckinghamshire, England, United Kingdom
Employment Type
Full-Time
Salary
£40,000 - £55,000 per annum
hands-on role in ensuring it is reliable, scalable, and observable. You will help establish and mature SRE practices, focusing on: Monitoring and observability Incident response Post-incident review Reliability testing and capacity planning Toil reduction Enabling development velocity We offer a hybrid working arrangement with one day per week … Build dashboards, alerts, and runbooks to improve visibility Automate repetitive tasks to reduce operational toil Collaborate with cross-functional teams to enhance reliability and observability Support performance testing and capacity planning Proactively identify and prioritise reliability improvements Experience & Skills Required: Hands-on experience with Azure Monitoring (Application Insights, Alerts, Action ...

Software Engineering Manager - Tooling and Optimisations

Hiring Organisation
Jobleads-UK
Location
Windsor, England, United Kingdom
practice, reduce duplication, and support maintainable, secure and high-performing systems. Improve delivery capability through platform reliability and DevOps maturity Continuously strengthen deployment pipelines, observability, alerting, incident response, recovery procedures and operational readiness across Field Ops engineering teams. Manage stakeholders and maintain clear communication Build trusted relationships across product, operations … data modelling and data quality controls. Ability to produce both high‐level and detailed design specifications. Experience leading DevOps practices, including CI/CD, observability, monitoring and incident management. Demonstrated capability leading multi‐squad engineering delivery in a product‐led organisation. Mindset & Ways of Working Comfortable working in iterative, outcome ...

SRE Technical Lead

Hiring Organisation
Capgemini
Location
Surrey, United Kingdom
Employment Type
Full Time
point for major incidents and high risk releases, protecting service stability and ensuring blameless post incident reviews lead to measurable improvement. • Define and govern observability and capacity practices so reliability risks are visible, actionable, and proactively managed. • Ensure SRE practices align with service governance, security, and compliance requirements, and contribute … including: • Strong expertise in Kubernetes and OpenShift. • Experience with multi cloud and hybrid architectures, including service mesh (e.g. Istio). • Hands on experience with observability platforms such as Prometheus, Grafana, Loki, Tempo, and OpenTelemetry. • Strong Infrastructure as Code and GitOps experience (Helm, Kustomize, ArgoCD, Tekton). • Experience with CI/ ...

SRE Technical Lead

Hiring Organisation
83zero Limited
Location
Wokingham, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
senior technical escalation point for major incidents and high-risk releases. Lead blameless post-incident reviews and ensure measurable service improvements. Define and oversee observability, monitoring and capacity management practices. Ensure SRE approaches align with security, governance and compliance requirements. Mentor and coach senior engineers, helping to improve SRE maturity … OpenShift. Experience designing and supporting hybrid and multi-cloud platforms. Experience with service mesh technologies such as Istio. Strong hands-on experience with observability tooling including Prometheus, Grafana, Loki, Tempo and OpenTelemetry. Infrastructure as Code and GitOps expertise using tools such as Helm, Kustomize, ArgoCD and Tekton. Experience building ...

Cloud Platform Engineer — Observability & Orchestration

Hiring Organisation
Jobleads-UK
Location
Reading, England, United Kingdom
ECMWF in Reading, UK, is seeking a Cloud Platform Engineer (A2) to help deliver platform engineering and observability for Destination Earth Data Bridges. You will join the Platform Engineering Team within the Application Delivery Section, building container orchestration, automation, and monitoring capabilities to ensure reliable deployments across Data Bridges. Candidates ...

Principal Software Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate and Banstead, England, United Kingdom
into production systems. ARTIFICIAL INTELLIGENCE AI Engineering Strategy: Pioneer AI integration patterns, publish on AI‐augmented development, and lead AI in engineering. AI Evaluation & Observability: Shape AI observability and evaluation practices, pioneer novel evaluation methodologies, contribute to industry standards for LLM evaluation, and lead AI quality assurance at scale. ...

Lead Data Engineer

Hiring Organisation
Tenth Revolution Group
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £110,000 per annum
Python SQL Data Warehousing Data Lakes Data Pipelines APIs & Data Services Machine Learning Infrastructure AI & Machine Learning Systems Analytics Engineering Data Governance Data Observability Cloud-Native Architecture You'll Be Responsible For Designing and building scalable cloud-based data platforms. Creating robust ETL/ELT pipelines and data services. Developing … APIs and trusted data products for enterprise clients. Supporting machine learning and AI initiatives through high-quality data architecture. Establishing governance, lineage, quality and observability standards across the data estate. Working closely with Product, Engineering, Data Science and Leadership teams to influence strategic decisions. Helping define the long-term data ...

Vice President, Build — Data, Engineering & AI

Hiring Organisation
Jobleads-UK
Location
Reigate and Banstead, England, United Kingdom
ready criteria through the Data Architect; run the Collibra dictionary, master & reference data operations and the governance council. Own data quality and observability: shift‐left checks, lineage and monitoring built in, not bolted on. Run data BAU — incidents, refreshes and access — baselined before any cost reduction is taken. Own data … access framework. Success measures Prioritized use cases delivered to production with named owner, evidence pack, run model and sunset criteria. Data quality, lineage and observability coverage across priority data products. Time from funded demand to production. Production reliability, incident rate and support performance. AI evaluation coverage, red‐team completion ...

AI Engineer - Contract

Hiring Organisation
Jobleads-UK
Location
Oxford, England, United Kingdom
execution Deploy AI systems into cloud, on-premises, and air-gapped environments Build production-ready pipelines from data ingestion through to inference Experience with observability for AI systems, including agent behaviour, model performance, and failure modes Collaborate with engineers, product leads, and customers to translate requirements into working systems Contribute … with edge or offline AI deployments Familiarity with Kubernetes (EKS/OpenShift) for monitoring and managing deployed applications MLOps experience - model evaluation, monitoring, reproducibility Observability tooling for agentic systems (model drift, agent behaviour, performance monitoring) Experience with agent orchestration patterns and inter‐agent communication protocols (e.g.A2A) Familiarity with MCPs ...

Head of AI Engineering

Hiring Organisation
Jobleads-UK
Location
Bexhill-on-Sea, England, United Kingdom
As a member of the CIO Technology Engineering senior leadership team, the Head of AI Engineering will lead the design, deployment, integration and continuous improvement of enterprise AI and machine learning capabilities, including ML platforms ...

Staff Engineer

Hiring Organisation
Stepstone UK
Location
South East London, London, United Kingdom
Employment Type
Permanent
engineering standards, best practices and reusable patterns while partnering with Enterprise Architecture and influencing technical direction Drive engineering excellence by improving code quality, testing, observability, reliability and operational practices Support end-to-end delivery by guiding teams through complex technical challenges, improving decision-making, and contributing to planning and risk … data lakes/lakehouse architectures, Iceberg or similar table formats, as well as batch and streaming processing Knowledge of data quality, governance, cataloguing and observability tools (e.g. Datadog), with DBT or AI-assisted engineering practices as a plus Additional Information Your benefits Werea community here that cares as much about ...

Senior Director, Technology Operations

Hiring Organisation
Jobleads-UK
Location
Guildford, England, United Kingdom
telephony lead and specialist team holding the hands‐on work. DevOps and developer experience (run side): CI/CD reliability, environments, deployment, and observability; currently delivered by the outsourced partner, to be shaped and, over time, selectively insourced. Run‐side Service Delivery, including telephony provisioning and change. Run‐side management … hours/on‐call model for the live service, acting as the senior escalation point. Replace firefighting with proactive reliability: root‐cause discipline, observability, and measurable reductions in downtime and change‐failure rate. Platform, telephony, and estate Own the evolution of the on‐prem and cloud estate against a quarterly ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Stansted, Essex, South East, United Kingdom
Employment Type
Permanent
Salary
£50,000
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, modern technology estate. Responsibilities … improve speed, accuracy and consistency Supporting major changes, deployments and post-incident reviews with data-driven evidence Qualifications Strong experience with monitoring and observability tools (LogicMonitor, Azure Monitor, App Insights, Log Analytics, Defender for Cloud) Excellent understanding of cloud performance, IaaS/PaaS, networking fundamentals, API performance and capacity modelling ...