6 of 6 Observability Jobs in South London

Senior Data Engineer

Hiring Organisation
FBI &TMT
Location
Kingston Upon Thames, Surrey, South East, United Kingdom
Employment Type
Permanent
Salary
£65,000
optimise batch and streaming workflows for reliability and performance * Contribute to lakehouse architecture and medallion pattern implementation * Implement data quality checks, monitoring and observability across pipelines * Apply platform security, access control and governance standards * Support code reviews and high engineering quality standards * Identify cloud cost optimisation opportunities * Translate business requirements ...

DevOps Team Manager

Hiring Organisation
Bromcom Computers Plc
Location
Bromley, London, United Kingdom
Employment Type
Permanent
technical quality while enabling engineers to own their work. Set and maintain engineering standards for Azure architecture, Azure DevOps, Bicep/ARM, deployment patterns, observability, resilience, security and operational support. Challenge designs and changes using risk, maintainability, failure-mode, rollback and supportability thinking; involve senior engineers and technical leadership where … access follows least-privilege principles, is reviewed regularly and is supported by effective joiner-mover-leaver, break-glass and segregation-of-duties controls. Own observability standards across Azure Monitor and Grafana, security and vulnerability follow-up, and cloud cost and FinOps accountability for the Azure estate. Stakeholder & Cross-Team Influence ...

Senior DevOps Engineer

Hiring Organisation
Bromcom Computers Plc
Location
Bromley, London, United Kingdom
Employment Type
Permanent
performance, and availability Investigate and resolve production incidents in a timely manner Perform root cause analysis and implement preventative measures Enhance logging, alerting, and observability across the platform Security & Governance Implement Azure security best practices and policies Manage identity and access controls in line with governance standards Ensure compliance with … modern development workflows Desirable Knowledge of scripting languages (PowerShell, Bash, or Python) Experience working with Azure Front Door, WAF, or CDN technologies Exposure to observability tooling (distributed tracing, metrics platforms) Experience with cost management and optimisation in Azure Understanding of DevOps principles and Agile delivery practices Personal Attributes Strong problem ...

Senior AI Software Engineer

Hiring Organisation
Futureheads
Location
Surbiton, Surrey, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
Reviewing, governing, and improving AI-generated code and outputs. Building scalable, resilient systems used by large numbers of users. Driving best practices across testing, observability, CI/CD, deployment, and platform reliability. Partnering with product leaders and engineers to solve complex business and technical challenges. Helping accelerate the organisation … sound engineering decisions. Experience working across multiple programming languages and technology stacks. Strong understanding of modern engineering practices including CI/CD, automated testing, observability, and cloud platforms. Excellent communication skills and the ability to explain technical concepts to both technical and non-technical audiences. A collaborative approach and experience ...

SRE & Reliability Lead — AI-Ops & Observability

Location
Carshalton, England, United Kingdom
services used by internal and external customers. You will drive reliability improvements, advance automation and AI-Ops capabilities, and lead a team focused on observability, incident response, operational excellence, and continuous improvement. Responsibilities include translating priorities into clear plans, line managing team leaders, ensuring RCAs and post-mortems are completed ...

Operations and SRE Manager

Location
Carshalton, England, United Kingdom
internal and external customers. You will be responsible for driving reliability improvements, advancing automation and AI-Ops capabilities, and leading a team focused on observability, incident response, operational excellence, and continuous improvement. Responsibilities Lead the implementation of the team’s strategic direction, translating priorities into clear operational plans, backlogs … improvement actions are owned, tracked and completed. Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective. Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support. Protect service levels by ensuring engineering effort is balanced across InfoSec commitments, operational tickets ...