10 of 10 Observability Jobs in Surrey

Senior Software Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
claims‐based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLI/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms. General knowledge ...

Senior DevOps Systems Administrator

Hiring Organisation
Mobilus Limited
Location
Guildford, Surrey, United Kingdom
Employment Type
Permanent
Salary
£55000 - £65000/annum + excellent benefits
support infrastructure across AWS and private cloud Automate infrastructure with Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement observability tools (Grafana, Prometheus, CloudWatch) Work across teams to improve reliability and security Manage hybrid networks, IAM, firewalls and VPNs The successful DevOps Systems Administrator will ...

Staff Machine Learing Engineer

Hiring Organisation
Jobleads-UK
Location
Sunbury-on-Thames, England, United Kingdom
innovations from experimentation through to productised, maintainable solutions that deliver measurable value.* Drive engineering excellence across ML systems, including CI/CD, testing, observability, reliability, and MLOps guidelines.* Define technical standards, patterns, and protocols for ML engineering and applied ML science across teams.* Lead complex, multi-team technical initiatives ...

Principal Software Engineer- Data Platforms

Hiring Organisation
Danaher
Location
Surrey, United Kingdom
Employment Type
Full Time
transformation, enrichment, indexing, and lifecycle management. Proven ability to deliver high‐quality, production-grade data systems with a focus on data quality, reliability, scalability, observability, and operational support. Experience enabling data for downstream AI and reporting use cases, including cross-entity queries, contextual linking, and performant data access patterns. Demonstrated ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

DevOps Engineer

Hiring Organisation
WeDo Technology Solutions Limited
Location
Croydon, Surrey, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
pipelines across multiple engineering teams• Automate infrastructure and deployments using Infrastructure as Code• Support Azure cloud infrastructure and AKS environments• Improve platform monitoring, observability, and operational efficiency• Troubleshoot production and deployment issues to maintain platform reliability• Collaborate closely with Software Engineers, Platform Engineers, and Security teams to improve delivery …/Kubernetes• Strong Terraform or Infrastructure as Code experience• Experience building and maintaining CI/CD pipelines• Good understanding of monitoring, logging, and observability tools• Strong troubleshooting and problem-solving skills• Experience working within Agile engineering teams Why Apply? You'll be joining an engineering organisation operating at genuine enterprise ...

Principal Software Development Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
pipelines, Infrastructure as Code, automation frameworks, and database-as-code practices using Redgate Flyway. Take ownership of critical customer systems, ensuring operational resilience, observability, performance optimisation, and rapid incident response. Collaborate closely with Product, Delivery, Operations, and Commercial teams to shape technical solutions, delivery plans, and strategic outcomes. Promote secure … Connect or Genesys Cloud. Proven ability to design and deliver secure, scalable, and resilient cloud-native solutions within complex enterprise environments. Strong understanding of observability, operational support, reliability engineering, and end-to-end ownership practices. Knowledge of regulated financial services environments, including UK GDPR and FCA Consumer Duty requirements. Excellent ...

SRE Technical Lead

Hiring Organisation
Capgemini
Location
Surrey, United Kingdom
Employment Type
Full Time
point for major incidents and high risk releases, protecting service stability and ensuring blameless post incident reviews lead to measurable improvement. • Define and govern observability and capacity practices so reliability risks are visible, actionable, and proactively managed. • Ensure SRE practices align with service governance, security, and compliance requirements, and contribute … including: • Strong expertise in Kubernetes and OpenShift. • Experience with multi cloud and hybrid architectures, including service mesh (e.g. Istio). • Hands on experience with observability platforms such as Prometheus, Grafana, Loki, Tempo, and OpenTelemetry. • Strong Infrastructure as Code and GitOps experience (Helm, Kustomize, ArgoCD, Tekton). • Experience with CI/ ...

Vice President, Build — Data, Engineering & AI

Hiring Organisation
Jobleads-UK
Location
Reigate and Banstead, England, United Kingdom
ready criteria through the Data Architect; run the Collibra dictionary, master & reference data operations and the governance council. Own data quality and observability: shift‐left checks, lineage and monitoring built in, not bolted on. Run data BAU — incidents, refreshes and access — baselined before any cost reduction is taken. Own data … access framework. Success measures Prioritized use cases delivered to production with named owner, evidence pack, run model and sunset criteria. Data quality, lineage and observability coverage across priority data products. Time from funded demand to production. Production reliability, incident rate and support performance. AI evaluation coverage, red‐team completion ...

Senior Director, Technology Operations

Hiring Organisation
Jobleads-UK
Location
Guildford, England, United Kingdom
telephony lead and specialist team holding the hands‐on work. DevOps and developer experience (run side): CI/CD reliability, environments, deployment, and observability; currently delivered by the outsourced partner, to be shaped and, over time, selectively insourced. Run‐side Service Delivery, including telephony provisioning and change. Run‐side management … hours/on‐call model for the live service, acting as the senior escalation point. Replace firefighting with proactive reliability: root‐cause discipline, observability, and measurable reductions in downtime and change‐failure rate. Platform, telephony, and estate Own the evolution of the on‐prem and cloud estate against a quarterly ...