326 to 350 of 502 Remote/Hybrid Observability Jobs

DevOps Engineer

Hiring Organisation
Eligo Recruitment
Location
Stockport, Cheshire, England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £80,000 per annum
Building and maintaining Infrastructure as Code using Terraform Automating infrastructure provisioning and deployment pipelines Managing Kubernetes and containerised workloads Implementing monitoring, logging and observability solutions Driving platform reliability, security and best practices Collaborating with engineering teams to improve developer experience Skills & Experience Essential: Strong commercial experience with Google Cloud Platform … container technologies Experience with Linux and scripting (Bash, Python or Go) Understanding of networking, IAM and cloud security principles Experience with monitoring and observability tooling Desirable: Experience with GitOps practices Knowledge of Prometheus, Grafana or similar tools Experience in a platform engineering or SRE environment Certifications in GCP are advantageous ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Southampton, Hampshire, United Kingdom
Employment Type
Permanent
Salary
£60000/annum Bonus, Pension, Healthcare
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

Principal Software Development Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
pipelines, Infrastructure as Code, automation frameworks, and database-as-code practices using Redgate Flyway. Take ownership of critical customer systems, ensuring operational resilience, observability, performance optimisation, and rapid incident response. Collaborate closely with Product, Delivery, Operations, and Commercial teams to shape technical solutions, delivery plans, and strategic outcomes. Promote secure … Connect or Genesys Cloud. Proven ability to design and deliver secure, scalable, and resilient cloud-native solutions within complex enterprise environments. Strong understanding of observability, operational support, reliability engineering, and end-to-end ownership practices. Knowledge of regulated financial services environments, including UK GDPR and FCA Consumer Duty requirements. Excellent ...

Staff Engineer - Data

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
create unnecessary complexity, risk, duplicated capability or long‐term support burden. Raise the quality bar for data products through clear ownership, robust testing, reconciliation, observability, lineage, documentation, performance and supportability. Collaborate with cross‐functional teams to address security, GDPR, PII handling, role‐based access, auditability and data governance are designed … services across batch, streaming and event‐driven patterns. Deep understanding of engineering practice: clean design, testing strategy, CI/CD, infrastructure as code, observability, performance, security, incident response and DevSecOps. Experience with cloud data services and modern data stacks. Relevant technologies may include Snowflake, Azure/AWS/GCP data ...

Senior Backend Engineer - Java

Hiring Organisation
Capco
Location
Borough of Tameside, United Kingdom
Employment Type
Full Time
This job is with Capco, an inclusive employer and a member of myGwork – the largest global platform for the LGBTQ+ business community. Please do not contact the recruiter directly. Senior Backend Engineer – Java Location: London ...

AVP Site Reliability Engineer - SRE/Infrastructure/Python/Powershell/AWS/Observability/ITIL - PERM

Hiring Organisation
Scope AT Limited
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
Site Reliability Engineer - SRE/Infrastructure/Python/Powershell/AWS/Observability/ITIL - PERM - Financial Services Job Purpose: The role is primarily responsible for developing SRE methodologies and ensuring they are applied to the Cloud hosted environment. In addition, the role will act as a central point … methodologies, collaborating closely with other infrastructure teams to optimize infrastructure and deployment processes, focusing on automation and operational excellence. Drives continuous improvement in system observability, alerting, and capacity planning through the definition and implementation of SLA, SLOs & SLIs Define and enhance frameworks for Toil identification, analysis & remediation to identify opportunities ...

MLOps Engineering Manager — Lead Scalable ML (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Trainline in London is seeking an experienced MLOps Engineering Manager to build and lead a new team of engineers. You will shape deployment, observability, and scalable machine learning systems across the platform. You will collaborate with ML Engineers, Data Engineers, Software Engineers, Data Scientists, Product Managers and stakeholders to deliver ...

Platform Engineer: Scale & Automate (Hybrid London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
accelerate value delivery, scale systems, and build the foundation that empowers our engineering organization to thrive. The role focuses on improving deployment speed, observability, automation, and platform health across production, with on-call responsibilities and a cloud-native mindset. #J-18808-Ljbffr ...

Staff Backend Engineer for AI-Driven Context Layer SaaS

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Grafana Labs, the company behind the open observability cloud, is seeking a Staff-level Backend Engineer to build production services for an AI-native context layer. This remote role offers autonomy, collaboration across a global team, and the opportunity to shape foundational architecture. You will design ingestion, storage, and retrieval ...

Hybrid Data Engineering Leader: Scale Data & AI

Hiring Organisation
Jobleads-UK
Location
Swindon, England, United Kingdom
delivered across the Society. You will drive modern data workflows and data products, establish architecture standards, promote Lakehouse patterns, automation, CI/CD and observability, and collaborate with governance, analytics and AI teams. We offer hybrid working and a Swindon office with at least two days per week on site. ...

Senior Backend Engineer - Remote UK (Python/Go)

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Python (Flask), Go, Ruby, and cloud services. You’ll collaborate with product and design in a remote-friendly environment while maintaining strong testing and observability practices. #J-18808-Ljbffr ...

Senior Backend Engineer — Platform Core APIs (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will join a hybrid-working environment, help shape long-term technical direction, and mentor engineers while delivering production-ready code with emphasis on safety, observability, and operational #J-18808-Ljbffr ...

Senior Product Manager, FS Resilience & Market Data

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ITRS is looking for a Senior Product Manager based in London to lead in delivering critical IT observability solutions. The role involves defining product strategy and engaging with Tier 1 financial institution customers to ensure the roadmap aligns with real needs. You will work on key projects including financial trading ...

FinCrime Software Engineer: Remote-First, Drive Impact

Hiring Organisation
Jobleads-UK
Location
United Kingdom
services that power our cross-border payments platforms, while ensuring robust AML and customer screening tools. We value pragmatic, well-documented code, strong observability, and a bias for action. This role offers global collaboration, fast iteration, and a culture focused on trust and #J-18808-Ljbffr ...

Senior Backend Engineer

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
product lives and dies by — the ingestion pipelines that turn warehouse data into a model, the APIs, durable storage, background work, and the observability that lets a small team operate them with confidence at 3 am. WareBee runs on two engines: Physical AI — a living, spatial model of the warehouse ...

IT Service Delivery Manager

Hiring Organisation
Opus Recruitment Solutions
Location
Newcastle upon Tyne, Tyne & Wear, United Kingdom
Employment Type
Contract
Contract Rate
£37000 - £55000/annum
Knowledge of SLA management and operational governance. Preferred Qualifications ITIL Foundation or higher certification. Experience within enterprise application support environments. Familiarity with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, SolarWinds, or similar. Experience working within large-scale managed services or consulting environments. Key Competencies Leadership & Team Management Incident ...

Senior Partner Manager - Channels (UKI)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
above may vary based on the country of your employment and the nature of your employment with Datadog. About Datadog: Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure ...

Senior Partner Manager - Channels (UKI)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
above may vary based on the country of your employment and the nature of your employment with Datadog. About Datadog: Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure ...

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
considering sustainability and cost. Bonus points if: You are familiar with search engine technology such as OpenSearch, ElasticSearch, or Vespa. You are familiar with observability, tracking, and data pipeline tools and methodologies. Additional Information Health & Mental Wellbeing: PMI and cash plan healthcare access with Bupa, subsidised counselling and coaching with ...

Senior Post-Purchase Systems Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Tulip apps), fulfilment, shipping and CS tooling. Scale, reliability and delivery: Lead cross‐team initiatives that increase throughput and reduce cost‐to‐serve. Improve observability and operability across the flow from “buy” to “delivered,” reducing WISMO and manual interventions. Data and tooling coherence: Assist in enabling a 360° order view ...

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
cost in mind through thinking thrift.Bonus points if:You are familiar with search engine technology such as OpenSearch, ElasticSearch, Vespa..You are familiar with observability, tracking and data pipeline tools and methodologies.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with BupaSubsidised counselling and coaching with Self SpaceCycle to Work ...

Senior or Staff Software Engineer, SRE/ Platform Team

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
code with Kubernetes and Terraform. You'll be at the forefront of shaping our foundational architecture, ensuring it’s both resilient and scalable. Drive Observability and Monitoring: Establish and maintain a state‐of‐the‐art observability and monitoring stack. Your insights will enable us to stay ahead of potential issues ...

Senior AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
developer tooling that enable self-service AI development across engineering teams. Design secure, scalable deployment pipelines for AI models and applications. Build AI observability capabilities including monitoring, tracing, evaluation, cost optimisation, and production quality measurement. Collaborate closely with AI Engineers, Backend Engineers and Engineering Leadership to define platform architecture … deployments. Have experience with solutions such as AWS Bedrock and AgentCore Understand how to deploy, monitor, and operate AI services in production. AI Operations & Observability Have experience implementing monitoring, tracing, evaluation, and cost optimisation for AI systems. Have experience with observability solutions such as Arize Phoenix, Langfuse, or Langsmith Understand ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, UK
embedded within the company’s engineering teams to co-deliver reusable, production-ready platform components, Infrastructure as Code (IaC), CI/CD automation, observability, resilience, and FinOps practices aligned to the client's backlog and sprint cadence. The role will support secure, scalable, and cost-effective cloud services while enabling …/CD & Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, England, United Kingdom
embedded within the company’s engineering teams to co-deliver reusable, production-ready platform components, Infrastructure as Code (IaC), CI/CD automation, observability, resilience, and FinOps practices aligned to the client's backlog and sprint cadence. The role will support secure, scalable, and cost-effective cloud services while enabling …/CD & Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog ...