251 to 275 of 392 Remote/Hybrid Observability Jobs

MLOps Engineering Manager — Lead Scalable ML (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Trainline in London is seeking an experienced MLOps Engineering Manager to build and lead a new team of engineers. You will shape deployment, observability, and scalable machine learning systems across the platform. You will collaborate with ML Engineers, Data Engineers, Software Engineers, Data Scientists, Product Managers and stakeholders to deliver ...

Platform Engineer: Scale & Automate (Hybrid London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
accelerate value delivery, scale systems, and build the foundation that empowers our engineering organization to thrive. The role focuses on improving deployment speed, observability, automation, and platform health across production, with on-call responsibilities and a cloud-native mindset. #J-18808-Ljbffr ...

Staff Backend Engineer for AI-Driven Context Layer SaaS

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Grafana Labs, the company behind the open observability cloud, is seeking a Staff-level Backend Engineer to build production services for an AI-native context layer. This remote role offers autonomy, collaboration across a global team, and the opportunity to shape foundational architecture. You will design ingestion, storage, and retrieval ...

Senior Backend Engineer — Platform Core APIs (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will join a hybrid-working environment, help shape long-term technical direction, and mentor engineers while delivering production-ready code with emphasis on safety, observability, and operational #J-18808-Ljbffr ...

Senior Product Manager, FS Resilience & Market Data

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ITRS is looking for a Senior Product Manager based in London to lead in delivering critical IT observability solutions. The role involves defining product strategy and engaging with Tier 1 financial institution customers to ensure the roadmap aligns with real needs. You will work on key projects including financial trading ...

Data Ops Manager

Hiring Organisation
Hunter Bond
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 120,000 Annual
Azure Synapse, Databricks, ADF, Power BI. Familiarity with CI/CD and automation. Strong FinOps mindset and cost management experience. Knowledge of monitoring and observability frameworks. Salary: Up to £120,000 + bonus + package Level: Manager Location: London (good work from home options available) If you are interested ...

Principal Data Engineer

Hiring Organisation
WRK DIGITAL LTD
Location
Skipton, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
systems) Modelling approaches (medallion, dimensional, event-driven, semantic layers) Data governance frameworks and metadata tooling Data lineage, cataloguing, discovery, and documentation Data quality and observability tooling Data access/security patterns Skilled in standardisationcreating reusable data frameworks, templates, modules. People leadership including performance management, wellbeing, and both personal and professional ...

Senior Backend Engineer

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
product lives and dies by — the ingestion pipelines that turn warehouse data into a model, the APIs, durable storage, background work, and the observability that lets a small team operate them with confidence at 3 am. WareBee runs on two engines: Physical AI — a living, spatial model of the warehouse ...

Lead Oracle Cloud Infrastructure Platform Engineer

Hiring Organisation
WRK DIGITAL LTD
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
Salary
£80,000
/subject matter expertise for OCI related matters and lead on root cause analysis with a focus on resilience and prevention Establish a proactive observability strategy - dashboards, metrics, logs, traces - for critical Oracle services Design and implement enterprise grade logging and monitoring solutions using OCI Logging, OCI Monitoring, Events ...

IT Service Delivery Manager

Hiring Organisation
Opus Recruitment Solutions
Location
Newcastle upon Tyne, Tyne & Wear, United Kingdom
Employment Type
Contract
Contract Rate
£37000 - £55000/annum
Knowledge of SLA management and operational governance. Preferred Qualifications ITIL Foundation or higher certification. Experience within enterprise application support environments. Familiarity with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, SolarWinds, or similar. Experience working within large-scale managed services or consulting environments. Key Competencies Leadership & Team Management Incident ...

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
considering sustainability and cost. Bonus points if: You are familiar with search engine technology such as OpenSearch, ElasticSearch, or Vespa. You are familiar with observability, tracking, and data pipeline tools and methodologies. Additional Information Health & Mental Wellbeing: PMI and cash plan healthcare access with Bupa, subsidised counselling and coaching with ...

Senior Post-Purchase Systems Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Tulip apps), fulfilment, shipping and CS tooling. Scale, reliability and delivery: Lead cross‐team initiatives that increase throughput and reduce cost‐to‐serve. Improve observability and operability across the flow from “buy” to “delivered,” reducing WISMO and manual interventions. Data and tooling coherence: Assist in enabling a 360° order view ...

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
cost in mind through thinking thrift.Bonus points if:You are familiar with search engine technology such as OpenSearch, ElasticSearch, Vespa..You are familiar with observability, tracking and data pipeline tools and methodologies.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with BupaSubsidised counselling and coaching with Self SpaceCycle to Work ...

Senior or Staff Software Engineer, SRE/ Platform Team

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
code with Kubernetes and Terraform. You'll be at the forefront of shaping our foundational architecture, ensuring it’s both resilient and scalable. Drive Observability and Monitoring: Establish and maintain a state‐of‐the‐art observability and monitoring stack. Your insights will enable us to stay ahead of potential issues ...

Senior AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
developer tooling that enable self‐service AI development across engineering teams. Design secure, scalable deployment pipelines for AI models and applications. Build AI observability capabilities including monitoring, tracing, evaluation, cost optimisation, and production quality measurement. Collaborate closely with AI Engineers, Backend Engineers and Engineering Leadership to define platform architecture … containerised deployments. Experience with solutions such as AWS Bedrock and AgentCore. Understand how to deploy, monitor, and operate AI services in production. AI Operations & Observability Experience implementing monitoring, tracing, evaluation, and cost optimisation for AI systems. Experience with observability solutions such as Arize Phoenix, Langfuse, or Langsmith. Understand the operational ...

Principal Platform Engineer

Hiring Organisation
SF Partners Admin
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
capabilities. Design and operate production-grade Kubernetes platforms, including EKS, AKS or OpenShift. Define engineering standards, golden paths, reusable modules and platform patterns. Build observability strategies using Prometheus, Grafana, OpenTelemetry and modern APM tooling. Improve reliability through SLOs, incident reviews and Site Reliability Engineering (SRE) practises. Embed DevSecOps, supply-chain … Infrastructure as Code (IaC). CI/CD automation. GitOps tools such as ArgoCD or Flux. Internal Developer Platforms or self-service engineering. Observability tools including Prometheus, Grafana, OpenTelemetry, ELK, Datadog, Dynatrace or New Relic. DevSecOps and supply-chain security. SRE practises, SLOs, SLIs and incident management. Platform governance, cloud ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Basingstoke, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Drive platform improvements and DevOps best practices. Design and implement self-service infrastructure and tooling. Deliver scalable, secure, and highly available systems. Enhance monitoring, observability, and operational performance. Support engineering teams with technical expertise and guidance. Skills & Experience Experience designing and implementing CI/CD pipelines and software delivery processes. … Infrastructure as Code experience using tools such as Terraform or Ansible. Experience with monitoring and observability tools. Strong knowledge of Docker, Kubernetes, AWS, and cloud technologies. Excellent communication skills and ability to collaborate across teams. A passion for automation, platform engineering, and continuous improvement. This is a full-time, permanent ...

Senior Tech Lead - FinTech

Hiring Organisation
Carousel Consultancy Ltd
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
Shaping the platform architecture Working closely with the tech team to evolve the platform Designing scalable backend systems and services Improving reliability, performance and observability Helping modernise legacy parts of the platform Enhancing the platform and DevOps - collaborating with AWS infrastructure and cloud-native services, refining CI/CD pipelines … Native Infrastructure: AWS (ECS, EKS, RDS, S3, Lambda) Containers and Orchestration: Docker, Kubernetes CI/CD: Jenkins, GitHub Actions Databases: MySQL, PostgreSQL Monitoring and Observability: Sentry, CloudWatch, Grafana Skills and experience required: Solid software engineering experience (c8+ years), working as a Senior, Staff, Principal Engineer or Tech Lead FinTech ...

Senior Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
health: tech debt, refactoring, security, and governance Take ownership of features from start to finish using agile methodologies, delivering with feature flags, tests, and observability Champion high‐quality technical communications: proposals, specs, testing reports, and release planning Drive AI‐first ways of working within the squad — embedding AI tooling into … relates to the Manage domain Contribute to CI/CD pipeline improvements and progressive delivery practices across squads Drive reliability monitoring and observability within Manage (Prometheus, Grafana, Sentry) Contribute to security posture improvements: vulnerability scanning, pen testing coordination, and enforcement of standards Cross‐Squad & Leadership Collaboration Work closely with ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£67,547 - £83,778 per annum, Pro-rata, Inc benefits
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
17918
Location
United Kingdom
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...