1,351 to 1,375 of 2,124 Remote/Hybrid Observability Jobs

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Helston, Cornwall, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Dingwall, Highland, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Esher, Surrey, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Morpeth, Northumberland, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Ferndown, Dorset, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Shropshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Kilwinning, North Ayrshire, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Huddersfield, West Yorkshire, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Callander, Perthshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Alfreton, Derbyshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Airdrie, Lanarkshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Thame, Oxfordshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Bedford, Bedfordshire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Location
Manchester, Lancashire, United Kingdom
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Remote Software Engineer

Hiring Organisation
Griffin Llc
Location
Ebbw Vale, Blaenau Gwent, UK
RFCs, build proof-of-concepts, and run experiments to validate approaches. Once we commit to building something, we care about maintainability more than cleverness, observability more than hoping it works, and scalability more than premature optimisation. We ship quality code over hitting arbitrary deadlines. Once something goes live, we refactor ...

Data Integration Lead

Location
Manchester, England, United Kingdom
control, monitoring, service discovery and deprecation. Experience with integration security, including authentication, authorisation, secrets management, certificate management, encryption and data protection. Understanding of integration observability, including logging, metrics, tracing, alerting, exception handling, dead-letter queues, error recovery and operational dashboards. Familiarity with Azure integration patterns, including Azure Service Bus, Azure ...

Data Integration Lead

Location
Greater London, England, United Kingdom
control, monitoring, service discovery and deprecation. Experience with integration security, including authentication, authorisation, secrets management, certificate management, encryption and data protection. Understanding of integration observability, including logging, metrics, tracing, alerting, exception handling, dead-letter queues, error recovery and operational dashboards. Familiarity with Azure integration patterns, including Azure Service Bus, Azure ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused on AWS, Kubernetes, Terraform, observability, monitoring, and automation, working closely with software engineering teams to improve platform reliability and developer experience. You'll have genuine ownership and the opportunity to influence … infrastructure using Terraform and Infrastructure as Code principles Develop and optimise CI/CD pipelines using GitHub Actions Build and improve comprehensive monitoring and observability across the platform Implement and maintain effective logging, metrics, tracing, alerting, and dashboards Define and improve SLIs, SLOs, and reliability metrics Proactively identify and resolve ...

Staff AI Engineer, Payments Intelligence

Location
Greater London, England, United Kingdom
70+ languages, adapting to context and intent. Evaluation, safety, and guardrails — define how the unit measures agent quality and safety, with rigorous evaluation, observability, and guardrails so agents behave reliably in a compliance‐sensitive, multi‐market environment. Intelligence and data products — architect the client‐intelligence systems that turn payment data … systems at scale. Model serving and LLM inference, with the cloud infrastructure to run AI workloads in production (AWS or similar). Evaluation and observability for AI — hands‐on with tools like LangFuse, LangSmith, Braintrust, or MLflow — plus solid automated testing and CI/CD. Proven technical leadership — mentoring engineers ...

Lead AI Software Engineer

Location
Greater London, England, United Kingdom
aligned to business requirements. Working closely with Technical Leads, Architects, Product Owners, Platform Engineers, and delivery teams, you will drive implementation quality, testing, observability, operational readiness, and continuous improvement throughout the software development lifecycle. What you will do Lead build execution within a squad, platform capability, or engineering domain. Translate … generated and engineer‐written code to ensure correctness, maintainability, security, and alignment with specifications. Drive engineering excellence through automated testing, contract testing, regression testing, observability, and production verification. Support CI/CD processes, deployment readiness, operational handover, runbooks, and service ownership. Ensure AI‐generated outputs are explainable, traceable, secure ...

Site Reliability Engineer (Remote/ EU based) - Chinese Speakers

Hiring Organisation
TrueWatch
Location
Dublin, City of Dublin, Republic of Ireland
Employment Type
Permanent
Salary
£77691 - £86324/annum bonus, benefits
business provides a modern unified observability platform, helping organisations monitor complex cloud environments through data collection, visualisation and security insights. This is a strong opportunity to join at an early stage of the European growth journey, where you'll have real visibility and impact as the team scales. What … Maintain system reliability, availability, and performance for cloud infrastructure and services to ensure continuous operations Monitor production environments and manage observability tools to track metrics, logs, and alerts for proactive issue detection Support incident response by troubleshooting issues, conducting root cause analysis, and leading post-incident reviews to prevent recurrence ...

Senior Platform Engineer - AI Native SaaS Platform

Location
City Of London, England, United Kingdom
someone who can design, build and operate scalable systems, not just manage cloud infrastructure. You'll work across infrastructure, automation, CI/CD, observability and production reliability, while writing high-quality software in Go and helping improve the way engineering teams build and run services at scale. What will … making builds, tests and deployments faster and safer Build automation and internal tooling that removes manual toil for engineering teams Improve observability across metrics, logging, tracing and production monitoring Define and manage SLOs across latency, availability and error rates Lead incident response, triage complex production issues and ship long-term ...

Senior Backend Engineer | AI Platform

Location
Greater London, England, United Kingdom
high degree of autonomy and ownership, as you'll be responsible for designing scalable solutions that empower multiple engineering teams while ensuring reliability, observability, and cost efficiency. What are we looking for: 5+ years of experience in Software Engineering, Backend Engineering, or Platform Engineering. Strong experience building and maintaining backend … LangChain, LangGraph, CrewAI, or similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. ...

Lead AI Software Engineer London, United Kingdom Value Stream Engineering Posted 12 hours ago

Location
Greater London, England, United Kingdom
aligned to business requirements. Working closely with Technical Leads, Architects, Product Owners, Platform Engineers, and delivery teams, you will drive implementation quality, testing, observability, operational readiness, and continuous improvement throughout the software development lifecycle.## **What you will do*** Lead build execution within a squad, platform capability, or engineering domain.* Translate … generated and engineer-written code to ensure correctness, maintainability, security, and alignment with specifications.* Drive engineering excellence through automated testing, contract testing, regression testing, observability, and production verification.* Support CI/CD processes, deployment readiness, operational handover, runbooks, and service ownership.* Ensure AI-generated outputs are explainable, traceable, secure ...

Staff Platform Engineer - AI Native SaaS Platform

Location
City Of London, England, United Kingdom
role for someone who can design and build scalable systems, not just manage cloud infrastructure. You'll work across infrastructure, automation, CI/CD, observability and production reliability, while writing high-quality software in Go and helping set the technical standard for how engineering teams build and operate services …/CD pipelines, making builds, tests and deployments faster and safer Build automation and internal tooling that removes manual toil for engineering teams Improve observability across metrics, logging, tracing and production monitoring Define and manage SLOs across latency, availability and error rates Lead incident response, resolve complex production issues ...