326 to 350 of 392 Remote Observability Jobs

Senior Software Engineer - Customer Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams, including … build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such ...

Senior Software Engineer - Customer Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams, including … build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such ...

Hybrid SRE Engineer — Observability & Cloud (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
help transform current workloads toward an SRE model while working in a hybrid setup, visiting the London office twice weekly. The role focuses on observability, high availability and incident management, with collaboration across Product Engineering and Infrastructure teams. Strong AWS, Terraform, Python and Kubernetes skills are valued. #J-18808-Ljbffr ...

Dynatrace/Observability Engineer - Remote and Telford - 6 months+

Hiring Organisation
Octopus Computer Associates
Location
Telford, Shropshire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
Dynatrace/Observability Engineer - Remote and Telford - 6 months+/RATE: £565 per day inside IR35 One of our Blue Chip Clients is urgently looking for a Dynatrace/Observability Engineer. Please find some details below: Clearance Required: SC Eligible Duration: 6 months Location: Telford - 2 days min per month … Description: Role Description: As a Dynatrace/Observability Engineer, you will be responsible for designing, implementing, and supporting monitoring solutions across a range of technologies and platforms, ensuring service stability, performance insight, and proactive incident management. Key Responsibilities: Translate high-level monitoring and non-functional requirements (NFRs) into actionable configurations ...

SRE Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£90000 - £100000/annum
hold currently). The Role As a Site Reliability Engineer (SRE) you will lead site reliability engineering initiatives with a strong emphasis on observability, ensuring high performance and reliability of applications & infrastructure. Provide strategic insights to shape the overall SRE strategy while collaborating on the design and implementation of scalable … solutions. Establish effective monitoring, alerting and incident response strategies to maintain system availability and promote continuous improvement by collaborating with team members to deliver observability best practices and SRE methodologies. The Responsibilities Define and implement Service Level Indicators (SLIs) and Service Level Objectives (SLOs) to measure and maintain system ...

Site Reliability Engineer SRE Azure SaaS

Hiring Organisation
Client Server
Location
Cambridge, Cambridgeshire, East Anglia, United Kingdom
Employment Type
Permanent, Work From Home
Site Reliability Engineer/SRE (Azure SaaS) Cambridge/WFH to £100k Do you have expertise with observability and monitoring within a SaaS environment? You could be progressing your career in a hands-on, influential Site Reliability Engineer SRE role at a global InsurTech business, working on a flagship product … happy to mentor and coach others, sharing your SRE expertise with software engineers and DevOps You have a strong knowledge of Azure including observability, monitoring, scaling, security and Azure DevOps pipelines You have experience with observability tools, Datadog preferred You have a good knowledge of automation, scripting (Python or PowerShell ...

Site Reliability Engineer SRE Azure SaaS

Hiring Organisation
17918
Location
London, United Kingdom
Site Reliability Engineer/SRE (Azure SaaS) Cambridge/WFH to £100k Do you have expertise with observability and monitoring within a SaaS environment? You could be progressing your career in a hands-on, influential Site Reliability Engineer SRE role at a global InsurTech business, working on a flagship product … happy to mentor and coach others, sharing your SRE expertise with software engineers and DevOps You have a strong knowledge of Azure including observability, monitoring, scaling, security and Azure DevOps pipelines You have experience with observability tools, Datadog preferred You have a good knowledge of automation, scripting (Python or PowerShell ...

DevOps Engineer

Hiring Organisation
Fruition Group
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Contract
Contract: Inside IR35 We're seeking an experienced Senior DevOps Engineer to join a small, highly skilled engineering team delivering a large-scale enterprise observability platform as they move away from Splunk This is an opportunity to work on a critical cloud platform supporting the migration of numerous services onto … modern monitoring and logging solution. What you'll be doing * Support and enhance a large-scale observability platform. * Help engineering teams onboard and migrate their services. * Build and maintain dashboards, log pipelines and alerting. * Develop and manage cloud infrastructure using Terraform across Azure and AWS. * Produce technical documentation and operational ...

Data Reliability Engineer

Hiring Organisation
Ashdown Group
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents, and building scalable observability across a modern enterprise data platform. Youll help ensure data across the organisation is accurate, reliable, and trusted for critical business decision-making. Youll take ownership … style roles, with strong SQL and Python skills and experience working in modern cloud-based data environments. Hands-on experience with data observability tools such as Grafana, Monte Carlo, or Acceldata, and data governance/quality platforms like Informatica, Collibra or Microsoft Purview is highly desirable. Experience within the Azure ...

AI Engineer

Hiring Organisation
Parkside
Location
London, United Kingdom
Employment Type
Permanent
Salary
£50000 - £90000/annum
regulated industry demands Implement retrieval systems from ingestion and chunking through to vector stores and retrieval optimisation Ship production-grade code with proper observability, error handling, testing and CI/CD Help design guardrails and failure handling so AI systems behave safely with real customers and real money involved … prompt engineering Exposure to agent frameworks (LangGraph, Claude Agent SDK, OpenAI SDK) or equivalent custom implementations An interest in LLM evaluation, debugging and observability Cloud platform experience (AWS, GCP or Azure) is a plus at junior level and expected at senior level The bar scales with the level. For senior ...

Machine Learning Engineer

Hiring Organisation
Spencer Rose Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 110,000 Annual
cycle Develop APIs and services for model training, inference and evaluation Deploy, version and manage machine learning models across production environments Design monitoring and observability for production ML systems Build automated workflows for model retraining, testing and deployment Collaborate with Applied Scientists to productionise new machine learning models Improve scalability … model serving, inference or training pipelines Building REST APIs (FastAPI or similar) Docker and containerisation CI/CD pipelines Linux Production monitoring, logging and observability Writing clean, maintainable and well-tested production code Desirable Experience Any exposure to the following would be beneficial: MLflow, Weights & Biases or similar experiment tracking ...

Digital Senior Full Stack Engineer

Hiring Organisation
Leeds Building Society
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
services. You'll lead complex technical delivery, champion modern engineering practices and help shape high-quality solutions through clean architecture, automation, CI/CD, observability and secure-by-default development. Just as importantly, you'll coach and mentor other engineers, raise standards across the squad and define ways of working. … leading code/design reviews; uplifting test automation and quality gates. Ability to influence stakeholders across Product, Architecture, InfoSec, Risk and Operations; governance experience. Observability experience: metrics, logs, traces; operational ownership of services. Experience of supporting UI/UX Design would be beneficial And in return ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Build and maintain CI/CD pipelines in GitHub Actions that are fast, reliable, and give engineers confidence to ship frequently and safely. Drive observability across the platform via instrumentation, alerting, dashboards, and on‐call tooling so the team can detect, diagnose, and resolve incidents quickly. Own security and compliance … Strong experience with CI/CD pipelines, particularly GitHub Actions, and a track record of improving deployment reliability and developer velocity. Deep familiarity with observability tooling and incident response practices. You've been on‐call and know what good looks like. Solid understanding of cloud security principles: IAM, least privilege ...

Sr. Cloud Platform Engineer (Hybrid, London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
culture emphasizes innovation, knowledge sharing, and technical excellence while maintaining a strong focus on security best practices. Our Tech Stack: Cloud Platforms: AWS, GCP Observability: LogScale (Humio), Grafana Infrastructure: Kubernetes, Kafka Programming: Golang (APIs), Python Data: PostgreSQL, Cassandra, OpenSearch Integration: REST, GraphQL, OAuth, JSON What You’ll Do: Design … processes, improve efficiency and drive business outcomes. Bonus Points: Experience working directly with customers or partners Security certifications Container orchestration experience Experience with observability tools Knowledge of the company Falcon platform #LI - GO1 Benefits of Working at the company: Market leader in compensation and equity awards Comprehensive physical and mental ...

AI Operations Engineer (Python)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
requirements evolve. Act as the point of contact for AI system issues, triaging, diagnosing, and resolving incidents while keeping stakeholders informed throughout. Own monitoring, observability, and quality across the AI estate, going beyond uptime to track health signals specific to agents, such as output quality, model and prompt regressions … understanding of Agile delivery in large‐scale enterprise environments. Experience supporting or operating production systems, ideally AI driven or data intensive, with strong monitoring, observability, and distributed systems diagnosis skills. Practical experience with cloud (particularly AWS), relational databases such as Postgres, and familiarity with container orchestration or PaaS such ...

Principal AI Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
captures and governs new categories of data (e.g. smart badge telemetry, location/proximity signals) in a privacy-compliant, lawful-basis-aware way Build observability into data systems — structured logging, freshness and quality monitoring, SLOs, and pipeline health dashboards Champion high-quality technical communication: proposals, specifications, and documentation that other … product engineering teams, even if your focus is data and AI infrastructure DevOps fluency: AWS, Kubernetes/EKS, Terraform, CI/CD pipelines Excellent observability practices — structured logging, metrics, distributed tracing, SLOs (Datadog, Sentry) Feature flags, canary deployments, and gradual rollout patterns Track record of driving data quality, governance ...

Principal Consultant - Cloud & Engineering

Hiring Organisation
Zhlke Engineering Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
The Role As a Principal Consultant for Cloud & Engineering, you will provide senior technical and engineering leadership across complex client engagements. You will help clients shape and deliver practical strategies for cloud adoption, optimisation and ...

Founding Infrastructure Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
# Founding Infrastructure EngineerReports toCTOTypeFull-timeLocationRemote — relocation to London or SF possible later## About Modern RelayWe're building a lakehouse-native graph engine with git-style workflows. Branch, commit, and merge typed graph data like ...

Principal Consultant - Cloud & Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
## Principal Consultant - Cloud & EngineeringApplylocations: Londontime type: Full timeposted on: Posted Todayjob requisition id: JR100890Founded in Switzerland in 1968, Zühlke is owned by its partners and located across Europe and Asia. We are a global ...

Principal Consultant - Cloud & Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Founded in Switzerland in 1968, Zühlke is a global transformation partner with engineering and innovation at its core. Role Summary Principal Consultant – Cloud & Engineering. Provide senior technical and engineering leadership across complex client engagements, shaping ...

Senior Platform Engineer - Developer Experience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
paved roads that other engineers use. You will work across the software development lifecycle, from creating a new service through to testing, deployment, observability and operating it in production.You will join an established Platform team and work alongside our existing Developer Experience Engineer. Your customers are 9fin’s software engineers … reliability and usability of our CI/CD systems. Developing reusable platform capabilities that product engineers can consume through self‐service. Helping engineers use observability effectively, with good defaults for logs, metrics, traces and service‐level indicators. Working directly with engineers to understand friction, test ideas and support adoption. Using ...

Backend Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
This position will take end-to-end ownership of product and feature delivery, from system design and backend architecture to integration with our infrastructure, observability of what gets shipped, and the best practices that keep the team moving fast and sustainably. You'll also help shape how we integrate … architecture patterns and best practices for designing highly available, scalable, and secure distributed systems; experience with event-driven architectures is an advantage Experience with observability and troubleshooting in production: log monitoring, error tracking, and APM Security-minded approach to development: secure coding practices and vulnerability remediation as part ...

DevSecOps Capability Manager

Hiring Organisation
WRK DIGITAL LTD
Location
Skipton, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
improvement Strategy, Governance & Technical Direction Set DevSecOps strategy across pipelines and security automation Establish governance for CI/CD, IaC, and cloud delivery Define observability standards (SLOs, tracing, dashboards) Embed security into pipelines (SAST, SCA, DAST, secrets, IaC scanning) Govern "Golden Path" templates and adoption Operational Oversight & Risk Management Oversee …/CD, DevSecOps, and security integration Strong cloud, containerisation, and IaC knowledge Proven ability to improve DORA and engineering performance metrics Experience with observability and monitoring frameworks Strong background in security tooling (SAST, SCA, DAST, scanning tools) Solid understanding of cloud security, IAM, and zero-trust principles Experience working ...

Senior Python Backend Engineer Fully Remote, UK

Hiring Organisation
Interact Consulting Limited
Location
South West London, London, United Kingdom
Employment Type
Permanent, Work From Home
APNs/FCM), user notification preferences, audience segmentation, and delivery tracking. Integrate with third-party data providers and external services, ensuring robust failure handling, observability, and system resilience. Design and support secure internal tooling APIs, including role-based access controls, audit trails, change history, and safe administrative workflows. Build … shape technical direction and the expectation to take real ownership of what you build. Scope: backend services, infra, event-driven systems, CI/CD, observability, all built for live-event traffic. Python-first, Postgres, Redis. You'll own your services fully: building them and keeping them running in production. Heavily ...

Senior SRE - Remote, Growth & Observability

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
highly available cloud platform supporting mission-critical services. You will work at the intersection of cloud infrastructure, software engineering, and operations, driving automation, observability, and reliability across the lifecycle. You will strengthen platform reliability, improve incident response, and embed SRE best practices. #J-18808-Ljbffr ...