301 to 325 of 392 Remote Observability Jobs

AI Platform engineer

Hiring Organisation
Nextech Group Limited
Location
East London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
strategies, embedding generation, and vector store management Architect event-driven microservices using Kafka/SQS for async processing of high-volume inference requests Implement observability and cost-tracking for token usage across multiple LLM providers (Anthropic, OpenAI, open-source models via vLLM) Own database performance for both relational (Postgres … pgvector Infra: AWS (ECS, Lambda, SQS/SNS), Docker, Kubernetes, Terraform AI/ML tooling: LangChain/LlamaIndex, vLLM, Anthropic & OpenAI APIs, embedding models Observability: Datadog, Grafana, OpenTelemetry CI/CD: GitHub Actions, ArgoCD Requirements: 4+ years backend development experience, ideally with at least 1 year working with LLM/ ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
persistence platforms Provide technical leadership and architectural guidance across multiple engineering teams Define engineering standards, platform roadmaps and best practices Drive automation, resilience, observability and operational excellence initiatives Support and mentor engineers through code reviews, coaching and technical leadership Collaborate with architects and stakeholders to translate business requirements into technical … automation and DevOps practices Experience mentoring engineers and providing technical leadership Key Technologies AWS Terraform Linux Cassandra Couchbase ScyllaDB Kafka CI/CD Pipelines Observability & Monitoring Platforms Distributed Database Technologies Nice to Have Experience with additional distributed persistence technologies Background in large-scale cloud-native environments Experience defining enterprise platform ...

DevSecOps Engineering Lead CGEMJP00346044

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
management, including allowlist processes and risk acceptance where required Secrets management and identity/access management Policy enforcement for workloads, container images and infrastructure Observability, monitoring, logging and audit controls Partner with developers to embed secure-by-design engineering and ensure compliance with CLIENT security standards. Enable and govern Infrastructure … compliance tooling (e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager) Containers and orchestration (e.g. Docker, AWS EKS) Infrastructure as Code (e.g. Terraform) Observability (e.g. Grafana, Loki) Scripting and automation (e.g. Python, Bash) Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies) Experience delivering within the UK Government ...

DevSecOps Engineering Lead CGEMJP00346044

Hiring Organisation
Experis
Location
London, United Kingdom
Employment Type
Contract, Work From Home
management, including allowlist processes and risk acceptance where required Secrets management and identity/access management Policy enforcement for workloads, container images and infrastructure Observability, monitoring, logging and audit controls Partner with developers to embed secure-by-design engineering and ensure compliance with CLIENT security standards. Enable and govern Infrastructure … compliance tooling (e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager) Containers and orchestration (e.g. Docker, AWS EKS) Infrastructure as Code (e.g. Terraform) Observability (e.g. Grafana, Loki) Scripting and automation (e.g. Python, Bash) Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies) Experience delivering within the UK Government ...

Azure Technical Architect

Hiring Organisation
JAM Recruitment Ltd
Location
United Kingdom
Employment Type
Contract
Contract Rate
Up to £507.51 per day
tagging standards, and cost management principles. Support CI/CD and DevOps patterns using GitHub Actions, DevOps pipelines, Infrastructure as Code, and automation. Define observability standards using Azure Monitor, Log Analytics, alerts, dashboards, and diagnostics. Review and validate engineering designs, ensuring alignment to standards and architectural patterns. Provide expert guidance … cloud, identity federation, Entra ID, service principals, and managed identities. Knowledge of storage architectures, resiliency patterns, backups, and Azure Site Recovery. Understanding of monitoring, observability, and operational readiness within Azure environments. Familiarity with DevOps, CI/CD, and Infrastructure-as-Code concepts and tooling. Knowledge of cloud security concepts including ...

Platform Engineer

Hiring Organisation
Hireful
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£80,000
We are recruiting founding Platform Engineers on behalf of a fast-growing enterprise level (global, 500+ staff) software business with a strong engineering culture and a genuine commitment to doing things the right way. They ...

Principal Engineer - Member Experience Platform

Hiring Organisation
Jobleads-UK
Location
Skipton, England, United Kingdom
Quality), and bar‐raising across squads: you shorten lead times, increase deployment frequency, hold change‐failure rate low, and improve MTTR through release‐linked observability - turning fast, safe flow into the default way of working. Operating at platform scale, you define cross‐cutting architecture and delivery standards (API/event … contracts, resilience, observability, language/dependency baselines) and drive adoption through the Golden Path: policy‐as‐code CI/CD, progressive delivery (feature flags, canary/blue‐green), automated rollback/forward‐fix, ephemeral, data‐ready environments, and guardrails that make security and compliance by design. You partner with Platform ...

Data Reliability Engineer

Hiring Organisation
Ashdown Group
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 95,000 Annual
work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents, and building scalable observability across a modern enterprise data platform click apply for full job details ...

Monitoring Engineer

Hiring Organisation
17918
Location
London, United Kingdom
days per month, and offers up to £580 per day (Inside IR35), 6 months rolling contract. We're looking for someone with strong observability and monitoring experience across technologies such as Dynatrace, Splunk, AppDynamics... WKCL1_UKTJ ...

AI-First Engineering Manager, Manage Squad Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reporting to the CPO, you will mentor four engineers, partner with Product and Design, own hiring, ensure quality and platform health, and drive security, observability and governance across the domain. #J-18808-Ljbffr ...

Data Services Sales Leader – Remote-First, EMEA/APJ

Hiring Organisation
Jobleads-UK
Location
Windsor, England, United Kingdom
Manager for our Data Services portfolio to lead a high-performing team of sales specialists across EMEA and APJ, driving the ARR target for Observability and Cyber Resilience solutions in large enterprise accounts. You will develop and execute a comprehensive sales strategy, oversee pipeline reviews and revenue forecasting, coach ...

Observability Engineer

Hiring Organisation
Tria
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£300 - £340/day
Observability Engineer 3-Month Contract £340/day (Inside IR35) Remote - Occasional travel to a London office Our client is an established media business with a UK operation of around 3,000 people, part of a larger international group. Its technology function is investing in observability and engineering, and this … contract role sits within the team responsible for its core monitoring platforms. You will operate and maintain the primary observability platforms, Splunk and SolarWinds, keeping them current and well configured, and work with technology and business teams to make sure the right monitoring and alerting is in place. The role ...

Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Reliability Engineers. Shape and deliver our Site Reliability Engineering roadmap alongside the Head of Platform. Champion modern engineering practices including SLIs, SLOs, error budgets, observability and automation. Improve the reliability, scalability and performance of our cloud platforms and digital services. Partner with Engineering, Security, Data and Product teams to embed … operational excellence from design through to production. Drive the adoption of our observability platform, helping teams gain deeper insight into the health and performance of their services. Lead incident learning, continuous improvement and automation initiatives that reduce operational toil. Provide technical leadership across AWS, Kubernetes, Infrastructure as Code, CI/ ...

Azure Platform Engineering Consultant

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
platform templates and landing zone patterns. CI/CD & Automation: Build and refine automated deployment pipelines, environment management, and release practices. Platform Quality: Embed observability (monitoring, logging, alerting), resilience, security, and FinOps principles directly into platform assets. Co-Delivery & Knowledge Transfer: Work closely alongside client engineering teams to pair, document … Core compute, networking, storage, identity, security, and platform services. Infrastructure as Code: Strong proficiency with Terraform AND Terragrunt using modular, reusable implementation patterns. DevOps & Observability: Strong experience with CI/CD tools (Azure DevOps/GitHub Actions) and monitoring stacks (Prometheus, Grafana, Azure Monitor, etc.). FinOps: Practical knowledge ...

Senior / Principal DevOps Engineer

Hiring Organisation
Hays Technology
Location
Bury, Greater Manchester, United Kingdom
Employment Type
Contract
Contract Rate
£700 - £800/day £700 - £800 p/d (depending on level)
best practices across engineering teams and onboard products onto shared platforms. Build and maintain secure, scalable, and high-performing cloud infrastructure in AWS. Implement observability, monitoring, and operational insights across multiple environments. Improve deployment processes, reduce friction, and enable self-service capabilities for development teams. Support cloud and infrastructure incident … focus on automation. Experience with containerisation and workload orchestration technologies. Scripting and programming experience with tools such as Python and Bash. Strong understanding of observability, reliability, and operational best practices. Knowledge of information security principles and experience embedding security throughout the software delivery lifecycle. If you're interested in this ...

Principal Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Head of Platform Engineering, you'll define and deliver robust platform services, establish the patterns for self-service infrastructure, CI/CD, and observability, and drive the integration of AI/ML capabilities across the organisation. You'll translate platform strategy into a coherent technical roadmap, setting standards that teams … these are sustained as the platform evolves. Defining and maintaining platform standards, patterns, and reusable components that drive consistency across teams. Leading improvements to observability, monitoring, and alerting capabilities across systems and services. Mentoring and developing engineers across the organisation, raising the collective standard of platform engineering practice. Identifying ...

Senior Java Engineer - FX eTrading

Hiring Organisation
Pontoon
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£800 - £900/day
data, order/risk workflows, and real-time streaming capabilities. Optimise Performance: Focus on improving latency, throughput, and reliability across the entire stack. Implement observability practises (metrics, tracing, logging) and conduct performance profiling. Establish Best Practises: Champion engineering excellence through code standards, testing strategies (unit/integration/… including market data, order flows, and execution workflows. Hands-On Skills: Proficiency with CI/CD, containerisation, cloud/on-prem deployments, and observability practises. AI Integration: Comfortable integrating AI coding tools into daily development workflows. Communication Skills: Excellent communication and stakeholder engagement abilities, with a track record of leading ...

Observability Engineer

Hiring Organisation
Hays Technology
Location
Telford, Shropshire, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £590/day Per Day
large-scale, complex programmes across the public sector. Working within a collaborative and forward-thinking environment, you will play a key role in enhancing observability capabilities and driving proactive service management through modern monitoring and automation practices. Your new role As an Observability Engineer, you will be responsible for designing … monitoring solutions across a range of technologies and platforms. You will work closely with architects, engineers, and project teams to deliver end-to-end observability solutions that provide valuable performance insights, improve service stability, and support proactive incident management. Key responsibilities include: - Configuring and optimising Dynatrace monitoring solutions - Translating monitoring ...

Site Reliability Engineer (SRE)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
approach. Key Responsibilities Integrating tightly with our Product Engineering teams Following SRE practices and maintaining high standards of compliance Implementing a new standard of observability utilising SLI/SLO/Error Budgets Continually evolving our observability platforms for greater coverage Using a code-first approach to build and changes … ongoing communication with stakeholders Skills Good experience in DevOps or SRE, with a keen interest to learn and grow as a Site Reliability Engineer Observability product experience (eg Datadog) Managing services using SLI/SLO & Error Budgets Experience with AWS or other cloud providers Experience in HA environments Automation skills ...

Global DevOps Lead

Hiring Organisation
Stott & May Professional Search Limited
Location
United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
with engineering, cloud, and operations teams to deliver a modern, automated, and scalable platform. You'll drive DevOps strategy across infrastructure, CI/CD, observability, SRE, and cloud optimisation while influencing senior stakeholders across the business. Key Responsibilities - Define and implement a global DevOps operating model, including governance, standards … initiatives. - Partner with engineering and cloud teams to establish clear ownership across DevOps and infrastructure. - Lead the implementation and optimisation of enterprise monitoring and observability using Datadog. - Build scalable deployment pipelines that improve release quality and speed. - Establish and monitor DORA metrics, driving improvements in deployment frequency, lead time, change ...

Full Stack Engineer (Contract) - Leeds

Hiring Organisation
Searchability (UK) Ltd
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£400 - £500 per day
secure, scalable and maintainable applications Create automated unit and integration tests Contribute to CI/CD pipelines and continuous delivery Implement logging, monitoring and observability best practices Support production issues and continuous improvement initiatives Participate in peer reviews and Agile ceremonies Produce and maintain technical documentation The Team … testing Strong troubleshooting and problem-solving skills Excellent communication and collaborative approach Nice to Have Cloud-native development experience Contract testing and UI automation Observability (logging, metrics and tracing) Financial Services experience JIRA or similar Agile tooling To Be Considered... Please either apply by clicking online or emailing your ...

Site Reliability Engineer

Hiring Organisation
Connells Limited
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
hands-on role in ensuring it is reliable, scalable, and observable. You will help establish and mature SRE practices, focusing on: Monitoring and observability Incident response Post-incident review Reliability testing and capacity planning Toil reduction Enabling development velocity We offer a hybrid working arrangement with one day per week … Build dashboards, alerts, and runbooks to improve visibility Automate repetitive tasks to reduce operational toil Collaborate with cross-functional teams to enhance reliability and observability Support performance testing and capacity planning Proactively identify and prioritise reliability improvements Experience & Skills Required: Hands-on experience with Azure Monitoring (Application Insights, Alerts, Action ...

Azure DevOps Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
modules aligned to business and security requirements Implement and enforce Azure security controls and policies Lead platform security initiatives across Azure environments Contribute to observability, alerting, and Site Reliability Engineering practices Collaborate with engineering teams to deliver resilient and scalable solutions Security & Platform Focus Areas Implement perimeter security using Azure … Terraform Kubernetes certification and experience with AKS Deep understanding of DevOps and platform engineering principles Strong knowledge of cloud security best practices Experience with observability, monitoring, and SRE concepts Why Apply Fully remote within the UK Work with a mission‐driven, highly respected health tech organisation Modern cloud environment with ...

Data Platform Lead

Hiring Organisation
Jobleads-UK
Location
Cardiff, Wales, United Kingdom
foundations that enable better decision-making across the business. From building resilient data pipelines and reusable data models to improving platform reliability, governance and observability, you'll create the capabilities that allow teams to move faster with confidence. You'll combine hands-on technical leadership with people management, building … future AI capabilities. Partner with Analytics, Product, Engineering and Architecture teams to ensure the platform supports current and future business needs. Improve platform reliability, observability, data quality and operational resilience. Ensure datasets are well-modeled, documented, discoverable and governed to support self-service analytics and AI readiness. Translate business ...

Senior Platform Engineer - Developer Experience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
paved roads that other engineers use. You will work across the software development lifecycle, from creating a new service through to testing, deployment, observability and operating it in production. You will join an established Platform team and work alongside our existing Developer Experience Engineer. You will speak directly with engineers … reliability and usability of our CI/CD systems. Developing reusable platform capabilities that product engineers can consume through self-service. Helping engineers use observability effectively, with good defaults for logs, metrics, traces and service‐level indicators. Working directly with engineers to understand friction, test ideas and support adoption. Using ...