876 to 900 of 3,664 Remote/Hybrid Observability Jobs

Senior Database Specialist

Location
Skipton, England, United Kingdom
Instance, Cosmos DB or Azure PostgreSQL, or other cloud environmentsFamiliarity with DevOps practices, CI/CD pipelines and modern engineering methodologies.Experience with monitoring and observability platforms.Knowledge of data platform technologies, analytics services or data engineering practices.What’s In It For YouYour work matters.And the way we reward you matters, too.At ...

R&D Software Engineer

Hiring Organisation
Aveva Group
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
with AVEVA CONNECT.Operate and improve cloud environments: support desktop streaming (Amazon WorkSpaces Applications) and Windows-centric infrastructure (EC2, FSx, Active Directory, DynamoDB) with strong observability and cost/performance focus. Deliver securely with automation and teamwork: write clean, tested, documented, deployable code; contribute to CI/CD and infrastructure automation ...

Senior Data Engineer II (Data Platform)

Location
Greater London, England, United Kingdom
batch and real-time workloads. Own the integration strategy across tools such asdbt, Spark and Beam Lead the implementation and adoption of data quality, observability, security, testing and deployment including creating and sharing standards Partner closely with Analytics Engineers, Data Scientists, ML Engineers, and product teams to ensure platform capabilities ...

R&D Software Engineer

Location
United Kingdom
AVEVA CONNECT. Operate and improve cloud environments: support desktop streaming (Amazon WorkSpaces Applications) and Windows-centric infrastructure (EC2, FSx, Active Directory, DynamoDB) with strong observability and cost/performance focus. Deliver securely with automation and teamwork: write clean, tested, documented, deployable code; contribute to CI/CD and infrastructure automation ...

R&D Software Engineer

Location
Cambridge, England, United Kingdom
AVEVA CONNECT. Operate and improve cloud environments: support desktop streaming (Amazon WorkSpaces Applications) and Windows-centric infrastructure (EC2, FSx, Active Directory, DynamoDB) with strong observability and cost/performance focus. Deliver securely with automation and teamwork: write clean, tested, documented, deployable code; contribute to CI/CD and infrastructure automation ...

ML/AI Engineer

Location
Manchester, England, United Kingdom
leverage TensorRT where appropriate. Operate scalable serving frameworks (NVIDIA Triton, TorchServe) with attention to latency, efficiency, resilience, and cost. Implement end‐to‐end observability for models and pipelines: drift, data quality, fairness signals, latency, GPU utilisation, error budgets, and SLOs/SLIs via Prometheus, Grafana, and Dynatrace. Establish actionable alerting ...

Embedded DevOps Engineer

Location
Harwell, England, United Kingdom
environments and automated test rigs. Familiarity with embedded Linux, cross-compilation toolchains, RTOS environments or FPGA development and deployment workflows. Experience with monitoring and observability platforms such asOpenTelemetry, Prometheus, Grafana or Loki. Knowledge of secure software supply-chain practices, including dependency scanning, artifact signing, software bills of materials, secrets management ...

Engineering Team Lead | Tech | London, UK

Location
City of Westminster, England, United Kingdom
other multi-agent/LLM orchestration frameworks Experience configuring Azure services (App Service, Container Registry, Blob Storage) or AWS equivalents Familiarity with LLM observability/evaluation tooling (e.g. Langfuse) and resilience patterns for third-party AI providers (e.g. circuit breakers, provider fallback) Awareness of data protection and AI transparency considerations ...

Embedded DevOps Engineer

Location
Kidlington, England, United Kingdom
environments and automated test rigs. Familiarity with embedded Linux, cross-compilation toolchains, RTOS environments or FPGA development and deployment workflows. Experience with monitoring and observability platforms such asOpenTelemetry, Prometheus, Grafana or Loki. Knowledge of secure software supply-chain practices, including dependency scanning, artifact signing, software bills of materials, secrets management ...

Senior software engineer - EU squad

Location
Greater London, England, United Kingdom
/statically typed language. Have a strong understanding of designing, building, and running high‐quality, standards‐compliant workflow APIs, with a focus on testing, observability, and performance. Have worked with a cloud provider (AWS/Azure/GCP). We use AWS. Have worked with distributed systems and are comfortable ...

Senior Mobile Solution Lead (iOS and Android) - London, UK

Location
Greater London, England, United Kingdom
services and enterprise systems.Apply appropriate architectural patterns, including Clean Architecture, modularisation, MVVM and platform-specific patterns, to support maintainability and scale.Architect for resilience, performance, observability, accessibility, privacy, offline capability and secure mobile lifecycle management.Ensure alignment with enterprise architecture standards, regulatory expectations and App Store/Google Play requirements.Architecture governance ...

Senior software engineer (Node.js/TypeScript)

Location
Bath, England, United Kingdom
/statically typed language. Have a strong understanding of designing, building, and running high‐quality, standards‐compliant workflow APIs, with a focus on testing, observability, and performance. Have worked with a cloud provider (AWS/Azure/GCP). We use AWS. Have worked with distributed systems and are comfortable ...

Senior software engineer - EU squad

Location
Greater London, England, United Kingdom
/statically typed language. Have a strong understanding of designing, building, and running high‐quality, standards‐compliant workflow APIs, with a focus on testing, observability, and performance. Have worked with a cloud provider (AWS/Azure/GCP). We use AWS. Have worked with distributed systems and are comfortable ...

Senior DevOps Engineer

Hiring Organisation
Informa Connect
Location
London, UK
Employment Type
Full-time
Company DescriptionDo you want to develop your career and make an impact in the fast-growth, fast-moving B2B technology space? At Informa TechTarget, you'll collaborate and grow alongside some of the industry's ...

Senior Software Engineer, Git Systems

Hiring Organisation
GitHub
Location
United Kingdom
Salary
£ 70 K
About GitHubGitHub is the world’s leading platform for agentic software development — powered by Copilot to build, scale, and deliver secure software. Over 180 million developers, including more than 90% of the Fortune 100 companies ...

IBM Netcool / Observability Technical Lead

Hiring Organisation
Deerfoot Recruitment Solutions Ltd
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£780.00 - £830.00 per day
Netcool/Observability Technical Lead Inside IR35 Contract -up to £827pd London Hybrid - 4 Days Onsite/1 Day WFH per Week Banking Are you the person who knows exactly why an ObjectServer failover didn't behave as expected, and how to stop a flood of duplicate events before anyone … shape how thousands of infrastructure and application events are detected, correlated and actioned across EMEA, and you'll have genuine scope to modernise observability capability rather than simply keep the lights on. This is a hands-on technical leadership role with no direct reports, so your influence comes from your ...

IBM Netcool / Observability Technical Lead

Hiring Organisation
Deerfoot Recruitment Solutions
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£780 - £830 per day
Netcool/Observability Technical Lead Inside IR35 Contract -up to £827pd London Hybrid - 4 Days Onsite/1 Day WFH per Week Banking Are you the person who knows exactly why an ObjectServer failover didn't behave as expected, and how to stop a flood of duplicate events before anyone … shape how thousands of infrastructure and application events are detected, correlated and actioned across EMEA, and you'll have genuine scope to modernise observability capability rather than simply keep the lights on. This is a hands-on technical leadership role with no direct reports, so your influence comes from your ...

Senior / Principal Applied AI Engineer (UK / Europe, Remote)

Location
United Kingdom
production-grade AI agents, agentic workflows, and LLM integrations for internal automation and in-product features for a web-native trading platform. Own implementation, observability, and security while partnering with Product and Platform teams to deliver measurable AI systems in production. Job Description Role Senior/Principal Applied AI Engineer … integrate LLMs into internal systems and the trading product. This is a hands-on engineering role: you will write production code, put evaluation and observability on everything shipped, and operate autonomously in a lean team. Key Responsibilities Build agent loops, tool-calling, structured outputs, planning/state management, retries, guardrails ...

SRE / Platform Engineer - Remote

Hiring Organisation
Genesis10
Location
New York, United States
Employment Type
Permanent
Salary
USD 105 Hourly
infrastructure and operational problems. Engineers on the team write code every day and work across application and infrastructure layers to improve reliability, performance, scalability, observability, and system integration. A major initiative for the team is establishing a centralized observability capability across an environment where monitoring and operational data have historically … been siloed. The organization is bringing telemetry together using Datadog and enterprise data lake capabilities, creating a common observability foundation that can ultimately support AIOps, agentic AI, automated remediation, and self-healing systems. This is not a traditional operations or Solutions Architecture position. The successful candidate will be expected ...

Senior DevOps Engineer

Hiring Organisation
Experis
Location
Derby, Derbyshire, United Kingdom
Employment Type
Contract
/CD pipelines and deployment orchestration. Support Kubernetes and OpenShift platform troubleshooting and optimisation. Deliver secure and compliant infrastructure solutions. Implement monitoring, logging, and observability tooling across environments. Collaborate with engineering, architecture, and delivery teams to improve deployment efficiency and platform reliability. Champion automation-first approaches to infrastructure and application … designing and maintaining enterprise-scale CI/CD pipelines . Strong understanding of cloud security and secure delivery practices. Experience implementing monitoring, logging, and observability solutions. Ability to define technical standards, governance, and reusable deployment frameworks. Experience working within large-scale enterprise transformation programmes. Desirable Skills Experience within highly regulated ...

Platform Engineer

Location
Milton Keynes, England, United Kingdom
Terraform, CloudFormation, or CDK), CI/CD pipelines, API Gateway, Lambda, and Aurora PostgreSQL — and confident applying fundamentals such as high availability, fault tolerance, observability, and cost control in a live environment. Payments or fintech exposure is an advantage but not required; what matters more is a practical, first-principles … rollback processes Support deployment of AI generated applications and tooling Support change control and release management alongside Engineering and the Information Security Officer Reliability, Observability & Data Infrastructure Design and maintain systems for high availability, fault tolerance, and resilience by default Implement logging, monitoring, and alerting (for example CloudWatch) across services ...

AI Platform & Site Reliability Engineering Managing Consultant

Location
United Kingdom
help clients design, build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services. You will work with technology, engineering, operations … operational requirements. AI Platform Engineering & LLMOps: Design and implement scalable AI platform capabilities including model deployment pipelines, prompt and model management, evaluation frameworks, AI observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including ...

Senior Workday Integrations Product Engineer

Location
Hook, England, United Kingdom
ecosystem supporting our global workforce. This is a hands‐on engineering role with a strong focus on Workday integration development, architecture, automation, reliability, security, observability, and operational excellence. Your Responsibilities: Design, develop, test, deploy, and support robust integrations between Workday and enterprise applications using Workday Studio, EIB, Core Connectors, Cloud … technology ecosystem. Apply strong software engineering principles including version control, peer reviews, automated testing, CI/CD, and structured release management. Build integrations with observability, operational resilience, proactive monitoring, alerting, logging, automated recovery, and self‐healing error handling by design. Troubleshoot complex production issues, conduct root‐cause analysis, and continuously ...

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
help clients design, build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering, operations … operational requirements.• AI Platform Engineering & LLMOps: Design and implement scalable AI platform capabilities including model deployment pipelines, prompt and model management, evaluation frameworks, AI observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£65,000 - £70,000 per annum
automate deployments using Terraform and Ansible, and build CI/CD pipelines using GitHub Actions. You'll also work across cloud security, networking and observability, while having the opportunity to develop your technical expertise through training and professional certifications. This role would suit an experienced DevOps Engineer looking to work … private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join ...