1 to 25 of 67 Remote/Hybrid OpenTelemetry Jobs in England

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud‐native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Cloud Native Specialist

Location
Greater London, England, United Kingdom
OpenShift) and container orchestration.* Demonstrate how Dynatrace provides automated, code-level visibility into microservices without manual instrumentation or sidecar overhead.* Advocate for OpenTelemetry (OTel) integration and explain how Dynatrace extends the value of open-source telemetry in a production-grade environment.2. Domain Execution:* Lead technical discovery and high-stakes Proof ...

Senior Backend Developer

Hiring Organisation
Protein Works
Location
Liverpool, Merseyside, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
OpenSearch). Advanced Engineering & Integrations: Production Go experience, GoF/Enterprise integration patterns, durable workflows, contract testing, ERP/WMS/e-commerce integrations, OpenTelemetry/Grafana/Datadog, and OWASP/ISO 27001 security standards. How we work Two-week sprints for planned platform work and a Kanban lane ...

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Location
Greater London, England, United Kingdom
Additional Technical Skills*** **Cloud Platforms:** *(e.g. AWS, GCP, Azure)** **CI/CD Tooling:** *(e.g. GitHub Actions, ArgoCD, Harness)** **Monitoring & Observability:** *(e.g. Prometheus, Grafana, Coralogix, OpenTelemetry)** **Infrastructure as Code:** *(e.g. Terraform, Pulumi)** **Config Management**: *(e.g. Ansible, Puppet, Chef)** **Programming/Scripting:** *(e.g. Python, Go, Bash)* **Soft Skills & General Requirements*** Strong problem ...

DevOps Engineer

Location
Greater London, England, United Kingdom
experience with policy‐as‐code frameworks for automated compliance and guardrails. Exposure to observability platforms such as Datadog, Prometheus/Grafana, or the OpenTelemetry ecosystem. Experience with container image hardening and scanning (Trivy, Grype, or similar). Experience with using AI tooling as well as a familiarity with security concerns ...

Logs Specialist

Location
Maidenhead, England, United Kingdom
powered observability platform.**Core Responsibilities****Domain Expertise:*** Act as the domain "subject matter expert" (SME) for Logs, staying ahead of industry trends like OpenTelemetry (OTel), log pipelines (Cribl/BindPane), and cloud-native logging (CloudWatch/Stackdriver).* Articulate the architectural superiority of Dynatrace Grail—specifically how its schema … optimize their SaaS consumption and maximize their Dynatrace investment.**Solution Architecture:*** Expertise in architecting resilient, vendor‐neutral log‐ingestion frameworks utilizing Fluentd, Logstash, and OpenTelemetry Collector pipelines, etc.* Help customers navigate complex log-routing scenarios, ensuring high-value data is prioritized for analytics while low-value data is archived cost ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi-step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root-cause analysis, anomaly detection, and semantic log clustering. Self-healing Infrastructure Engineering: Experience designing closed-loop, self-healing ...

Senior Software Engineer

Location
Reading, England, United Kingdom
Kubernetes and cloud‐native deployments CI/CD pipeline experience (GitLab, GitHub or similar) Infrastructure as Code (Terraform or similar) Experience with observability tooling (OpenTelemetry, Prometheus, Grafana, etc.) Strong testing mindset (TDD, automated testing, contract testing) Highly Desirable Experience with API Gateway technologies (e.g. AWS API Gateway) Experience building ...

Principal DevOps Engineer

Location
Nottingham, England, United Kingdom
KEDA or Karpenter Falco or Kyverno Go development Azure Front Door Azure Firewall Entra ID and PIM Observability platforms such as Grafana, Loki or OpenTelemetry Experience working within regulated or compliance-focused environments #J-18808-Ljbffr ...

SC Cleated DevOps Engineer

Hiring Organisation
IO Associates
Location
Nottingham, Nottinghamshire, East Midlands, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£500 - 550 per day
environments. Python: Comfortable using Python for scripting, automation and supporting platform engineering activities. Monitoring & Observability: Experience working with tools such as Grafana, Prometheus and OpenTelemetry to monitor, troubleshoot and improve the reliability of cloud and platform environments. Development Practices: Familiarity with modern software development practices, including Conventional Commits , version control ...

Site Reliability Engineer

Hiring Organisation
E-Solutions IT Services UK Ltd
Location
Leeds, West Yorkshire, United Kingdom
Employment Type
Full-Time
Salary
£280.00 - £300.00 per day
Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux). ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Manchester, England, United Kingdom
keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Greater London, England, United Kingdom
keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing ...

Staff Cloud SRE – AI/ML Platform & GPU Compute London, United Kingdom on-site

Location
Greater London, England, United Kingdom
toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skills Familiarity with infrastructure-as-code (e.g. Terraform ...

Platform Engineer

Location
Greater London, England, United Kingdom
datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest, SAST/DAST Developer Portal & Self … cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow‐up actions ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
Linux environments Hands-on experience operating containerised workloads using Docker and Kubernetes Experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, OpenTelemetry or the ELK Stack Experience with Infrastructure as Code and automation tooling such as Terraform and Ansible Experience building and managing CI/CD pipelines … Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft Azure/Azure/Prometheus/Grafana/Loki/OpenTelemetry/ELK Stack/GitLab CI/CD/GitHub Actions/Jenkins/ArgoCD/Helm/AWS/GCP/Networking/Observability ...

Web: Full Stack Tech Lead

Location
Greater London, England, United Kingdom
/Lambda or Cloud Run/GKE), containerized with Docker. Own CI/CD (GitHub Actions), IaC (Terraform), logging/metrics/tracing ( OpenTelemetry , CloudWatch/Stackdriver, Grafana/Prometheus), and SLOs . Optimize p95 latency, throughput, and cost ; manage secrets, networking, VPCs, and build resilient retries/backoffs. ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
delivery lifecycles. An understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Stoke-On-Trent, Staffordshire, West Midlands, United Kingdom
Employment Type
Permanent, Work From Home
delivery lifecycles. An understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge ...

Java (Kotlin) Developer (Agile, Test-Driven) AVP

Location
Greater London, England, United Kingdom
Kotlin Cloud Technologies (Kubernetes, Open Shift) Messaging Technologies (Kafka, Solace, TIBCO) Database/Data Store/Data Query Technologies (SQL Server, S3) Observability Technologies (OpenTelemetry, Elastic Stack/ELK, Grafana) Qualifications Proven experience in an App Dev role. Demonstrated execution capabilities. Education Bachelor’s/University degree or equivalent experience ...

Lead DevSecOps Engineer

Location
Greater London, England, United Kingdom
platforms (e.g., EC2 to EKS, or cross-cloud) with a focus on data integrity and minimal downtime Ability to implement standardized telemetry pipelines (e.g., OpenTelemetry, Prometheus, or ELK) that provide developers with out-of-the-box visibility into their services Familiarity with automated policy enforcement and compliance-as-code (e.g. ...

typescript developer for authentication platforms

Location
Greater London, England, United Kingdom
GitHub Actions, and writing unit and integration tests with Jest Familiarity with security principles including IAM, encryption and networking, alongside observability tools such as OpenTelemetry, Honeycomb or Grafana Nice to have: Exposure to identity or MFA platforms such as Auth0 or Transmit Security, knowledge of microservices architecture, API gateways such ...

Senior Architect Private Cloud

Hiring Organisation
Randstad Technologies Recruitment
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £550/day Inside IR35 via Umbrella
Desirable Skills: Familiarity with Service Mesh (e.g., Istio), API Gateways, Policy as Code (e.g., OPA), developer portals (Platform-as-a-Product), and Observability stacks (OpenTelemetry, Prometheus, ELK). Randstad Technologies is acting as an Employment Business in relation to this vacancy. ...

Senior .NET Backend Developer

Location
York and North Yorkshire, England, United Kingdom
software handling sensitive or clinically important data. PostgreSQL, Redis, Elasticsearch or other data and caching technologies. GraphQL, including schema design and gateway patterns. Grafana, OpenTelemetry, Prometheus or equivalent observability tooling. CI/CD, production services, microservices and message-driven systems such as RabbitMQ. How we work We value engineers ...

Senior AI Engineer - Agentic AI

Location
Greater London, England, United Kingdom
governance practices including RBAC, prompt safety checks, traceability and secrets management Implement evaluation pipelines and observability frameworks using tools such as Langfuse, Arize or OpenTelemetry Contribute to architectural design decisions, code reviews and engineering standards for platform development Requirements Bachelor's or Master's degree in Computer Science, Engineering ...