1 to 25 of 104 OpenTelemetry Jobs in England

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands‐on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct technical … transparency, innovation, and technical excellence that encourages continuous improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
telemetry to track unit economics and throughput for training and serving AI models. Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Senior Cloud Engineer, AI Platform SRE

Location
Leeds, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Manchester, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Greater London, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Cloud Native Specialist

Location
Greater London, England, United Kingdom
OpenShift) and container orchestration.* Demonstrate how Dynatrace provides automated, code-level visibility into microservices without manual instrumentation or sidecar overhead.* Advocate for OpenTelemetry (OTel) integration and explain how Dynatrace extends the value of open-source telemetry in a production-grade environment.2. Domain Execution:* Lead technical discovery and high-stakes Proof ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
systems at scale. Familiarity with infrastructure-as-code tools such as Terraform or Ansible. Experience with log aggregation and analysis platforms such as the OTEL Stack or AWS CloudWatch Logs Insights. Exposure to Kubernetes or other container orchestration platforms. Experience working in an Agile or DevOps team environment. EEO Statement ...

Database Platform Engineer

Location
Greater London, England, United Kingdom
services relevant to data platforms such as RDS, Aurora, S3, EC2 or EKS Familiarity with modern observability stacks such as Prometheus, Grafana, Elk or OTel Desirable: experience with cloud‐native and distributed SQL databases such as Aurora, YugabyteDB or TiDB Desirable: knowledge of data streaming and integration tools such ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi-step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root-cause analysis, anomaly detection, and semantic log clustering. Self-healing Infrastructure Engineering: Experience designing closed-loop, self-healing ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi‐step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root‐cause analysis, anomaly detection, and semantic log clustering. Self‐Healing Infrastructure Engineering: Experience designing closed‐loop, self‐healing ...

Senior Backend Developer (Python) - Remote

Hiring Organisation
IO
Location
Bristol, Gloucestershire, United Kingdom
Employment Type
Contract
Contract Rate
GBP 45 - 50 Hourly
Databricks Microsoft Azure Event-driven architectures (Kafka, MQTT, Azure Event Hubs, etc.) Geospatial or time-series data Observability tools such as Prometheus, Grafana or OpenTelemetry Eligibility Due to project requirements, applicants must be citizens of a NATO member country and currently reside within a NATO member country. Contract Details Fully ...

Senior Software Engineer (Infrastructure)

Location
Greater London, England, United Kingdom
Experience developing production‐ready infrastructure management tooling with either Python or Golang Familiarity with at least one of the following: Observability Tools (e.g. Prometheus, OpenTelemetry, Grafana) Databases (e.g. Postgres, DuckDB) Event Streaming platforms (e.g. Kafka) Container Orchestration (e.g. Docker, Kubernetes) Familiarity with cloud platforms such as AWS, Azure ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
working knowledge of AWS services including ECS, EC2, Lambda, VPC, IAM, Route53, CloudFront, S3, RDS Good understanding of monitoring and logging solutions. We use OpenTelemetry, AWS Cloudwatch and SigNoz so experience with them is a bonus. Basic SRE knowledge, and experience in alerting and incident management platforms (eg. incident.io, Pagerduty ...

Staff / Lead Software Engineer - Backend Platform

Location
Greater London, England, United Kingdom
Spring Boot framework internals Experience with other scripting languages (e.g. Python, Groovy, Bash) and IaC tools (e.g. Terraform). Knowledge of observability tools (e.g. OTEL, Grafana, Dynatrace etc). Experience working in financial services, investment platforms, or similar regulated organisations. Hands-on experience building and maintaining CI/CD platforms ...

Expert Forward Deployment Engineer

Location
Greater London, England, United Kingdom
failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. - Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. - Experience modernizing monolithic or legacy enterprise applications into maintain #J-18808-Ljbffr ...

Enterprise Architect - AI

Location
Greater London, England, United Kingdom
latency, cost, hallucination/quality metrics) using tools such as Arize, WhyLabs, or Langfuse. Logging & tracing: centralized logging (ELK/OpenSearch) and distributed tracing (OpenTelemetry) across data, training, and inference pipelines for end-to-end root‐cause analysis. Integration — AI Stack, Enterprise Networks & Service Provider Environments Platform integration: API-based ...

Devops SRE

Location
Greater London, England, United Kingdom
Performance Strong security mindset with a proven track record of designing secure, resilient cloud‐native systems. Experience implementing observability stacks including Prometheus , Dynatrace , and OpenTelemetry . Deep understanding of Linux internals , system performance tuning, and troubleshooting. Familiarity with Aqua Security for container runtime protection. CI/CD & Automation Tooling Hands ...

Senior DevOps Engineer

Hiring Organisation
Hackajob Ltd
Location
Leicester, Leicestershire, East Midlands, United Kingdom
Employment Type
Permanent
Salary
£70,000
scripting Networking skills/fundamentals? SQL Server/NoSQL databases Windows & Linux servers Source Control Management (Git) Docker containers Kubernetes Microservices Monitoring tool - Dynatrace, OTel, Grafana Docker Compose/Helm charts DevSecOps Tooling - SonarCloud/PrismaCloud/CrowdStrike ...

Engineer C# (Full Stack)

Location
Greater London, England, United Kingdom
collaboration skills.Desired* Experience with microservices and event-driven architectures.* Experience with AI assisted design and coding.* GraphQL, and WebSockets.* Knowledge of observability tools (e.g., OpenTelemetry, Grafana).* Familiarity with Infrastructure as Code (Terraform).* Understanding of financial markets or trading systems.* Contribution to open-source projects.* Awareness of security principles ...

Senior Software Engineer, Inference Platform

Location
Greater London, England, United Kingdom
event‐driven architectures Knowledge of GPU computing, model serving optimizations (batching, quantization, multi‐tenancy), and resource allocation Experience with observability tools (Prometheus, Grafana, OpenTelemetry) and distributed tracing Understanding of API design, rate limiting, authentication/authorization, and security best practices Exposure to AI model deployment workflows and model lifecycle management ...

Lead Observability Engineer / Senior Software Engineer

Location
Nottingham, England, United Kingdom
based culture, so you’ll have plenty of opportunity to talk, coach, and learn with many great and diverse individuals. Observability & Telemetry tools , including OTel, Grafana, and Prometheus Good understanding of programming languages used for high-performance engineering such as Golang and Java AWS , leveraging cloud-based services such ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
systems (Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline ...

AI Platform Engineer

Hiring Organisation
The Portfolio Group
Location
Manchester, United Kingdom
Employment Type
Permanent
Salary
£80000 - £100000/annum
OpenSearch. Strong Python skills. Experience with containerisation and Terraform. Solid grasp of microservices, API design, and cloud-native architecture. Observability tooling experience, such as OpenTelemetry, is a plus. Background in platform or infrastructure engineering within production environments. Track record owning cloud platforms with accountability for reliability and scale. Experience with ...