51 to 75 of 213 OpenTelemetry Jobs

Lead Observability Engineer / Senior Software Engineer

Location
United Kingdom
based culture, so you’ll have plenty of opportunity to talk, coach, and learn with many great and diverse individuals. Observability & Telemetry tools , including OTel, Grafana, and Prometheus Good understanding of programming languages used for high-performance engineering such as Golang and Java AWS , leveraging cloud-based services such ...

Lead Observability Engineer / Senior Software Engineer

Location
Nottingham, England, United Kingdom
based culture, so you’ll have plenty of opportunity to talk, coach, and learn with many great and diverse individuals. Observability & Telemetry tools , including OTel, Grafana, and Prometheus Good understanding of programming languages used for high-performance engineering such as Golang and Java AWS , leveraging cloud-based services such ...

Principal Site Reliability Engineer

Location
Greater London, England, United Kingdom
GitLab Pipelines, ArgoCD, Octopus Deploy Data - ElasticSearch hosted with Kubernetes Operator, PostgreSQL, SQL Server, BigQuery Monitoring and Security - Splunk, Grafana/Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry AI Tools - Claude, Amazon Bedrock, Gemini, Vertex AI Key Responsibilities Design, implement, and operate scalable, reliable, and secure infrastructure across ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
bias toward automation.Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale.Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements.Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform) and secure cloud production ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform) and secure ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
systems (Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline ...

UK Sr Software Engineer

Location
United Kingdom
Infrastructure automation technologies such as Kubernetes, Helm and Terraform. Developing software using cloud services such as AWS, GCP and Azure. Monitoring tools such as OpenTelemetry and Grafana Build automations such as Gradle, Maven and Github Actions Building APIs using REST Attachments (1) UK Sr Software Engineer job description.pdf #J ...

Lead DevOps Engineer - Real Time Platform

Location
Greater London, England, United Kingdom
Aurora, VPC, IAM, CloudFront, WAF, and others as required). CI/CD: GitLab CI, Terraform, AWS CLI. Observability: CloudWatch, Datadog, OpenTelemetry Languages: Python, Bash, TypeScript. AI Development: Agentic AI tooling (e.g. GitHub Copilot, Claude Code). WHAT YOU'LL BRING Strong experience designing and operating cloud‐native infrastructure ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Location
Greater London, England, United Kingdom
/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cilium). Demonstrated success ...

Director of Software Engineering (AIOps) - Executive Director

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 120 K
Platforms (example products IBM Netcool, (Watson CloudPak AiOps), Tivoli, SCOM, SMARTS, Dynatrace, Splunk, Elastic, Prometheus, Grafana or Messaging systems for telemetry transport like Kafka, OTEL) Formal training or certification on software engineering concepts and expert applied experience. In addition, advanced experience leading technologists to manage, anticipate and solve complex technical ...

Director of Software Engineering (AIOps) - Executive Director

Location
City Of London, England, United Kingdom
Platforms (example products IBM Netcool, (Watson CloudPak AiOps), Tivoli, SCOM, SMARTS, Dynatrace, Splunk, Elastic, Prometheus, Grafana or Messaging systems for telemetry transport like Kafka, OTEL) Formal training or certification on software engineering concepts and expert applied experience. In addition, advanced experience leading technologists to manage, anticipate and solve complex technical ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 100 K
knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps).Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL).Expertise in performance optimisation of distributed systems (e.g., caching, network latency).Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium).Demonstrated success in Mentorship ...

Staff Software Engineer - International Pricing

Hiring Organisation
Marks & Spencer
Location
United Kingdom
Salary
£ 80 K
Messaging: Kafka/Confluent, IBM MQ/MQFTE and SFTP.Databases: MongoDB and SQL Server on Azure.API & Integration: Apigee, REST APIs and Windows Services.Observability: Dynatrace, OpenTelemetry, PagerDuty and Helix.What’s in it for you Working at M&S means being part of something bigger – helping to deliver quality, value and service ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Location
London, United Kingdom
ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium). Demonstrated ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium). Demonstrated ...

Staff Software Engineer - Commercial Planning

Hiring Organisation
Marks & Spencer
Location
United Kingdom
Salary
£ 60 K
Python (desirable).Cloud & Infrastructure: Azure IaaS and PaaS services, Docker, Kubernetes, Helm and Terraform.Data & Messaging: Kafka/Confluent.Databases: MongoDB and Azure SQL Server.Observability: Dynatrace, OpenTelemetry and PagerDuty.What’s in it for you Working at M&S means being part of something bigger – helping to deliver quality, value and service ...

Staff Software Engineer - Commercial Planning

Location
Greater London, England, United Kingdom
Infrastructure: Azure IaaS and PaaS services, Docker, Kubernetes, Helm and Terraform. Data & Messaging: Kafka/Confluent. Databases: MongoDB and Azure SQL Server. Observability: Dynatrace, OpenTelemetry and PagerDuty. What’s in it for you? Working at M&S means being part of something bigger – helping to deliver quality, value and service ...

Senior Software Engineer – Live & VOD Video Infrastructure

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
technologiesExperience with GPU-accelerated encoding or hardware media pipelinesFamiliarity with Kubernetes, ECS, Nomad, or other orchestration platformsExperience with observability stacks such as Prometheus, Grafana, OpenTelemetry, ELK, or DatadogExperience building fault-tolerant ingest or transcoding platforms operating across multiple regions#LI-JC5What's Roku's approach to hybrid working Roku fosters ...

Platform Engineer

Location
Greater London, England, United Kingdom
datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest, SAST/DAST Developer Portal & Self … cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow‐up actions ...

Site Reliability Engineer III

Hiring Organisation
CME- Group
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
Google Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.Incident Response & Operations: Engage with urgency … cross-functional teams, coupled with an eagerness to learn independently and collaboratively.Preferred Qualifications/DesirableObservability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified ...

Site Reliability Engineer III

Location
Belfast City District, Northern Ireland, United Kingdom
Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement. Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection. Incident Response & Operations: Engage with … teams, coupled with an eagerness to learn independently and collaboratively. Preferred Qualifications/Desirable Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana. Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles. Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator ...

Senior Kotlin Developer

Location
Greater London, England, United Kingdom
years of Java development experience. Nice to have Kafka, Solace, or TIBCO EMS. SQL Server, MongoDB, and S3/Object Storage. Kubernetes and OpenShift. OpenTelemetry, ELK Stack, Grafana. HashiCorp Vault and CyberArk. Prime Brokerage, Synthetic Swaps, or Cash Equities domain experience. Regulatory reporting exposure (EMIR, MiFID, 871(m)). Experience ...

Senior Full Stack Engineer (Kotlin + React)

Location
Greater London, England, United Kingdom
independently in complex production environments. Nice to have - Kafka, Solace, or TIBCO EMS.- SQL Server, MongoDB, and S3/Object Storage.- Kubernetes and OpenShift.- OpenTelemetry, ELK Stack, Grafana.- HashiCorp Vault and CyberArk.- Prime Brokerage, Synthetic Swaps, or Cash Equities domain experience.- Regulatory reporting exposure (EMIR, MiFID, 871(m)).- Experience ...

Technical Architect

Hiring Organisation
CBSbutler Holdings Limited trading as CBSbutler
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£700 - £800/day
/Next.js | TypeScript/Node.js | Azure Functions | Azure App Services | Cosmos DB | Contentful | Algolia | Talon.One | Adyen | Cloudflare | GitHub CI/CD | Infrastructure as Code | OpenTelemetry | Splunk Equivalent technologies may also be considered where the candidate has strong composable commerce experience. What We're Looking For We're looking for someone ...

Lead Network Operations Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
with automation and CI/CD tooling (e.g. Ansible, Python, Jenkins) and Infrastructure as Code Experience with observability and monitoring tools (e.g. Prometheus, Grafana, OpenTelemetry, ELK) Strong ability to analyse and troubleshoot distributed systems end-to-end Proactive, self-starting mindset with a strong sense of ownership Comfortable engaging with ...