26 to 50 of 299 OpenTelemetry Jobs in England

Site Reliability Engineer (SRE)

Location
Cambridge, England, United Kingdom
Software development experience(ideally working with and as a .NET developer) Strong understanding of SDLC, microservice and HA architecture Observability - NewRelic, ELK, Grafana, PagerDuty, OTEL or similar Experience with Kubernetes clusters in production setting, AWS, IOC Experience with operational tasks Knowledge of CI-CD tooling Jenkins, Gitlab, GitHub, ArgoCD ...

Lead Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
Southampton, Hampshire, South East, United Kingdom
Employment Type
Permanent
Salary
£85,000
AKS. Extensive experience in platform engineering, cloud provisioning and observability. Strong monitoring, alerting and dashboarding experience using technologies such as: Azure Monitor, Grafana, Prometheus, OpenTelemetry, Elasticsearch Experience creating custom metrics, queries, dashboards and alerts for microservices. Advanced scripting or software development skills using PowerShell, Python, C# or a comparable language. ...

Lead Software Engineer

Location
Greater London, England, United Kingdom
customer support requirements. Experience with Kafka and event-driven architectures. Experience with OAuth2, OpenID Connect, SAML, and Keycloak. Experience with Datadog, Grafana, Elastic, OpenTelemetry, or similar observability tooling. Experience implementing AI-assisted software development practices. WSD is an employer that values diversity.We highly encourage applications from appropriately qualified and eligible ...

Lead Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
Northam, Devon, UK
AKS. Extensive experience in platform engineering, cloud provisioning and observability. Strong monitoring, alerting and dashboarding experience using technologies such as: Azure Monitor, Grafana, Prometheus, OpenTelemetry, Elasticsearch Experience creating custom metrics, queries, dashboards and alerts for microservices. Advanced scripting or software development skills using PowerShell, Python, C# or a comparable language. ...

Database Platform Engineer

Location
Greater London, England, United Kingdom
services relevant to data platforms such as RDS, Aurora, S3, EC2 or EKS Familiarity with modern observability stacks such as Prometheus, Grafana, Elk or OTel Desirable: experience with cloud‐native and distributed SQL databases such as Aurora, YugabyteDB or TiDB Desirable: knowledge of data streaming and integration tools such ...

Senior Software Engineer in Test (SET) New London

Location
Greater London, England, United Kingdom
experience with infrastructure-as-code (Terraform), Kubernetes,cloud‐native platforms (AWS), service meshes, CI/CD (e.g. GitHub Actions) and observability tooling (e.g. OpenTelemetry, Grafana) Strong programming skills in a modern backend language (e.g. Kotlin, Java, Go, or Python) with a test automation mindsetm Familiarity with resilience patterns, chaos testing ...

Senior Forward Deployment Engineer

Location
Slough, England, United Kingdom
failover, and production‐recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain #J-18808-Ljbffr ...

Senior Forward Deployment Engineer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain OtherLanguages English: C1 Advanced Seniority Senior London ...

Expert Forward Deployment Engineer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain OtherLanguages English: C1 Advanced Seniority Senior London ...

SC Cleated DevOps Engineer

Location
England, United Kingdom
environments. Python: Comfortable using Python for scripting, automation and supporting platform engineering activities. Monitoring & Observability: Experience working with tools such as Grafana, Prometheus and OpenTelemetry to monitor, troubleshoot and improve the reliability of cloud and platform environments. Development Practices: Familiarity with modern software development practices, including Conventional Commits, version control ...

Site Reliability Engineer

Hiring Organisation
E-Solutions IT Services UK Ltd
Location
Leeds, West Yorkshire, United Kingdom
Employment Type
Full-Time
Salary
£280.00 - £300.00 per day
Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux). ...

Platform Engineer – Monitoring, Observability & SIEM (MONSO)

Location
Greater London, England, United Kingdom
pipelines, code reviews, and self-service enablement. Modern Observability & Telemetry: Strong background in Splunk (SPL, dashboards, alerts, data ingestion, forwarders) and also configuring OpenTelemetry collectors and pipelines; knowledge of Prometheus and Grafana or similar tools. Kubernetes (Power User): Strong, hands-on experience deploying and operating workloads, stateful appli-cations, Helm ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
systems (Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job Responsibilities Design and build reusable CI/CD templates, pipeline ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Manchester, England, United Kingdom
keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Greater London, England, United Kingdom
keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing ...

Devops SRE

Location
Greater London, England, United Kingdom
Performance Strong security mindset with a proven track record of designing secure, resilient cloud‐native systems. Experience implementing observability stacks including Prometheus , Dynatrace , and OpenTelemetry . Deep understanding of Linux internals , system performance tuning, and troubleshooting. Familiarity with Aqua Security for container runtime protection. CI/CD & Automation Tooling Hands ...

Engineer C# (Full Stack)

Location
City Of London, England, United Kingdom
with microservices and event‐driven architectures. Experience with AI‐assisted design and coding. Knowledge of GraphQL and WebSockets. Familiarity with observability tools such as OpenTelemetry and Grafana. Experience with Infrastructure as Code (Terraform). Understanding of financial markets or trading systems. Contribution to open‐source projects. Awareness of security principles ...

Infrastructure Engineer

Location
Greater London, England, United Kingdom
Experience developing production-ready infrastructure management tooling with either Python or Golang Familiarity with at least one of the following: Observability Tools (e.g. Prometheus, OpenTelemetry, Grafana) Databases (e.g. Postgres, DuckDB) Event Streaming platforms (e.g. Kafka) Container Orchestration (e.g. Docker, Kubernetes) Familiarity with cloud platforms such as AWS, Azure ...

Engineer C# (Full Stack)

Location
Greater London, England, United Kingdom
collaboration skills.Desired* Experience with microservices and event-driven architectures.* Experience with AI assisted design and coding.* GraphQL, and WebSockets.* Knowledge of observability tools (e.g., OpenTelemetry, Grafana).* Familiarity with Infrastructure as Code (Terraform).* Understanding of financial markets or trading systems.* Contribution to open-source projects.* Awareness of security principles ...

Platform Engineer

Location
Manchester, England, United Kingdom
Kubernetes/EKS, Terraform, Helm and GitHub Actions. Experience with service discovery, secrets management, networking, monitoring or scripting. Observability tooling such as Honeycomb, OpenTelemetry, Prometheus, Splunk. Pragmatic use of AI-assisted engineering tools. #LI-AP1 At Zopa we value flexible ways of working. We value face-to-face collaboration ...

Lead Site Reliability Engineer

Location
Nottingham, England, United Kingdom
Docker). Proficiency in CI/CD pipelines and infrastructure‐as‐code tools (Terraform, GitHub Actions, Jenkins). Familiarity with observability platforms (Datadog, BigPanda, OpenTelemetry). Experience working with identity platforms and/or fraud detection systems. Excellent communication and stakeholder management skills; ability to influence across technical and business ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium). Demonstrated ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
Hackajob Ltd
Location
Westminster, Greater London, UK
ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium). Demonstrated ...

Staff Cloud SRE – AI/ML Platform & GPU Compute London, United Kingdom on-site

Location
Greater London, England, United Kingdom
toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skills Familiarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer

Location
Greater London, England, United Kingdom
GitLab Pipelines, ArgoCD, Octopus Deploy Data - ElasticSearch hosted with Kubernetes Operator, PostgreSQL, SQL Server, BigQuery Monitoring and Security - Splunk, Grafana/Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry AI Tools - Claude, Amazon Bedrock, Gemini, Vertex AI Key Responsibilities Design, implement, and operate scalable, reliable, and secure infrastructure across ...