101 to 125 of 201 Permanent OpenTelemetry Jobs

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph. Experience with every technology is not required. Who we're looking forSuccess in this role is not measured ...

Python Backend Developer

Location
Greater London, England, United Kingdom
frontend work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Staff Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
with Python; TypeScript/Node a plus.Hands-on work with agent frameworks and eval tooling.Track record of debugging distributed systems from traces; experience with OpenTelemetry’s GenAI semantic conventions a plus.Cloud-native development, automating deployments, familiarity with container systems; AWS experience a plus.Excellent time management skills, comfortable coordinating work with ...

Senior AI Engineer| London

Hiring Organisation
Infosys Technologies
Location
London, United Kingdom
Salary
£ 80 K
Kubeflow, Docker, Kubernetes).Preferred•Delivered AI projects within Agile frameworks•Experience on Gen AI Feedback Analysis, topic modelling, sentiment analysis•Knowledge of AgentOps and OpenTelemetry•Understanding of Network Security Concepts, Network Telemetry and Analytics•Understanding of Cloud computing and Virtualization•Exposure to APM/Observability tools (Dynatrace, AppDynamics, Datadog, Splunk ...

SRE Observability Technical Lead - Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
Observability Engineering, or platform infrastructure roles focused on operational telemetry. Hands‐on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms. Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high‐availability environments. Proven ability to troubleshoot integration issues ...

SRE Observability Technical Lead - Vice President

Hiring Organisation
Citigroup
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
Experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry.Hands-on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments.Proven ability to troubleshoot integration issues ...

Senior Lead SRE: Reliability, Observability & Resiliency

Location
Auchentibber, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 80 K
team members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgetsDesigns, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearchDrives the assessment … refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stabilityActively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application of leading ...

Hybrid Lead Observability Engineer | Platform & SRE

Location
United Kingdom
seeking a Lead Observability Engineer/Senior Software Engineer in Nottingham, hybrid role. You will monitor and enhance our observability estate, designing platforms with OTel, Grafana and Prometheus while collaborating across teams. You'll work with Golang/Java, AWS services (Fargate/Lambda), and Kubernetes, driving best practices, code ...

Network Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
Network Engineer in enterprise or large-scale environmentsExperience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and JenkinsExperience with Cisco and Arista switching and routing, alongside network security infrastructure, including firewalls, IDS/IPS and network segmentationExpertise in Cisco and Arista ...

Senior Software Engineer - iCloud Platform - Observability

Location
Greater London, England, United Kingdom
asynchronous network application, I/O frameworks Experience with AWS, GCP, and cloud native technologies (Containers, Kubernetes, gRPC) Experience with observability technologies such as OpenTelemetry, Prometheus, Jaeger, or similar At Apple, we're not all the same. And that's our greatest strength. We draw on the differences ...

Software Engineer, Growth

Location
Greater London, England, United Kingdom
that both humans and AI love: Frontend: React Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, Many LLM providers DevOps & Tooling: Github Actions, Grafana, OTEL, infra-as-code (Terraform) #J-18808-Ljbffr ...

Associate Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
learn enterprise platform-level tradeoffsSolid working experience with Python; TypeScript/Node a plus.Comfortable debugging services and reading logs and traces; exposure to OpenTelemetry a plus.Familiarity with cloud-native development and container systems; AWS experience a plus.Excellent time management skills and willingness to coordinate work with colleagues ...

Software Architect

Hiring Organisation
Autodesk
Location
United Kingdom
Salary
£ 70 K
Service Design: REST API design and evolution; GraphQL (including federation). Database Architecture: Relational and NoSQL (e.g., DynamoDB, PostgreSQL). Observability/SRE: OpenTelemetry, distributed tracing, metrics/SLOs for data services. Bonus: Semantic Technologies: Knowledge graphs, RDF/OWL, or property graphs — particularly relevant to the connected/semantic ...

Database Reliability Engineer

Hiring Organisation
Starling Bank
Location
London, United Kingdom
Salary
£ 80 K
ensuring rigorous data integrity and mobilityA Security & Observability Mindset: You believe security is paramount. You focus on building deep observability (Prometheus/Grafana/OpenTelemetry/Humio) and automated guardrails so the fleet is secure by design without requiring manual interventionEngineering via Code: While you are a systems expert, your ...

Senior Lead Site Reliability Engineer

Location
Glasgow, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Platform Engineer

Hiring Organisation
itecopeople
Location
London, United Kingdom
Employment Type
Permanent
Salary
£54000 - £65000/annum
Code Develop and maintain GitOps CI/CD pipelines Manage Kubernetes networking, service mesh and gateway technologies Improve platform observability using Grafana, Prometheus and OpenTelemetry Maintain platform security, resilience and automation Troubleshoot production platform issues and drive continuous improvement Work closely with architects to turn high-level designs into robust … platform engineering within production environments Terraform and Infrastructure as Code Docker, GitOps and CI/CD pipelines Linux and Bash scripting Grafana, Prometheus and OpenTelemetry Kubernetes networking and service mesh technologies Production platform operations, troubleshooting and automation You'll also be able to demonstrate: Experience owning technical implementation decisions rather ...

Lead Observability Engineer - Hybrid Cloud Telemetry

Location
United Kingdom
will design and operate the observability estate across our UK platforms, mentoring peers and shaping architecture. The role emphasizes building secure, scalable systems with OTel, Grafana, Prometheus, Golang, Java, AWS, Fargate, and Kubernetes. You will collaborate across teams in an outcome-based culture with flexible work-from-home days. #J ...

Staff / Principal Design Engineer

Location
Greater London, England, United Kingdom
both humans and AI love: Frontend: React, Tailwind Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, many LLM providers DevOps & Tooling: GitHub Actions, Grafana, OTEL, infra‐as‐code (Terraform) And always on the lookout for what’s next! We treat all candidates equally – if you’re interested please apply through ...

Staff Software Engineer – Identity and Access, Identity Squads | UK | Remote

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
working with OpenFGA and KiFederate IDP.Experience working in security critical environments.Experience contributing to or maintaining Open Source projects.Familiarity with observability tooling (e.g., Grafana, Prometheus, OpenTelemetry).Compensation & Rewards:In the UK, the Base compensation range for this role is 100,000 - 124,000. Actual compensation may vary based on level, experience ...

AVP, Observability & SRE Engineer

Location
Greater London, England, United Kingdom
Reliability Engineer - Assistant Vice President in London to drive end‐to‐end observability, migrate legacy monitoring to Google Cloud Observability and Grafana, and implement OpenTelemetry instrumentation across OpenShift/Kubernetes environments. The role emphasizes hands‐on deployment, automation (Ansible/Terraform), and collaboration with application teams to ensure resilience ...

Senior AI Infrastructure Engineer – AWS, Terraform, Multi-Region

Location
Greater London, England, United Kingdom
across regions. You’ll maintain an AWS environment with Terraform, design CI/CD pipelines using GitHub Actions and Docker, and implement observability with OpenTelemetry and Datadog to detect issues early. AI workloads demand ultra-low latency and robust autoscaling. #J-18808-Ljbffr ...

Senior Backend Engineer

Location
Manchester, England, United Kingdom
Quality & Collaboration Investigate and resolve production data issues in collaboration with Support and Engineering. Maintain strong test coverage: unit and integration tests. Improve observability (OpenTelemetry) and operational stability. Document architecture decisions and write clear, actionable PRs and commit messages. Mentor other engineers - in backend craft and in effective agent-driven ...

SRE Managing Consultant - Cloud Operating Model

Location
Manchester, England, United Kingdom
/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms.* Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision‐making.## **Security Check ...

SRE Managing Consultant - Cloud Operating Model

Location
Greater London, England, United Kingdom
/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms.* Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision‐making.## **Security Check ...