101 to 125 of 206 OpenTelemetry Jobs

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
across our hybrid and cloud-native environments.* Innovate & Automate: Spearhead the evaluation, selection, and implementation of cutting-edge observability tools and platforms (e.g., Dynatrace, OpenTelemetry, Prometheus, Grafana). Architect and build robust, automated observability pipelines. Take an active part in documenting and defining processes and best practice.* Optimize & Analyze: Conduct … large-scale monitoring and observability solutions.* Expert-Level Tooling: Deep expertise with modern observability platforms (e.g., Dynatrace, AWS Cloudwatch, Prometheus, Grafana, ELK Stack, Splunk, OpenTelemetry).* Cloud & Infrastructure: Advanced knowledge of major cloud platforms (AWS, Azure, GCP), containerization (Docker, Kubernetes), and Infrastructure as Code (Terraform, Ansible).* Programming & Automation: Strong ...

typescript developer for authentication platforms

Location
Greater London, England, United Kingdom
GitHub Actions, and writing unit and integration tests with Jest Familiarity with security principles including IAM, encryption and networking, alongside observability tools such as OpenTelemetry, Honeycomb or Grafana Nice to have: Exposure to identity or MFA platforms such as Auth0 or Transmit Security, knowledge of microservices architecture, API gateways such ...

Site Reliability Engineer

Location
West of England, England, United Kingdom
systems integration Version-controlled automation and operational tooling Experience with any of the following would be particularly useful: ServiceNow, Halo, Jira Service Management, OpenTelemetry, distributed tracing, Slack/Teams automation, datacentre or colocation environments, GPU infrastructure, DCIM, IPAM, virtualisation platforms or LLM-assisted operational automation. This ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 60 K
data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph.Experience with every technology is not required.Who we’re looking forSuccess in this role is not measured by the number ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph. Experience with every technology is not required. Who we're looking forSuccess in this role is not measured ...

Python Backend Developer

Location
Greater London, England, United Kingdom
frontend work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Staff Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
with Python; TypeScript/Node a plus.Hands-on work with agent frameworks and eval tooling.Track record of debugging distributed systems from traces; experience with OpenTelemetry’s GenAI semantic conventions a plus.Cloud-native development, automating deployments, familiarity with container systems; AWS experience a plus.Excellent time management skills, comfortable coordinating work with ...

Senior AI Engineer| London

Hiring Organisation
Infosys Technologies
Location
London, United Kingdom
Salary
£ 80 K
Kubeflow, Docker, Kubernetes).Preferred•Delivered AI projects within Agile frameworks•Experience on Gen AI Feedback Analysis, topic modelling, sentiment analysis•Knowledge of AgentOps and OpenTelemetry•Understanding of Network Security Concepts, Network Telemetry and Analytics•Understanding of Cloud computing and Virtualization•Exposure to APM/Observability tools (Dynatrace, AppDynamics, Datadog, Splunk ...

SRE Observability Technical Lead - Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
Observability Engineering, or platform infrastructure roles focused on operational telemetry. Hands‐on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms. Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high‐availability environments. Proven ability to troubleshoot integration issues ...

SRE Observability Technical Lead - Vice President

Hiring Organisation
Citigroup
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
Experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry.Hands-on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments.Proven ability to troubleshoot integration issues ...

Senior Lead SRE: Reliability, Observability & Resiliency

Location
Auchentibber, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 80 K
team members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgetsDesigns, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearchDrives the assessment … refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stabilityActively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application of leading ...

Hybrid Lead Observability Engineer | Platform & SRE

Location
United Kingdom
seeking a Lead Observability Engineer/Senior Software Engineer in Nottingham, hybrid role. You will monitor and enhance our observability estate, designing platforms with OTel, Grafana and Prometheus while collaborating across teams. You'll work with Golang/Java, AWS services (Fargate/Lambda), and Kubernetes, driving best practices, code ...

Network Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
Network Engineer in enterprise or large-scale environmentsExperience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and JenkinsExperience with Cisco and Arista switching and routing, alongside network security infrastructure, including firewalls, IDS/IPS and network segmentationExpertise in Cisco and Arista ...

Senior Software Engineer - iCloud Platform - Observability

Location
Greater London, England, United Kingdom
asynchronous network application, I/O frameworks Experience with AWS, GCP, and cloud native technologies (Containers, Kubernetes, gRPC) Experience with observability technologies such as OpenTelemetry, Prometheus, Jaeger, or similar At Apple, we're not all the same. And that's our greatest strength. We draw on the differences ...

Software Engineer, Growth

Location
Greater London, England, United Kingdom
that both humans and AI love: Frontend: React Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, Many LLM providers DevOps & Tooling: Github Actions, Grafana, OTEL, infra-as-code (Terraform) #J-18808-Ljbffr ...

Associate Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
learn enterprise platform-level tradeoffsSolid working experience with Python; TypeScript/Node a plus.Comfortable debugging services and reading logs and traces; exposure to OpenTelemetry a plus.Familiarity with cloud-native development and container systems; AWS experience a plus.Excellent time management skills and willingness to coordinate work with colleagues ...

Software Architect

Hiring Organisation
Autodesk
Location
United Kingdom
Salary
£ 70 K
Service Design: REST API design and evolution; GraphQL (including federation). Database Architecture: Relational and NoSQL (e.g., DynamoDB, PostgreSQL). Observability/SRE: OpenTelemetry, distributed tracing, metrics/SLOs for data services. Bonus: Semantic Technologies: Knowledge graphs, RDF/OWL, or property graphs — particularly relevant to the connected/semantic ...

Database Reliability Engineer

Hiring Organisation
Starling Bank
Location
London, United Kingdom
Salary
£ 80 K
ensuring rigorous data integrity and mobilityA Security & Observability Mindset: You believe security is paramount. You focus on building deep observability (Prometheus/Grafana/OpenTelemetry/Humio) and automated guardrails so the fleet is secure by design without requiring manual interventionEngineering via Code: While you are a systems expert, your ...

Senior Lead Site Reliability Engineer

Location
Glasgow, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Platform Engineer

Hiring Organisation
itecopeople
Location
London, United Kingdom
Employment Type
Permanent
Salary
£54000 - £65000/annum
Code Develop and maintain GitOps CI/CD pipelines Manage Kubernetes networking, service mesh and gateway technologies Improve platform observability using Grafana, Prometheus and OpenTelemetry Maintain platform security, resilience and automation Troubleshoot production platform issues and drive continuous improvement Work closely with architects to turn high-level designs into robust … platform engineering within production environments Terraform and Infrastructure as Code Docker, GitOps and CI/CD pipelines Linux and Bash scripting Grafana, Prometheus and OpenTelemetry Kubernetes networking and service mesh technologies Production platform operations, troubleshooting and automation You'll also be able to demonstrate: Experience owning technical implementation decisions rather ...

Lead Observability Engineer - Hybrid Cloud Telemetry

Location
United Kingdom
will design and operate the observability estate across our UK platforms, mentoring peers and shaping architecture. The role emphasizes building secure, scalable systems with OTel, Grafana, Prometheus, Golang, Java, AWS, Fargate, and Kubernetes. You will collaborate across teams in an outcome-based culture with flexible work-from-home days. #J ...

Staff / Principal Design Engineer

Location
Greater London, England, United Kingdom
both humans and AI love: Frontend: React, Tailwind Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, many LLM providers DevOps & Tooling: GitHub Actions, Grafana, OTEL, infra‐as‐code (Terraform) And always on the lookout for what’s next! We treat all candidates equally – if you’re interested please apply through ...

Staff Software Engineer – Identity and Access, Identity Squads | UK | Remote

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
working with OpenFGA and KiFederate IDP.Experience working in security critical environments.Experience contributing to or maintaining Open Source projects.Familiarity with observability tooling (e.g., Grafana, Prometheus, OpenTelemetry).Compensation & Rewards:In the UK, the Base compensation range for this role is 100,000 - 124,000. Actual compensation may vary based on level, experience ...

AVP, Observability & SRE Engineer

Location
Greater London, England, United Kingdom
Reliability Engineer - Assistant Vice President in London to drive end‐to‐end observability, migrate legacy monitoring to Google Cloud Observability and Grafana, and implement OpenTelemetry instrumentation across OpenShift/Kubernetes environments. The role emphasizes hands‐on deployment, automation (Ansible/Terraform), and collaboration with application teams to ensure resilience ...