101 to 125 of 213 OpenTelemetry Jobs

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
across our hybrid and cloud-native environments.* Innovate & Automate: Spearhead the evaluation, selection, and implementation of cutting-edge observability tools and platforms (e.g., Dynatrace, OpenTelemetry, Prometheus, Grafana). Architect and build robust, automated observability pipelines. Take an active part in documenting and defining processes and best practice.* Optimize & Analyze: Conduct … large-scale monitoring and observability solutions.* Expert-Level Tooling: Deep expertise with modern observability platforms (e.g., Dynatrace, AWS Cloudwatch, Prometheus, Grafana, ELK Stack, Splunk, OpenTelemetry).* Cloud & Infrastructure: Advanced knowledge of major cloud platforms (AWS, Azure, GCP), containerization (Docker, Kubernetes), and Infrastructure as Code (Terraform, Ansible).* Programming & Automation: Strong ...

Site Reliability Engineer

Location
United Kingdom
delivery lifecycles. An understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge ...

Network Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
Network Engineer in enterprise or large-scale environmentsExperience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and JenkinsExperience with Cisco and Arista switching and routing, alongside network security infrastructure, including firewalls, IDS/IPS and network segmentationExpertise in Cisco and Arista ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 60 K
data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph.Experience with every technology is not required.Who we’re looking forSuccess in this role is not measured by the number ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph. Experience with every technology is not required. Who we're looking forSuccess in this role is not measured ...

Lead AI Engineer

Location
Greater London, England, United Kingdom
experience using frameworks like Autogen and LangGraph . Solid grounding in MLOps , containerisation (Docker, Kubernetes), and vector databases. Understanding of agent monitoring tools (Langfuse, OpenTelemetry). Strong software engineering best practices (testing, CI/CD, code reviews). Excellent communicator able to work with cross‐functional teams and clients. Desire ...

Python Backend Developer

Location
Greater London, England, United Kingdom
frontend work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Senior AI Engineer| London

Hiring Organisation
Infosys Technologies
Location
London, United Kingdom
Salary
£ 80 K
Kubeflow, Docker, Kubernetes).Preferred•Delivered AI projects within Agile frameworks•Experience on Gen AI Feedback Analysis, topic modelling, sentiment analysis•Knowledge of AgentOps and OpenTelemetry•Understanding of Network Security Concepts, Network Telemetry and Analytics•Understanding of Cloud computing and Virtualization•Exposure to APM/Observability tools (Dynatrace, AppDynamics, Datadog, Splunk ...

Senior Backend Engineer - ClickStack

Hiring Organisation
ClickHouse
Location
United Kingdom
Salary
£ 70 K
Points:Experience with ClickHouse, PostgreSQL, or other analytical/OLTP databases.Strong opinions on observability tools and a vision for making them 10x better.Familiarity with OpenTelemetry, logging pipelines, or metrics infrastructure.Experience with SDKs or client library design.Exposure to message queues, streaming systems (Kafka, NATS), or event-driven architectures.CompensationFor roles based ...

Staff Software Engineer - International Pricing

Location
Greater London, England, United Kingdom
Confluent, IBM MQ/MQFTE and SFTP. Databases: MongoDB and SQL Server on Azure. API & Integration: Apigee, REST APIs and Windows Services. Observability: Dynatrace, OpenTelemetry, PagerDuty and Helix. What’s in it for you? Working at M&S means being part of something bigger – helping to deliver quality, value ...

Database Reliability Engineer

Hiring Organisation
Starling Bank
Location
London, United Kingdom
Salary
£ 80 K
ensuring rigorous data integrity and mobilityA Security & Observability Mindset: You believe security is paramount. You focus on building deep observability (Prometheus/Grafana/OpenTelemetry/Humio) and automated guardrails so the fleet is secure by design without requiring manual interventionEngineering via Code: While you are a systems expert, your ...

Senior Lead SRE: Reliability, Observability & Resiliency

Location
Auchentibber, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 80 K
team members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgetsDesigns, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearchDrives the assessment … refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stabilityActively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application of leading ...

Site Reliability Engineer

Location
West of England, England, United Kingdom
systems integration Version-controlled automation and operational tooling Experience with any of the following would be particularly useful: ServiceNow, Halo, Jira Service Management, OpenTelemetry, distributed tracing, Slack/Teams automation, datacentre or colocation environments, GPU infrastructure, DCIM, IPAM, virtualisation platforms or LLM-assisted operational automation. This ...

Staff Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
with Python; TypeScript/Node a plus.Hands-on work with agent frameworks and eval tooling.Track record of debugging distributed systems from traces; experience with OpenTelemetry’s GenAI semantic conventions a plus.Cloud-native development, automating deployments, familiarity with container systems; AWS experience a plus.Excellent time management skills, comfortable coordinating work with ...

Staff Software Engineer - AI Agents (Satori)

Location
Belfast City District, Northern Ireland, United Kingdom
TypeScript/Node a plus. Hands‐on work with agent frameworks and eval tooling. Track record of debugging distributed systems from traces; experience with OpenTelemetry’s GenAI semantic conventions a plus. Cloud‐native development, automating deployments, familiarity with container systems; AWS experience a plus. Excellent time management skills, comfortable coordinating ...

Software Engineer - AI Agents (Satori)

Location
Belfast City District, Northern Ireland, United Kingdom
level usage at enterprise scale. Development experience with Python; TypeScript/Node a plus. Track record of debugging distributed systems from traces; experience with OpenTelemetry's GenAI semantic conventions a plus. Cloud-native development, automating deployments, familiarity with container systems; AWS experience a plus. Excellent time management skills, comfortable coordinating ...

Software Architect

Hiring Organisation
Autodesk
Location
United Kingdom
Salary
£ 70 K
Service Design: REST API design and evolution; GraphQL (including federation). Database Architecture: Relational and NoSQL (e.g., DynamoDB, PostgreSQL). Observability/SRE: OpenTelemetry, distributed tracing, metrics/SLOs for data services. Bonus: Semantic Technologies: Knowledge graphs, RDF/OWL, or property graphs — particularly relevant to the connected/semantic ...

Senior Lead Site Reliability Engineer

Location
Glasgow, Scotland, United Kingdom
members and stakeholders to define comprehensive service level indicators, service level objectives, and error budgets Designs, implements, and maintains operational reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives … assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system stability Actively contributes to the engineering community as an advocate of firmwide frameworks, tools, and practices, and influences peers and project decision-makers to consider the use and application ...

Platform Engineer

Hiring Organisation
itecopeople
Location
London, United Kingdom
Employment Type
Permanent
Salary
£54000 - £65000/annum
Code Develop and maintain GitOps CI/CD pipelines Manage Kubernetes networking, service mesh and gateway technologies Improve platform observability using Grafana, Prometheus and OpenTelemetry Maintain platform security, resilience and automation Troubleshoot production platform issues and drive continuous improvement Work closely with architects to turn high-level designs into robust … platform engineering within production environments Terraform and Infrastructure as Code Docker, GitOps and CI/CD pipelines Linux and Bash scripting Grafana, Prometheus and OpenTelemetry Kubernetes networking and service mesh technologies Production platform operations, troubleshooting and automation You'll also be able to demonstrate: Experience owning technical implementation decisions rather ...

Hybrid Lead Observability Engineer | Platform & SRE

Location
United Kingdom
seeking a Lead Observability Engineer/Senior Software Engineer in Nottingham, hybrid role. You will monitor and enhance our observability estate, designing platforms with OTel, Grafana and Prometheus while collaborating across teams. You'll work with Golang/Java, AWS services (Fargate/Lambda), and Kubernetes, driving best practices, code ...

Staff / Principal Design Engineer

Location
Greater London, England, United Kingdom
both humans and AI love: Frontend: React, Tailwind Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, many LLM providers DevOps & Tooling: GitHub Actions, Grafana, OTEL, infra‐as‐code (Terraform) And always on the lookout for what’s next! We treat all candidates equally – if you’re interested please apply through ...

Senior Software Engineer - iCloud Platform - Observability

Location
Greater London, England, United Kingdom
asynchronous network application, I/O frameworks Experience with AWS, GCP, and cloud native technologies (Containers, Kubernetes, gRPC) Experience with observability technologies such as OpenTelemetry, Prometheus, Jaeger, or similar At Apple, we're not all the same. And that's our greatest strength. We draw on the differences ...

Software Engineer, Growth

Location
Greater London, England, United Kingdom
that both humans and AI love: Frontend: React Backend: Golang and Rust Cloud: Cloudflare, GCP, AWS, Many LLM providers DevOps & Tooling: Github Actions, Grafana, OTEL, infra-as-code (Terraform) #J-18808-Ljbffr ...

Associate Software Engineer - AI Agents (Satori)

Hiring Organisation
Proofpoint
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
learn enterprise platform-level tradeoffsSolid working experience with Python; TypeScript/Node a plus.Comfortable debugging services and reading logs and traces; exposure to OpenTelemetry a plus.Familiarity with cloud-native development and container systems; AWS experience a plus.Excellent time management skills and willingness to coordinate work with colleagues ...