851 to 875 of 1,512 Observability Jobs

Director, Solutions Engineering Splunk UKI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
region. Experience working across multiple customer segments (Enterprise, Public Sector, Service Provider, Commercial). Strong domain expertise in enterprise software (e.g., Cybersecurity, Observability, Cloud & AI, IT Operations, Application Performance Management, or Big Data). Exceptional communication and articulation skills; ability to translate complex technical ideas into clear business value ...

DWIT Technical Architect

Hiring Organisation
Hays
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Contract
Contract Rate
Up to £700.0 per day + up to 700 per day inside IR35
delivery, operations and governance teams.* Supporting the definition of target-state AWS transformation approaches for existing feeds, including data movement, processing, integration, security, resilience, observability and support considerations.* Contributing to decommissioning strategies for legacy services by documenting dependencies, migration sequencing, risks, parallel running needs, cutover considerations and retirement evidence.* Using ...

Senior Software Engineer - Marketing Platform

Hiring Organisation
Wise
Location
Greater London, United Kingdom
Employment Type
Full Time
with Java or Kotlin (or another JVM language). Experience with Spring (or similar JVM frameworks). Experience operating data-heavy systems with strong observability (monitoring, alerting, data quality checks). Experience with measurement/attribution, marketing platforms, or comparison products (not required). Fintech or regulated-environment experience ...

Solutions Architect

Hiring Organisation
Genesis10
Location
Columbus, Ohio, United States
Employment Type
Permanent
Salary
USD 1,000 Annual
system design for transaction processing Horizontal scalability and throughput optimization Multi-region deployments and failover strategies Experience implementing: Disaster recovery (DR) and business continuity Observability (monitoring, tracing, alerting) in payment environments Working knowledge of: PCI-DSS , KYC/AML, FFIEC, NACHA rules Experience integrating: Fraud detection and risk scoring systems ...

Sr./Staff Software Engineer (Data Team)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
batch and streaming processing patterns, data quality, and schema evolution. Workflow orchestration at scale - designing and operating multi-step automated pipelines with retries, observability, and graceful failure handling. Distributed systems fundamentals - you understand how things break at scale: eventual consistency, idempotency, backpressure, job scheduling, and failure modes in distributed compute ...

Kotlin Engineer Contractor (GenieConnect)

Hiring Organisation
Jobleads-UK
Location
United Kingdom
coding practices (Claude Code, Copilot, Cursor) so a small team delivers more than its size suggests. Keep the lights on: CI/CD, testing, observability, AWS. Small team, so you’ll touch everything. Qualifications and Skills The codebase is Kotlin Multiplatform and that’s where we want the leverage ...

Technical Program Manager - Compute Systems Engineering

Hiring Organisation
Jobleads-UK
Location
United Kingdom
based on both business requirements and engineering needs. Building and operating reliable, comprehensive deployment automation and validation processes. Ensuring that our systems provide sufficient observability across the entire stack. Supporting customers with complex technical issues that prevent them from fully utilizing Nebius Cloud GPU and CPU capacity. We participate ...

Engineering Manager - Stream Aligned

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
that owns delivery end to end, from discovery through to production operation. Strong understanding of modern software delivery practices, including CI/CD, testing, observability, deployability, and operational readiness. Evidence of improving delivery flow, DORA metrics, release confidence, incident response, or sustainable engineering practices. Experience partnering closely with Product Managers ...

Software Engineer, Codex Enterprise

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
user experience. Are proficient in one or more backend languages (e.g., Python, Go, Rust) and distributed systems concepts, with a focus on reliability, observability, and security. Enjoy building cross‐cutting platform capabilities that unlock product velocity, and you’re comfortable working across services, APIs, end‐user product surfaces, and extensibility ...

Software Engineer, Agents & Automations

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
engineer on this team, you will help build the core platform that makes this possible: the workflow builder, execution engine, integrations, debugging tools, observability, evaluation systems, and feedback loops that help customers understand and improve what they deploy. This is broad product engineering work across frontend, backend, and AI-powered ...

Identity & Access Lead - BPL

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
bound strictly to our core ecosystem. To be successful in this role, you will need the following: Extensive experience of Identity & Access Management (IAM) Observability Pipeline experience partnering with peers and CISO Experience of Joiner-Mover-Leaver pipeline creation and innovation Some other highly valued skills may include: Tailscale – ability ...

Google AI FDE

Hiring Organisation
Tata Consultancy Services
Location
City of London, London, United Kingdom
between Google’s AI products and customer's live infrastructure, including APIs, legacy data silos, and security perimeters. Build high-performance evaluation pipelines and observability frameworks to ensure agentic systems meet rigorous requirements for accuracy, safety, and latency. Identify repeatable field patterns and technical friction points in Google ...

Senior Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
code quality, the maturity of our agentic SDLC, and individual engineer performance. Set and enforce a high bar for engineering standards code review, testing, observability, and architectural review and hold every team to it. Agentic and Spec‐Driven Development Champion agentic and spec‐driven development practices across every team. Roll ...

Senior Software Engineer, Inference Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
workloads Optimize inference pipelines for latency, throughput, batching efficiency, and resource utilization Design fault‐tolerant systems with graceful degradation and automatic recovery mechanisms Observability & Engineering Excellence Build high‐performance telemetry and observability stack for inference metrics, performance tracking, and debugging Implement comprehensive monitoring for model latency, throughput, error rates … RabbitMQ), databases (PostgreSQL, Redis), and event‐driven architectures Knowledge of GPU computing, model serving optimizations (batching, quantization, multi‐tenancy), and resource allocation Experience with observability tools (Prometheus, Grafana, OpenTelemetry) and distributed tracing Understanding of API design, rate limiting, authentication/authorization, and security best practices Exposure to AI model deployment ...

Principal SRE Engineer / Grafana Specialist - (Outside IR35)

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £750 per day + Outside IR35
Lead SRE/Observability Engineering Lead - (Outside IR35 Contract/Remote) Location: Bristol/London HQ - Largely Remote (Occasional Travel) Day Rate: Outside IR35 - £700 p/d Duration: 3-6 Months Initial - with intention to extend Payment Terms: Monthly Our client is a FTSE100 Wealth/Asset Management firm … seeking to engage a Lead SRE Engineer (Observability SME) to support the implementation and instrumentation of their new Observability solution. This role will be critical in delivering against our Digital OKRs by embedding observability best practices, frameworks, and tooling across digital platforms and engineering teams. Key Responsibilities: Strategy & Roadmap: Define ...

Senior DevOps & Infrastructure Engineer

Hiring Organisation
Jobleads-UK
Location
United Kingdom
monitoring, release management, and platform operations. Help define practical, secure, and scalable ways of using AI in DevOps processes, whilemaintainingstrong engineering standards and governance. Observability, Reliability & Continuous Improvement Improve monitoring, alerting, logging, and observability practices to help teams detect issues earlier and resolve incidents faster. Analyze platform performance, deployment quality ...

AIOps Product Lead: Observability & Automation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
S&P Global, Inc. is seeking an experienced AIOps Product Manager to own the enterprise AIOps roadmap, translating complex operational needs into concrete product requirements and backlog items. You will lead high-impact capabilities like ...

AWS SRE - Datadog

Hiring Organisation
Vallum Associates
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£475 - £500/day
role in our technology team, contributing to the development, deployment, and maintenance of our monitoring and alerting infrastructure. Key Responsibilities Architect, implement, and maintain observability platforms using Datadog and Geneos to ensure comprehensive monitoring and alerting. Design and manage scalable infrastructure using Terraform and Infrastructure as Code (IaC) principles. Champion … actionable insights. Mentor junior engineers and contribute to the evolution of SRE best practices. Oversee and continuously optimize cloud cost management strategies for our Observability infrastructure in line with Cloud FinOps principles. Collaborate with executive leadership and cross-functional stakeholders to align infrastructure strategy with long-term business objectives. Lead ...

GenAI Engineer - SRE

Hiring Organisation
Dimension Consulting
Location
Phoenix, Arizona, United States
Employment Type
Permanent
Salary
USD 50 Annual
Health, and Engineering Productivity initiatives. The candidate will leverage Generative AI technologies, automation frameworks, and cloud-native tooling to improve operational efficiency, incident reduction, observability, code quality, and developer productivity across enterprise platforms. The role requires close collaboration with SRE teams, platform engineering teams, application development teams, and business stakeholders … Code Monitoring tools such as Splunk, Dynatrace, Prometheus, Grafana, Datadog, New Relic SRE Knowledge Incident Management Problem Management Service Reliability Availability Management Operational Excellence Observability Principles Production Support ...

Senior Consultant Snowflake

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ecosystem. The ideal candidate will combine deep technical expertise with strong delivery and people leadership capabilities to drive platform reliability, operational excellence, governance, automation, observability, and stakeholder management. The role requires hands‐on expertise in at least two platform technologies, with mandatory expertise in either Snowflake or Confluent Kafka. … KPIs, operational metrics, and customer commitments. Proactively manage risks, dependencies, and technical blockers. Drive compliance, security, audit readiness, and cost optimization. Implement effective monitoring, observability, and service management processes. Hands‐on experience with Snowflake, Kafka, Cloud, DevOps, and Platform engineering. Mentor and guide Snowflake, Kafka, Cloud, DevOps, and Platform engineers. ...

Senior DevOps Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
support cloud-native platforms across AWS and Azure Implement DevSecOps and Infrastructure as Code best practices Drive platform reliability using SRE principles and observability tools Support incident management and continuous improvement initiatives Implement Terraform-based infrastructure solutions Leverage automation and AI-assisted engineering tools to improve delivery efficiency Support … delivery teams What We're Looking For in a Senior DevOps Engineer Hands-on DevSecOps and platform engineering expertise Advanced Terraform knowledge Experience with observability tools such as Dynatrace, Grafana or similar Understanding of Site Reliability Engineering principles Experience supporting production environments and incident management Strong stakeholder management and communication ...

DevSecOps Automation Lead London, United Kingdom SMA DevSecOps Posted 3 hours ago

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
automation, and deployment issues, and contribute to root cause analysis and service improvement actions.* Contribute to documentation, engineering standards, and best practices for automation, observability, and platform operations.## **Who you are*** Hands-on experience with Elastic Stack (Elasticsearch, Logstash, Kibana, Beats) or similar observability platforms.* Strong scripting or development skills ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT
Location
Yorkshire, United Kingdom
Employment Type
Permanent
Salary
GBP 65,000 - 75,000 Annual
improvements across platform reliability, automation and infrastructure as code. Lead the implementation of CI/CD best practices to improve software delivery. Enhance monitoring, observability and incident management across cloud environments. Collaborate with engineering teams to improve performance, resilience and operational efficiency. Mentor and coach engineers, promoting SRE and DevOps … . Experience building and maintaining CI/CD pipelines. Knowledge of containerisation technologies such as Kubernetes, Amazon EKS or ECS. Experience with monitoring and observability tooling such as Grafana, Prometheus, OpenSearch or similar. Strong understanding of cloud security, resilience and infrastructure automation. Previous experience mentoring engineers or providing technical leadership. ...

Site Reliability Engineer (DV Security Clearance)

Hiring Organisation
CGI
Location
Gloucestershire, United Kingdom
Employment Type
Full Time
mission-critical services supporting national security programmes. You will work closely with software engineers, platform teams and stakeholders to automate operational processes, improve observability and enhance system reliability. In this role, you will take ownership of service health, contribute to platform evolution and help create scalable solutions that enable teams … effectiveness. Key responsibilities: ~Improve & Enhance service reliability, availability and operational resilience ~Automate & Optimise infrastructure, operational workflows and platform processes ~Design & Implement monitoring, alerting and observability solutions ~Support & Scale Kubernetes and containerised environments ~Develop & Deliver CI/CD pipelines and deployment automation capabilities ~Investigate & Resolve incidents, conducting root cause analysis ...

Java Engineer

Hiring Organisation
Response Informatics
Location
Belfast, County Antrim, Northern Ireland, United Kingdom
Employment Type
Contract
Contract Rate
From £250 to £300 per day
delivering APIs using Java, Go, or both. Strong understanding of cloud-native application architecture, including containerized services, service-to-service communication, configuration management, observability, and resilience patterns. Experience delivering production services deployed on Kubernetes. Experience building highly available, scalable, and reliable distributed systems. Strong knowledge of API design principles, including … Helm, Operators, Custom Resource Definitions, or GitOps workflows. Experience with distributed systems, stateful services, or high-volume transactional platforms. Experience with cloud-native observability tools such as Prometheus, Grafana, OpenTelemetry, ELK/OpenSearch, or similar. Experience implementing security best practices for APIs and cloud-native services, including OAuth2, mTLS, secrets ...