126 to 150 of 457 Datadog Jobs

Site Reliability Engineer - Frontend

Hiring Organisation
Capital On Tap
Location
London, United Kingdom
Salary
£ 80 K
great features, by owning the architecture and paved paths they build on top of.What you’ll be doing ð️Manage and automate resources in Azure, Datadog, NGINX & Cloudflare.Develop, deploy and monitor Kubernetes and Serverless resources.Build, manage, and evolve IAC using Terraform, Helm and Go CRDs.Improving systems, processes, and technologies; consulting stakeholders ...

Senior Software Engineer - Risk & FX

Hiring Organisation
Visa
Location
London, United Kingdom
Salary
£ 80 K
microservices architecture Cloud-related tools, services, and distributed system observability to support these applications, such as Docker, Kubernetes, ElasticSearch, log management systems, and Datadog APM, to name but a few API specifications, conforming to the OpenAPI (Swagger) standard, provide a clean boundary both externally between our customers and our product ...

Manager, System and Platform Operations

Location
Greater London, England, United Kingdom
understanding of networking, security, and system architecture. Proficient in scripting languages (Java, Golang, Python, Bash, or similar). Experience with monitoring and observability tools (DataDog, Prometheus, Grafana). Knowledge of database management systems (PostgreSQL, Bigtable). Understanding of API and microservices architecture. Strong people leadership skills with at least ...

Service Reliability Engineer - London

Hiring Organisation
Fitch Ratings
Location
London, United Kingdom
Salary
£ 80 K
architect and govern GitHub Actions CI/CD with quality gates, canary/blue‐green strategies, and AI‐assisted redeploy checksOwn observability in Datadog—define SLIs/SLOs, dashboards, alerting, and MS Teams integrations—and reduce incidents via telemetry-driven automation and blameless postmortems. Champion AI‐enabled operations using ...

Service Reliability Engineer - Manchester

Hiring Organisation
Fitch Ratings
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 60 K
architect and govern GitHub Actions CI/CD with quality gates, canary/blue‐green strategies, and AI‐assisted redeploy checksOwn observability in Datadog—define SLIs/SLOs, dashboards, alerting, and MS Teams integrations—and reduce incidents via telemetry-driven automation and blameless postmortems. Champion AI‐enabled operations using ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline components ...

Lead DevOps Engineer

Hiring Organisation
Collinson Group
Location
London, United Kingdom
Salary
£ 80 K
comfortable in security audits and risk conversations. You're comfortable with security tooling such as CrowdStrike and Rapid7, SIEM/SOC.Observability ownership with Datadog (or equivalent) - you've defined SLOs, built the dashboards, and set the alerting culture.Experience leading or mentoring a DevOps team - you can set direction, manage technical ...

Service Reliability Engineer - London

Hiring Organisation
Fitch Group
Location
Greater London, United Kingdom
Employment Type
Full Time
architect and govern GitHub Actions CI/CD with quality gates, canary/blue‐green strategies, and AI‐assisted redeploy checks Own observability in Datadog—define SLIs/SLOs, dashboards, alerting, and MS Teams integrations—and reduce incidents via telemetry-driven automation and blameless postmortems. Champion AI‐enabled operations using ...

Service Reliability Engineer - Manchester

Hiring Organisation
Fitch Group
Location
Manchester, United Kingdom
Employment Type
Full Time
architect and govern GitHub Actions CI/CD with quality gates, canary/blue‐green strategies, and AI‐assisted redeploy checks Own observability in Datadog—define SLIs/SLOs, dashboards, alerting, and MS Teams integrations—and reduce incidents via telemetry-driven automation and blameless postmortems. Champion AI‐enabled operations using ...

SRE Engineer

Location
Greater London, England, United Kingdom
good at coding in terraform CICD Tools hands on : Jenkins , GitHub , GitHub Actions, Cloud Deployment pipelines Observability Tools - Splunk//Graphana/Datadog and Distributed Tracing, ELF, & Dynatrace Problem-Solving : Proven ability to troubleshoot complex issues in distributed systems and debug problems effectively. #J-18808-Ljbffr ...

DevOps, AI and Automation Engineer

Location
Greater London, England, United Kingdom
DevSecOps: Azure DevOps, GitHub Actions, CheckMarx, JFrog Artifactory, Octopus Deploy* Infrastructure & Automation: Terraform, Ansible* Containers & Platforms: Docker, Kubernetes, Rancher, Azure, Windows, Linux* Observability: Datadog, Grafana, Prometheus, OpenTelemetry* AI Enablement: Azure AI Foundry, Model Context Protocol (MCP)**Expected Outcomes*** Establish scalable and secure CI/CD processes for AI-enabled platforms ...

Lead Site Reliability Engineer

Location
City of Westminster, England, United Kingdom
Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.), experience with container and container orchestration (e.g., ECS, Kubernetes, Docker ...

Platform Lead: Fintech Architecture & SRE

Location
United Kingdom
cloud provider platforms.* Strong hands-on experience with container orchestration (e.g., Kubernetes), Infrastructure-as-Code (e.g., Terraform), Python, and Observability tools (e.g., Prometheus, Grafana, Datadog, Sentry)Benefits* 26 vacation days + 2 duvet days, so you can truly recharge and enjoy life* Comprehensive health and dental care coverage* Central location ...

Platform Lead - UK

Location
United Kingdom
cloud provider platforms.* Strong hands-on experience with container orchestration (e.g., Kubernetes), Infrastructure-as-Code (e.g., Terraform), Python, and Observability tools (e.g., Prometheus, Grafana, Datadog, Sentry)Benefits* 26 vacation days + 2 duvet days, so you can truly recharge and enjoy life* Comprehensive health and dental care coverage* Central location ...

Lead Software Engineer - Java, AI

Hiring Organisation
Appcast
Location
London, UK
Terraform or AWS CloudFormationFamiliarity with containerization and orchestration technologies such as Docker and KubernetesExposure to observability and monitoring practices using tools such as Datadog, Splunk, or equivalent platformsExperience contributing to or leading architectural decisions in a large, matrixed enterprise environmentKnowledge of security best practices in cloud and application development contextsJ.P. ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.), experience with container and container orchestration (e.g., ECS, Kubernetes, Docker ...

Lead Software Engineer - Java, AI

Hiring Organisation
JP Morgan Chase
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 80 K
Terraform or AWS CloudFormationFamiliarity with containerization and orchestration technologies such as Docker and KubernetesExposure to observability and monitoring practices using tools such as Datadog, Splunk, or equivalent platformsExperience contributing to or leading architectural decisions in a large, matrixed enterprise environmentKnowledge of security best practices in cloud and application development contextsJ.P. ...

Senior Software Engineer, Full-Stack Applications (Python)

Hiring Organisation
Fitch Group
Location
Manchester, United Kingdom
Employment Type
Full Time
building interactive data applications • Advanced Data Management – Strong SQL design, query optimization, and database architecture expertise • Observability – Experience with observability patterns and tools like Datadog, distributed tracing, monitoring, and logging best practices • DevOps and Infrastructure – Familiarity with ArgoCD for GitOps and Security/Access Management (IAM federation access via Entra ...

Lead Site Reliability Engineer

Hiring Organisation
London Stock Exchange Group
Location
Nottingham, Nottinghamshire, United Kingdom
Salary
£ 70 K
with Kubernetes and containerised platformsStrong background in Linux systems administrations.Proven experience designing and operating observability platforms, including monitoring, logging, and alertingHands-on experience with Datadog for metrics, logs, APM, and alertingStrong understanding of SRE principles, including SLOs, error budgets, incident management, and reliability engineeringExperience working closely with architecture and engineering ...

Lead Site Reliability Engineer

Hiring Organisation
London Stock Exchange Group
Location
Nottingham, UK
Employment Type
Full-time
Kubernetes and containerised platformsStrong background in Linux systems administrations. Proven experience designing and operating observability platforms, including monitoring, logging, and alertingHands-on experience with Datadog for metrics, logs, APM, and alertingStrong understanding of SRE principles, including SLOs, error budgets, incident management, and reliability engineeringExperience working closely with architecture and engineering ...

Platform Engineer

Hiring Organisation
London Stock Exchange Group
Location
London, United Kingdom
Salary
£ 80 K
PracticesCI/CD tools (GitLab, GitHub, Azure DevOps, or similar)Version control systems (Git)Infrastructure automation and configuration managementMonitoring, logging, and alerting tools like Datadog, Splunk etc.Nice to HaveDevelopment experience with streaming data architectures and event-driven systemsFamiliarity with data lake/Lakehouse based big data systems and architecturesExperience ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
C++) with a bias toward automation.Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale.Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements.Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Platform Engineer

Location
Greater London, England, United Kingdom
/CD tools (GitLab, GitHub, Azure DevOps, or similar) Version control systems (Git) Infrastructure automation and configuration management Monitoring, logging, and alerting tools like Datadog, Splunk etc. Nice to Have Development experience with streaming data architectures and event-driven systems Familiarity with data lake/Lakehouse based big data systems ...

Team Lead - Platform Engineering (London)

Hiring Organisation
Fresha
Location
London, United Kingdom
Salary
£ 80 K
Infrastructure as Code (Terraform) to build repeatable, low-risk systemsPractical experience building observability systems across metrics, logs, and traces, using tools such as Datadog, Grafana, ELK, Sentry, and OpsGenieCreative problem-solving mindset, with the ability to simplify complex systems and unlock scalable solutionsComfortable operating in a fast-paced, rapidly evolving ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
with a bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. ...