1 to 25 of 115 Datadog Jobs in the UK

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
fundamentals (DNS, routing, load balancing, VPNs, firewalls). Familiarity with monitoring and logging tools (e.g. Prometheus, Grafana, ELK/EFK stack, CloudWatch, Azure Monitor, Datadog, etc.). Good understanding of security best practices in cloud and Linux environments (IAM, least privilege, secrets management, patching). Experience presenting technical concepts ...

Cloud Platform Architect (68020)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Azure DevOps for infrastructure pipelines Policy‐as‐code frameworks: OPA/Rego, HashiCorp Sentinel, Azure Policy, AWS Config Rules Monitoring and observability: Prometheus, Grafana, Datadog, CloudWatch, or Dynatrace Networking fundamentals: VPC/VNet design, load balancers, DNS, CDN, and hybrid connectivity Soft Skills & Competencies Strong leadership and mentoring ability – coaches ...

Cloud Platform Architect (68020)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Azure DevOps for infrastructure pipelines Policy‐as‐code frameworks: OPA/Rego, HashiCorp Sentinel, Azure Policy, AWS Config Rules Monitoring and observability: Prometheus, Grafana, Datadog, CloudWatch, or Dynatrace Networking fundamentals: VPC/VNet design, load balancers, DNS, CDN, and hybrid connectivity SOFT SKILLS & COMPETENCIES Strong leadership and mentoring ability – coaches ...

Principal/Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Welwyn, England, United Kingdom
chaos engineering experiments that validate systems and surface weaknesses before incidents. Build deep observability with monitoring, logging, and alerting frameworks such as Prometheus, Grafana, Datadog, and ELK. Provide technical leadership to a team of engineers, fostering collaboration, innovation, and continuous improvement. Partner across teams to align infrastructure with ...

DevOps / Cloud / Platform Engineer (All Levels) - UK Wide

Hiring Organisation
describe.me
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£50,000 - £130,000 per annum
production—deployment, scaling, networking, troubleshooting CI/CD platforms (GitHub Actions, GitLab CI, Jenkins, CircleCI, ArgoCD or equivalent) Observability stack experience (Prometheus, Grafana, Datadog, ELK, OpenTelemetry, New Relic) Scripting and automation in Bash, Python or Go Linux fundamentals, networking basics and cloud security concepts Familiarity with secret management (Vault ...

DevOps Engineer - DV Cleared

Hiring Organisation
CBSbutler Holdings Limited trading as CBSbutler
Location
Worcestershire, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £600/day
cloud (OpenStack) Containerization Docker, Podman Orchestration Kubernetes (EKS, AKS, GKE), Helm, OpenShift Version Control Git, GitLab, Bitbucket Monitoring & Logging Prometheus, Grafana, ELK Stack, Splunk, Datadog Security & Compliance HashiCorp Vault, Snyk, SonarQube, Trivy, AWS IAM, CIS Benchmarks Configuration Mgmt. Ansible, Puppet, Chef Build Tools Maven, Gradle, NPM, Webpack Testing Tools Selenium ...

Principal/Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
that validate systems and surface weaknesses before they become incidents. You build deep observability with monitoring, logging, and alerting frameworks such as Prometheus, Grafana, Datadog, and ELK. You provide technical leadership to a team of engineers, fostering collaboration, innovation, and continuous improvement. You partner across teams to align infrastructure with ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Bash automation GitHub and CI/CD pipelines AWS and cloud‐native infrastructure Kubernetes/Amazon EKS Grafana stack, Prometheus, Loki or Datadog JFrog Artifactory or artifact management Docker or container runtime experience AWS Batch, Step Functions, IAM and Karpenter Hybrid cloud platform migration or modernisation Secure platform design ...

SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cloud teams to embed reliability into infrastructure and deployment pipelines TECHNICAL SKILLS & EXPERTISE Expert-level observability: Prometheus, Grafana, ELK/OpenSearch, Jaeger/Zipkin, Datadog, or Dynatrace Strong experience with AIOps and ML-driven monitoring: PagerDuty, Moogsoft, BigPanda, or custom ML pipelines Deep knowledge of FMEA, fault tree analysis ...

Senior Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £65000/annum
cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.) Experience with container and container orchestration (e.g., ECS, Kubernetes ...

Site Reliability Engineer - Comcast Technology Solutions

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
with infrastructure‐as‐code tools (e.g., Terraform, Ansible). Familiarity with containerization and orchestration tools (e.g., Docker, Kubernetes). Experience with monitoring tools (e.g., Datadog, Splunk). Experience with database performance monitoring and tuning (e.g., NoSQL, SQL). Experience with Kubernetes performance monitoring and tuning. Excellent problem‐solving skills ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Experience owning, managing, and maintaining mission‐critical operational tooling. Desirable: Proven background in implementing and managing centralised logging solutions or similar platforms (e.g., Splunk, DataDog). Desirable: Familiarity with distributed tracing tools (e.g., Jaeger, Zipkin) and Application Performance Monitoring (APM) solutions. What we offer Competitive salary ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum It Recruitment Limited
Location
Southampton, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Birmingham, West Midlands, West Midlands (County), United Kingdom
Employment Type
Permanent
cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure ...

MLOps Engineer

Hiring Organisation
Jobleads-UK
Location
United Kingdom
focus on maintainable, scalable, and production‐quality software. Knowledge of AI system security, model governance, compliance, monitoring, and observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry. Experience with FastAPI, Databricks, Snowflake, SRE practices, or cloud security certifications is considered an advantage. Benefits Competitive salary and equity package. Comprehensive healthcare ...

Architect & Delivery Lead (68018)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
delivery models Financial modelling: FinOps, TCO analysis, business case development, and outcome‐based commercial models Tooling breadth: ITSM (ServiceNow), monitoring (Grafana/ELK/Datadog/Dynatrace), automation (Ansible/Terraform), and collaboration platforms Soft Skills & Competencies Exceptional leadership presence — inspires confidence at C‐level while earning respect from engineering ...

SVP of Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
engineering, function calling, agent frameworks) and graph/knowledge graph technologies. DevOps/SRE practices at scale: CI/CD, IaC (Terraform, Pulumi), observability (Datadog, Grafana), incident management. Leadership Qualities Builder mentality with hands-on orientation; executive presence and strong communication skills; collaborative and bias for action; comfort with ambiguity. ...

Platform Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Salary
£80,000
Secrets Operator, Vault, or similar secret sync patterns is beneficial Familiarity with Terraform, multi cloud environments (GCP and Azure), observability tooling (e.g. New Relic, Datadog, Prometheus, Grafana), or contract testing with Pact would be valuable What you can expect from us We won't just meet your expectations. ...

Sr. Software Engineer - Data Platform (London, Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
multiple cloud providers (AWS, GCP, Azure, OCI) Knowledge of data serialization formats (Avro, Protobuf, Parquet) and schema management Experience with observability platforms (Prometheus, Grafana, Datadog, Splunk) Understanding of data governance, security, and compliance requirements Contributions to open-source streaming or distributed systems projects Experience with CI/CD pipelines ...

AI Platform engineer

Hiring Organisation
Nextech Group Limited
Location
East London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
Infra: AWS (ECS, Lambda, SQS/SNS), Docker, Kubernetes, Terraform AI/ML tooling: LangChain/LlamaIndex, vLLM, Anthropic & OpenAI APIs, embedding models Observability: Datadog, Grafana, OpenTelemetry CI/CD: GitHub Actions, ArgoCD Requirements: 4+ years backend development experience, ideally with at least 1 year working with LLM/ ...

Head of Infrastructure & Security

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ecosystems and Kubernetes orchestration. You must be an expert systems administrator with hands‐on experience in Infrastructure as Code (Terraform) and modern observability tools (Datadog, Prometheus). Problem‐Solving: You are a data‐driven strategist who can devise high‐level strategy and is equally comfortable rolling up your sleeves ...

Principal SRE Engineer / Grafana Specialist - (Outside IR35)

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £750 per day + Outside IR35
10+ years in engineering roles, with at least 5 years in SRE, Observability, or DevOps functions. Hands-on proficiency with observability tools such as Datadog, Grafana, Prometheus, OpenTelemetry. Strong knowledge of distributed systems, microservices, and container orchestration (Kubernetes, Docker). Experience with automation and Infrastructure as Code (Terraform, Ansible ...

Lead DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
with ArgoCD or FluxCD for progressive delivery, drift correction, and multi‐environment releases. Building containerized, serverless, or event‐driven systems with strong observability using DataDog, Splunk, or OpenTelemetry. Strengthening platform security through Vault‐based secret management, least privilege access, and compliance automation. Designing CI/CD workflows that include SAST ...

Staff Implementation Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
reliability & incident response (incl. AI‐assisted): incident management and postmortems (PagerDuty, Opsgenie, incident.io, Rootly, FireHydrant, Blameless); AIOps/AI SRE (Resolve.ai, Cleric, Traversal, Neubird, Datadog Bits AI, BigPanda, Moogsoft). Observability: Datadog, Prometheus/Grafana, New Relic, Splunk, Dynatrace, Honeycomb, OpenTelemetry. Containers, cloud & foundations: Docker and Kubernetes (Helm, manifests, operators ...