26 to 50 of 288 Datadog Jobs in the UK

Senior Solutions Engineer

Location
United Kingdom
experience with: Cloud platforms (AWS, Azure, or GCP) Infrastructure as Code (Terraform, CloudFormation, etc.) Containers and orchestration (Docker, Kubernetes) Monitoring and observability platforms (Datadog, Prometheus, Grafana, etc.) Deep understanding of release engineering and modern cloud‐native architectures. Strong automation and scripting skills (Python, Bash, etc.). Experience contributing to architectural ...

DevOps Engineer

Location
Greater London, England, United Kingdom
related tooling. Manage and improve Kubernetes platforms, GitOps workflows, Helm charts, and application onboarding. Develop monitoring, logging, alerting, and observability capabilities using Grafana, Prometheus, DataDog, Loki, and CloudWatch. Support cloud networking and connectivity across AWS accounts, regions, on-prem environments, and other cloud platforms we use. Implement and maintain secure ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
Azure), specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures ...

Lead Dev Ops Engineer

Location
Luton, England, United Kingdom
architect secure, performant, and highly available cloud solutions. Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence. Demonstrated leadership skills in managing complex, high‐pressure situations and guiding teams through incident resolution. Exceptional ...

Lead Dev Ops Engineer

Location
Greater London, England, United Kingdom
architect secure, performant, and highly available cloud solutions. Proficiency with monitoring and log analytics tools such as AWS CloudWatch, ELK Stack, Prometheus, Grafana, Datadog, or New Relic, to maintain observability and ensure operational excellence. Demonstrated leadership skills in managing complex, high‐pressure situations and guiding teams through incident resolution. Exceptional ...

Software Engineer III – DevOps & AWS

Location
Glasgow, Scotland, United Kingdom
resolve them. Maintain and monitor asset inventory across environments. Monitor, troubleshoot, and remediate issues using Splunk and observability/monitoring platforms such as Datadog, Dynatrace, or Grafana. Support cost rationalization efforts; partner with architects to gather, clarify, and translate technical requirements. Proficient with engineering workflow tools such as Git/ ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Azure), specifically building and operating highly resilient cloud‐native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures ...

Strategic DevSecOps Consultant

Hiring Organisation
CloudBees
Location
London, UK
Employment Type
Full-time
enabled software development, agentic workflows, large language models (LLMs), or AI governance practices. Experience with observability and telemetry platforms such as OpenTelemetry, Splunk, Dynatrace, Datadog, AppDynamics, Grafana, or similar technologies. Experience working with large-scale enterprise architecture, governance, compliance, and regulated environments. Thought leadership experience through technical publications, conference speaking ...

Lead Software Engineer - Cloud/Java-Python/AI-ML

Hiring Organisation
JP Morgan Chase
Location
Bournemouth, Dorset, UK
Employment Type
Full-time
serverless architectures. Experience with Git, CI/CD pipelines, automated testing frameworks and DevOps practices. Experience with monitoring and logging tools like CloudWatch, Dynatrace, Datadog, Prometheus, Grafana, or ELK stack. Strong understanding of cloud security best practices. Proficiency in data warehousing solutions like Snowflake and Iceberg and relational databases such ...

Senior Site Reliability Engineer

Hiring Organisation
Carta
Location
London, UK
Employment Type
Full-time
service mesh is a big plus. Monitoring and Observability: Strong knowledge of monitoring tools and practices, such as Prometheus, Grafana, ELK Stack, or Datadog, and the ability to set up and maintain comprehensive monitoring solutions. Software Development: Proficiency in Python, with the ability to write efficient, maintainable, and scalable code. ...

SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Location
Greater London, England, United Kingdom
Cloud teams to embed reliability into infrastructure and deployment pipelines TECHNICAL SKILLS & EXPERTISE Expert-level observability: Prometheus, Grafana, ELK/OpenSearch, Jaeger/Zipkin, Datadog, or Dynatrace Strong experience with AIOps and ML-driven monitoring: PagerDuty, Moogsoft, BigPanda, or custom ML pipelines Deep knowledge of FMEA, fault tree analysis ...

Senior Cloud Engineer, AI Platform SRE

Location
Leeds, England, United Kingdom
practice around it. Pipelines: Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing ...

Senior Cloud Engineer, AI Platform SRE

Location
Manchester, England, United Kingdom
practice around it. Pipelines: Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing ...

Senior Cloud Engineer, AI Platform SRE

Location
City of Edinburgh, Scotland, United Kingdom
practice around it. Pipelines: Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing ...

Senior Cloud Engineer, AI Platform SRE

Location
Greater London, England, United Kingdom
practice around it. Pipelines: Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing ...

AWS Cloud Engineer

Location
Belfast City District, Northern Ireland, United Kingdom
TypeScript Experience with GraphQL and REST APIs Experience running production workloads on EKS Experience with serverless technologies such as AWS Lambda Experience with Datadog, Prometheus, Grafana or New Relic Exposure to distributed systems and event‐driven architectures Experience working within consulting or client‐facing environments Familiarity with GenAI, machine learning ...

Senior Lead SRE: Reliability, Observability & Resiliency

Location
Auchentibber, Scotland, United Kingdom
proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.) Experience with container and container orchestration (e.g., ECS, Kubernetes ...

Senior Lead Software Engineer - Python / Go

Location
Auchentibber, Scotland, United Kingdom
software development best practices. Proficiency with cloud infrastructure provisioning tools (Terraform, KRO, Crossplane, etc.). Experience with logging and monitoring tools (Splunk, Grafana, Datadog, Prometheus, etc.). Deep understanding of cloud infrastructure design, architecture, and migration strategies. Proficiency with AI-assisted coding workflows, including LLM-powered development tools, spec-driven ...

Senior Lead Software Engineer - Python / Go

Location
Glasgow, Scotland, United Kingdom
software development best practices. Proficiency with cloud infrastructure provisioning tools (Terraform, KRO, Crossplane, etc.). Experience with logging and monitoring tools (Splunk, Grafana, Datadog, Prometheus, etc.). Deep understanding of cloud infrastructure design, architecture, and migration strategies. Proficiency with AI‐assisted coding workflows, including LLM‐powered development tools, spec‐driven ...

Full Stack Engineer, Platform Reliability

Location
East Midlands, England, United Kingdom
Smart Desirable Skills Experience working to formal incident management and SLA frameworks, and improving them Experience with observability and monitoring tooling (CloudWatch, Grafana, Datadog, or similar) Experience picking up an unfamiliar codebase and making it dependable Experience in rail testing, NDT, or sensor-based inspection industries (ultrasound, eddy current, electromagnetic ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.)Experience with container and container orchestration (e.g., ECS, Kubernetes ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
Bash automation GitHub and CI/CD pipelines AWS and cloud‐native infrastructure Kubernetes/Amazon EKS Grafana stack, Prometheus, Loki or Datadog JFrog Artifactory or artifact management Docker or container runtime experience AWS Batch, Step Functions, IAM and Karpenter Hybrid cloud platform migration or modernisation Secure platform design ...

Senior Lead Platform Engineer - Python

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
understanding of software development best practices. Proficiency with cloud infrastructure provisioning tools (Terraform, KRO, Crossplane, etc.).Experience with logging and monitoring tools (Splunk, Grafana, Datadog, Prometheus, etc.).Deep understanding of cloud infrastructure design, architecture, and migration strategies. Proficiency with AI-assisted coding workflows, including LLM-powered development tools, spec-driven ...

Lead DevOps Engineer (AWS/Snowflake)

Location
Greater London, England, United Kingdom
data engineering teamso Understanding of ELT/ETL pipelines and data workflows Monitoring & Operationso Experience with monitoring and observability tools (e.g., Prometheus, Grafana, Datadog, or cloud-native services)o Strong troubleshooting and incident management skills Securityo Knowledge of cloud security best practices, IAM, and secrets managemento Awareness of compliance requirements ...

Infrastructure / DevOps Engineer

Location
Birmingham, England, United Kingdom
equivalent) Solid understanding of networking, security groups, load balancing, and DNS Experience with container orchestration (Docker, Kubernetes, or ECS) Familiarity with observability tooling (Datadog, Grafana, CloudWatch, or equivalent) Understanding of HIPAA infrastructure requirements (encryption at rest/in transit, audit trails, access controls) Nice to have Site Reliability Engineering background ...