1 to 25 of 91 Permanent Datadog Jobs

Network Engineer

Hiring Organisation
Armor Defense Inc
Location
Plano, Texas, United States
Employment Type
Permanent
Salary
USD Annual
peering, Transit Gateway, and Virtual WAN hub-and-spoke architectures, and cross-region or cross-account patterns. Experience with network observability and monitoring tools (Datadog, vRealize Network Insight, Grafana, Prometheus, ELK Stack). Understanding of Disaster Recovery (DR) and Business Continuity Planning (BCP) strategies for networking. Experience with firewall platforms ...

DevOps / Cloud / Platform Engineer (All Levels) - UK Wide

Hiring Organisation
describe.me
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£50,000 - £130,000 per annum
production—deployment, scaling, networking, troubleshooting CI/CD platforms (GitHub Actions, GitLab CI, Jenkins, CircleCI, ArgoCD or equivalent) Observability stack experience (Prometheus, Grafana, Datadog, ELK, OpenTelemetry, New Relic) Scripting and automation in Bash, Python or Go Linux fundamentals, networking basics and cloud security concepts Familiarity with secret management (Vault ...

Senior Systems Engineer, Production

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
scale. Strong expertise with Terraform and infrastructure automation. Solid understanding of AWS services such as ECS/EKS, RDS, Lambda. Familiarity with observability platforms (Datadog, Prometheus, Grafana, etc.). Proficiency in containerization (Docker, Kubernetes). Experience with CI/CD systems such as Buildkite, GitHub Actions, etc. Familiarity with Linux ...

Senior Site Reliability Engineer - JA London

Hiring Organisation
Spectrum It Recruitment Limited
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.) Experience with container and container orchestration (e.g., ECS, Kubernetes ...

Senior Systems Engineer, Production

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
infrastructure at scale .Strong expertise with Terraform and infrastructure automation.Solid understanding of AWS services such as ECS/EKS, RDS, LambdaFamiliarity with observability platforms (Datadog, Prometheus, Grafana, etc.).Proficiency in containerization (Docker, Kubernetes).Experience with CI/CD systems such as Buildkite, GitHub Actions, etc.Familiarity with Linux systems administration , networking ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Experience owning, managing, and maintaining mission‐critical operational tooling. Desirable: Proven background in implementing and managing centralised logging solutions or similar platforms (e.g., Splunk, DataDog). Desirable: Familiarity with distributed tracing tools (e.g., Jaeger, Zipkin) and Application Performance Monitoring (APM) solutions. What we offer Competitive salary ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum It Recruitment Limited
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent, Work From Home
cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure ...

Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
Basingstoke, Hampshire, United Kingdom
Employment Type
Permanent
reliability objectives (SLIs/SLOs). Reduce alert fatigue through continuous tuning and optimisation. Build and maintain dashboards using technologies such as: Grafana Prometheus Datadog Splunk AWS CloudWatch Reliability Engineering & Automation Automate repetitive operational tasks to minimise manual effort. Improve Mean Time to Detect (MTTD) and Mean Time to Resolve ...

Principal Site Reliability Engineer (SRE)

Hiring Organisation
INFINITE CHOICE LLC
Location
Dallas, Texas, United States
Employment Type
Permanent
Salary
USD Annual
cost optimization and resource management Technical Skills Strong programming skills in Python, Go, Java, or similar languages Experience with monitoring tools (Prometheus, Grafana, Datadog, New Relic, or similar) Proficiency with containerization (Docker, Kubernetes) and orchestration platforms Knowledge of CI/CD pipelines, automated testing, and deployment strategies Understanding of database ...

SVP of Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
engineering, function calling, agent frameworks) and graph/knowledge graph technologies. DevOps/SRE practices at scale: CI/CD, IaC (Terraform, Pulumi), observability (Datadog, Grafana), incident management. Leadership Qualities Builder mentality with hands-on orientation; executive presence and strong communication skills; collaborative and bias for action; comfort with ambiguity. ...

Staff Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
builders. Bonus Points Deep experience with Google Cloud Platform (GCP) services and tools. Expert-level knowledge of modern observability platforms (e.g., Prometheus, Grafana, Datadog, OpenTelemetry). Experience designing and building reliable systems capable of handling high throughput and low latency. Significant experience with Go and Terraform. Familiarity with working ...

AI Platform engineer

Hiring Organisation
Nextech Group Limited
Location
East London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
Infra: AWS (ECS, Lambda, SQS/SNS), Docker, Kubernetes, Terraform AI/ML tooling: LangChain/LlamaIndex, vLLM, Anthropic & OpenAI APIs, embedding models Observability: Datadog, Grafana, OpenTelemetry CI/CD: GitHub Actions, ArgoCD Requirements: 4+ years backend development experience, ideally with at least 1 year working with LLM/ ...

Head of Infrastructure & Security

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ecosystems and Kubernetes orchestration. You must be an expert systems administrator with hands‐on experience in Infrastructure as Code (Terraform) and modern observability tools (Datadog, Prometheus). Problem‐Solving: You are a data‐driven strategist who can devise high‐level strategy and is equally comfortable rolling up your sleeves ...

Senior DevSecOps Engineer

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Infrastructure as Code: OpenTofu, Terragrunt, CloudFormation Security & Identity: Microsoft Entra, AWS IAM, OIDC, secrets management, policy‐as‐code Observability: Centralised logging, metrics, tracing (e.g. Datadog, OpenTelemetry) Platform Automation: Declarative configuration and infrastructure management Internal Tooling: Developer‐facing tools and services built with Python, Go, and modern frontend frameworks Version Control ...

Staff Implementation Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
reliability & incident response (incl. AI‐assisted): incident management and postmortems (PagerDuty, Opsgenie, incident.io, Rootly, FireHydrant, Blameless); AIOps/AI SRE (Resolve.ai, Cleric, Traversal, Neubird, Datadog Bits AI, BigPanda, Moogsoft). Observability: Datadog, Prometheus/Grafana, New Relic, Splunk, Dynatrace, Honeycomb, OpenTelemetry. Containers, cloud & foundations: Docker and Kubernetes (Helm, manifests, operators ...

DevOps Engineer

Hiring Organisation
Noir
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£60,000 - £75,000 per annum
pipelines, and be confident scripting in Python, C# or similar scripting languages. You'll also be comfortable working with monitoring and performance tools like Datadog or Prometheus, and ideally, you'll have worked in a fast-moving SaaS or product-led business before. Bonus points if you've helped shape ...

Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools such as Jenkins, GitLab, Terraform, etc. Experience with container and container orchestration such ...

Principal Engineer, CoinDesk Data Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
investment platforms. Containerization & Deployment: Proficiency with containerization technologies such as Docker or Kubernetes. Observability: Hands-on experience with modern observability tooling (e.g., Prometheus, DataDog, Jaeger, OpenTelemetry). Data Governance: Experience with data privacy (GDPR/CCPA) and security compliance in a regulated financial environment. Please note you will need ...

Senior Site Reliability Engineer - UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
preferred). Working knowledge of Kubernetes and containerised workloads. Infrastructure as Code experience (Terraform or similar). Familiarity with monitoring and alerting tools (Datadog, Prometheus, etc). Scripting or automation experience (Python, Bash, or similar). Nice to have: Experience leading incidents or mentoring others during on‐call. Experience ...

Principal Engineer - Java - £110,000

Hiring Organisation
IntecSelect
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
using tools like Apache Kafka, Snowflake, or Databricks. Monitoring and Performance Tuning: Implement advanced monitoring and observability solutions using tools like Prometheus, Grafana, or Datadog to proactively identify and resolve performance bottlenecks. Code and System Optimisation: Proactively analyse and optimise existing systems for improved performance, scalability, and maintainability. Core skill ...

Principal Engineer - Java - £110,000

Hiring Organisation
Intec Select Ltd
Location
London, Broad Street, United Kingdom
Employment Type
Permanent
Salary
£100000 - £110000/annum Hybrid + 25% bonus
using tools like Apache Kafka, Snowflake, or Databricks. Monitoring and Performance Tuning: Implement advanced monitoring and observability solutions using tools like Prometheus, Grafana, or Datadog to proactively identify and resolve performance bottlenecks. Code and System Optimisation: Proactively analyse and optimise existing systems for improved performance, scalability, and maintainability. Core skill ...

Principal Engineer (Payments)

Hiring Organisation
Intec Select Ltd
Location
Wolverhampton, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£90000 - £115000/annum Hybrid + 25% bonus
using tools like Apache Kafka, Snowflake, or Databricks. Monitoring and Performance Tuning: Implement advanced monitoring and observability solutions using tools like Prometheus, Grafana, or Datadog to proactively identify and resolve performance bottlenecks. Code and System Optimisation: Proactively analyse and optimise existing systems for improved performance, scalability, and maintainability. Core skill ...

Principal Engineer (Payments)

Hiring Organisation
Intec Select Ltd
Location
London, Fitzrovia, United Kingdom
Employment Type
Permanent
Salary
£90000 - £115000/annum Hybrid + 25% bonus
using tools like Apache Kafka, Snowflake, or Databricks. Monitoring and Performance Tuning: Implement advanced monitoring and observability solutions using tools like Prometheus, Grafana, or Datadog to proactively identify and resolve performance bottlenecks. Code and System Optimisation: Proactively analyse and optimise existing systems for improved performance, scalability, and maintainability. Core skill ...

Site Reliability Engineer (SRE) - Cloud & Automation

Hiring Organisation
Spencer Rose Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 60,000 - 70,000 Annual
environment, multi-region cloud platforms (AWS or GCP), using IaC and GitOps workflows. Hands-on experience with observability/APM tooling such as Grafana, Datadog or Dynatrace. Background working in regulated financial services or banking environments. Excellent troubleshooting, analytical and communication skills, able to work effectively with both technical ...