76 to 100 of 297 Remote/Hybrid Grafana Jobs

Forward Deployed Engineer - Infrastructure

Location
Greater London, England, United Kingdom
Anywhere, AWS Outposts A strong background in Go, Python or Java Experience with CockroachDB, AlloyDB, Aurora Experience with observability tools, e.g. Prometheus, Grafana Benefits Highly competitive salary Pension plan (match up to 5%) Life insurance - three times annual salary Competitive maternity (six months fully paid) and paternity leave (four weeks ...

Principal AI Quality Engineer

Hiring Organisation
Fourth
Location
London, UK
Employment Type
Full-time
leading edge. Experience with CI/CD tooling, e.g. Jenkins, Azure DevOps, and Octopus Deploy. Experience with observability tooling such as Prometheus, Grafana, and Sumo Logic. Benefitsð Holidays. We all need to rest so you get 25 basic holidays with the option to grow up to 30 with service + ...

VodafoneThree - Senior SRE

Location
Greater London, England, United Kingdom
processes. – Essential Ability to write production-quality code and automation in languages such as Python, Go, – Essential Java or similar-– Essential Experience with Datadog, Grafana, Prometheus or OpenTelemetry; product development; platform engineering/internal developer platforms; FinOps and cloud cost optimisation- Desirable What we offer We care about our people ...

Embedded DevOps Engineer

Location
Kidlington, England, United Kingdom
Familiarity with embedded Linux, cross-compilation toolchains, RTOS environments or FPGA development and deployment workflows. Experience with monitoring and observability platforms such asOpenTelemetry, Prometheus, Grafana or Loki. Knowledge of secure software supply-chain practices, including dependency scanning, artifact signing, software bills of materials, secrets management and vulnerability management. Experience administering ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
pragmatically. Strong system design fundamentals across scalability, performance, and distributed systems, including API design (REST, GraphQL). Hands‐on experience with observability tooling (Datadog, Grafana, or similar) and a data‐informed approach to system health and reliability. Solid SQL and data management skills, with an appreciation for AI-enabled, data ...

Senior Software Engineer, Web Development – AI

Location
Manchester, England, United Kingdom
good understanding of Agile practices. Ability to accurately estimate software tasks and work to schedule. Exposure to observability tools and practices (e.g., Prometheus, Grafana) are preferred. Understanding of databases and data architecture choices (SQL vs NoSQL, caching layers, and data access patterns) Track record of tuning performance and reliability ...

Senior Data Engineer

Location
Altrincham, England, United Kingdom
first‐class concerns. Ensure data quality and observability: Embed data testing and validation processes alongside monitoring and alerting systems such as Prometheus and Grafana to ensure reliability, performance, and operational insight across data services. Work with metadata and standards: Implement metadata capture and validation within data pipelines, including working with ...

Staff Cloud Security Engineer (GCP) - Engine by Starling

Location
Greater London, England, United Kingdom
native Microservice-based architecture Kubernetes (GKE on GCP, EKS on AWS) TeamCity for CI/CD (with multiple production releases per day) Terraform and Grafana RDS and CloudSQL for PostgreSQL Our Interview Process: Interviewing is a two-way process and we want you to have the time and opportunity ...

Senior Cloud Security Engineer (GCP) - Engine by Starling

Location
City Of London, England, United Kingdom
native Microservice-based architecture Kubernetes (GKE on GCP, EKS on AWS) TeamCity for CI/CD (with multiple production releases per day) Terraform and Grafana RDS and CloudSQL for PostgreSQL Our Interview Process: Interviewing is a two-way process and we want you to have the time and opportunity ...

SC Cleared AWS DevOps & Platform Engineer (Remote)

Location
England, United Kingdom
build automated pipelines with GitLab CI/CD and ArgoCD, and manage Kubernetes, Docker, Terraform and related tools while ensuring observability with Grafana/Prometheus. Strong hands-on AWS skills, container orchestration, IaC and CI/CD expertise are essential #J-18808-Ljbffr ...

Senior Backend Developer for Video on Demand and Live Streaming

Hiring Organisation
Develop
Location
Twickenham, London, United Kingdom
Employment Type
Contract
queues (Kafka, RabbitMQ). Strong DevOps experience, including: CI/CD pipelines (GitHub Actions). Containerization technologies (Docker, Kubernetes). Monitoring & logging tools (Prometheus, Grafana, ELK stack). Strong problem-solving and debugging skills. Excellent communication and collaboration abilities. Experience working in Agile development environments. Fluent written & spoken English. Details ...

Principal DevSecOps Engineer - LONDON

Hiring Organisation
83zero Limited
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Compliance: Trivy, vulnerability management, HashiCorp Vault, cert-manager * Containers & Cloud: Docker, AWS EKS, AWS IAM, S3 and network policies * Infrastructure as Code: Terraform * Observability: Grafana, Loki * Automation: Python and Bash * Experience delivering within the UK Government Digital Service (GDS) lifecycle on a public sector engagement Why join? You'll work ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Code (Terraform, Ansible, Puppet or similar) Hands-on experience with Kubernetes, Docker, and cloud platforms (AWS preferred) Experience with monitoring/observability tools (Prometheus, Grafana, ELK, APM tools) Solid understanding of system performance, scalability, and resilience Strong collaboration and communication skills within cross-functional product teams Desirable: Experience working ...

Principal DevSecOps Engineer

Hiring Organisation
83zero Limited
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
compliance - Trivy, vulnerability management, HashiCorp Vault, cert-manager * Containers & cloud - Docker, AWS EKS, AWS IAM, S3 and network policies * Infrastructure as Code - Terraform * Observability - Grafana, Loki * Automation - Python and Bash * Experience delivering within the UK Government Digital Service (GDS) lifecycle on a public sector engagement Why join? You'll work ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
focus on observability for large-scale, distributed compute or network systems. Deep, hands-on expertise building, scaling, and managing observability platforms (e.g., Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger, Honeycomb, etc.). You have proven experience using these tools to support performance analysis and debugging of complex distributed ...

Data Platform Engineer

Location
Greater London, England, United Kingdom
some GCP Warehouse & Storage: Snowflake, S3/Parquet Data & ETL: dbt, Fivetran Platform & Infra: Kubernetes, Kafka, RabbitMQ, Argo, GitHub Actions, HashiCorp Vault Observability: Datadog, Grafana Dashboarding: Preset Other: Claude What we’re looking forStrong fundamentals and the ability to apply them pragmatically: Solid programming ability (Python or similar) Strong experience ...

Staff Cloud SRE – AI/ML Platform & GPU Compute London, United Kingdom on-site

Location
Greater London, England, United Kingdom
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skills Familiarity with infrastructure-as-code (e.g. ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Senior Infrastructure Engineer

Hiring Organisation
AJ BELL BUSINESS SOLUTIONS LIMITED
Location
Salford, Greater Manchester, North West, United Kingdom
Employment Type
Permanent
FortiClient VPN or equivalent. Microsoft Windows Server and Red Hat Enterprise Linux server configuration and management. Monitoring and observability tooling such as SolarWinds, Opsgenie, Grafana, Prometheus or equivalent. Storage, backup and disaster recovery technologies, including SAN/NAS, replication, snapshotting and recovery testing concepts. Identity and access technologies such ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
analyze internet traffic patterns across multiple dimensions using flow-based tools. Experience working with alerting, monitoring and visibility tools (such as Graphite/Grafana, Prometheus, or Splunk). Knowledge across cloud hosting solutions (i.e., GCP, AWS and Azure). Knowledge of DevOps practices and CI/CD pipelines (ie. ...

AWS DevOps Engineer

Location
Bolsterstone, England, United Kingdom
maintain database change management using Liquibase or similar tooling. Implement monitoring, logging, alerting and distributed tracing using technologies such as Splunk, ELK, Prometheus and Grafana . Troubleshoot production issues, perform root cause analysis and implement solutions to improve reliability and performance. Work closely with architects, developers, infrastructure and cybersecurity teams … . Experience working within financial services or other regulated environments . Strong software engineering background combined with infrastructure experience. Experience with Splunk, ELK, Prometheus, Grafana or distributed tracing . Knowledge of DevSecOps, cloud governance and regulatory controls . Experience with non‐functional testing and continuous testing practices. Experience with messaging ...

GO Backend Developer

Hiring Organisation
itecopeople
Location
London, United Kingdom
Employment Type
Permanent
Salary
£55000 - £61000/annum OUTSIDE IR35
tested code and experience working within Agile teams. Strong communication skills and a collaborative approach to technical problem-solving. Experience with Python, Docker, Kubernetes, Grafana or OpenTelemetry would be beneficial. This is a hands-on backend development role; deep AI or LLM expertise isn't required. Benefits include 30 days ...

Remote Senior Site Reliability Engineer Manager (Remote)

Location
Cambourne, England, United Kingdom
services. Expertise in incident management, including incident response, resolution, and post-mortem analysis. Proficiency in monitoring, alerting, and observability tools such as Prometheus, Grafana, ELK stack or Datadog. Experience with cloud platforms such as AWS, Azure, or GCP, including infrastructure as code tools like Terraform or CloudFormation. Strong scripting ...

Infrastructure / DevOps Engineer

Location
Birmingham, England, United Kingdom
equivalent) Solid understanding of networking, security groups, load balancing, and DNS Experience with container orchestration (Docker, Kubernetes, or ECS) Familiarity with observability tooling (Datadog, Grafana, CloudWatch, or equivalent) Understanding of HIPAA infrastructure requirements (encryption at rest/in transit, audit trails, access controls) Nice to have Site Reliability Engineering background ...