101 to 125 of 360 Grafana Jobs in London

Enterprise Architect - AI

Location
Greater London, England, United Kingdom
gates, and promotion pipelines (continuous training/continuous delivery) that move models safely from experimentation to production. Infrastructure & GPU observability: NVIDIA DCGM, Prometheus/Grafana, and related telemetry stacks for GPU utilization, thermal, and cluster health monitoring. Model & LLM observability: production model performance monitoring, data/concept drift detection ...

Enterprise Architect - AI

Hiring Organisation
World Wide Technology
Location
London, UK
Employment Type
Full-time
gates, and promotion pipelines (continuous training/continuous delivery) that move models safely from experimentation to production. Infrastructure & GPU observability: NVIDIA DCGM, Prometheus/Grafana, and related telemetry stacks for GPU utilization, thermal, and cluster health monitoring. Model & LLM observability: production model performance monitoring, data/concept drift detection ...

Head of Infrastructure – Market making

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 120 K
identity (AD/Entra ID)CI/CD (GitLab CI or similar)IaC and automation (CloudFormation, Ansible, Terraform)Monitoring/Observability (CloudWatch, Prometheus, Grafana)Scripting (Python, Bash)This is a fantastic opportunity for a Senior Infrastructure Engineer with some light management and leadership skills to step into a Head role ...

Senior Infrastructure Engineer – Trading Technology

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 100 K
identity (AD/Entra ID)CI/CD (GitLab CI or similar)IaC and automation (CloudFormation, Ansible, Terraform)Monitoring/Observability (CloudWatch, Prometheus, Grafana)Scripting (Python, Bash)This is a fantastic opportunity for a Senior Infrastructure Engineer with some light management and leadership skills to step into a Head role ...

Lead Infrastructure Engineer

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 100 K
identity (AD/Entra ID)CI/CD (GitLab CI or similar)IaC and automation (CloudFormation, Ansible, Terraform)Monitoring/Observability (CloudWatch, Prometheus, Grafana)Scripting (Python, Bash)This is a fantastic opportunity for a Senior Infrastructure Engineer with some light management and leadership skills to step into a Head role ...

Senior Python Developer - up to £90,000 + Bonus - Hybrid

Hiring Organisation
Involved Solutions
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £90,000 per annum
RabbitMQ or Kafka and event driven architectures OpenStack experience Infrastructure as Code knowledge Test automation experience using Pytest, Selenium or Playwright Monitoring experience using Grafana, ELK Stack or Prometheus Experience working within Agile/Scrum environments Python Developer, Software Developer, Software Engineer, Python, Python Programmer, Senior Python Developer ...

Platform Application Specialist

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 100 K
Bash, Containerization).Infrastructure automation and configuration management experience (e.g., Ansible, Terraform, Puppet).Nice to have:Deep knowledge of a modern observability stack (e.g., Prometheus, Grafana, Elastic Stack, Vector, AlertManager).Experience with a variety of database platforms (e.g., PostgreSQL, ClickHouse, MSSQL, Redis, FoundationDB).Familiarity with specific middleware (e.g., Kafka, Consul ...

IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]

Location
Greater London, England, United Kingdom
code (e.g. Terraform). Experience with containerization and orchestration tools (e.g. Kubernetes). Knowledge of observability tools for monitoring and logging (e.g. Prometheus, Grafana, and BQL). Problem‐Solving and Analytical Skills Excellent problem‐solving and analytical skills to identify and resolve complex issues. Ability to proactively identify potential operational ...

Lead Site Reliability Engineer

Location
City of Westminster, England, United Kingdom
more technical disciplines Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.), experience with container and container orchestration (e.g. ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
focus on observability for large-scale, distributed compute or network systems. Deep, hands-on expertise building, scaling, and managing observability platforms (e.g., Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger, Honeycomb, etc.). You have proven experience using these tools to support performance analysis and debugging of complex distributed ...

IAM Secrets Management Engineering - SRE Platform Engineer - VP - London

Hiring Organisation
Goldman Sachs
Location
London, UK
Employment Type
Full-time
with infrastructure as code (i.e. Terraform).Experience with containerization and orchestration tools (i.e. Kubernetes).Knowledge of observability tools for monitoring and logging (e.g. Prometheus, Grafana and BQL).Problem-Solving and Analytical Skills: Excellent problem-solving and analytical skills to identify and resolve complex issues. Ability to proactively identify potential operational ...

IAM Secrets Management Engineering - SRE Platform Engineer - VP - London

Hiring Organisation
Goldman Sachs
Location
London, United Kingdom
Salary
£ 100 K
with infrastructure as code (i.e. Terraform).Experience with containerization and orchestration tools (i.e. Kubernetes).Knowledge of observability tools for monitoring and logging (e.g. Prometheus, Grafana and BQL).Problem-Solving and Analytical Skills:Excellent problem-solving and analytical skills to identify and resolve complex issues.Ability to proactively identify potential operational risks ...

Cloud Infrastructure Engineer - Digital Assets

Hiring Organisation
Optiver
Location
London, United Kingdom
Salary
£ 80 K
workflows.Build and maintain CI/CD and automated testing for changes minimizing risk in always-on markets.Implement and enhance observability using CloudWatch, Prometheus, Grafana, and similar tools.Participate in operational ownership through incident response, post-incident improvement, and continuous hardening.Who you areDemonstrated expertise in owning production infrastructure as an engineer (Platform ...

Architect & Delivery Lead (68018)

Location
Greater London, England, United Kingdom
SAFe, LeSS), waterfall, hybrid delivery models Financial modelling: FinOps, TCO analysis, business case development, and outcome‐based commercial models Tooling breadth: ITSM (ServiceNow), monitoring (Grafana/ELK/Datadog/Dynatrace), automation (Ansible/Terraform), and collaboration platforms Soft Skills & Competencies Exceptional leadership presence — inspires confidence at C‐level while ...

Professional Services Consultant - AI Security

Hiring Organisation
Cato Networks
Location
London, United Kingdom
Salary
£ 70 K
plusFamiliarity with container security, runtime protection, and service mesh architectures (Istio, App Mesh)Practical experience with observability and monitoring stacks (e.g., CloudWatch, Prometheus, Grafana, Datadog)Knowledge of data sovereignty, residency requirements and compliance frameworks relevant to AI workloads (e.g., FedRAMP, SOC 2, ISO 27001, NIST 800-53, HIPPA, PCI, HITRUST ...

Site Reliability Engineer

Hiring Organisation
Lloyds Banking Group
Location
London, United Kingdom
Salary
£ 70 K
with Amazon Web Services and Google Cloud PlatformExperience supporting Kubernetes (EKS) environments and service mesh technologies such as IstioKnowledge of observability tooling including Prometheus, Grafana or CoralogixExperience with PostgreSQL, MongoDB or HashiCorp VaultExperience using GitLab, Flux or Helm within CI/CD pipelinesKnowledge of PCI-compliant infrastructure designExperience working within ...

Sr. Software Engineer

Hiring Organisation
Meltwater Group
Location
London, United Kingdom
Salary
£ 80 K
Tech Stack:Languages: Golang (primary) with some TypeScriptInfrastructure: AWS, S3, Lambda, SQS, SNS, CloudFront, Kubernetes (Helm), Kong API GatewayDatabases: Postgres, Redis, DynamoDB, OpenSearchMonitoring: Coralogix, Grafana, CloudWatchCI/CD & IaC: GitHub Actions, TerraformWhat We Offer:Generous paid time off and flexible working arrangements.Comprehensive health insurance and wellness benefits.Family leave program that ...

Software Engineering Team Lead

Location
Greater London, England, United Kingdom
microservice development - we use Azure Service Bus, and welcome experience with similar messaging technologies such as Kafka or RabbitMQ Infrastructure: Kubernetes, Docker Observability: Prometheus, Grafana Engineering culture: DevOps, infrastructure as code, automated testing across all environments including production, continuous delivery Our Engineering Approach Full ownership: Teams own their solutions ...

Sr. Software Engineer

Hiring Organisation
Meltwater Group
Location
London, UK
Employment Type
Full-time
Tech Stack: Languages: Golang (primary) with some TypeScriptInfrastructure: AWS, S3, Lambda, SQS, SNS, CloudFront, Kubernetes (Helm), Kong API GatewayDatabases: Postgres, Redis, DynamoDB, OpenSearchMonitoring: Coralogix, Grafana, CloudWatchCI/CD & IaC: GitHub Actions, TerraformWhat We Offer: Generous paid time off and flexible working arrangements. Comprehensive health insurance and wellness benefits. Family leave ...

Site Reliability Engineer

Hiring Organisation
Trainline
Location
London, United Kingdom
Salary
£ 70 K
have...Experience of SRE concepts such as SLI, SLO and error budgets.Hands-on experience with observability tooling such as New Relic, Elastic (ELK Stack), Influx, Grafana or similarExperience working with cloud providers (preferably AWS).Experience troubleshooting Linux operating systems.Experience of scripting in at least one language (preferably Python)Understanding of load ...

Software Engineering Team Lead

Location
Greater London, England, United Kingdom
microservice development - we use Azure Service Bus, and welcome experience with similar messaging technologies such as Kafka or RabbitMQ Infrastructure: Kubernetes, Docker Observability: Prometheus, Grafana Engineering culture: DevOps, infrastructure as code, automated testing across all environments including production, continuous delivery Our Engineering Approach Full ownership: Teams own their solutions ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
with a bias toward automation.Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale.Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements.Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform) and secure ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline components ...