276 to 300 of 360 Grafana Jobs in London

Platform Engineer

Hiring Organisation
itecopeople
Location
London, United Kingdom
Employment Type
Permanent
Salary
£54000 - £65000/annum
Infrastructure as Code Develop and maintain GitOps CI/CD pipelines Manage Kubernetes networking, service mesh and gateway technologies Improve platform observability using Grafana, Prometheus and OpenTelemetry Maintain platform security, resilience and automation Troubleshoot production platform issues and drive continuous improvement Work closely with architects to turn high-level designs … essential) Kubernetes platform engineering within production environments Terraform and Infrastructure as Code Docker, GitOps and CI/CD pipelines Linux and Bash scripting Grafana, Prometheus and OpenTelemetry Kubernetes networking and service mesh technologies Production platform operations, troubleshooting and automation You'll also be able to demonstrate: Experience owning technical implementation ...

Vice President, Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Objectives, and service health measures aligned to operational and business priorities. Build and optimize monitoring, observability, and alerting capabilities using tools such as Prometheus, Grafana, AppDynamics, and Splunk. Apply AIOps capabilities to improve event correlation, anomaly detection, root cause analysis, predictive insights, and proactive issue prevention. Partner with engineering, infrastructure … Demonstrated ability to define and operationalize SLIs, SLOs, dashboards, alerts, and health indicators. Hands-on experience with enterprise monitoring and observability platforms including Prometheus, Grafana, AppDynamics, and Splunk. Strong troubleshooting, analytical, and problem-solving skills in complex distributed or production environments. Strong verbal and written communication skills, with the ability ...

Vice President, Site Reliability Engineering

Location
Westminster, West End, United Kingdom
Objectives, and service health measures aligned to operational and business priorities. Build and optimize monitoring, observability, and alerting capabilities using tools such as Prometheus, Grafana, AppDynamics, and Splunk. Apply AIOps capabilities to improve event correlation, anomaly detection, root cause analysis, predictive insights, and proactive issue prevention. Partner with engineering, infrastructure … Demonstrated ability to define and operationalize SLIs, SLOs, dashboards, alerts, and health indicators. Hands-on experience with enterprise monitoring and observability platforms including Prometheus, Grafana, AppDynamics, and Splunk. Strong troubleshooting, analytical, and problem-solving skills in complex distributed or production environments. Strong verbal and written communication skills, with the ability ...

Vice President, Site Reliability Engineering

Hiring Organisation
The Bank of New York Mellon
Location
London, United Kingdom
Salary
£ 80 K
Level Objectives, and service health measures aligned to operational and business priorities.Build and optimize monitoring, observability, and alerting capabilities using tools such as Prometheus, Grafana, AppDynamics, and Splunk.Apply AIOps capabilities to improve event correlation, anomaly detection, root cause analysis, predictive insights, and proactive issue prevention.Partner with engineering, infrastructure, production support … production environments.Demonstrated ability to define and operationalize SLIs, SLOs, dashboards, alerts, and health indicators.Hands-on experience with enterprise monitoring and observability platforms including Prometheus, Grafana, AppDynamics, and Splunk.Strong troubleshooting, analytical, and problem-solving skills in complex distributed or production environments.Strong verbal and written communication skills, with the ability to collaborate ...

Vice President, Site Reliability Engineering

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
Objectives, and service health measures aligned to operational and business priorities. Build and optimize monitoring, observability, and alerting capabilities using tools such as Prometheus, Grafana, AppDynamics, and Splunk. Apply AIOps capabilities to improve event correlation, anomaly detection, root cause analysis, predictive insights, and proactive issue prevention. Partner with engineering, infrastructure … Demonstrated ability to define and operationalize SLIs, SLOs, dashboards, alerts, and health indicators. Hands-on experience with enterprise monitoring and observability platforms including Prometheus, Grafana, AppDynamics, and Splunk. Strong troubleshooting, analytical, and problem-solving skills in complex distributed or production environments. Strong verbal and written communication skills, with the ability ...

Observability Engineer (Dynatrace) — Telemetry & Performance

Location
Greater London, England, United Kingdom
customer IT estates in the UK. You will collect telemetry, diagnose issues and drive proactive improvements across teams. The role requires strong Dynatrace/Grafana/Splunk experience, scripting skills, cloud familiarity (Azure/AWS) and Agile delivery experience. Travel across UK sites is possible in a dynamic IT consulting ...

Technology Integration Specialist

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 100 K
MySQL, MS SQL)Experience with the Atlassian stack (JIRA, Service Desk, Confluence)Nice to have:Project management experienceExposure to reporting tools like Tableau Dashboard, Grafana ...

automation qa

Location
Greater London, England, United Kingdom
soak, and spike tests; Hands‐on depth with k6 for building and running performance tests; Monitoring and observability know‐how using tools like Datadog, Grafana, or CloudWatch; Strong programming ability in at least one language, with TypeScript used across the stack; Understanding of microservices, distributed systems, and testing services ...

Platform Engineer

Location
Greater London, England, United Kingdom
from recurring Cloud infrastructure: experience on Google Cloud and other major providers Observability that actually helps: log management and monitoring with tools like QuickWit, Grafana or Datadog Developer environments people love: tooling like Tailscale, Workbrew and dev containers that make everyone's day-to-day faster Clean, well-crafted code ...

Software Developer - Data Reliability

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 100 K
vendorsMindset: Proactive, detail-oriented, and self-driven with a strong sense of ownership and accountabilityNice to haveExperience with observability and monitoring tools such as Grafana, Kibana, or PrometheusExperience developing automation tooling and implementing configuration managementExperience with cloud platforms such as Google Cloud or AWSExperience operating job orchestration or workload scheduling ...

Senior Network Engineer IT Infrastructure London

Location
Greater London, England, United Kingdom
tooling (Ansible, Python, NetBox). Experience with ITSM platforms (Jira Service Management, ServiceNow or similar). Experience with network monitoring tools (LogicMonitor, SolarWinds, Elasticsearch, Grafana, Splunk or Datadog). Familiarity with data centre networking concepts and their integration with corporate office networks. Benefits Hybrid working model offers flexibility, with three ...

Staff Machine Learning Engineer - Ops

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
model deploymentStrong CI/CD and Github Actions experienceStrong communications skills with a collaborative mindsetDesirable Experience with Pytorch, TensorRT, quantisation and model deploymentExperience with Grafana monitoring and production observabilityThis is a full-time role based in our office in London. At Wayve we want the best of all worlds ...

Mission Operations Engineer

Location
Greater London, England, United Kingdom
success leader, and/or technical program manager. Data Engineering Proficiency : Expertise in SQL, Python, Pandas, Matlab, or R. Bonus if familiarity with Influx, Grafana, Spark, Polars, Arrow, Kafka, Beam, and Flink. [Bonus] In-Field Hardware Mastery : Proficiency with hardware engineering tools such as DAQs, DDS, Nix, RTOS, MCAP ...

Support engineer (UK) Writer London, UK

Location
Greater London, England, United Kingdom
technical support for an enterprise B2B SaaS organization Demonstrate deep technical proficiency navigating cloud tech (AWS/GCP), Python, SDKs, SSO/SCIM, Jira, Grafana and Datadog Possess strong skills in RESTful API debugging, integration and usage to troubleshoot complex customer environments Enjoy working on-screen with customers to overcome ...

Senior Network Engineer

Hiring Organisation
Checkout.com
Location
London, United Kingdom
Salary
£ 80 K
network automation tooling (Ansible, Python, NetBox)Experience with ITSM platforms (Jira Service Management, ServiceNow or similar)Experience with network monitoring tools (LogicMonitor, SolarWinds, Elasticsearch, Grafana, Splunk or Datadog)Familiarity with data centre networking concepts and their integration with corporate office networks.Additional InformationBring all of you to workWe create the conditions ...

Network Engineer

Hiring Organisation
Indigo Vision
Location
London, United Kingdom
Salary
£ 80 K
IPSEC, IKEV2, route based).Monitoring & SecurityExperience with taking packet captures and using analysis tools such as wireshark to troubleshoot connectivity issues.Experience with network monitoring (Grafana/Prometheus/Solarwinds/Fortianalyser).Experience in identifying security best practices when deploying solutions.Desirable Experience:Experience with SD-WAN solutions (Fortinet SD-WAN)Basic ...

FinOps Architect (68021) (DEAI DS) Cloud & Data Engineering United Kingdom

Location
Greater London, England, United Kingdom
analytics: SQL, Python/Pandas for cost data analysis; experience building normalisation pipelines and cost allocation engines Dashboard and reporting tools: Power BI, Tableau, Grafana, or custom BI solutions for real-time cost visibility and executive dashboards Understanding of ITIL asset and configuration management, hardware lifecycle management, and capacity planning ...

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
will drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands-on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct … improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply with relevant enterprise standards and security requirements. Resiliency ...

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
will drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands‐on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct … improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply with relevant enterprise standards and security requirements. Resiliency ...

DevOps and Automation Engineer (Contract)

Location
Greater London, England, United Kingdom
self-service environment provisioning using Infrastructure-as-Code (Terraform) and pipeline-driven automation.* Implement advanced observability and monitoring: Use platforms such as Datadog, Prometheus, Grafana, and OpenTelemetry to provide real-time insights into system health, deployments, and business metrics.* Embed security and compliance by design: Integrate security into every stage … generate new automation ideas and create user stories for rapid prototyping.* Use technologies like Terraform, Ansible, Azure DevOps, Github, OctopusDeploy, Kubernetes, OpenTelemetry, Datadog, Grafana, Prometheus, low-code automation platforms (e.g., Power Automate, UiPath) to help evolve the team's capabilities.* Coordinate planned outage and environment refreshes in collaboration with project ...

Site Reliability Engineer- London

Location
Greater London, England, United Kingdom
operational visibility capabilities across critical engineering systems.This role requires hands-on expertise in the deployment, administration, and optimisation of OpenSearch, alongside experience with Grafana, Geneos, and automation/scripting technologies. Particular emphasis will be placed on the candidate's ability to design, deploy, and support enterprise-grade OpenSearch environments. Responsibilities … analytics, and observability use cases using OpenSearch. Knowledge of OpenSearch security, access controls, backups, upgrades, and operational best practices. Monitoring & Observability Strong experience with Grafana, including dashboard development, alerting, and data source integration. Experience with enterprise monitoring platforms, specifically Geneos. Understanding of modern observability principles, including metrics, logs, traces, alerting ...

Site Reliability Engineer

Hiring Organisation
GoCardless
Location
London, United Kingdom
Salary
£ 70 K
maintenance, improvements, and support for audits and compliance.Tech stack and tools:Python, Ruby, Golang;Terraform;Atlantis AWS, GCP;Kubernetes, GKE;Github, GitHub Actions, ArgoCD;Grafana, Prometheus;DatadogWhat excites you:We use a wide range of technologies, and will never expect you to have experience with all of them. … considered an advantage;Experience with CI/CD tooling such as Github, GitHub Actions, ArgoCD, etc.;Experience working with monitoring tooling such as Grafana, Prometheus, etc.;Experience with relational databases and other datastores, especially around high availability and performance optimisation;Awareness of DevOps and Agile principles;Fluency in English;Excellent ...

Observability SRE AVP: Cloud Observability & Migrations

Location
Greater London, England, United Kingdom
lead end-to-end observability and resiliency initiatives in a large-scale environment. You will migrate monitoring tooling to Google Cloud Observability and Grafana, implement OpenTelemetry instrumentation, and author reusable deployment solutions for OpenShift/Kubernetes and VM environments. The role requires deep SRE knowledge, strong collaboration with application teams ...

C# Developer

Hiring Organisation
Meraki Talent Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £75,000 per annum
Services SQL Server (Including T-SQL) Angular (with Typescript) RabbitMQ/Kafka Azure Git Snowflake Nuget (Producing and Consuming)Azure DevOps (CI) Prometheus & Grafana (Monitoring & Alerting) ELK Stack/Azure Log Analytics (Logging) This role sits in the IT Development team, and its day-to-day activities include The development ...

Python Developer

Location
Greater London, England, United Kingdom
childhood spent hacking away in 8-bit assembly language Python 3.10+, JavaScript, TypeScript, and Go RabbitMQ, Kafka, PostgreSQL, Redis, Linux servers, OpenTelemetry, Prometheus, Grafana, and Zabbix Nice to have: Interest in functional programming and its application in the real world Условия: Extremely lucrative salary and significant bonus Greenfield Python ...