776 to 800 of 827 Grafana Jobs

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge of Cloudflare … operational consistency. Write and contribute to code, telemetry and instrumentation that improve service reliability and observability. Build dashboards and operational views using telemetry from Grafana, Splunk, New Relic and related platforms. Configure and manage Cloudflare edge services using Infrastructure as Code and integrate edge telemetry with observability platforms. Diagnose incidents ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Stoke-On-Trent, Staffordshire, West Midlands, United Kingdom
Employment Type
Permanent, Work From Home
understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge of Cloudflare … operational consistency. Write and contribute to code, telemetry and instrumentation that improve service reliability and observability. Build dashboards and operational views using telemetry from Grafana, Splunk, New Relic and related platforms. Configure and manage Cloudflare edge services using Infrastructure as Code and integrate edge telemetry with observability platforms. Diagnose incidents ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
Staff level. Embed DataOps best practices across the team: CI/CD, testing, observability, data drift monitoring, and incident response using tools like Grafana and Datadog. Collaborate with ML engineers and data scientists to productionise models and agentic tooling built on internal AI platforms, with hands … build expertise in stream processing technologies such as Apache Flink, Kafka, or Spark. (Nice to have: ClickHouse.) Familiarity with observability tooling such as Grafana or Datadog, and a good instinct for keeping systems healthy and well-monitored. Active engagement with the evolving AI/ML tooling landscape — you integrate ...

Senior Backend Engineer - Asset Sales

Location
Greater London, England, United Kingdom
C# stack : Distributed C# and .NET microservices Cloud & orchestration : Hosted on Azure using Kubernetes Architecture : Event-driven, supporting products used at significant scale Observability : Grafana, Azure Application Insights, logs, traces, and metrics AI tooling : Claude and other AI tools used throughout the engineering workflow — design exploration, code generation and review … have experience with similar messaging technology (Kafka, RabbitMQ, etc.) Microservices architecture experience, ideally on Azure and Kubernetes Experience using observability tooling (e.g. Grafana, Application Insights) to understand production behaviour Comfort using AI tools as part of a daily engineering workflow (design, code review, testing, incident investigation) DevOps culture mindset, including ...

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will leverage Java-based data collection pipelines and metrics platforms (Prometheus, Grafana, Dash0, Sentry, Humio) to detect performance regressions, slow queries, and health issues before they impact customers Fortify Business Continuity (BCP): Design and implement rigorous Business … region data platforms—while ensuring data integrity, clean relational modeling, and mobility A Security & Observability Mindset: You focus on building deep observability (Prometheus/Grafana/Dash0/Sentry/Humio) and automated security guardrails directly into your Java services so the fleet is secure and monitored by design Interview ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will leverage Java-based data collection pipelines and metrics platforms (Prometheus, Grafana, Dash0, Sentry, Humio) to detect performance regressions, slow queries, and health issues before they impact customers Fortify Business Continuity (BCP): Design and implement rigorous Business … region data platforms—while ensuring data integrity, clean relational modeling, and mobility A Security & Observability Mindset: You focus on building deep observability (Prometheus/Grafana/Dash0/Sentry/Humio) and automated security guardrails directly into your Java services so the fleet is secure and monitored by design Interview ...

Database Reliability Engineer

Location
Cardiff, Wales, United Kingdom
Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will leverage Java-based data collection pipelines and metrics platforms (Prometheus, Grafana, Dash0, Sentry, Humio) to detect performance regressions, slow queries, and health issues before they impact customers Fortify Business Continuity (BCP): Design and implement rigorous Business … region data platforms—while ensuring data integrity, clean relational modeling, and mobility A Security & Observability Mindset: You focus on building deep observability (Prometheus/Grafana/Dash0/Sentry/Humio) and automated security guardrails directly into your Java services so the fleet is secure and monitored by design Interview ...

Database Reliability Engineer

Location
Greater London, England, United Kingdom
Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will leverage Java-based data collection pipelines and metrics platforms (Prometheus, Grafana, Dash0, Sentry, Humio) to detect performance regressions, slow queries, and health issues before they impact customers Fortify Business Continuity (BCP): Design and implement rigorous Business … region data platforms—while ensuring data integrity, clean relational modeling, and mobility A Security & Observability Mindset: You focus on building deep observability (Prometheus/Grafana/Dash0/Sentry/Humio) and automated security guardrails directly into your Java services so the fleet is secure and monitored by design Interview ...

Senior Java Developer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
Project descriptionJoin the Equity Accelerator programme to modernize critical Prime Finance Equities trading platforms. The role focuses on building and evolving highly scalable backend systems using Kotlin and Java, with event-driven, cloud-native architectures. ...

Technical Product Manager (Superapp)

Location
Greater London, England, United Kingdom
About LendableLendable is on a mission to build the world's best technology to help people get credit and save money. We're building one of the world’s leading fintech companies and are off ...

Solutions Integration Engineer

Hiring Organisation
Christy Media Solutions
Location
London, UK
Employment Type
Full-time
Reference: JOB-8544Christy Media Solutions is recruiting for a newly created Solutions Integration (DevOps) engineer position supporting our client's transition towards IP-centric, Software-as-a-Service broadcast services. Working within the architecture function ...

Senior Kotlin Developer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
Project descriptionJoin the Equity Accelerator programme to modernize critical Prime Finance Equities trading platforms. The role focuses on building and evolving highly scalable backend systems using Kotlin and Java, with event-driven, cloud-native architectures. ...

Site Reliability Engineer

Location
United Kingdom
understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge of Cloudflare … operational consistency. Write and contribute to code, telemetry and instrumentation that improve service reliability and observability. Build dashboards and operational views using telemetry from Grafana, Splunk, New Relic and related platforms. Configure and manage Cloudflare edge services using Infrastructure as Code and integrate edge telemetry with observability platforms. Diagnose incidents ...

Site Reliability Engineer

Location
Manchester, England, United Kingdom
Service Level Objectives (SLO's) for reliability and customer satisfaction. Knowledge of contemporary observability tools, techniques and best practice including Splunk, New Relic, Grafana and PagerDuty. Proficiency in shell scripting for automation and system management tasks. Experience with Infrastructure as Code (IaC), automation and orchestration tools such as Ansible … observability of services, including telemetry, operational APIs and tooling. Building sophisticated dashboards using a range of telemetry data and dash boarding technologies like Grafana, Splunk and New Relic. Actively participating in live incident resolution and post‐mortem analysis, providing effective remediation strategies to improve overall system health and prevent future ...

Software Engineer, SRE

Location
United Kingdom
Service Level Objectives (SLO's) for reliability and customer satisfaction. Knowledge of contemporary observability tools, techniques and best practice including Splunk, New Relic, Grafana and PagerDuty. Proficiency in shell scripting for automation and system management tasks. Experience with Infrastructure as Code (IaC), automation and orchestration tools such as Ansible … observability of services, including telemetry, operational APIs and tooling. Building sophisticated dashboards using a range of telemetry data and dash boarding technologies like Grafana, Splunk and New Relic. Actively participating in live incident resolution and post-mortem analysis, providing effective remediation strategies to improve overall system health and prevent future ...

Software Engineer, SRE

Location
Manchester, England, United Kingdom
Service Level Objectives (SLO's) for reliability and customer satisfaction. Knowledge of contemporary observability tools, techniques and best practice including Splunk, New Relic, Grafana and PagerDuty. Proficiency in shell scripting for automation and system management tasks. Experience with Infrastructure as Code (IaC), automation and orchestration tools such as Ansible … observability of services, including telemetry, operational APIs and tooling. Building sophisticated dashboards using a range of telemetry data and dash boarding technologies like Grafana, Splunk and New Relic. Actively participating in live incident resolution and post-mortem analysis, providing effective remediation strategies to improve overall system health and prevent future ...

ServiceNow Technology Support II - Production Support at JPMorganChase

Location
Bournemouth, England, United Kingdom
operations. Install, upgrade, and manage APIs and integrations between ServiceNow and other enterprise systems. Monitor ServiceNow environments for anomalies using observability tools (e.g., Splunk, Grafana, APICA), and proactively address issues. Apply ServiceNow best practices to optimize performance, configuration, and workflow automation. Collaborate with product owners, developers, and support teams … JavaScript, Glide API) and web development frameworks. Familiarity with ITIL processes and ServiceNow ITSM workflows. Experience with production management and monitoring tools (e.g., Splunk, Grafana, APICA). Ability to prioritize and manage workload, communicate issues clearly, and take ownership of assigned tasks. Strong analytical and problem-solving skills. Preferred qualifications ...

Network Voice enginer

Hiring Organisation
CACI Network Services
Location
London, UK
Employment Type
Full-time
/configuration Infrastructure support - Flexpod, NettApps Other responsibilities could include: Faults and service requests Technical escalations Contribute to the Internal monitoring and management tools (Grafana, Splunk and python based tools, Ansible) Core skills, knowledge and experience required Network Infrastructure LAN/WAN fundamentals, QoS for video and voice traffic. VLAN … replication, and backup solutions. Cisco UCS, Hyperflex, FlexPOD Netapp Infrastructure Automation & Monitoring Scripting and automation Infrastructure as Code (IaC) experience (Terraform, Ansible). Splunk, Grafana Video Conferencing Systems Deploy, configure, and troubleshoot Cisco VTC endpoints (Room Kits, Webex Boards, SX/MX series, Desk series). Knowledge of Zoom Rooms ...

Data Scientist

Location
Oxford, England, United Kingdom
well as other data sources that are currently not databased Python (3+ for production level code, including pydantic, liniting, type hinting etc) Dashboard building (Grafana, Streamlit, Superset, Plotly Dash etc) Development Lifecycle Skills Requirements/Request elicitation and logging (e.g. understanding user needs for models/analysis/outputs … likely be using: Programming: Python, polars, pydantic, uv, SQLAlchemy, Streamlit Infrastructure and DB: Postgres, Warehousing (if we did it), prefect, Kubernetes, SQLAlchemy, AWS Visualisation: Grafana, Superset, Marimo Forecasting + general DS: lightgbm, xgboost, numpy, scipy, scikit-learn Claude AI Ultimately we are looking for someone who is a great ...

Senior Backend Engineer - Databases - Analytics | UK | Remote

Location
Pathhead, Scotland, United Kingdom
United Kingdom (Remote) Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimise telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud … their disparate data, wherever it lives, and move at the speed of their ambitions. Customers, including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce, rely on Grafana Labs. We are a 100% remote company with team members across 40+ countries, backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue ...

Network Specialist - Core

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, UK
Employment Type
Full-time
Position Overview: Squarepoint is looking for a highly skilled and detail-oriented Network Core Specialist to architect, develop, optimize and secure scalable networks for Datacenter, Campus and Cloud infrastructures. The Network Core Specialist is part ...

ServiceNow Technology Support II - Production Support

Location
Bournemouth, England, United Kingdom
operations. Install, upgrade, and manage APIs and integrations between ServiceNow and other enterprise systems. Monitor ServiceNow environments for anomalies using observability tools (e.g., Splunk, Grafana, APICA), and proactively address issues. Apply ServiceNow best practices to optimize performance, configuration, and workflow automation. Collaborate with product owners, developers, and support teams … JavaScript, Glide API) and web development frameworks. Familiarity with ITIL processes and ServiceNow ITSM workflows. Experience with production management and monitoring tools (e.g., Splunk, Grafana, APICA). Ability to prioritize and manage workload, communicate issues clearly, and take ownership of assigned tasks. Strong analytical and problem-solving skills. Preferred qualifications ...

Senior Systems Reliability Engineer (SRE), Edge

Location
Greater London, England, United Kingdom
Load balancing and reverse proxies such as Nginx, Varnish, HAProxy, Squid or Apache SQL databases Time series databases such as OpenTSDB, Graphite, Prometheus or Grafana Bonus Points Experience with continuous/rapid release engineering Strong tooling and automation development experience Experience working in a 24/7/365 service … environment Experience working with large scale production distributed systems A history of contributing to Open Source Software Some tools that we use Docker Grafana Consul Nomad Salt What Makes Cloudflare Special? We’re not just a highly ambitious, large-scale technology company. We’re a highly ambitious, large-scale technology ...

Senior Database Administrator London ·

Location
Greater London, England, United Kingdom
developers on integration between application code and database layers, assisting with connectivity, driver configuration, and troubleshooting. Write and optimise SQL queries for use in Grafana dashboards and reporting front‐ends, surfacing operational and business metrics from QuestDB and MySQL. Work with the Engineering team to ensure database metrics and logs … execution plans, transactions, and locking. Ability to write complex SQL queries, both for operational and reporting purposes. Experience writing and integrating database queries with Grafana or similar observability and dashboarding tools is desirable. Experience with or interest in version‐controlled or branching database systems (e.g. Dolt). Familiarity with Redis ...

Senior DevOps Engineer - Hybrid (3/2) | Kubernetes & CI/CD

Location
Belfast City District, Northern Ireland, United Kingdom
Citi is seeking a Senior DevOps Engineer to join its Technology team in Belfast, delivering scalable CI/CD pipelines, container orchestration with Kubernetes/OpenShift, and robust observability using Splunk, Prometheus, and Grafana. You ...