851 to 875 of 964 Grafana Jobs

DevOps Engineer

Hiring Organisation
BAE SYSTEMS
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 50 K
Data Engineers to ensure reproducible, reliable, and auditable infrastructure across environmentsMonitoring system health, logging, and metrics using platform observability tools (e.g. Azure Monitor, Prometheus, Grafana where applicable)Essential Skills:You will have experience of writing, maintaining and debugging CI/CD pipelines, ideally in GitLab CIYou will have experience … sign-on, ideally in restricted or air gapped environmentsExperience of operating and troubleshooting services in production, using monitoring and logging tools such as Prometheus, Grafana and LokiThe Shared Services Data & Analytics team:The Shared Services Data & Analytics Team is at the heart of our company data landscape, bringing together data ...

DevOps Engineer

Hiring Organisation
BAE SYSTEMS
Location
Christchurch, Dorset, United Kingdom
Salary
£ 50 K
Data Engineers to ensure reproducible, reliable, and auditable infrastructure across environmentsMonitoring system health, logging, and metrics using platform observability tools (e.g. Azure Monitor, Prometheus, Grafana where applicable)Essential Skills:You will have experience of writing, maintaining and debugging CI/CD pipelines, ideally in GitLab CIYou will have experience … sign-on, ideally in restricted or air gapped environmentsExperience of operating and troubleshooting services in production, using monitoring and logging tools such as Prometheus, Grafana and LokiThe Shared Services Data & Analytics team:The Shared Services Data & Analytics Team is at the heart of our company data landscape, bringing together data ...

DevOps Engineer

Hiring Organisation
BAE SYSTEMS
Location
Todmorden, Lancashire, United Kingdom
Salary
£ 50 K
Data Engineers to ensure reproducible, reliable, and auditable infrastructure across environmentsMonitoring system health, logging, and metrics using platform observability tools (e.g. Azure Monitor, Prometheus, Grafana where applicable)Essential Skills:You will have experience of writing, maintaining and debugging CI/CD pipelines, ideally in GitLab CIYou will have experience … sign-on, ideally in restricted or air gapped environmentsExperience of operating and troubleshooting services in production, using monitoring and logging tools such as Prometheus, Grafana and LokiThe Shared Services Data & Analytics team:The Shared Services Data & Analytics Team is at the heart of our company data landscape, bringing together data ...

Platform Engineer

Location
Greater London, England, United Kingdom
huge, distributed scale efficiently Monitoring and alerting: Measuring application performance and delivering insights, metrics and relevant alerts to the engineering teams with ELK, Grafana and New Relic Ownership: Driving engineering teams to own their infrastructure and costs by building great tooling, visibility and documentation Security: Setting the standards for fine … understanding of CI/CD and relevant tooling (we use GitHub Actions and Argo CD) Expertise in logging and monitoring at scale (e.g.S3, Graphite, Grafana, ELK, NewRelic, Datadog) Knowledge of a DevOps toolchain to drive ownership of a self-hosted platform Competent in Git and the GitOps philosophy Familiarity with ...

Senior Infrastructure Engineer (GCP) - Engine by Starling

Location
Cardiff, Wales, United Kingdom
workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing CI/… native Container-based architecture Kubernetes (GKE on GCP, EKS on AWS) TeamCity for CI/CD (with multiple production releases per day) Terraform and Grafana RDS and CloudSQL for PostgreSQL Our Interview Process Interviewing is a two-way process and we want you to have the time and opportunity ...

Senior Infrastructure Engineer (GCP) - Engine by Starling

Location
Greater London, England, United Kingdom
workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing CI/… native Container-based architecture Kubernetes (GKE on GCP, EKS on AWS) TeamCity for CI/CD (with multiple production releases per day) Terraform and Grafana RDS and CloudSQL for PostgreSQL Our Interview Process Interviewing is a two-way process and we want you to have the time and opportunity ...

Senior Java Full Stack Engineer

Hiring Organisation
Pinnacle Technical Resources
Location
Tempe, Arizona, United States
Employment Type
Permanent
Salary
USD 55 Hourly
Position: Sr. Java Full Stack Developer Location: Tempe, Arizona Duration: Contract Job ID: 180092 Interview: F2F required Job Overview: We are seeking a highly skilled and experienced Sr. Java Full Stack Developer to join our ...

Senior Network Engineer- IP

Location
Greater London, England, United Kingdom
Job Description Job Title: Senior Network Engineer- IP Req ID: 62900 Job Function: Engineering Posting Start Date: 24/09/2026 Posting End Date: 07/10/2026 Division: Networks Job Location: GBR ...

Senior Network Engineer- IP

Location
Birmingham, England, United Kingdom
Job Description Job Title: Senior Network Engineer- IP Req ID: 62900 Job Function: Engineering Posting Start Date: 24/09/2026 Posting End Date: 07/10/2026 Division: Networks Job Location: GBR ...

Senior Network Engineer- IP

Location
Ipswich, England, United Kingdom
Job Description Job Title: Senior Network Engineer- IP Req ID: 62900 Job Function: Engineering Posting Start Date: 24/09/2026 Posting End Date: 07/10/2026 Division: Networks Job Location: GBR ...

Engineering Manager, App Security (Cloud & OSS) - Remote

Location
United Kingdom
Grafana Labs is seeking an Engineering Manager for Application Security in the EMEA region (Remote, UK). You will lead security engineering squads, drive automation, and shape security posture across cloud and on‐premise components. The role emphasizes coaching, risk‐based prioritization, and collaboration with product and engineering teams. ...

ServiceNow Production Support Engineer II

Location
Bournemouth, England, United Kingdom
troubleshoot, maintain, and optimize ServiceNow configurations and integrations, while collaborating with product owners and developers to implement enhancements. Responsibilities include monitoring environments with Splunk, Grafana, and APICA, scripting in JavaScript, and driving standard operating procedures for the team. Weekend on-call rotation is required. #J-18808-Ljbffr ...

SRE-NOC Engineer: Master Incident Response & Reliability

Location
United Kingdom
responsibilities with reliability engineering, focusing on 24/7 service reliability, incident response, and automation. You’ll own runbooks, design alerting, build dashboards with Grafana, and work with cross‐functional teams to reduce toil. The role suits those who engineer solutions rather than only respond to alerts, with a focus ...

Remote UKI Regional Enterprise Growth Director

Location
United Kingdom
Grafana Labs is seeking a Regional Sales Director, Enterprise Growth to lead a team of Enterprise Growth Account Executives across the UK & Ireland. Drive revenue growth, attract and retain talent, and expand the customer base in the region. You will mentor the team, shape strategy, and partner with multiple groups ...

Senior DBA: Lead High-Perf MySQL/PostgreSQL & QuestDB

Location
Greater London, England, United Kingdom
successful candidate will work closely with Engineering and Operations to ensure high-performing, well-managed databases that support our exchange technology products and Grafana dashboards. #J-18808-Ljbffr ...

Security Platform Engineer, UK Security Operations

Hiring Organisation
Google
Location
London, UK
Employment Type
Full-time
Experience with Kubernetes security, including workload isolation, Role-Based Access Control (RBAC), and network policies, containerisation, orchestration, and Kubernetes observability tools (e.g., Falco, Prometheus, Grafana).Experience with infrastructure-as-code and configuration management tools (e.g., Terraform, Helm, ArgoCD).Active, or the ability to obtain, a Developed Vetting (DV) UK security … Experience with Kubernetes security, including workload isolation, Role-Based Access Control (RBAC), and network policies, containerisation, orchestration, and Kubernetes observability tools (e.g., Falco, Prometheus, Grafana).Experience with infrastructure-as-code and configuration management tools (e.g., Terraform, Helm, ArgoCD).Active, or the ability to obtain, a Developed Vetting (DV) UK security ...

HPC Support Engineer

Hiring Organisation
Hays
Location
London, United Kingdom
Salary
£ 45 K
make effective use of HPC resources. You will also improve automation, monitoring and deployment processes using technologies such as Python, Bash, Ansible, Docker, Kubernetes, Grafana and CI/CD tooling.Working closely with researchers and technical teams, you will:Administer and support HPC clusters, GPU systems and specialist research hardwareManage workload … automation skills using Python, Bash or AnsibleExperience with Docker, Kubernetes, Git and CI/CD practicesKnowledge of monitoring and observability tools such as Grafana, ELK Stack or SplunkExperience using HPC software build frameworks such as Spack or EasyBuildKnowledge of scientific software stacks, compilers and librariesExperience managing large-scale storage ...

Kubernetes Platform Engineer

Location
Greater London, England, United Kingdom
FluxCD) for safe, auditable changes. Drive Infrastructure as Code practices with Terraform and Helm for reliable and repeatable builds. Heavily embed observability using Prometheus, Grafana, and OpenTelemetry to make systems measurable and reliable. Stay ahead of Kubernetes evolution by testing and adopting new versions and features early. Collaborate with teams … Open Policy Agent. Proven ability to troubleshoot complex performance and reliability issues across infrastructure and workloads. Experience with observability tools such as Prometheus, Grafana, and OpenTelemetry to monitor cluster metrics and health. Great communication skills, with experience collaborating with internal platform users to gather feedback and deliver improvements. Experience writing ...

Kubernetes Platform Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
FluxCD) for safe, auditable changes. Drive Infrastructure as Code practices with Terraform and Helm for reliable and repeatable builds. Heavily embed observability using Prometheus, Grafana, and OpenTelemetry to make systems measurable and reliable. Stay ahead of Kubernetes evolution by testing and adopting new versions and features early. Collaborate with teams … Open Policy Agent. Proven ability to troubleshoot complex performance and reliability issues across infrastructure and workloads. Experience with observability tools such as Prometheus, Grafana and OpenTelemetry to monitor cluster metrics and health. Great communication skills, with experience collaborating with internal platform users to gather feedback and deliver improvements. Experience writing ...

Network Automation & OSS Designer

Location
Greater London, England, United Kingdom
Architect AIOps capabilities including closed‐loop automation, anomaly detection, and predictive analytics for network operations. Integrate OSS observability with cloud‐native monitoring stacks (Prometheus, Grafana, OpenTelemetry, Elasticsearch). Lead design of intent‐based networking and policy‐driven automation frameworks. Collaborate with product managers, network engineers, platform teams, and DevOps practitioners …/CNF lifecycle management). Understanding of AIOps platforms and closed‐loop automation design for network operations. Experience with observability tooling: OpenTelemetry, Prometheus, Grafana, Jaeger, Loki, ELK Stack. Knowledge of ML/AI model integration for anomaly detection, root‐cause analysis, and predictive network management. Strong grasp of TM Forum ...

ML/AI Engineer

Hiring Organisation
Lloyds Banking Group
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 70 K
observability for models and pipelines: drift, data quality, fairness signals, latency, GPU utilisation, error budgets, and SLOs/SLIs via Prometheus, Grafana, and Dynatrace.Establish actionable alerting and runbooks for on‐call operations; drive incident reviews and reliability improvements.Operate a model registry (e.g., MLflow) with experiment tracking, versioning, lineage, and environment … multi‐stage pipelines; experience with GitOps, artefact repositories, and environment promotion.Practical experience with CUDA, TensorRT, Triton, TorchServe, and GPU scheduling/optimisation.Proficiency in Prometheus, Grafana, Dynatrace defining SLIs/SLOs and alert thresholds for ML systems.Experience operating MLflow (or equivalent) for experiment tracking, model bundling, and deployments.Expert ...

ML/AI Engineer

Location
Manchester, England, United Kingdom
observability for models and pipelines: drift, data quality, fairness signals, latency, GPU utilisation, error budgets, and SLOs/SLIs via Prometheus, Grafana, and Dynatrace. Establish actionable alerting and runbooks for on‐call operations; drive incident reviews and reliability improvements. Operate a model registry (e.g., MLflow) with experiment tracking, versioning, lineage … pipelines; experience with GitOps, artefact repositories, and environment promotion. Practical experience with CUDA, TensorRT, Triton, TorchServe, and GPU scheduling/optimisation. Proficiency in Prometheus, Grafana, Dynatrace defining SLIs/SLOs and alert thresholds for ML systems. Experience operating MLflow (or equivalent) for experiment tracking, model bundling, and deployments. Expert ...

Site Reliability Engineer with Python

Hiring Organisation
BC Forward
Location
Charlotte, North Carolina, United States
Employment Type
Permanent
Salary
USD Hourly
Job Title: Site Reliability Engineer with Python Location: Pennington, NJ Duration: Contract - 9 months Pay Range: $73.67/hr (W2) Job ID: 410592 About BCforward BCforward is a leading global IT consulting and workforce solutions ...

Site Reliability Engineer

Hiring Organisation
Proactive Appointments
Location
Gloucester, Gloucestershire, UK
Employment Type
Full-time
Job Description Site Reliability Engineer - DV Cleared Our client is urgently looking for an experienced Site Reliability Engineer to join their team on a contract basis, initially for 6 months with a view to extend. ...

Software Engineer, Security Rules

Location
Greater London, England, United Kingdom
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world's largest networks that powers millions of websites and other Internet properties ...