526 to 550 of 971 Grafana Jobs

Senior Site Reliability Engineer

Hiring Organisation
Oracle Corporation
Location
United Kingdom
Salary
£ 60 K
Must support network segmentation (e.g., security lists, network security groups, or firewalls).-Deep Understanding of manipulating telemetry data (traffic flows, health status) using Grafana dashboards and MQL.-Experience with major public cloud providers (e.g., Oracle Cloud Infrastructure OCI, or equivalent).-Experience using Jira and Confluence for incident tracking ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless or inference serving infrastructure. ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
cloud platforms (GCP and/or AWS) and infrastructure as code (e.g., Terraform). Experience with monitoring/observability tooling (e.g., Datadog, Monte Carlo, Grafana) for proactive detection of data quality and pipeline issues. Familiarity with CI/CD practices applied to data workflows (e.g., automated testing for pipelines, version ...

Staff Quality Engineer - Mobile

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
push back on ambiguous acceptance criteria, and surface risk before code is writtenClose the loop on production issues using our observability stack (Datadog, Sentry, Grafana) - tying test coverage back to real customer impactEnsure teams have Service Level Objectives set up and are achieving themRun targeted exploratory testing on high-risk ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Senior Engineering Manager - 9-10 month FTC

Location
Greater London, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Site Reliability Engineer

Location
Slough, England, United Kingdom
environments Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps) Package and deploy applications to customer environments Implement and optimise observability tooling (Prometheus, Grafana, Loki) Support incident response, monitoring, and backup/recovery planning Mentor project teams in DevSecOps practices and environment management Ensure cloud environments are secure, performant … cost-optimised Tech Environment & Skills: Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC) Containers: Kubernetes, Docker, Helm (OpenShift a plus) Observability: Prometheus, Grafana, Loki (network visualisation desirable) CI/CD & GitOps: GitHub Actions, ArgoCD/Flux Security: Cloud access models, Zero Trust principles Programming: Python or Golang preferred ...

Platform Engineer

Location
Greater London, England, United Kingdom
Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest … supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow ...

Test Environment Manager

Location
Greater London, England, United Kingdom
Objectives (SLOs) and key Service Level Indicators (SLIs), such as environment availability, provisioning time, and stability metrics. Monitor environment health using observability tools (Prometheus, Grafana, Splunk, etc.) and proactively identify and resolve performance issues or bottlenecks. Incident & Problem Management Lead incident response for environment-related issues, driving quick resolution … teams to ensure test data is consistent, compliant, refreshed automatically, and aligned with environment provisioning needs. Technical Skills & Experience Monitoring & Observability: Expertise with Prometheus, Grafana, Splunk, ELK/EFK, or similar platforms. CI/CD & Automation Tools: Strong experience with Jenkins, GitLab CI, GitHub Actions, and configuration management tools (Terraform ...

AI Platform & DevSecOps Engineer

Hiring Organisation
Credence
Location
McLean, Virginia, United States
Employment Type
Permanent
Salary
USD 190,000 Annual
dashboards, and alerting across cloud infrastructure, Kubernetes, applications, and platform services using technologies such as AWS CloudWatch, Security Hub, GuardDuty, Splunk/ELK, Prometheus, Grafana, OpenTelemetry, or comparable solutions. Participate in incident response, root-cause analysis, vulnerability remediation, and continuous reliability improvements. Automation & Scripting: Develop automation and platform tooling using … federal or regulated environments. Experience with Helm, Argo CD, Flux, or other Kubernetes/GitOps technologies. Experience with observability technologies such as CloudWatch, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK. Experience with software supply-chain security practices such as SBOM generation, artifact signing, provenance, container hardening, and policy enforcement . click ...

Senior Software Engineer- Product Reliability Engineering

Hiring Organisation
Visa
Location
Austin, Texas, United States
Employment Type
Permanent
Salary
USD Annual
intelligent and autonomous engineering organization. Own Observability & Platform Health Design and build dashboards, alerts, telemetry pipelines, and health indicators using tools such as Prometheus, Grafana, Splunk, or ELK to provide visibility across globally distributed systems. Analyze platform performance, reliability, utilization, and availability data to identify trends and implement long-term … more logging or search platforms such as Splunk, ClickHouse, OpenSearch, or Elasticsearch. Hands-on experience with metrics and visualization technologies such as Prometheus, Thanos, Grafana, or Bosun. Experience developing backend services, APIs, command-line tools, integrations, or operational automation. Experience with cloud platforms, preferably AWS or GCP, and with cloud ...

AWS Cloud Architect - eSC or eDV Clearance Required

Location
Leicester, England, United Kingdom
Introduction At IBM Consulting UK FutureNow, you’ll build a career at the forefront of hybrid cloud and AI, working with leading clients across the public and private sectors. At IBM Consulting UK FutureNow, you ...

Test Environment Manager

Hiring Organisation
Euroclear
Location
United Kingdom
Salary
£ 70 K
Job description:We are seeking an experienced Test Environment Manager to drive the test environment management to support a large-scale transformation programme moving from a legacy monolithic setup to a microservices-based set up ...

Senior Cloud Operations Engineer

Hiring Organisation
NICE Systems
Location
United Kingdom
Salary
£ 70 K
At NiCE, we don’t limit our challenges. We challenge our limits. Always. We’re ambitious. We’re game changers. And we play to win. We set the highest standards and execute beyond them. And ...

Cloud SRE — IaC, Kubernetes & CI/CD

Location
England, United Kingdom
engineering teams. You will manage Kubernetes clusters (EKS/AKS), build CI/CD pipelines with GitHub Actions, and implement monitoring with Prometheus and Grafana to ensure observability, security and cost-efficiency across platforms. #J-18808-Ljbffr ...

Senior Platform Engineer: AI-Ready Infra & Security

Location
Slough, England, United Kingdom
automate operations, and build self-service workflows. The role emphasizes strong Python, Linux, Terraform/Ansible, Docker and Kubernetes proficiency, plus observability with Prometheus, Grafana and OpenTelemetry. #J-18808-Ljbffr ...

Platform Engineer: AI-Driven Cloud and Tooling Advocate

Location
St Albans, England, United Kingdom
platforms and the tools that support software engineers to develop, test and deploy applications. We work with AWS, Linux, Kubernetes, Kafka, Redis, HAProxy, Consul, Grafana and OpenSearch, plus Terraform and more. Our stack includes C#/.NET, Node.js and Python, with Azure DevOps Pipelines and ArgoCD. #J-18808-Ljbffr ...

Senior Scala Engineer

Location
Greater London, England, United Kingdom
Essential: Scala Play or other MVC/Rest API frameworks SQL AWS Suite Continuous Integration Agile methodologies Desirable: Containerisation principles/Docker Jenkins Kibana Grafana Airflow #J-18808-Ljbffr ...

DevOps Engineer

Location
United Kingdom
Lambda, CloudFormation) Terraform or similar IaC tools Comfortable picking up new tools quickly and figuring things out independently Nice to have Observability tools (Prometheus, Grafana, Datadog) Security automation What we won’t ask you to do Sit through pointless meetings that could have been a Slack message Fill in timesheets ...

EKS Engineer

Location
Greater London, England, United Kingdom
manage CI/CD pipelines for containerized workloads • Ensure security, compliance, and governance across EKS environments • Monitor cluster performance using tools like Prometheus, Grafana, CloudWatch • Manage networking components (VPC, load balancers, ingress controllers) • Optimize cost, performance, and resource utilization • Troubleshoot cluster, networking, and application issues • Collaborate with development teams ...

Lead Cloud Infrastructure Engineer

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
Configuration Management experience, as well as networking experience and an understanding of DevOps principles and practices. Experience with observability tools such as Prometheus and Grafana and working with Kubernetes in production environments at scale is a plus. If you're open to hearing further details about this opportunity, then apply ...

Agile Software Engineer - Java - Greenfield Multi-Asset Trading Platform

Hiring Organisation
Stanford Black
Location
London, UK
Employment Type
Full-time
Scala and Kotlin, with everything being built in AWS.Not only this, you could be cross-trained in modern open source technology such as Helm, Grafana, and GraphQL. In terms of salary, this particular client offers unmatched compensation packages and pay above the market standard, providing incredible increases on bonuses annually. ...

DevOps Engineer DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
Romsey, Hampshire, United Kingdom
Salary
£ 50 K
/CD pipeline development and optimisationKubernetes platform administration and automationInfrastructure as Code using Terraform and PackerGitOps implementation with ArgoCDMonitoring, observability and performance optimisation using Grafana, Prometheus and OpenTelemetryContinuous improvement of engineering standards, automation and release processesRequirementsActive SC or DV ClearanceUK-basedAble to attend site in Romsey approximately once per monthProven ...

Senior .NET Backend Developer

Hiring Organisation
Oscar Associates (UK) Limited
Location
York, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Kubernetes. PostgreSQL, Redis or Elasticsearch. GraphQL. RabbitMQ or other messaging technologies. CI/CD pipelines and modern DevOps practices. Observability tooling such as Grafana, OpenTelemetry or Prometheus. Experience using AI-assisted development tools within the software development lifecycle. What's on Offer Hybrid working (2 days per week in York ...

Software Engineer - Infrastructure and Automation

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
firmwideComfortable working across technical domains and collaborating with peersNice to HaveExperience with Docker, KVM, or other container/virtualisation toolsFamiliar with observability stacks (Prometheus, Grafana)TerraformPractical knowledge of networking protocols and hardware environmentsUnderstanding of low-latency or post-trade systemsWhy Apply? Work on the internal infrastructure that makes the firm ...