501 to 525 of 938 Grafana Jobs

Lead Site Reliability Engineer

Location
Greater London, England, United Kingdom
experience in front office trading environments or similarly high pressure, low latency domains. Proficiency with SRE tooling and techniques, including FIX messaging, Kafka, Grafana, Splunk, ITRS Geneos, Dynatrace, InfluxDB, MQ (IBM MQ or similar), Oracle DB Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve ...

Software Engineering Specialist

Hiring Organisation
BT Group
Location
Cheltenham, Gloucestershire, United Kingdom
Salary
£ 70 K
Elastic stack technologies and Kibana plugin developmentHave used DevOps tools and principles like Git, Jenkins, GitOpsUsed system monitoring tools such as CheckMK, Prometheus, Grafana or LokiOur PackageTailored benefits make a real difference. That’s why we offer a comprehensive range to support your growth, wellbeing, and everyday life. ...

Platform Engineer

Location
Greater London, England, United Kingdom
with Kubernetes, and enough Terraform to have opinions about how to structure it. Familiarity with CI/CD pipelines and observability tooling (Prometheus, Grafana, Datadog, or similar). Strong technical foundation: Demonstrated ability to write production-quality code and solve hard technical problems (experience with Python, TypeScript/JavaScript, systems ...

Managing Engineer - Observability, Pipeline & Analytics (Hybrid)

Location
Belfast City District, Northern Ireland, United Kingdom
Kusto Query Language (KQL) and large‐scale analytics platforms. Experience with Cribl, ADX, Datadog Pipelines, Splunk, Sentinel, Kafka, Event Hubs, Open Telemetry, Elastic, Grafana, or similar observability ecosystems. Ability to leverage AI assisted development tools (e.g., Copilot, Cursor) responsibly to improve developer productivity and code quality. Supervisory Responsibilities: This role ...

Senior Site Reliability Engineer

Hiring Organisation
Oracle Corporation
Location
United Kingdom
Salary
£ 60 K
Must support network segmentation (e.g., security lists, network security groups, or firewalls).-Deep Understanding of manipulating telemetry data (traffic flows, health status) using Grafana dashboards and MQL.-Experience with major public cloud providers (e.g., Oracle Cloud Infrastructure OCI, or equivalent).-Experience using Jira and Confluence for incident tracking ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless or inference serving infrastructure. ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
cloud platforms (GCP and/or AWS) and infrastructure as code (e.g., Terraform). Experience with monitoring/observability tooling (e.g., Datadog, Monte Carlo, Grafana) for proactive detection of data quality and pipeline issues. Familiarity with CI/CD practices applied to data workflows (e.g., automated testing for pipelines, version ...

Staff Quality Engineer - Mobile

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
push back on ambiguous acceptance criteria, and surface risk before code is writtenClose the loop on production issues using our observability stack (Datadog, Sentry, Grafana) - tying test coverage back to real customer impactEnsure teams have Service Level Objectives set up and are achieving themRun targeted exploratory testing on high-risk ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Senior Engineering Manager - 9-10 month FTC

Location
Greater London, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Site Reliability Engineer

Location
Slough, England, United Kingdom
environments Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps) Package and deploy applications to customer environments Implement and optimise observability tooling (Prometheus, Grafana, Loki) Support incident response, monitoring, and backup/recovery planning Mentor project teams in DevSecOps practices and environment management Ensure cloud environments are secure, performant … cost-optimised Tech Environment & Skills: Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC) Containers: Kubernetes, Docker, Helm (OpenShift a plus) Observability: Prometheus, Grafana, Loki (network visualisation desirable) CI/CD & GitOps: GitHub Actions, ArgoCD/Flux Security: Cloud access models, Zero Trust principles Programming: Python or Golang preferred ...

Platform Engineer

Location
Greater London, England, United Kingdom
Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest … supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow ...

Test Environment Manager

Location
Greater London, England, United Kingdom
Objectives (SLOs) and key Service Level Indicators (SLIs), such as environment availability, provisioning time, and stability metrics. Monitor environment health using observability tools (Prometheus, Grafana, Splunk, etc.) and proactively identify and resolve performance issues or bottlenecks. Incident & Problem Management Lead incident response for environment-related issues, driving quick resolution … teams to ensure test data is consistent, compliant, refreshed automatically, and aligned with environment provisioning needs. Technical Skills & Experience Monitoring & Observability: Expertise with Prometheus, Grafana, Splunk, ELK/EFK, or similar platforms. CI/CD & Automation Tools: Strong experience with Jenkins, GitLab CI, GitHub Actions, and configuration management tools (Terraform ...

AI Platform & DevSecOps Engineer

Hiring Organisation
Credence
Location
McLean, Virginia, United States
Employment Type
Permanent
Salary
USD 190,000 Annual
dashboards, and alerting across cloud infrastructure, Kubernetes, applications, and platform services using technologies such as AWS CloudWatch, Security Hub, GuardDuty, Splunk/ELK, Prometheus, Grafana, OpenTelemetry, or comparable solutions. Participate in incident response, root-cause analysis, vulnerability remediation, and continuous reliability improvements. Automation & Scripting: Develop automation and platform tooling using … federal or regulated environments. Experience with Helm, Argo CD, Flux, or other Kubernetes/GitOps technologies. Experience with observability technologies such as CloudWatch, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK. Experience with software supply-chain security practices such as SBOM generation, artifact signing, provenance, container hardening, and policy enforcement . click ...

Senior Software Engineer- Product Reliability Engineering

Hiring Organisation
Visa
Location
Austin, Texas, United States
Employment Type
Permanent
Salary
USD Annual
intelligent and autonomous engineering organization. Own Observability & Platform Health Design and build dashboards, alerts, telemetry pipelines, and health indicators using tools such as Prometheus, Grafana, Splunk, or ELK to provide visibility across globally distributed systems. Analyze platform performance, reliability, utilization, and availability data to identify trends and implement long-term … more logging or search platforms such as Splunk, ClickHouse, OpenSearch, or Elasticsearch. Hands-on experience with metrics and visualization technologies such as Prometheus, Thanos, Grafana, or Bosun. Experience developing backend services, APIs, command-line tools, integrations, or operational automation. Experience with cloud platforms, preferably AWS or GCP, and with cloud ...

AWS Cloud Architect - eSC or eDV Clearance Required

Location
Leicester, England, United Kingdom
Introduction At IBM Consulting UK FutureNow, you’ll build a career at the forefront of hybrid cloud and AI, working with leading clients across the public and private sectors. At IBM Consulting UK FutureNow, you ...

Test Environment Manager

Hiring Organisation
Euroclear
Location
United Kingdom
Salary
£ 70 K
Job description:We are seeking an experienced Test Environment Manager to drive the test environment management to support a large-scale transformation programme moving from a legacy monolithic setup to a microservices-based set up ...

Senior Cloud Operations Engineer

Hiring Organisation
NICE Systems
Location
United Kingdom
Salary
£ 70 K
At NiCE, we don’t limit our challenges. We challenge our limits. Always. We’re ambitious. We’re game changers. And we play to win. We set the highest standards and execute beyond them. And ...

Cloud SRE — IaC, Kubernetes & CI/CD

Location
England, United Kingdom
engineering teams. You will manage Kubernetes clusters (EKS/AKS), build CI/CD pipelines with GitHub Actions, and implement monitoring with Prometheus and Grafana to ensure observability, security and cost-efficiency across platforms. #J-18808-Ljbffr ...

Senior Platform Engineer: AI-Ready Infra & Security

Location
Slough, England, United Kingdom
automate operations, and build self-service workflows. The role emphasizes strong Python, Linux, Terraform/Ansible, Docker and Kubernetes proficiency, plus observability with Prometheus, Grafana and OpenTelemetry. #J-18808-Ljbffr ...

Platform Engineer: AI-Driven Cloud and Tooling Advocate

Location
St Albans, England, United Kingdom
platforms and the tools that support software engineers to develop, test and deploy applications. We work with AWS, Linux, Kubernetes, Kafka, Redis, HAProxy, Consul, Grafana and OpenSearch, plus Terraform and more. Our stack includes C#/.NET, Node.js and Python, with Azure DevOps Pipelines and ArgoCD. #J-18808-Ljbffr ...

Senior Scala Engineer

Location
Greater London, England, United Kingdom
Essential: Scala Play or other MVC/Rest API frameworks SQL AWS Suite Continuous Integration Agile methodologies Desirable: Containerisation principles/Docker Jenkins Kibana Grafana Airflow #J-18808-Ljbffr ...

DevOps Engineer

Location
United Kingdom
Lambda, CloudFormation) Terraform or similar IaC tools Comfortable picking up new tools quickly and figuring things out independently Nice to have Observability tools (Prometheus, Grafana, Datadog) Security automation What we won’t ask you to do Sit through pointless meetings that could have been a Slack message Fill in timesheets ...

EKS Engineer

Location
Greater London, England, United Kingdom
manage CI/CD pipelines for containerized workloads • Ensure security, compliance, and governance across EKS environments • Monitor cluster performance using tools like Prometheus, Grafana, CloudWatch • Manage networking components (VPC, load balancers, ingress controllers) • Optimize cost, performance, and resource utilization • Troubleshoot cluster, networking, and application issues • Collaborate with development teams ...

Lead Cloud Infrastructure Engineer

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
Configuration Management experience, as well as networking experience and an understanding of DevOps principles and practices. Experience with observability tools such as Prometheus and Grafana and working with Kubernetes in production environments at scale is a plus. If you're open to hearing further details about this opportunity, then apply ...