326 to 350 of 835 Permanent Grafana Jobs

Safety Engineer - Free Tier Abuse

Location
Greater London, England, United Kingdom
Infrastructure & DevOps proficiency: cloud platforms (AWS/GCP), containerization (Docker/K8s), CI/CD pipelines Observability mindset with experience in monitoring tools (Prometheus, Grafana) and building observable systems Track record of taking products or systems from 0→1 with measurable impact, including deploying or working alongside ML/ ...

Database Engineer (MongoDB, Postgres)

Location
Crawley, England, United Kingdom
Support PostgreSQL environments, including day-to-day administration, backup, and recovery.**-Performance & Optimisation*** Monitor database health, performance, and availability using tools such as Prometheus, Grafana, and MongoDB Ops Manager.* Analyse and optimise database performance, queries, and configurations.* Develop strategies to support data growth, scalability, and high-traffic workloads.**-Backup, Recovery ...

DevOps Engineer

Location
Belfast City, Northern Ireland, United Kingdom
Kubernetes) Experience setting up CI/CD pipelines using GitHub Actions or similar tools Familiarity with monitoring and alerting tools (e.g. Prometheus, Grafana, CloudWatch, Sentry, DataDog) A security‐first mindset when designing and managing infrastructure Nice to haves: Experience working in regulated or high‐trust environments Knowledge of zero‐downtime ...

Lead Java Developer

Location
Greater London, England, United Kingdom
Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc. Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches Hands‐on knowledge of Linux/ ...

Platform Engineer

Location
Greater London, England, United Kingdom
security fundamentals Experience troubleshooting production systems and working through operational issues Useful experience Terragrunt or Helm GitHub Actions, ArgoCD or Flux Datadog, Prometheus or Grafana PostgreSQL or MySQL Supporting developer-experience or internal-platform improvements Working with security tooling or policy-as-code We don’t expect every ...

Lead Java Developer

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
gRPC etcProficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuningExperience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc. Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed cachesHands-on knowledge of Linux/UnixExperience ...

Principal DevSecOps Engineer

Location
Filton, England, United Kingdom
e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager)Containers and orchestration (e.g. Docker, AWS EKS) ; Infrastructure as Code (e.g. Terraform)Observability (e.g. Grafana, Loki) ;Scripting and automation (e.g. Python, Bash)Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies)Experience delivering within the UK Government Digital Service ...

Lead Site Reliability Engineer

Hiring Organisation
London Stock Exchange Group
Location
Nottingham, UK
Employment Type
Full-time
cloud security principles and experience collaborating with security teamsExperience with cloud cost optimisation strategies and toolingHands-on experience integrating AI with observability stacks (Prometheus, Grafana, ELK, OpenTelemetry) for proactive issue detection. Good to have SkillsExperience or working knowledge of Microsoft AzureExperience supporting multi-cloud or hybrid environmentsExposure to Infrastructure ...

Enablement Pod Platform SME IRC304283

Location
Glasgow, Scotland, United Kingdom
Helm, Terraform modules) Background in developer experience research - understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) - particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline components ...

Foundation Engineering - SRE Platforms - Site Reliability Engineer – Associate - London

Location
Greater London, England, United Kingdom
systems, data structures, algorithms and software design fundamentals. Hands‐on experience with observability tooling, including metrics, logging, tracing and dashboarding platforms such as Prometheus, Grafana, ELK or OpenTelemetry. Proven ability to investigate production issues, identify root causes and deliver durable engineering fixes that improve system behaviour and reduce repeat incidents. ...

Enablement Pod Platform SME IRC304282

Location
Greater London, England, United Kingdom
Helm, Terraform modules) Background in developer experience research - understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) - particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline components ...

Lead Java Developer

Location
Greater London, England, United Kingdom
Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning* Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc.* Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches* Hands-on knowledge of Linux/ ...

Senior Software Development Engineer

Location
Reading, England, United Kingdom
.NET (latest versions) Kubernetes & Docker Azure (SQL, CosmosDB, cloud services) PostgreSQL TypeScript/modern web frameworks (e.g. Vue.js) Observability tooling (e.g. Azure Monitor, Prometheus, Grafana) Azure DevOps/CI-CD pipelines Holidays: 25 days per annum + 8 days bank holidays (options to buy/sell days) 37.5 hour working ...

Lead SRE - Chase UK

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
service discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive ...

Software Engineer (ML Projects)

Location
Greater London, England, United Kingdom
cloud‐native TeamCity for CI/CD (lots of teams are releasing code 15-20 times per day!) Terraform Prometheus and Grafana If you have built and deployed complex Python applications or have hands‐on experience with generative AI and LLMs, we would be especially keen to talk. ...

Senior SRE

Location
Horsell, England, United Kingdom
operating within a complex multi-cloud/hybrid ecosystem, with a solid understanding of distributed systems. Proficient in observability stacks such as Prometheus, Grafana, Loki, and Tempo. Hands on experience with EDB Postgres for enterprise-grade database solutions. Ability to develop and maintain Infrastructure as Code (IaC) using tools like ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Hiring Organisation
Starling Bank
Location
London, UK
Employment Type
Full-time
workloads and CI/CDExperience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana)Experience setting up Google Workspace/Google Cloud IdentityExperience with automation using a scripting language like Python or GoExperience implementing CI/CD pipelinesAn understanding ...

Senior Infrastructure Engineer (GCP) - Engine by Starling

Hiring Organisation
Starling Bank
Location
London, UK
Employment Type
Full-time
workloads and CI/CDExperience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana)Experience setting up Google Workspace/Google Cloud IdentityExperience with automation using a scripting language like Python or GoExperience implementing CI/CD pipelinesAn understanding ...

Network Automation Engineer

Location
Greater Manchester, England, United Kingdom
administration and Bash/Python scripting AWS and/or Azure networking Kubernetes or container platform experience Git and version control Monitoring using Prometheus, Grafana or similar Strong understanding of NetDevOps principles Agile/Scrum delivery experience Desirable Cisco networking Palo Alto or Fortinet OCI Helm GitOps ArgoCD ServiceNow Tech … Stack Terraform, Ansible, AWS, Azure, Kubernetes, Docker, Git, GitLab CI, Jenkins, GitHub Actions, Linux, Bash, Python, Prometheus, Grafana, Helm, GitOps #J-18808-Ljbffr ...

Senior Site Reliability Engineer

Location
Southampton, England, United Kingdom
/CD, or CircleCI. Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch). Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems. Experience of Incident management and blameless postmortems that … have an advantage if you also have: Handson experience of working with large Kubernetes Cluster. Certification will be an added plus. Working experience of Grafana Observability Suite (Loki, Mimir, Tempo). Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
/CD, or CircleCI. Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch). Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems. Experience of Incident management and blameless postmortems that … have an advantage if you also have: Handson experience of working with large Kubernetes Cluster. Certification will be an added plus. Working experience of Grafana Observability Suite (Loki, Mimir, Tempo). Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity ...

Senior Software Engineer

Hiring Organisation
Visa
Location
Basingstoke, Hampshire, UK
Employment Type
Full-time
experience with cloud platforms, AI technologies, middleware systems, and modern engineering practicesCore TechnologiesPythonJavaGoAI - LLM Frameworks (LangGraph, LangChain, Agent Frameworks)AWS - GCP - AzureKubernetes - DockerTerraform - AnsibleJenkinsPrometheus - Grafana - SplunkKafka - FlinkSpring BootTomcatIBM MQDataPowerLinux - UnixWhat You'll DoBuild Production-Grade SoftwareDesign and develop scalable automation platforms and reliability toolingWrite clean, maintainable code in Python, Java … Docker and KubernetesExperience with CI-CD pipelines and automation tools including Jenkins, Terraform, or AnsibleWorking knowledge of observability and monitoring tools such as Prometheus, Grafana, or SplunkStrong troubleshooting, debugging, and problem-solving skillsExperience working in Linux-Unix environmentsStrong written and verbal communication skillsNice to Have (Not Required)Experience with middleware ...

Platform Engineer

Location
Greater London, England, United Kingdom
Platform Engineer Job Reference Number: PG1 Our client, a leading global supplier for IT services, requires Platform Engineer to be based at their client’s office in London, UK. This is a hybrid role – you ...

DV Cleared SRE — On-Site in Gloucester (6m)

Location
Greater London, England, United Kingdom
week. You will work with Terraform, Ansible/Chef, Docker and Kubernetes, CI/CD through Jenkins, monitoring with InfluxDB/Prometheus/Grafana, RabbitMQ messaging, SQL databases, Linux CLI, and AWS cloud services (EC2, RDS, S3, Lambda). #J-18808-Ljbffr ...

QA Automation Lead

Hiring Organisation
Nitya Software
Location
Grapevine, Texas, United States
Employment Type
Any
Salary
USD Annual
Must Have: 6 8 years of Automation Development 3+ years of experience in the Security. Strong Python scripting and automation framework development Hands-on Grafana experience for dashboards, alerts, and reporting Strong understanding of IAM, Security Monitoring, and Compliance Experience designing modular and scalable automation frameworks Strong experience with REST ...