451 to 475 of 875 Grafana Jobs

Site Reliability Engineer

Location
Greater London, England, United Kingdom
reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full ...

Lead Network Operations Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
Familiarity with automation and CI/CD tooling (e.g. Ansible, Python, Jenkins) and Infrastructure as Code Experience with observability and monitoring tools (e.g. Prometheus, Grafana, OpenTelemetry, ELK) Strong ability to analyse and troubleshoot distributed systems end-to-end Proactive, self-starting mindset with a strong sense of ownership Comfortable engaging ...

Core AI Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
with sandboxing and workload isolation technologiesExperience in quantitative finance or low-latency systemsAWS experience particularly in hybrid environmentsExperience with observability tooling such as Prometheus, Grafana or OpenTelemetryContributions to open-source projects in relevant domainsWhy join us? Highly competitive compensation plus annual discretionary bonusLunch provided (via Just Eat for Business ...

Software Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full ...

Software Engineer (EMS)

Location
Greater London, England, United Kingdom
tools for multi-region deployments . A working knowledge of databases and caches (PostgreSQL, SQLite, Redis, Zookeeper). Experience with observability tools (e.g. Prometheus, Grafana, ELK stack). Soft Skills: Strong analytical and problem-solving abilities. Excellent communication and collaboration skills, with a track record of working in cross-functional ...

Senior Software Developer

Location
City Of London, England, United Kingdom
services on dedicated hardware and on AWS We use Azure DevOps for build/deploy pipelines and story management Our observability stack is Splunk, Grafana, Pyroscope and Prometheus For this position You like working in a cross-functional team with Product Managers, Testers and DevOps engineers You aim to deliver ...

Platform Engineer (Barcelona)

Location
Cambourne, England, United Kingdom
workloads on Kubernetes (device plugins, node scheduling, NVIDIA GPU Operator) and with LLM serving tools (vLLM, Triton, NIM). Familiarity with observability tooling (Prometheus, Grafana, OpenTelemetry) and with exposing it for third‐party components. Exposure to ML orchestration tooling (Flyte, Airflow, MLflow, SkyPilot). Go, and a track record ...

Software Engineer

Location
Uxbridge, England, United Kingdom
Boot, JUnit Client-Side: Typescript, Next.js, React and various React ecosystem tools and libraries Infrastructure: AWS, Kubernetes, Terraform, Kafka, DynamoDB, PostgreSQL, Redis, ElasticSearch, Kibana, Grafana, and Prometheus. Be comfortable using a variety of frameworks, languages, and tools and be happy to learn new skills when the need arises. Key responsibilities ...

Feature Lead - Technology

Location
Greater London, England, United Kingdom
Linux based environments including shell basics as well as process and network diagnostics. Exposure to monitoring, metrics and tracing tooling - ELK stack, Splunk, Prometheus, Grafana, Graphite, OpenTSDB, OpenTrace, Jaeger Exposure to message-oriented architecture - ZeroMQ, JMS, AMPS, RabbitMQ, Kafka, Google pub/sub Exposure to process/container orchestration technologies ...

Software Engineer

Location
United Kingdom
development or related role. Broad experience of modern software delivery including cloud (e.g. AWS). Familiarity with common tools; Jira, Jenkins, Git, GitHub Actions, Grafana, Terraform, Kubernetes, Playwright B2B software company experience. Specific experience in SQL, Postgres, Data Science Artificial Intelligence and Machine Learning(AI & ML), and/or performance ...

Test Engineer

Location
United Kingdom
engineer or related role. Broad experience of modern software delivery including cloud (e.g. AWS). Familiarity with common tools; Jira, Jenkins, Git, GitHub Actions, Grafana, Terraform, Kubernetes, Playwright B2B software company experience. Competitive salary with regular pay reviews 25 days holiday per year The chance to work with some ...

Principal Architect — Distributed Systems & Event Streaming (Director I - Product Architect)

Hiring Organisation
UST Global
Location
London, UK
Employment Type
Full-time
Spark Streaming, ksqlDB, or similar technologies. Financial services, banking, lending, payments, or regulated industry experience. Kubernetes and cloud-native platforms. Observability stacks including Prometheus, Grafana, OpenTelemetry, and distributed tracing. CI/CD, contract testing, performance testing, chaos engineering, and production-readiness practices. Event governance, auditability, data lineage, and security ...

Software Engineer, AI Libraries

Location
Greater London, England, United Kingdom
large GPU clusters or distributed training environments. Familiarity with distributed training techniques such as DDP or FSDP. Experience with observability tools such as Prometheus, Grafana, Datadog, or OpenTelemetry. Experience with data pipeline orchestration tools such as Airflow, Flyte, Ray, Metaflow, or Argo Workflows. Experience with containerisation and infrastructure tooling such ...

Neo4j Platform Consultant

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
Security (RBAC, authentication, data protection)Preferred SkillsExperience with Graph RAG workloads and traversal optimizationDevOps tools (Docker, Kubernetes, CI/CD pipelines)Monitoring tools (Prometheus, Grafana, etc.)Experience with other graph databases (Neptune, TigerGraph)Job ExpectationsEnsure stable, secure, and high-performing Neo4j platform operationsEnable efficient graph query execution ...

Senior Software Engineer

Hiring Organisation
ASOS
Location
London, UK
Employment Type
Full-time
Policy and MonitorAzure Functions and Bicep — serverless pipelines and infrastructure as code across multiple environmentsEntra ID — managed identities and RBAC across subscriptionsApplication Insights and Grafana — telemetry, monitoring and alertingPower BI — semantic modelling, DAX, composite and DirectQuery models, TMDL and XMLADatabricks — Delta Lake and SQL warehousesKQL — Azure Monitor, Log Analytics ...

Lead Site Reliability Engineer

Location
Greater London, England, United Kingdom
experience in front office trading environments or similarly high pressure, low latency domains. Proficiency with SRE tooling and techniques, including FIX messaging, Kafka, Grafana, Splunk, ITRS Geneos, Dynatrace, InfluxDB, MQ (IBM MQ or similar), Oracle DB Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve ...

Platform Engineer

Location
Greater London, England, United Kingdom
with Kubernetes, and enough Terraform to have opinions about how to structure it. Familiarity with CI/CD pipelines and observability tooling (Prometheus, Grafana, Datadog, or similar). Strong technical foundation: Demonstrated ability to write production-quality code and solve hard technical problems (experience with Python, TypeScript/JavaScript, systems ...

Managing Engineer - Observability, Pipeline & Analytics (Hybrid)

Location
Belfast City District, Northern Ireland, United Kingdom
Kusto Query Language (KQL) and large‐scale analytics platforms. Experience with Cribl, ADX, Datadog Pipelines, Splunk, Sentinel, Kafka, Event Hubs, Open Telemetry, Elastic, Grafana, or similar observability ecosystems. Ability to leverage AI assisted development tools (e.g., Copilot, Cursor) responsibly to improve developer productivity and code quality. Supervisory Responsibilities: This role ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless or inference serving infrastructure. ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
cloud platforms (GCP and/or AWS) and infrastructure as code (e.g., Terraform). Experience with monitoring/observability tooling (e.g., Datadog, Monte Carlo, Grafana) for proactive detection of data quality and pipeline issues. Familiarity with CI/CD practices applied to data workflows (e.g., automated testing for pipelines, version ...

Staff Quality Engineer - Mobile

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
push back on ambiguous acceptance criteria, and surface risk before code is writtenClose the loop on production issues using our observability stack (Datadog, Sentry, Grafana) - tying test coverage back to real customer impactEnsure teams have Service Level Objectives set up and are achieving themRun targeted exploratory testing on high-risk ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Senior Engineering Manager - 9-10 month FTC

Location
Greater London, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Site Reliability Engineer

Location
Slough, England, United Kingdom
environments Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps) Package and deploy applications to customer environments Implement and optimise observability tooling (Prometheus, Grafana, Loki) Support incident response, monitoring, and backup/recovery planning Mentor project teams in DevSecOps practices and environment management Ensure cloud environments are secure, performant … cost-optimised Tech Environment & Skills: Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC) Containers: Kubernetes, Docker, Helm (OpenShift a plus) Observability: Prometheus, Grafana, Loki (network visualisation desirable) CI/CD & GitOps: GitHub Actions, ArgoCD/Flux Security: Cloud access models, Zero Trust principles Programming: Python or Golang preferred ...