151 to 175 of 297 Remote/Hybrid Grafana Jobs

Software Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full ...

Platform Engineer (Barcelona)

Location
Cambourne, England, United Kingdom
workloads on Kubernetes (device plugins, node scheduling, NVIDIA GPU Operator) and with LLM serving tools (vLLM, Triton, NIM). Familiarity with observability tooling (Prometheus, Grafana, OpenTelemetry) and with exposing it for third‐party components. Exposure to ML orchestration tooling (Flyte, Airflow, MLflow, SkyPilot). Go, and a track record ...

Software Engineer

Location
Uxbridge, England, United Kingdom
Boot, JUnit Client-Side: Typescript, Next.js, React and various React ecosystem tools and libraries Infrastructure: AWS, Kubernetes, Terraform, Kafka, DynamoDB, PostgreSQL, Redis, ElasticSearch, Kibana, Grafana, and Prometheus. Be comfortable using a variety of frameworks, languages, and tools and be happy to learn new skills when the need arises. Key responsibilities ...

Neo4j Platform Consultant

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
Security (RBAC, authentication, data protection)Preferred SkillsExperience with Graph RAG workloads and traversal optimizationDevOps tools (Docker, Kubernetes, CI/CD pipelines)Monitoring tools (Prometheus, Grafana, etc.)Experience with other graph databases (Neptune, TigerGraph)Job ExpectationsEnsure stable, secure, and high-performing Neo4j platform operationsEnable efficient graph query execution ...

Managing Engineer - Observability, Pipeline & Analytics (Hybrid)

Location
Belfast City District, Northern Ireland, United Kingdom
Kusto Query Language (KQL) and large‐scale analytics platforms. Experience with Cribl, ADX, Datadog Pipelines, Splunk, Sentinel, Kafka, Event Hubs, Open Telemetry, Elastic, Grafana, or similar observability ecosystems. Ability to leverage AI assisted development tools (e.g., Copilot, Cursor) responsibly to improve developer productivity and code quality. Supervisory Responsibilities: This role ...

Staff Quality Engineer - Mobile

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
push back on ambiguous acceptance criteria, and surface risk before code is writtenClose the loop on production issues using our observability stack (Datadog, Sentry, Grafana) - tying test coverage back to real customer impactEnsure teams have Service Level Objectives set up and are achieving themRun targeted exploratory testing on high-risk ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Senior Engineering Manager - 9-10 month FTC

Location
Greater London, England, United Kingdom
content management, search and recommendations capabilities. Terraform for infrastructure as code. GitHub/GitLab and CI/CD practices supporting frequent, reliable delivery. Grafana and AWS CloudWatch for monitoring and observability. Automated testing and engineering practices focused on building quality into the development lifecycle. Production health, incident management and strong ...

Platform Engineer

Location
Greater London, England, United Kingdom
Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest … supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow ...

Senior .NET Backend Developer

Hiring Organisation
Oscar Associates (UK) Limited
Location
York, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Kubernetes. PostgreSQL, Redis or Elasticsearch. GraphQL. RabbitMQ or other messaging technologies. CI/CD pipelines and modern DevOps practices. Observability tooling such as Grafana, OpenTelemetry or Prometheus. Experience using AI-assisted development tools within the software development lifecycle. What's on Offer Hybrid working (2 days per week in York ...

Cloud Operations Engineer (remote - London)

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
will be: Experienced and Certified in cloud computing with AWSExperience in a public cloud such as AWSKnowledge of monitoring and alerting technologies such as Grafana, Prometheus,Expereince of working with of Docker & Kubernetes and Container technology in productionWindows and Linux Operating System Management TechniquesSolid understanding of the OSI ModelExperience ...

SRE Technical Lead

Hiring Organisation
83zero Limited
Location
Wokingham, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
supporting hybrid and multi-cloud platforms. Experience with service mesh technologies such as Istio. Strong hands-on experience with observability tooling including Prometheus, Grafana, Loki, Tempo and OpenTelemetry. Infrastructure as Code and GitOps expertise using tools such as Helm, Kustomize, ArgoCD and Tekton. Experience building and improving CI/ ...

Site Reliability Engineer

Location
Milton Keynes, England, United Kingdom
below would be considered an advantage for any potential candidate, but are not essential: Experience or exposure of Terraform IAC software tooling Understanding of Grafana Analytics and Monitoring Solution Nessus Vulnerability Assessment Solution Knowledge of ITIL Foundation principles and application Knowledge of ISO27001, ISO27018, ISO9001 and CE+ Certifications Your security ...

Senior Scala Engineer CGEMJP00355784

Location
Greater London, England, United Kingdom
Essential: Scala Play or other MVC/Rest API frameworks SQL AWS Suite Continuous Integration Agile methodologies Desirable: Containerisation principles/Docker Jenkins Kibana Grafana Airflow If you receive suspicious outreach claiming to be from us, please contact us via the ManpowerGroup website. #J-18808-Ljbffr ...

Python Backend Developer

Location
Greater London, England, United Kingdom
work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Lead Software Developer ( SC Cleared )

Location
Swansea, Wales, United Kingdom
senior on-call responder and drive long-term resilience improvements. Required Skills & Experience You'll be working with: Spring Boot Terraform GitHub Actions Splunk Grafana Strong experience designing, building, and operating production software systems. Confident with modern development standards, including test-driven development, continuous integration/continuous deployment, and code ...

Java (Kotlin) Developer (Agile, Test-Driven) AVP

Location
Greater London, England, United Kingdom
Shift) Messaging Technologies (Kafka, Solace, TIBCO) Database/Data Store/Data Query Technologies (SQL Server, S3) Observability Technologies (OpenTelemetry, Elastic Stack/ELK, Grafana) Qualifications Proven experience in an App Dev role. Demonstrated execution capabilities. Education Bachelor’s/University degree or equivalent experience in a similar role This ...

Java (Kotlin) Developer (Agile, Test-Driven) AVP

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
Shift) Messaging Technologies (Kafka, Solace, TIBCO) Database/Data Store/Data Query Technologies (SQL Server, S3) Observability Technologies (OpenTelemetry, Elastic Stack/ELK, Grafana) Qualifications: Proven experience in an App Dev role. Demonstrated execution capabilities. Education: Bachelor's/University degree or equivalent experience in a similar roleThis ...

Senior Backend Engineer

Hiring Organisation
Plum Fintech
Location
London, UK
Employment Type
Full-time
life cycle of the softwareOur Tech Stack: Tech Stack: Kotlin with Spring Boot & Python with FastAPIDatastores: Postgres, BigQuery, Infrastructure: GCP (kubernetes, docker), RabbitMQ, TerraformMonitoring: Grafana, Prometheus, Datadog, Incident.io, SentryWhat to Expect from Our Hiring ProcessAt Plum, we value a lot the time you devote to the hiring process, this ...

System Test Engineer London, United Kingdom

Location
Greater London, England, United Kingdom
/CD principles and related tooling (e.g., GitLab CI, Buildkite, Bazel). Experience with test automation dashboards, logging, and reporting infrastructure (e.g., Grafana, Looker, Datadog, JIRA). Experience developing or testing automotive software across embedded and system levels. Strong cross‐functional communication skills; able to work collaboratively across software, hardware ...

AI Platform Support Engineer (EMEA)

Location
Greater London, England, United Kingdom
including networking, storage, process management, and performance tuning Experience with cloud infrastructure and distributed systems Experience with observability and debugging tools such as Prometheus, Grafana, or OpenTelemetry ML Infrastructure Experience Hands on experience operating machine learning workloads in production or research environments Experience with distributed ML systems and tooling such ...

Product Manager - Data

Location
Greater London, England, United Kingdom
standards such as ISO 26262 and ISO 21434 Application Management - Open source solutions in the enterprise including Observability, IAM, App Stores and technologies such Grafana, GitOps, and Juju Charms We will route you to the most suitable team. Location: These roles are home based in the EMEA time zone. ...

Senior Software Engineer II, Developer Experience / Operational Excellence

Location
Greater London, England, United Kingdom
metrics and data analysis Proven track record architecting monitoring frameworks, SLO platforms, and automated response workflows Datadog (or equivalent observabilty tooling like New Relic, Grafana). Proven experience working on large-scale enterprise software applications Experience in Developer Experience (DevEx) & Internal Portals: Designing and implementing solutions/tools that centralise ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
Stand Out If You Have: Practical experience managing large-scale Kubernetes clusters; certifications in Kubernetes are a strong bonus Hands-on familiarity with the Grafana Observability Suite, including tools like Loki, Mimir, and Tempo Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty …/CD, or CircleCI Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch Excellent analytical and troubleshooting abilities, especially within complex distributed systems Proven experience handling incident management and conducting blameless postmortems, including ...