201 to 225 of 269 Remote/Hybrid Grafana Jobs

Platform Engineer

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£580 - £610 per day
Platform/Observability Engineer | £610p/day (Outside IR35) | Remote | 6 months (Initially) | Grafana Cloud/Grafana Dashboard Our client is looking for an experienced Platform or Observability Engineer to focus on building and maturing their observability stack, with particular emphasis on Grafana Cloud, Alloy agent management, and dashboarding. …/day (Outside IR35) Duration: 6 months initial term, with scope for extension Working arrangement: Fully remote Key Responsibilities Set up and configure a Grafana Cloud test environment, ensuring it is production-representative and suitable for validating observability changes before rollout. Design, configure, and manage Alloy agent configurations, including planning ...

Platform Engineer

Hiring Organisation
Oscar Technology
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£580.00 - £610.00 per day
Platform/Observability Engineer | £610p/day (Outside IR35) | Remote | 6 months (Initially) | Grafana Cloud/Grafana Dashboard Our client is looking for an experienced Platform or Observability Engineer to focus on building and maturing their observability stack, with particular emphasis on Grafana Cloud, Alloy agent management, and dashboarding. …/day (Outside IR35) Duration: 6 months initial term, with scope for extension Working arrangement: Fully remote Key Responsibilities Set up and configure a Grafana Cloud test environment, ensuring it is production-representative and suitable for validating observability changes before rollout. Design, configure, and manage Alloy agent configurations, including planning ...

Site Reliability Engineer

Location
West of England, England, United Kingdom
production infrastructure operations, together with strong hands-on Python automation skills. You'll also need experience with: Monitoring and observability tools such as Prometheus, Grafana or similar Production incident management and/or on-call environments Automating operational runbooks and repetitive infrastructure processes APIs and systems integration Version-controlled automation ...

Staff Machine Learning Engineer - Ops

Location
Greater London, England, United Kingdom
/CD and Github Actions experience Strong communications skills with a collaborative mindset Desirable Experience with Pytorch, TensorRT, quantisation and model deployment Experience with Grafana monitoring and production observability This is a full-time role based in our office in London. At Wayve we want the best of all worlds ...

Support Engineer

Location
United Kingdom
Previous experience working in a startup or small team Experience with workflow orchestration systems (e.g. Temporal) Familiarity with monitoring and observability tools such as Grafana or Prometheus Familiarity with ML/AI systems at an operational level; you do not need to build models, but understanding how they fit into ...

Consulting Principal - Solution Architect

Location
Greater London, England, United Kingdom
controlled artefact management using tools such as Jenkins, GitLab and AWS CodePipeline. Establish effective monitoring, logging and incident-management capabilities using CloudWatch, Prometheus, Grafana and the ELK stack, while advising stakeholders on AWS container best practices. Work model We believe hybrid work is the way forward as we strive … implementing secure secrets management, Kubernetes security policies, network policies and controls for multi-tenant platforms. Knowledge of observability and centralized logging using CloudWatch, Prometheus, Grafana and ELK in high-security environments. Strong analytical and problem-solving skills, with the ability to create resilient solutions and manage technical ambiguity with accountability. ...

Platform Metal Engineer - Remote Infra & Kubernetes

Location
United Kingdom
Grafana Labs is seeking a Software Engineer to join the Platform Metal squad, focusing on on-premises resources and Kubernetes control plane. The role involves building and provisioning hardware-backed infrastructure and operating in a remote-first team across the UK. You will work with Go, Python and Shell, manage ...

Senior Performance Engineer | AI Infrastructure | Cambridge (Hybrid) |

Hiring Organisation
Pure Resourcing Solutions
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Full-Time
Salary
£90,000 - £120,000 per annum
inference, matrix multiplication, KV-caching, that level of detail Comfortable in profiling tools like Nsight or PyTorch Profiler, and monitoring stacks like Prometheus and Grafana Python for data work, Pandas and NumPy, plus general scripting Nice to have rather than essential: a postgraduate degree and research background (publications welcome), real ...

Site Reliability Engineer, Big Data (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
capabilities and observability. TechnologyApache Kafka for messaging layerHadoop and Ceph as distributed storage layerSQL Server backup and recoveryTerraform, Ansible, Puppet, ArgoCD for operational automationPrometheus, Grafana, Icinga and PagerDuty as observability layerBare-metal servers and hybrid cloud/on-prem infrastructureWhat matters: understand distributed systems, failure recovery, and operational patterns ...

Data Analyst

Location
Greater London, England, United Kingdom
count) Solid SQL skills (BigQuery experience is a plus) Familiarity with Python or R for data analysis Experience building dashboards or visualisations (e.g., Looker, Grafana, Tableau) Understanding of basic A/B testing concepts and experimentation frameworks Ability to transform data into clear insights and actionable recommendations Strong communication skills ...

Database Administrator

Location
Greater London, England, United Kingdom
replication and performance tuning. Experience in Linux environment Scripting & Automation: Shell scripting or Ansible for automation. Monitoring Tools: Experience with monitoring tools like PMM, Grafana, Nagios, or similar. Security & Compliance: Understanding of database security, access controls, and compliance standards. Problem-Solving: Strong troubleshooting skills and a proactive approach to issue ...

Staff Software Developer (IAM)

Location
Greater London, England, United Kingdom
experienced in building and operating SaaS production services in a distributed, multi-tenant environment Can fluently navigate a modern devops stack: K8S, OpenTelemetry, Grafana Have a track record of taking over and improving existing production services Nice to have: Prior experience with identity management protocols, such as OAuth 2.0, OIDC ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment
Location
North London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join … certification desirable Reference: BBBH27029A DevOps, DevOps Engineer, AWS, Cloud Security, Cyber, DevSecOps, Terraform, Ansible, Linux, Networking, GitHub Actions, CI/CD, Bash, Python, Go, Grafana, Prometheus, Woking, Remote, Surrey, London If you're interested in this role, click 'apply now' to forward an up-to-date copy of your CV. ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
Camden, London, Camden Town, United Kingdom
Employment Type
Permanent
Salary
£65000 - £70000/annum + Remote + Progression
private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join … certification desirable Reference: BBBH27029A DevOps, DevOps Engineer, AWS, Cloud Security, Cyber, DevSecOps, Terraform, Ansible, Linux, Networking, GitHub Actions, CI/CD, Bash, Python, Go, Grafana, Prometheus, Woking, Remote, Surrey, London If you're interested in this role, click 'apply now' to forward an up-to-date copy of your CV. ...

Cloud Infrastructure Engineer

Location
Greater London, England, United Kingdom
monitoring spend, eliminating waste, and rightsizing resources to balance performance and cost. Monitoring & Incident Management Monitor and manage platform activity using tools like Prometheus , Grafana , or AWS CloudWatch Respond quickly to alerts and incidents, independently resolving issues and ensuring service uptime. Conduct post‐incident reviews and help improve system resiliency … experience with AWS services and containerised applications Strong experience operating operational data stores (Aurora MySQL, DynamoDB). Expertise in using monitoring tools(e.g. Prometheus, Grafana, CloudWatch) for real‐time platform performance insights. Strong understanding of network security and Cloudflare, VPC and networking fundamentals, with a clear grasp of how traffic ...

Senior DevSecOps Engineer

Location
Manchester, England, United Kingdom
Lambda, RDS/Postgres, DynamoDB, SQS, Kinesis, S3, Cognito, Route53, VPC, EC2) Terraform, Kubernetes, Helm, Ansible, Puppet GitLab CI Observability & Security Prometheus, Grafana, OpenSearch, CloudWatch Okta (SSO/IdP) Vulnerability management, secrets management, penetration test tooling Application layer (context, not expectation) Scala, Kotlin, TypeScript, Python, built by the engineering team … stable. Own our platform security controls: vulnerability management, penetration test remediation, secrets management, and least-privilege access. Run our observability and monitoring platforms (Prometheus, Grafana, OpenSearch, CloudWatch), tuning alerts and driving improvements into the delivery pipeline. Continuously review and optimise our AWS infrastructure: Cost monitoring, right-sizing, capacity planning, reporting ...

Senior Azure DevOps Engineer

Location
Milton Keynes, England, United Kingdom
Senior Azure DevOps Engineer – Shared Records Reporting To: Chief Software Engineer Department: Development Location: Milton Keynes/Homebased - with occasional travel to the office in Milton Keynes Overview Graphnet Health is the leading UK supplier ...

Backend Engineer - Platform - Stacks | UK | Remote

Location
United Kingdom
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand … speed of their ambitions. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company ...

Application Support Analyst (Data Ops)

Hiring Organisation
Erin Associates
Location
Chorley, Lancashire, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£40,000
skills with excellent attention to detail. Ability to manage multiple priorities and work effectively in a fast-paced environment. Desirable Skills PowerShell, AWS, MySQL, Grafana, ITIL knowledge If you have not heard back from us within 5 working days, please assume that your application has been unsuccessful on this occasion. ...

Performance Engineer (Junior) | AI Infrastructure | Cambridge (Hybrid)

Hiring Organisation
Pure Resourcing Solutions Limited
Location
Linton, Dry Drayton, Cambridgeshire, United Kingdom
Employment Type
Permanent
Salary
£55000 - £70000/annum
work with GPU or accelerator code, CUDA or similar Familiarity with profiling tools (Nsight, PyTorch Profiler) and ideally some exposure to monitoring stacks (Prometheus, Grafana) Strong Python for data work, Pandas and NumPy, genuine scripting ability Nice to have: exposure to inference serving frameworks like vLLM, published research, or open ...

Tech Lead, Autonomy Performance - Robotaxi

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
including defining release processes, gates, rollback plans, and regression handlingFamiliarity with modern data/observability stacks used in autonomy programs (e.g., Databricks/SQL, Grafana, internal eval/sim tooling, dashboarding/reporting)Track record of building scalable triage systems and performance programs (metrics taxonomy, prioritisation frameworks, quality bars ...

Engineering Manager, Runtime Platform, Robot Software

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
automotive, robotics, or another safety-relevant real-time domain. Familiarity with profiling toolchains (pprof/gperftools, perf, Nsight, NVLumo) and observability stacks (OpenTelemetry, Grafana, Datadog).Experience with NVIDIA (Orin/Thor) and/or Qualcomm compute platforms. Experience with a micro-kernel and/or real-time OS (e.g. ...

Senior Software Engineer - Zero Gravity

Location
Reading, England, United Kingdom
EDAs (TypeScript,Node, Nest, Kafka, Rabbit, PHP,Symfony) Reverse-engineer complex code to solve problems Lead implementingSLIs and SLOs to better understand our services(Grafana) Develop andmaintainuser interfaces and web applications (React) ImproveCI/CD pipeline integrationswith wider-team and colleagues Collaborate with your colleagues and bea strong teamplayer Perform ...

Platform Engineer

Location
Greater London, England, United Kingdom
huge, distributed scale efficiently Monitoring and alerting: Measuring application performance and delivering insights, metrics and relevant alerts to the engineering teams with ELK, Grafana and New Relic Ownership: Driving engineering teams to own their infrastructure and costs by building great tooling, visibility and documentation Security: Setting the standards for fine … understanding of CI/CD and relevant tooling (we use GitHub Actions and Argo CD) Expertise in logging and monitoring at scale (e.g.S3, Graphite, Grafana, ELK, NewRelic, Datadog) Knowledge of a DevOps toolchain to drive ownership of a self-hosted platform Competent in Git and the GitOps philosophy Familiarity with ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Manchester, England, United Kingdom
workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python or Go Experience implementing CI/… native Container-based architecture Kubernetes (GKE on GCP, EKS on AWS) TeamCity for CI/CD (with multiple production releases per day) Terraform and Grafana RDS and CloudSQL for PostgreSQL Our Interview Process Interviewing is a two-way process and we want you to have the time and opportunity ...