576 to 600 of 608 Prometheus Jobs in the UK

Cloud Infrastructure Engineer (Open LMS) UK, Remote

Hiring Organisation
Learning Technologies Group
Location
United Kingdom
Salary
£ 70 K
configuration management (etcd)Managing and tuning a multi-tier caching strategy (Varnish, Redis/Valkey, PHP OPcache)Running and scaling our observability stack (Prometheus, Grafana, Loki, Fluentd, PagerDuty) and participating in on-call rotationsEvaluating and implementing distributed storage solutions as the platform evolvesImproving deployment workflows and release processesCollaborating with internal … networking fundamentalsUnderstanding of distributed systems concepts: consensus, leader election, distributed locking, eventual consistency, and the tradeoffs involvedProficiency in building and maintaining observability pipelines (Prometheus, Grafana, Loki, or equivalent) in productionComfortable working in a GitLab-based CI/CD workflowClear communicator who can document architectural decisions and explain technical tradeoffs ...

Cloud Infrastructure Engineer (Open LMS) UK, Remote

Hiring Organisation
Learning Technologies Group
Location
Moffat, Dumfries & Galloway, UK
Employment Type
Full-time
configuration management (etcd)Managing and tuning a multi-tier caching strategy (Varnish, Redis/Valkey, PHP OPcache)Running and scaling our observability stack (Prometheus, Grafana, Loki, Fluentd, PagerDuty) and participating in on-call rotationsEvaluating and implementing distributed storage solutions as the platform evolvesImproving deployment workflows and release processesCollaborating with internal … networking fundamentalsUnderstanding of distributed systems concepts: consensus, leader election, distributed locking, eventual consistency, and the tradeoffs involvedProficiency in building and maintaining observability pipelines (Prometheus, Grafana, Loki, or equivalent) in productionComfortable working in a GitLab-based CI/CD workflowClear communicator who can document architectural decisions and explain technical tradeoffs ...

Cloud Infrastructure Engineer (Open LMS) UK, Remote

Location
United Kingdom
configuration management (etcd) Managing and tuning a multi-tier caching strategy (Varnish, Redis/Valkey, PHP OPcache) Running and scaling our observability stack (Prometheus, Grafana, Loki, Fluentd, PagerDuty) and participating in on-call rotations Evaluating and implementing distributed storage solutions as the platform evolves Improving deployment workflows and release processes … fundamentals Understanding of distributed systems concepts: consensus, leader election, distributed locking, eventual consistency, and the tradeoffs involved Proficiency in building and maintaining observability pipelines (Prometheus, Grafana, Loki, or equivalent) in production Comfortable working in a GitLab-based CI/CD workflow Clear communicator who can document architectural decisions and explain ...

Network Automation & OSS Designer

Location
Greater London, England, United Kingdom
pipelines. Architect AIOps capabilities including closed‐loop automation, anomaly detection, and predictive analytics for network operations. Integrate OSS observability with cloud‐native monitoring stacks (Prometheus, Grafana, OpenTelemetry, Elasticsearch). Lead design of intent‐based networking and policy‐driven automation frameworks. Collaborate with product managers, network engineers, platform teams, and DevOps … MANO, VNF/CNF lifecycle management). Understanding of AIOps platforms and closed‐loop automation design for network operations. Experience with observability tooling: OpenTelemetry, Prometheus, Grafana, Jaeger, Loki, ELK Stack. Knowledge of ML/AI model integration for anomaly detection, root‐cause analysis, and predictive network management. Strong grasp ...

DevOps & Infrastructure Engineer

Location
Gloucester, England, United Kingdom
solutions. Develop and maintain CI/CD pipelines, GitOps workflows and automated deployment approaches using tools such as ArgoCD. Implement and improve observability using Prometheus, Grafana, logging and alerting to support resilient platform operations. Use infrastructure-as-code and platform automation with Helm, Go and Terraform to deliver repeatable, assured …/CD and GitOps tooling experience, ideally including ArgoCD and automated deployment pipelines. Good understanding of observability, monitoring and alerting using tools such as Prometheus and Grafana, alongside security, networking, logging, secrets management and operational assurance. Able to learn new technologies quickly and help others adopt them safely and effectively. ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
integrations with third-party custodians. Build out Kubernetes and Docker deployments as we containerize more of the custody stack. Set up monitoring and alerting (Prometheus, Grafana, or equivalent) so we know about problems before customers do. Apply security controls and standards throughout - access control, key management, incident response. Provide … skills, with real production experience. A strong security focus - you've worked on systems where key management and access control are critical. Experience with Prometheus, Grafana, or an equivalent monitoring stack. Experience integrating with third-party custodians like BitGo or Fireblocks - this is a key differentiator for this role. What ...

Senior Software Engineer, Custody Services

Hiring Organisation
Robinhood Financial
Location
London, UK
Employment Type
Full-time
integrations with third-party custodians. Build out Kubernetes and Docker deployments as we containerize more of the custody stack. Set up monitoring and alerting (Prometheus, Grafana, or equivalent) so we know about problems before customers do. Apply security controls and standards throughout — access control, key management, incident response. Provide … skills, with real production experience. A strong security focus — you've worked on systems where key management and access control are critical. Experience with Prometheus, Grafana, or an equivalent monitoring stack. Bonus pointsExperience integrating with third-party custodians like BitGo or Fireblocks — this is a key differentiator for this role. ...

Business Consultant - Databricks

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
Collaborate with data scientists and platform teams to improve query performance on Delta Lake. Develop automated monitoring and alerting mechanisms using Databricks REST APIs, Prometheus, or Azure Monitor. Maintain best practices for data engineering performance, testing, and documentation. Technical Skills Required Strong proficiency with Databricks, Apache Spark (SQL, PySpark, Scala … offs. Version control (Git, GitHub Actions, DevOps pipelines) and CI/CD practices for Databricks. Desirable Skills Experience with Databricks monitoring and observability (Ganglia, Prometheus, or Datadog). Understanding of serverless Databricks clusters. Familiarity with Terraform and infrastructure as code for Databricks resource management. Exposure to MLflow and model-serving ...

Database Engineer (MongoDB, Postgres)

Hiring Organisation
CGG
Location
Oxford, Oxfordshire, United Kingdom
Salary
£ 70 K
replica sets.Support PostgreSQL environments, including day-to-day administration, backup, and recovery.-Performance & OptimisationMonitor database health, performance, and availability using tools such as Prometheus, Grafana, and MongoDB Ops Manager.Analyse and optimise database performance, queries, and configurations.Develop strategies to support data growth, scalability, and high-traffic workloads.-Backup, Recovery & ReliabilityImplement backup … production environments, including sharded clusters and replica sets.Experience supporting PostgreSQL or other enterprise relational databases.Experience with MongoDB Ops Manager and monitoring tools such as Prometheus and Grafana.Strong knowledge of database backup, recovery, PITR, and High Availability strategies.Experience troubleshooting and optimising database performance.Strong analytical and problem-solving skills.Ability to work effectively ...

Database Engineer (MongoDB, Postgres)

Hiring Organisation
CGG
Location
Oxford, Oxfordshire, UK
Employment Type
Full-time
replica sets. Support PostgreSQL environments, including day-to-day administration, backup, and recovery.-Performance & OptimisationMonitor database health, performance, and availability using tools such as Prometheus, Grafana, and MongoDB Ops Manager. Analyse and optimise database performance, queries, and configurations. Develop strategies to support data growth, scalability, and high-traffic workloads.-Backup … including sharded clusters and replica sets. Experience supporting PostgreSQL or other enterprise relational databases. Experience with MongoDB Ops Manager and monitoring tools such as Prometheus and Grafana. Strong knowledge of database backup, recovery, PITR, and High Availability strategies. Experience troubleshooting and optimising database performance. Strong analytical and problem-solving skills. ...

Database Engineer (MongoDB, Postgres)

Location
Crawley, England, United Kingdom
sets.* Support PostgreSQL environments, including day-to-day administration, backup, and recovery.**-Performance & Optimisation*** Monitor database health, performance, and availability using tools such as Prometheus, Grafana, and MongoDB Ops Manager.* Analyse and optimise database performance, queries, and configurations.* Develop strategies to support data growth, scalability, and high-traffic workloads.**-Backup … including sharded clusters and replica sets.* Experience supporting PostgreSQL or other enterprise relational databases.* Experience with MongoDB Ops Manager and monitoring tools such as Prometheus and Grafana.* Strong knowledge of database backup, recovery, PITR, and High Availability strategies.* Experience troubleshooting and optimising database performance.* Strong analytical and problem-solving skills. ...

Platform Site Reliability Engineer

Location
Gloucester, England, United Kingdom
tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance Radiant’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root cause analysis, document … routing, switching Strong experience with API interrogation Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana preferred) Strong grasp of ITSM and service operation best practices Excellent communication and mentorship skills Comfortable interfacing with internal stakeholders and external customers Bonus ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
provision of tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance ’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root cause analysis …/IP, DNS, DHCP, VLANs, routing, switching Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana) Hands-on experience operating orchestration platforms (Kubernetes, MAAS, Tinkerbell) Strong grasp of ITSM and service operation best practices Excellent communication and mentorship skills Comfortable ...

Forward Deployed Engineer - Lead Platform Engineer

Hiring Organisation
Kyndryl
Location
London, UK
Employment Type
Full-time
with security as a first-class concern (policy-as-code/OPA, access controls, secrets management, compliance-driven engineering) Instrument platforms for observability (Grafana, Prometheus, OpenTelemetry) Provide architectural oversight across multi-disciplinary workstreams, staying close enough to unblock the team directly Capture field learnings, codify reusable patterns and blueprints … endpoints Hands-on experience with CI/CD pipelines, Git-based workflows, and microservices/API architectures Practical experience with observability stacks (Grafana, Prometheus, OpenTelemetry) Experience with generative AI platforms: LLM hosting, LLM gateways (e.g. LiteLLM, Portkey, Kong AI Gateway), and MLOps/LLMOps practices for deploying and monitoring ...

Cloud Engineer

Location
Greater London, England, United Kingdom
Winton is a research-based investment management company with a specialist focus on statistical and mathematical inference in financial markets. The firm researches and trades quantitative investment strategies, which are implemented systematically via thousands of ...

Senior Software Engineer

Hiring Organisation
Ask4.com
Location
Sheffield, South Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
Salary
£60,000
environments Integrate with network management systems, message brokers, and third-party APIs Instrument and monitor applications using observability tooling (metrics, logs, and traces Grafana, Prometheus, or similar) Provide technical mentorship to mid-level engineers and act as an escalation point for complex problems Contribute to development standards, code review practices … Exposure to OpenWiFi, OpenWrt or similar open-source network controller frameworks Observability experience. Use of metrics, logs, and traces using tools such as Grafana, Prometheus and Sentry Experience with AI agent development or LLM integration Agile/Scrum working practices ISP or managed connectivity background Automated configuration management (Ansible, Puppet ...

Kubernetes Platform Engineer

Location
Greater London, England, United Kingdom
ArgoCD, FluxCD) for safe, auditable changes. Drive Infrastructure as Code practices with Terraform and Helm for reliable and repeatable builds. Heavily embed observability using Prometheus, Grafana, and OpenTelemetry to make systems measurable and reliable. Stay ahead of Kubernetes evolution by testing and adopting new versions and features early. Collaborate with … policies, and Open Policy Agent. Proven ability to troubleshoot complex performance and reliability issues across infrastructure and workloads. Experience with observability tools such as Prometheus, Grafana, and OpenTelemetry to monitor cluster metrics and health. Great communication skills, with experience collaborating with internal platform users to gather feedback and deliver improvements. ...

Kubernetes Platform Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
ArgoCD, FluxCD) for safe, auditable changes. Drive Infrastructure as Code practices with Terraform and Helm for reliable and repeatable builds. Heavily embed observability using Prometheus, Grafana, and OpenTelemetry to make systems measurable and reliable. Stay ahead of Kubernetes evolution by testing and adopting new versions and features early. Collaborate with … policies, and Open Policy Agent. Proven ability to troubleshoot complex performance and reliability issues across infrastructure and workloads. Experience with observability tools such as Prometheus, Grafana and OpenTelemetry to monitor cluster metrics and health. Great communication skills, with experience collaborating with internal platform users to gather feedback and deliver improvements. ...

Senior HPC Engineer

Hiring Organisation
Hays
Location
London, UK
Employment Type
Full-time
Reference: 4546575Job ID: 5416319Posted: 2026-09-23Closing date: 2026-12-21Location: LondonSalary: Up to 70k (per annum)Job type: PermanentWorking pattern: Full timeIndustry: Scientific and R&D/School TechniciansCompany: HaysConsultant: Lorenz PaschHays ...

Senior HPC Engineer

Hiring Organisation
Hays
Location
London, United Kingdom
Salary
£ 70 K
Reference: 4546575Job ID: 5416319Posted: 2026-09-23Closing date: 2026-12-21Location: LondonSalary: Up to 70k (per annum)Job type: PermanentWorking pattern: Full timeIndustry: Scientific and R&D/School TechniciansCompany: HaysConsultant: Lorenz PaschHays ...

Senior HPC Engineer

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£60,000 - £70,000 per annum
Senior HPC Engineer Please read the advert below carefully, and if you are a good match I want to speak to you ASAP. Please call or email Lorenz Pasch at Hays Recruitment - my contact details ...

Network Engineer

Location
Greater London, England, United Kingdom
security posture of network and network security platforms end to end Enhancing observability across telemetry, alerting and performance monitoring using tools such as Grafana, Prometheus and ThousandEyes to enable proactive operations Implementing and maintaining scalable, resilient datacentre network infrastructure built on Cisco and Arista technologies Championing automation of operational tasks … Network Engineer in enterprise or large-scale environments Experience applying SRE, observability and automation principles to networking, using technologies such as Python, Prometheus, Grafana, OpenTelemetry, Ansible and Jenkins Experience with Cisco and Arista switching and routing, alongside network security infrastructure, including firewalls, IDS/IPS and network segmentation Expertise ...

Senior Network Engineer

Location
Greater London, England, United Kingdom
Using BGP, EVPN-VXLAN, JunOS (Juniper QFX/MX), OSPF, Spine-Leaf/IP Fabric, VLANs, VRFs, Linux networking, Ansible, Terraform, Prometheus/Grafana, Observability tooling The adventures that await you after becoming Senior Network Engineer at Hack The Box: Design and implement spine-leaf network architectures across multiple data … running JunOS Develop and maintain network automation using Ansible, Terraform or similar infrastructure‐as‐code tooling Establish and improve network monitoring, alerting and observability (Prometheus, Grafana, SNMP, streaming telemetry) Plan and execute network capacity upgrades, site bring‐ups and hardware refresh cycles Collaborate with the platform/systems engineering team ...

Junior Network and Automation Engineer

Hiring Organisation
Venn Telecom UK
Location
Exeter, Devon, South West, United Kingdom
Employment Type
Permanent
Salary
£35,000
Venn Telecom UK is looking for a curious and motivated Junior Network and Automation Engineer to join our infrastructure team. This is an entry-level role suited to a recent university graduate, or someone with ...

Software Engineer

Location
Uxbridge, England, United Kingdom
Full/part time : Full time Business area : Tech, Software Location : Uxbridge (can be remote with monthly office visits) A bit about giffgaff Do you want to join a connectivity provider that’s up to ...