76 to 100 of 163 Remote Prometheus Jobs

Senior Infrastructure Engineer (GCP) - Engine by Starling

Location
Cardiff, Wales, United Kingdom
Identity Federation for keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python ...

Senior Infrastructure Engineer (GCP) - Engine by Starling

Location
Greater London, England, United Kingdom
Identity Federation for keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Manchester, England, United Kingdom
Identity Federation for keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python ...

Staff Infrastructure Engineer (GCP) - Engine by Starling

Location
Southampton, England, United Kingdom
Identity Federation for keyless authentication of workloads and CI/CD Experience with observability tooling — Cloud Monitoring, Cloud Logging, Cloud Trace, Managed Service for Prometheus and OpenTelemetry (we also use Grafana) Experience setting up Google Workspace/Google Cloud Identity Experience with automation using a scripting language like Python ...

Software engineering specialist

Hiring Organisation
Randstad Digital
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £700 per day
deployments, and tool setups directly into developer workflows. Tech Stack Delivery: Automate the provisioning and configuration of our operational ecosystem, including tools across observability ( Prometheus, Elastic, Checkmk ), security/compliance ( Tenable, Red Hat Satellite ), service registry/IPAM ( NetBox ), container orchestration ( ArgoCD ), and artifact management ( Artifactory, GitLab ). Engineering Standards ...

Java Software Engineer - VP

Hiring Organisation
Henderson Scott
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
Docker, Kubernetes, and relational databases (MS SQL, Sybase). Delivery Automation & Monitoring: Strong focus on automated testing, automated release pipelines, and observability tools (Grafana, Prometheus). Mindset: Delivery-focused problem solver with a hands-on approach and strong stakeholder communication skills. Desirable Experience Background in Equity Swaps processing, Equity Derivatives ...

Platform Engineer AWS IaC

Location
Greater London, England, United Kingdom
seeking to improve operational efficiency through scripting, tooling and automation. You'll monitor and optimise AWS environments for performance, reliability and cost using CloudWatch, Prometheus, Grafana and other observability tools, implement AWS security best practice, including IAM, roles, policies and least privilege access. As a senior member of the team ...

Infrastructure Python Developer

Hiring Organisation
Hays
Location
Sheffield, South Yorkshire, Yorkshire, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
Up to £400.0 per day + Inside IR35
software delivery processes. Experience working within Agile environments and using tools such as JIRA. Desirable skills and experience Knowledge of observability tools such as Prometheus, Grafana and OpenTelemetry. Experience with enterprise tooling including Control-M, TrueSight, Guardium, Tenable Nessus or Delinea. Exposure to HashiCorp Vault for secrets management. Knowledge ...

Infrastructure Python Developer

Hiring Organisation
Hays Specialist Recruitment Limited
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Full-Time
Salary
£400.00 per day
software delivery processes. Experience working within Agile environments and using tools such as JIRA. Desirable skills and experience Knowledge of observability tools such as Prometheus, Grafana and OpenTelemetry. Experience with enterprise tooling including Control-M, TrueSight, Guardium, Tenable Nessus or Delinea. Exposure to HashiCorp Vault for secrets management. Knowledge ...

Solution Engineer (Pre-Sales)

Location
Greater London, England, United Kingdom
technical architecture.* Exceptional communication and presentation skills.* Proven ability in technical integrations and conducting POCs.* In-depth knowledge of Kubernetes, AWS, Azure, GCP, Docker, Prometheus, OpenTelemetry.* Background in Engineering/DevOps will be considered an advantage.* Previous experience in Technical Sales of Observability, Monitoring, APM, RUM, SIEM is desirable.* Proficiency ...

Platform Support Engineer

Location
Greater London, England, United Kingdom
clusters, including deployments, pods, services and Helm. ArgoCD – GitOps-based deployment and release management using ArgoCD. Loki – log aggregation and querying with Grafana Loki. Prometheus – metrics collection, querying and alerting with Prometheus. Grafana – building and maintaining dashboards and alerts in Grafana. Bash scripting – automating operational tasks and tooling with Bash. ...

Lead Java Developer

Location
Greater London, England, United Kingdom
gRPC etc. Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc. Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches Hands‐on knowledge ...

Lead Java Developer

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
Kafka, JMS, gRPC etcProficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuningExperience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc. Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed cachesHands-on knowledge of Linux ...

Principal AI Quality Engineer

Location
Greater London, England, United Kingdom
practices at the leading edge. Experience with CI/CD tooling, e.g. Jenkins, Azure DevOps, and Octopus Deploy. Experience with observability tooling such as Prometheus, Grafana, and Sumo Logic. Benefits Holidays. We all need to rest so you get 25 basic holidays with the option to grow ...

Site Reliability Manager - Environment Strategy

Location
Greater London, England, United Kingdom
complex environments Experience implementing automation using Terraform or CloudFormation Experience with CI/CD pipelines and operational monitoring with tools such as CloudWatch, Prometheus and Grafana Understanding of OS patching for Windows Server and RHEL as part of governance practices Background in risk management, compliance frameworks and incident response processes ...

Observability SME/Architect/Consultant

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
Making Cross-functional Collaboration Technical ExpertiseCandidates should demonstrate experience with one or more of the following technologies and platforms:Observability Platforms Dynatrace Splunk Grafana Prometheus Elastic/ELK Stack AppDynamics New Relic Service Management & IT Operations ServiceNow ServiceNow Event Management ServiceNow ITOM CMDB and Dependency Mapping Solutions Cloud Monitoring Azure ...

Python Backend Developer

Location
Greater London, England, United Kingdom
frontend work, and Go for select infrastructure Tools: RabbitMQ and Kafka for messaging, PostgreSQL and Redis for data storage Environment: Linux servers Observability: OpenTelemetry, Prometheus, Grafana and Zabbix Must-Haves: Strong background in software development, with strong experience with Python. A degree in Computer Science or a numerical subject from ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
Fargate clusters in AWS, creating common tooling to aid in development tasks, and running shared services such as Opensearch, Envoy, Vault and Prometheus to name a few. We are committed to Open Source software and give back to the community by open sourcing interesting projects. What you’ll be doing ...

Senior Backend Engineer | AI Platform

Location
Greater London, England, United Kingdom
organization. Tech Stack: Backend Python FastAPI Agent Development Kit (ADK) Datastores PostgreSQL BigQuery Firestore Infrastructure Google Cloud Platform (GCP) RabbitMQ Terraform Monitoring & Observability Grafana Prometheus Langfuse incident.io Sentry What to Expect from Our Hiring Process At Plum, we value a lot the time you devote to the hiring process, this ...

Senior Software Engineer, Video Encoding

Location
United Kingdom
Experience with GPU‐accelerated encoding or hardware media pipelines Familiarity with Kubernetes, ECS, Nomad, or other orchestration platforms Experience with observability stacks such as Prometheus, Grafana, OpenTelemetry, ELK, or Datadog Experience building fault‐tolerant ingest or transcoding platforms operating across multiple regions Our Hybrid Work Approach Roku fosters an inclusive ...

Staff Platform Engineer

Location
Greater London, England, United Kingdom
Fargate clusters in AWS, creating common tooling to aid in development tasks, and running shared services such as Opensearch, Envoy, Vault and Prometheus to name a few. The team has also expanded its scope to simplify Data engineering in the organisation using the same techniques we used to ease creating ...

Platform Engineer

Location
Greater London, England, United Kingdom
Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest … supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents and helping restore service. Contribute to blameless post‐incident reviews and implement follow ...

Site Reliability Engineer III

Location
Belfast City District, Northern Ireland, United Kingdom
Manage cluster lifecycles, data replication, RBAC, and workload placement. Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection. Incident Response & Operations: Engage with urgency … with an eagerness to learn independently and collaboratively. Preferred Qualifications/Desirable Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana. Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles. Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified ...

Cloud Operations Engineer (remote - London)

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Experienced and Certified in cloud computing with AWSExperience in a public cloud such as AWSKnowledge of monitoring and alerting technologies such as Grafana, Prometheus,Expereince of working with of Docker & Kubernetes and Container technology in productionWindows and Linux Operating System Management TechniquesSolid understanding of the OSI ModelExperience in database technology ...

Site Reliability Engineer

Location
City Of London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...