276 to 300 of 875 Grafana Jobs

Lead Software Engineer - Platform Engineering - Chase UK

Location
Greater London, England, United Kingdom
Debugging expertise in Kubernetes. Understanding of Infrastructure Tools, preferably Terraform, and AWS services understanding. Experience with any GitHub, GitHub Actions, Artifactory, Terraform Cloud, Slack, Grafana, SonarQube is considered a plus. Proficient in coding in one or more languages. Equal Employment Opportunity Statement We recognize that our people are our strength ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
focus on observability for large-scale, distributed compute or network systems. Deep, hands-on expertise building, scaling, and managing observability platforms (e.g., Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger, Honeycomb, etc.). You have proven experience using these tools to support performance analysis and debugging of complex distributed ...

Senior Backend Engineer (.NET & Python)

Location
Greater London, England, United Kingdom
hands‐on experience working with AI coding assistants. Bonus points Experience with Kubernetes, GCP, Google Pub/Sub, Kafka or Redis. Familiarity with Langfuse, Grafana, Prometheus or Terraform. Experience mentoring or supporting other engineers. Even if you don't meet all of the requirements for this role, we encourage ...

Deployed Architect, Professional Services (London)

Location
Greater London, England, United Kingdom
sizing Experience designing high-availability and disaster recovery solutions Strong understanding of networking, security (SSO/RBAC, TLS, secrets management), and observability (Prometheus, Grafana, Datadog) Experience with CI/CD pipelines for infrastructure and applications Agent Engineering & Development: 1+ years of experience building production AI/ML applications or agents ...

Principal / Sr. Principal DevOps Engineer (AHT)

Hiring Organisation
Northrop Grumman
Location
Sacramento, California, United States
Employment Type
Permanent
Salary
USD Annual
with at least 3 of the preferred qualifications Preferred Qualifications: Current Security+ Terraform Kubernetes administration AWS administration Flux Helm DynamoDB NATS configuration Big Bang (Grafana, Prometheus, Loki) Jenkins/GitLab/Bamboo Primary Level Salary Range: $114,000.00 - $163,200.00 Secondary Level Salary Range: $135,800.00 - $203,600.00 The above ...

Professional Services Consultant - AI Security

Hiring Organisation
Cato Networks
Location
London, UK
Employment Type
Full-time
plusFamiliarity with container security, runtime protection, and service mesh architectures (Istio, App Mesh)Practical experience with observability and monitoring stacks (e.g., CloudWatch, Prometheus, Grafana, Datadog)Knowledge of data sovereignty, residency requirements and compliance frameworks relevant to AI workloads (e.g., FedRAMP, SOC 2, ISO 27001, NIST 800-53, HIPPA, PCI, HITRUST ...

Data Platform Engineer

Location
Greater London, England, United Kingdom
some GCP Warehouse & Storage: Snowflake, S3/Parquet Data & ETL: dbt, Fivetran Platform & Infra: Kubernetes, Kafka, RabbitMQ, Argo, GitHub Actions, HashiCorp Vault Observability: Datadog, Grafana Dashboarding: Preset Other: Claude What we’re looking forStrong fundamentals and the ability to apply them pragmatically: Solid programming ability (Python or similar) Strong experience ...

Professional Services Consultant - AI Security

Location
Greater London, England, United Kingdom
plus Familiarity with container security, runtime protection, and service mesh architectures (Istio, App Mesh) Practical experience with observability and monitoring stacks (e.g., CloudWatch, Prometheus, Grafana, Datadog) Knowledge of data sovereignty, residency requirements and compliance frameworks relevant to AI workloads (e.g., FedRAMP, SOC 2, ISO 27001, NIST 800‐53, HIPPA ...

Platform Engineer – Monitoring, Observability & SIEM (MONSO)

Location
Greater London, England, United Kingdom
Modern Observability & Telemetry: Strong background in Splunk (SPL, dashboards, alerts, data ingestion, forwarders) and also configuring OpenTelemetry collectors and pipelines; knowledge of Prometheus and Grafana or similar tools. Kubernetes (Power User): Strong, hands-on experience deploying and operating workloads, stateful appli-cations, Helm charts, and manifests on K8s (cluster administration ...

Development Lead - Payments (ICON IPF)

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
Development (BDD), with the ability to promote quality engineering practices across the development lifecycle. Familiarity with Bitbucket, Jira, Confluence, Docker, Kubernetes, Observability tools like (Grafana, Dynatrace, Tivoli) and CI/CD pipelines such as Jenkins. Proven experience working closely with major banking clientsKnowledge of ISO 20022 payments, including payment types ...

Sr. Software Engineer

Hiring Organisation
Meltwater Group
Location
London, UK
Employment Type
Full-time
Tech Stack: Languages: Golang (primary) with some TypeScriptInfrastructure: AWS, S3, Lambda, SQS, SNS, CloudFront, Kubernetes (Helm), Kong API GatewayDatabases: Postgres, Redis, DynamoDB, OpenSearchMonitoring: Coralogix, Grafana, CloudWatchCI/CD & IaC: GitHub Actions, TerraformWhat We Offer: Generous paid time off and flexible working arrangements. Comprehensive health insurance and wellness benefits. Family leave ...

Platform Engineer - Monitoring, Observability & SIEM (MONSO)

Hiring Organisation
Berenberg
Location
London, UK
Employment Type
Full-time
Modern Observability & Telemetry: Strong background in Splunk (SPL, dashboards, alerts, data ingestion, forwarders) and also configuring OpenTelemetry collectors and pipelines; knowledge of Prometheus and Grafana or similar tools. Kubernetes (Power User): Strong, hands-on experience deploying and operating workloads, stateful appli-cations, Helm charts, and manifests on K8s (cluster administration ...

Corporate KYC : Lead Software Engineer- Python/ PySpark

Location
Glasgow, Scotland, United Kingdom
cloud, AI/ML, or data engineering Experience in large-scale data processing, microservices, API design, Kafka, Redis, MemCached, observability tools (Dynatrace, Splunk, Grafana), and orchestration frameworks (Airflow, Temporal) Advanced working knowledge of relational and NoSQL databases, vector stores, data lake architectures, and data governance Practical cloud-native experience ...

Senior Forward Deployment Engineer

Location
Slough, England, United Kingdom
recovery, controlled failover, and production‐recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain #J-18808-Ljbffr ...

Senior Research HPC Engineer

Location
Greater London, England, United Kingdom
Python programming Experience with GPU‐focused hardware and software Experience monitoring and optimising research workloads and system performance, using tools such as Grafana, Prometheus, Arbiter2 People Leadership Desirable Motivating and developing staff to utilise HPC effectively through HPC onboarding and training Project managing work packages within and between groups Resolving ...

Senior Forward Deployment Engineer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
recovery, controlled failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain OtherLanguages English: C1 Advanced Seniority ...

Expert Forward Deployment Engineer

Hiring Organisation
Luxoft
Location
London, UK
Employment Type
Full-time
recovery, controlled failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain OtherLanguages English: C1 Advanced Seniority ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
skillsAdvanced knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps).Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL).Expertise in performance optimisation of distributed systems (e.g., caching, network latency).Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium).Demonstrated success ...

Staff Cloud SRE – AI/ML Platform & GPU Compute London, United Kingdom on-site

Location
Greater London, England, United Kingdom
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skills Familiarity with infrastructure-as-code (e.g. ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Senior Infrastructure Engineer

Hiring Organisation
AJ BELL BUSINESS SOLUTIONS LIMITED
Location
Salford, Greater Manchester, North West, United Kingdom
Employment Type
Permanent
FortiClient VPN or equivalent. Microsoft Windows Server and Red Hat Enterprise Linux server configuration and management. Monitoring and observability tooling such as SolarWinds, Opsgenie, Grafana, Prometheus or equivalent. Storage, backup and disaster recovery technologies, including SAN/NAS, replication, snapshotting and recovery testing concepts. Identity and access technologies such ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline components ...

Technical Site Reliability Engineer

Location
Greater London, England, United Kingdom
just running them. Infrastructure-as-code and configuration management (Terraform, Ansible, Docker, Kubernetes). On‐prem and cloud deployment experience. Observability tooling: Prometheus, Grafana, ELK, or equivalent. Linux systems administration depth; comfort in mixed Linux/Windows environments. Prior work in a defense, aerospace, or classified environment. Active security clearance. ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
analyze internet traffic patterns across multiple dimensions using flow-based tools. Experience working with alerting, monitoring and visibility tools (such as Graphite/Grafana, Prometheus, or Splunk). Knowledge across cloud hosting solutions (i.e., GCP, AWS and Azure). Knowledge of DevOps practices and CI/CD pipelines (ie. ...