251 to 275 of 835 Permanent Grafana Jobs

Senior DevOps Systems Administrator

Hiring Organisation
Mobilus Limited
Location
Guildford, Surrey, United Kingdom
Employment Type
Permanent
Salary
£55000 - £65000/annum + excellent benefits
across AWS and private cloud Automate infrastructure with Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement observability tools (Grafana, Prometheus, CloudWatch) Work across teams to improve reliability and security Manage hybrid networks, IAM, firewalls and VPNs The successful DevOps Systems Administrator will have ...

AWS Architect

Hiring Organisation
vaaridatech
Location
Texas, United States
Employment Type
Permanent
Salary
USD Annual
pipelines to enable automated build, testing, deployment, and release processes. • Utilize observability and monitoring tools such as Dynatrace, Amazon CloudWatch, AWS X-Ray, Grafana, and Splunk to monitor application health and troubleshoot production issues. • Participate in on-call support, incident resolution, root cause analysis, and continuous improvement initiatives to ensure ...

Senior MLOps Engineer 201043

Hiring Organisation
Harnham - Data & Analytics Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
Knowledge of infrastructure-as-code approaches using Terraform, Bicep, Pulumi, or equivalent technologies. Experience implementing monitoring and observability solutions using tools such as Prometheus, Grafana, or similar. Familiarity with orchestration platforms including Dagster, Airflow, Prefect, or related technologies. Desirable experience includes: Model serving infrastructure for real-time or batch inference ...

Site Reliability Engineer

Hiring Organisation
JAM Recruitment Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£700 - £750 per day
system integration. Experience managing Windows IIS servers. Experience using Atlassian tools (Jira, Confluence) to support Agile development and documentation practices. Knowledge of Keycloak, Grafana, Elasticsearch and NiFi. ...

Sr. DevOps Engineer

Location
Greater London, England, United Kingdom
TeamCity to achieve continuous integration and continuous delivery (CI/CD). Set up robust tracing and observability tools, such as Stackdriver, Prometheus, Grafana, and Jaeger, to monitor system health, performance, and reliability. Collaborate closely with cross‐functional teams including development, QA, and operations to troubleshoot issues, optimize workflows ...

Senior Engineer, Systematic Data

Hiring Organisation
Balyasny Asset Management
Location
London, UK
Employment Type
Full-time
platforms (e.g. AWS, GCP, Azure) Experience with orchestration and container technologies (e.g. Airflow, Kubernetes, Docker) Experience with monitoring and alerting tools (e.g. CloudWatch, Prometheus, Grafana, Sentry/OTel ...

Software Engineer III

Location
Christchurch, England, United Kingdom
with Kubernetes — deploying, operating, and troubleshooting containerized workloads (EKS preferred) Experience with Infrastructure-as-Code tooling (Terraform, CloudFormation, CDK) Familiarity with observability tooling (Prometheus, Grafana, Datadog) AWS certifications (Developer Associate, Solutions Architect, or higher) #J-18808-Ljbffr ...

Senior DevOps Engineer - AWS - Manchester

Hiring Organisation
Circle Group
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£70,000
cloud platform - Solid scripting and automation skills, using languages like Python, Bash, or PowerShell. - Experience with monitoring and logging tools (e.g., ELK Stack, Prometheus, Grafana) to ensure system reliability and performance. - Any Linux experience would be a bonus Key Responsibilities: - Drive the strategy and implementation for DevOps and SRE practices. ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
balancers, DNS, security groups/NSGs) Experience with secrets management and identity/access control (IAM, OIDC, Azure AD) Familiarity with observability tooling (Prometheus, Grafana, CloudWatch, Azure Monitor) Risk Benefit Statement Learn more about the LexisNexis Risk team and how we work here We know your well-being and happiness ...

Infrastructure and MLOps Engineer

Location
Greater London, England, United Kingdom
Experience with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu) Experience with GitHub Actions Experience with modern observability tooling (e.g. Prometheus) Experience with Grafana Knowledge of Go/Java/C++ (or similar language) Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave ...

Staff Software Infrastructure Engineer

Location
West of England, England, United Kingdom
using Kubernetes (k8s) or OpenStack Experience with GitHub Actions Experience with build tools (e.g. CMake) Experience with modern observability tooling (e.g. Prometheus) Experience with Grafana Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental ...

Senior Software Infrastructure Engineer

Location
Greater London, England, United Kingdom
using Kubernetes (k8s) or OpenStack Experience with GitHub Actions Experience with build tools (e.g. CMake) Experience with modern observability tooling (e.g. Prometheus) Experience with Grafana Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental ...

Senior Build & Cloud Infra Engineer

Location
Cambridge, England, United Kingdom
using Kubernetes (k8s) or OpenStack Experience with GitHub Actions Experience with build tools (e.g. CMake) Experience with modern observability tooling (e.g. Prometheus) Experience with Grafana Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental ...

Software Engineer

Location
Greater London, England, United Kingdom
Typescript, or Java. CI/CD and DevOps: Knowledge of building pipelines, automated testing and deployment strategies. Monitoring and observability: Familiarity with tools like Grafana, Prometheus or other observability platforms. Containerisation and orchestration: Docker, Kubernetes. We are open to a wide range of backgrounds, but some examples we expect ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
while keeping us in control. What you'll be doing from day one Owning our infrastructure as code in Terraform, plus alerting, observability (Prometheus, Grafana) and reliability, including load testing and disaster recovery exercises. Making CI/CD faster (GitHub Actions, ArgoCD) and taking obstacles out of engineers’ way, from ...

Site Reliability Engineer

Location
England, United Kingdom
balancers, DNS, security groups/NSGs)* Experience with secrets management and identity/access control (IAM, OIDC, Azure AD)* Familiarity with observability tooling (Prometheus, Grafana, CloudWatch, Azure Monitor)**Risk benefit statement**Learn more about the LexisNexis Risk team and how we work##**We know your well-being and happiness ...

Lead Software Engineer - Platform Engineering - Chase UK

Location
Greater London, England, United Kingdom
Debugging expertise in Kubernetes. Understanding of Infrastructure Tools, preferably Terraform, and AWS services understanding. Experience with any GitHub, GitHub Actions, Artifactory, Terraform Cloud, Slack, Grafana, SonarQube is considered a plus. Proficient in coding in one or more languages. Equal Employment Opportunity Statement We recognize that our people are our strength ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
focus on observability for large-scale, distributed compute or network systems. Deep, hands-on expertise building, scaling, and managing observability platforms (e.g., Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger, Honeycomb, etc.). You have proven experience using these tools to support performance analysis and debugging of complex distributed ...

Senior Backend Engineer (.NET & Python)

Location
Greater London, England, United Kingdom
hands‐on experience working with AI coding assistants. Bonus points Experience with Kubernetes, GCP, Google Pub/Sub, Kafka or Redis. Familiarity with Langfuse, Grafana, Prometheus or Terraform. Experience mentoring or supporting other engineers. Even if you don't meet all of the requirements for this role, we encourage ...

Deployed Architect, Professional Services (London)

Location
Greater London, England, United Kingdom
sizing Experience designing high-availability and disaster recovery solutions Strong understanding of networking, security (SSO/RBAC, TLS, secrets management), and observability (Prometheus, Grafana, Datadog) Experience with CI/CD pipelines for infrastructure and applications Agent Engineering & Development: 1+ years of experience building production AI/ML applications or agents ...

Principal / Sr. Principal DevOps Engineer (AHT)

Hiring Organisation
Northrop Grumman
Location
Sacramento, California, United States
Employment Type
Permanent
Salary
USD Annual
with at least 3 of the preferred qualifications Preferred Qualifications: Current Security+ Terraform Kubernetes administration AWS administration Flux Helm DynamoDB NATS configuration Big Bang (Grafana, Prometheus, Loki) Jenkins/GitLab/Bamboo Primary Level Salary Range: $114,000.00 - $163,200.00 Secondary Level Salary Range: $135,800.00 - $203,600.00 The above ...

Professional Services Consultant - AI Security

Hiring Organisation
Cato Networks
Location
London, UK
Employment Type
Full-time
plusFamiliarity with container security, runtime protection, and service mesh architectures (Istio, App Mesh)Practical experience with observability and monitoring stacks (e.g., CloudWatch, Prometheus, Grafana, Datadog)Knowledge of data sovereignty, residency requirements and compliance frameworks relevant to AI workloads (e.g., FedRAMP, SOC 2, ISO 27001, NIST 800-53, HIPPA, PCI, HITRUST ...

Data Platform Engineer

Location
Greater London, England, United Kingdom
some GCP Warehouse & Storage: Snowflake, S3/Parquet Data & ETL: dbt, Fivetran Platform & Infra: Kubernetes, Kafka, RabbitMQ, Argo, GitHub Actions, HashiCorp Vault Observability: Datadog, Grafana Dashboarding: Preset Other: Claude What we’re looking forStrong fundamentals and the ability to apply them pragmatically: Solid programming ability (Python or similar) Strong experience ...

Professional Services Consultant - AI Security

Location
Greater London, England, United Kingdom
plus Familiarity with container security, runtime protection, and service mesh architectures (Istio, App Mesh) Practical experience with observability and monitoring stacks (e.g., CloudWatch, Prometheus, Grafana, Datadog) Knowledge of data sovereignty, residency requirements and compliance frameworks relevant to AI workloads (e.g., FedRAMP, SOC 2, ISO 27001, NIST 800‐53, HIPPA ...

Platform Engineer – Monitoring, Observability & SIEM (MONSO)

Location
Greater London, England, United Kingdom
Modern Observability & Telemetry: Strong background in Splunk (SPL, dashboards, alerts, data ingestion, forwarders) and also configuring OpenTelemetry collectors and pipelines; knowledge of Prometheus and Grafana or similar tools. Kubernetes (Power User): Strong, hands-on experience deploying and operating workloads, stateful appli-cations, Helm charts, and manifests on K8s (cluster administration ...