1 to 25 of 34 Permanent Observability Jobs in Bristol

AI Technical Lead | Staff Software Engineer | Principal AI Engineer | GenAI | £80k-£130k

Hiring Organisation
167 Solutions Ltd
Location
Bristol, Avon, England, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £120,000 per annum
LLMs Large Language Models Foundation Models Agentic AI AI Agents Autonomous Agents Multi-Agent Systems RAG Retrieval Augmented Generation Prompt Engineering AI Evaluation AI Observability AI Governance Responsible AI AI Security Guardrails Model Context Protocol (MCP) Function Calling Tool Use AI Orchestration AI Platforms Experience with one or more: OpenAI ...

AWS DevOps Engineer

Hiring Organisation
Leidos Innovations UK Limited
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
disaster recovery strategies to ensure data protection and business continuity] Ability to implement robust monitoring and logging solutions e.g., CloudWatch, to ensure system reliability, observability, and proactive incident response Comfortable working in Agile development teams, translating business requirements into technical solutions, and actively participating in sprint planning, retrospectives, and daily ...

Software Engineer (Mid and Senior Levels)

Hiring Organisation
One Big Circle Ltd
Location
City Of Bristol, England, United Kingdom
Engineering Cloud infrastructure, AWS, distributed systems, backend services, RESTful APIs, microservices architecture, service integration, data pipelines, storage systems, relational databases (MySQL), scalability, system reliability, observability, infrastructure as code, serverless deployments, containerisation (Docker), AWS services (Lambda, SQS, EC2, S3), IAM permissions (AWS), DevOps security, CI/CD, CI/CD tools ...

Senior Platform Engineer

Hiring Organisation
SF Partners Admin
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent
Salary
£90,000
Infrastructure as Code using Terraform. Support Internal Developer Platforms and self-service engineering capabilities. Build and improve CI/CD pipelines and automation. Implement observability solutions using Prometheus, Grafana and OpenTelemetry. Support reliability engineering initiatives including SLOs, SLIs and incident response improvements. Collaborate with architects, security teams and software engineers ...

DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
maintaining GitLab CI/CD pipelines Knowledge of containerisation technologies such as Docker and orchestration platforms such as ECS Experience with monitoring and observability tools (e.g. CloudWatch, Grafana, Prometheus) Strong understanding of security best practices and cloud governance Excellent troubleshooting and problem-solving skills Experience using AI-assisted development tools ...

Full Stack Engineer

Hiring Organisation
Newpage Solutions
Location
City Of Bristol, England, United Kingdom
least one of AWS, Azure, Cloudflare, or Vercel—with Docker, Kubernetes, and GitHub Actions. A no-compromise attitude on clean code, TDD, security, observability, scalability, performance, and cost. A deep working understanding of how LLMs behave—and where they break—and how to optimize accuracy, latency, and cost. Clear writing ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
systems. Infrastructure monitoring, performance optimisation and operational support. Infrastructure security, vulnerability management and remediation. Troubleshooting complex infrastructure, platform and application deployment issues. Monitoring and observability tooling such as Grafana, Prometheus or equivalent technologies. Configuration management and automation tooling such as Chef, Ansible or equivalent technologies. PowerShell, Bash or similar scripting ...

Senior Integration Engineer

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
data transformations, testing, and documentation Familiarity with Snowflake data modelling, optimisation, and cost‐aware design Understanding of data quality, lineage, and observability concepts CI/CD pipeline implementation for applications, infrastructure, and data workloads Strong focus on test automation and engineering quality Ability to operate production systems, including incident support ...

Linux Engineering Lead

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
goals, and achieve results Desirable Experience managing or coordinating incident response activities Experience working alongside DevOps, platform, or infrastructure engineering teams Experience with monitoring, observability, and logging systems Experience supporting AI/ML or high‐performance computing environments Understanding of identity and access management concepts Experience building or scaling operational ...

System Engineer

Hiring Organisation
Signature Recruitment Limited
Location
Bristol, Avon, England, United Kingdom
Employment Type
Full-Time
Salary
£50,000 - £60,000 per annum
infrastructure-as-code. Manage and support Linux-based servers and build containerised applications. Monitor and maintain edge devices and laboratory data collection systems. Build observability and monitoring capabilities to proactively identify issues. Support integrations between control systems, software applications and data infrastructure. Develop and maintain documentation, operational procedures and runbooks. ...

Principal Data Architect

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
curated, versioned datasets with clear data contracts and lineage. Enable feature creation, reuse, and publication for low‐latency serving and batch inference. Improve data observability, quality monitoring, alerting, and health checks across the platform. Architect for Scale & Cost Efficiency Work with Toumetis’ Principal Cloud Engineer to evaluate and advise ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

AWS Principal Systems & Security Architect

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
scale. • Streaming & Lakehouse Architectures: Hands-on experience with Kinesis/MSK, EKS/ECS/Fargate, Lambda, and Step Functions. • AIOps & Observability Tooling: Familiarity with CloudWatch, OpenSearch, Grafana, and SageMaker for real-time analytics and anomaly detection. • Performance & Cost Trade-offs: Track record of optimising systems for reliability, latency ...

Software Engineering Team Lead

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
software lifecycle from design ideation through to production and eventual decommissioning. Our engineering teams work under a true DevOps culture — with infrastructure as code, observability, automated testing, and continuous delivery treated as first-order concerns, not afterthoughts. You'll set architectural direction, partner closely with your Product Manager counterpart … systems and microservice development - we use Azure Service Bus, and welcome experience with similar messaging technologies such as Kafka or RabbitMQ Infrastructure: Kubernetes, Docker Observability: Prometheus, Grafana Engineering culture: DevOps, infrastructure as code, automated testing across all environments including production, continuous delivery Our Engineering Approach Full ownership: Teams own their ...

Senior Cloud Platform Engineer—AI Infra & Automation

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
storage solutions. This is a hand‐on technical role requiring a solid background in the use of cloud infrastructure, deployment using Infrastructure‐as‐Code, observability, high‐performance networking and storage systems. You may have been working in an IT organisation, a datacentre, a cloud provider or as a developer … internal users in their use. Turn end‐user and product requirements into deployed services. Help build automation to collect and analyse metrics and other observability data from the cloud services to support clear identification and reporting of any issues. Work with users to provide information on any product‐related issues ...

Platform Engineer

Hiring Organisation
IC Resources
Location
Greater Bristol Area, United Kingdom
Apache NiFi Elasticsearch/Logstash/Kibana (ELK) Data Pipelines/Data Integration Linux Python, Bash or similar scripting SQL Enterprise data platforms or observability solutions Experience within secure or highly regulated environments would be advantageous 📍 Bristol (Hybrid) 🔒 Due to the nature of the work, candidates must be eligible ...

Principal Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
storage solutions. This is a hand‐on technical role requiring a solid background in the use of cloud infrastructure, deployment using Infrastructure‐as‐Code, observability, high‐performance networking and storage systems. You may have been working in an IT organisation, a datacentre, a cloud provider or as a developer … users in their use. Turn end‐user and product requirements into deployed services. Help to build automation to collect and analyse metrics and other observability data from the cloud services to support clear identification and reporting of any issues. Work with users to provide information of any product‐related issues ...

Data Platform Architect & Growth Lead

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
7+ years of experience, strong skills in Python and SQL, and familiarity with modern cloud data platforms. Join us to enhance data governance, improve observability, and craft innovative solutions for our evolving business needs. #J-18808-Ljbffr ...

Operate Service Manager

Hiring Organisation
Apto Solutions
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent
Summary Apto Operate is our managed service for telemetry, security, and observability platform management and it is growing. The Operate Service Manager owns the service itself: the quality and consistency of delivery across every customer account, the performance and development of our engineering team, and the continuous improvement ...

Principal Software Architect

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
/software ecosystem. Assess the architectural impact of new technologies. Be aware of the usability, performance, reliability, maintainability, testability, security and observability constraints on the software architecture. Prototyping and validating architectural concepts through proof-of-concept implementations. Contribute to future and/or related product definitions with a forward-looking ...

Principal Platform Engineer

Hiring Organisation
SF Partners Admin
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
capabilities. Design and operate production-grade Kubernetes platforms, including EKS, AKS or OpenShift. Define engineering standards, golden paths, reusable modules and platform patterns. Build observability strategies using Prometheus, Grafana, OpenTelemetry and modern APM tooling. Improve reliability through SLOs, incident reviews and Site Reliability Engineering (SRE) practises. Embed DevSecOps, supply-chain … Infrastructure as Code (IaC). CI/CD automation. GitOps tools such as ArgoCD or Flux. Internal Developer Platforms or self-service engineering. Observability tools including Prometheus, Grafana, OpenTelemetry, ELK, Datadog, Dynatrace or New Relic. DevSecOps and supply-chain security. SRE practises, SLOs, SLIs and incident management. Platform governance, cloud ...

Observability Architect - 12 Month FTC

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You’re comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with … grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate ...

Operate Service Manager

Hiring Organisation
17918
Location
Bristol, Gloucestershire, United Kingdom
Summary Apto Operate is our managed service for telemetry, security, and observability platform management and it is growing. The Operate Service Manager owns the service itself: the quality and consistency of delivery across every customer account, the performance and development of our engineering team, and the continuous improvement ...

Global Head of SRE & Reliability – Hybrid Role

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
office and collaboration across Engineering, Infrastructure Operations and Security to boost reliability and performance of critical platforms. You will define reliability strategy, drive automation, observability, and incident maturity, and scale the organization while embedding reliability into the #J-18808-Ljbffr ...

Head of Site Reliability Engineering – SRE

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Define, lead, and evolve our global reliability strategy. Drive operational excellence, service reliability, observability, automation, and continuous improvement across our technology landscape. Work closely with Engineering, Infrastructure, Security, and Technology Operations teams to establish and embed modern SRE practices that enable highly reliable, scalable, and resilient services while fostering … culture of shared ownership and continuous learning. Drive adoption of SRE principles (SLOs, error budgets, toil reduction). Establish observability and monitoring standards. Lead automation-first operations. Improve incident and problem management maturity. Partner with software and infrastructure engineering teams to embed reliability into the product lifecycle. Establish SRE governance ...