26 to 39 of 39 Observability Jobs in Gloucestershire

Technologist

Location
Cheltenham, England, United Kingdom
wide variety of complementary themes. If you're interested in the below, you'll be interested in this role: ‘Under the hood’ AI engineering Observability Technologies Message Broker Technologies NLP and Image Processing Technologies The Skills: As a Technologist, we understand you’ll have different areas of expertise, passion ...

Senior Platform Engineer Cheltenham

Location
Cheltenham, England, United Kingdom
delivery pipelines. Work across greenfield and brownfield platform engineering projects. Develop infrastructure-as-code solutions using modern tooling and engineering practices. Improve platform reliability, observability and operational maturity through automation and engineering excellence. Work closely with developers, architects and security teams to understand challenges and deliver pragmatic solutions. Champion platform … Experience with container technologies such as Docker and Kubernetes. Experience with cloud platforms such as AWS, Azure or GCP. Experience implementing monitoring, logging and observability solutions. Ability to work effectively across engineering, security and customer teams. Strong problem-solving skills with the ability to operate in complex and ambiguous environments. ...

Fullstack Engineer (DV Clearance) - Cloud Native, 37.5h

Location
Cheltenham, England, United Kingdom
Java applications, with cloud native deployments in OpenShift and Kubernetes. You will work across distributed architectures, implement Web API integrations, and contribute to observability and security practices. A DV clearance or recent status is required. #J-18808-Ljbffr ...

Lead .Net Software Engineer

Location
Cheltenham, England, United Kingdom
technical risks, bottlenecks, dependencies, and opportunities for improvement Contribute to AWS‐based architecture and engineering practices, including environments using Lambda, ECS, and EC2 Improve observability, reliability, performance, security, and operational readiness across the platform Contribute to Jenkins pipelines and CI/CD practices to improve consistency and delivery efficiency Help … with DevOps practices and infrastructure‐aware development Experience with messaging, event streaming, or related event‐driven technologies Experience with Blazor and MudBlazor Familiarity with observability and monitoring tooling in distributed systems Experience defining governance, controls, or operating models for AI agents or AI‐enabled internal tools Experience working ...

HPC Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
role in continuous service improvement (CSI)—reducing operational toil, increasing automation, and improving reliability, consistency, and operational efficiency across the platform. This includes strengthening observability, refining operational workflows, and eliminating repetitive or failure‐prone processes. Over time, you will help shape future infrastructure design and deployment approaches, feeding operational insight ...

Cross Domain Systems Integration Engineers

Hiring Organisation
Hackajob Ltd
Location
Gloucester, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
networking, and infrastructure. Strong knowledge in at least one core area: Networking (routing, switching, firewalls, protocols) Infrastructure (virtualisation, cloud platforms) Experience with: Monitoring and observability tools System audit and compliance frameworks Build pipelines and deployment processes (CI/CD) Familiarity with scripting/automation (e.g., PowerShell, Python, Bash) is advantageous. ...

24/7 HPC Infra SRE for AI & GPU Compute

Location
Gloucester, England, United Kingdom
work across network, storage, virtualization and orchestration with hands‐on Linux expertise, NVIDIA GPU ecosystems, RoCE/InfiniBand, and performance benchmarking. This role champions observability, automation and on‐call reliability, shaping next‐gen HPC platforms within a globally distributed team. #J-18808-Ljbffr ...

AI Product Engineer – Shape Next-Gen National Security

Location
Cheltenham, England, United Kingdom
driven team building AI capabilities for national security and government customers, shaping product direction, and applying modern engineering practices including testing, CI/CD, observability and robust #J-18808-Ljbffr ...

Senior Elixir Engineer (Sovereign Territory Surveillance)

Location
Gloucester, England, United Kingdom
telemetry Developing rich LiveView interfaces Working with large geospatial datasets Designing fault‐tolerant distributed services Helping shape technical strategy across the platform Improving deployment, observability and developer tooling You’ll ideally have experience with: Elixir and OTP Phoenix and LiveView Designing distributed systems Building APIs and backend services Distributed relational ...

Site Reliability Engineer

Hiring Organisation
Anson Mccade
Location
Gloucester, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
reliability and performance. Automating repetitive operational tasks and reducing manual intervention wherever possible. Monitoring and troubleshooting systems across the application and infrastructure stack. Improving observability and instrumentation to identify issues and measure system performance. Working alongside development and product teams to build scalable and resilient services. Responding to production incidents … including Bash or PowerShell. Cloud platforms such as AWS, Azure or OpenStack. Infrastructure automation and configuration management. CI/CD and deployment tooling. Monitoring, observability and troubleshooting of production systems. Docker, containers and/or microservices. Diagnosing issues across different levels of the technology stack. Working within Agile engineering teams. ...

DevOps & Infrastructure Engineer

Location
Gloucester, England, United Kingdom
Security customers, spanning both on-premise environments and cloud-based solutions. You’ll lead hands-on DevOps and infrastructure engineering across CI/CD, observability, infrastructure-as-code and platform automation, helping teams build secure, reliable and scalable services in demanding environments. What you’ll be doing: You’ll lead … cloud-based solutions. Develop and maintain CI/CD pipelines, GitOps workflows and automated deployment approaches using tools such as ArgoCD. Implement and improve observability using Prometheus, Grafana, logging and alerting to support resilient platform operations. Use infrastructure-as-code and platform automation with Helm, Go and Terraform to deliver ...

Platform Site Reliability Engineer

Location
Gloucester, England, United Kingdom
provision of tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance Radiant’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root cause … DHCP, VLANs, routing, switching Strong experience with API interrogation Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana preferred) Strong grasp of ITSM and service operation best practices Excellent communication and mentorship skills Comfortable interfacing with internal stakeholders and external ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
resolution and provision of tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance ’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root … Strong networking fundamentals: TCP/IP, DNS, DHCP, VLANs, routing, switching Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana) Hands-on experience operating orchestration platforms (Kubernetes, MAAS, Tinkerbell) Strong grasp of ITSM and service operation best practices Excellent communication ...

QA Test Infrastructure Engineer

Location
Cheltenham, England, United Kingdom
QA Test Infrastructure Engineer - Chelmsford, Onsite - Outside IR35 - Highest Security Clearance As a QA Test Infrastructure Engineer, you'll help design, build, and deliver secure digital solutions in highly secure environments. You'll work alongside ...