12 of 12 Remote/Hybrid Observability Jobs in the East of England

Senior Private Cloud Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
ability to debug compute, networking, storage in large production environments."Nice To Have" Skills and Experience:Exposure to modern cloud native principles.Familiarity with observability tools (Prometheus, Grafana)!Exposure to large-scale or multi-tenant environments.Exposure to GitOps driven and CI/CD pipelines (Jenkins, ArgoCD. FluxCD, etc)Contribution to Open ...

Staff Full Stack Software Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
delivers a high-quality user experience.Design, build, and maintain our developer portal including CI/CD pipelines, documentation, automated testing, security upgrades, and observability integrations.Partner closely with platform, software and hardware teams to integrate services, tooling, and policies into the portal in a user-centric and automated manner.We invest significant ...

Senior Software Engineer - Live & VOD Video Infrastructure

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
GStreamer, FFmpeg, MediaMTX, or similar technologiesExperience with GPU-accelerated encoding or hardware media pipelinesFamiliarity with Kubernetes, ECS, Nomad, or other orchestration platformsExperience with observability stacks such as Prometheus, Grafana, OpenTelemetry, ELK, or DatadogExperience building fault-tolerant ingest or transcoding platforms operating across multiple regions#LI-JC5What's Roku's approach ...

AI Platform Engineer - Cambridge

Hiring Organisation
Nextech Group Limited
Location
Cambridgeshire, East Anglia, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
/conversation systems Collaborate with data, product, and engineering teams to translate business requirements into technical solutions Implement CI/CD pipelines, monitoring, and observability for AI-driven services Own the full lifecycle: experimentation, prototyping, and production deployment What We're Looking For: Strong commercial experience with Python and/ ...

R&D Software Engineer

Hiring Organisation
Aveva Group
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
with AVEVA CONNECT.Operate and improve cloud environments: support desktop streaming (Amazon WorkSpaces Applications) and Windows-centric infrastructure (EC2, FSx, Active Directory, DynamoDB) with strong observability and cost/performance focus.Deliver securely with automation and teamwork: write clean, tested, documented, deployable code; contribute to CI/CD and infrastructure automation (Azure ...

Service Design Specialist

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
needed to support the service, including incident, major incident, request, problem, change, release, knowledge and continual improvement activities.Embed availability, capacity, resilience, continuity, performance, security, observability, user experience and supportability requirements into service designs.Help define service levels and service health measures, including SLAs, SLOs, SLIs, operational metrics, monitoring, alerting, dashboards ...

Senior CPU Performance Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
ResponsibilitiesDevice Performance AnalysisAnalyse performance on real devices using real-world workloadsSupport investigation of system-level performance issues, identifying CPU-related factorsOperate effectively in low-observability environments (e.g. no waveforms, partial counters, noisy systems)CPU-Focused AnalysisUse PMU counters and profiling tools to understand CPU behaviourContribute to identifying performance bottlenecks (e.g. ...

Site Reliability Engineer SRE Kubernetes

Hiring Organisation
Client Server
Location
Cambridge, Cambridgeshire, East Anglia, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
embed reliability best practices, improve system resilience and solve complex operational challenges. You will take a proactive approach to identifying risks, improving system observability and enhancing the developer experience through automation and tooling. Key responsibilities will include developing tooling, frameworks and best practices to improve reliability and operational efficiency, monitoring … similar You have experience with cloud platforms: AWS, Azure or GCP You have experience with containerisation and orchestration including Kubernetes You're familiar with observability practices; monitoring, logging, tracing You have strong problem solving and critical thinking skills You're collaborative with great communication skills You are degree educated, having ...

Cloud HPC Network Engineer

Hiring Organisation
Hays Specialist Recruitment Limited
Location
Cambridge, Cambridgeshire, England, United Kingdom
Employment Type
Contractor
Contract Rate
£600 - £650 per day
with OpenStack technologies including Neutron, OVN and ML2 Define network architectures supporting virtual machines, bare-metal platforms and Kubernetes environments Perform network optimisation, benchmarking, observability and performance tuning across HPC environments Work closely with compute, storage and platform engineering teams to resolve complex cross-domain infrastructure challenges Key Requirements: Strong … ECMP, Layer 2/Layer 3 networking and Leaf-Spine architectures Experience with OpenStack networking services including Neutron, OVN and ML2 Strong troubleshooting, network observability and performance tuning experience within HPC or data centre environments Experience with network automation, API-driven infrastructure and software-defined networking Additional Information: Exposure ...

IT Infrastructure Solutions Architect

Hiring Organisation
Aveva Group
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
strategic roadmap.Key responsibilitiesDefine and maintain reference architectures and target-state designs for VMware VCF 9.0 platform architecture and lifecycle patterns.Define Aria Operations observability strategy (telemetry standards, alert philosophy, capacity/performance governance, service reporting) and ensure operational adoption.Define VCF Automation platform approach (catalog/service design, templates/guardrails, governance …/or iSCSI), VSAN, NAS, and software-defined storage concepts.Experience or exposure to infrastructure-as-codeProven capability to architect and operationalize enterprise monitoring/observability standards (Logic Monitor and Aria Operations).Proven capability to architect, govern, and troubleshoot provisioning automation (VCF Automation).Proven backup/recovery architecture and operational assurance ...

Efficiency Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
Job Overview:We are looking for an Efficiency Engineer to support the data-driven optimization, analysis and maintenance of Arm's engineering platforms. This role will work as part of a team effort to improve ...

SIEM Engineer (Elastic)

Hiring Organisation
Searchability NS&D
Location
Hemel Hempstead, England, United Kingdom
environments. This is an opportunity to take technical ownership of a large-scale, on-premises Elastic estate, supporting critical monitoring, logging, security analytics and observability capabilities. Due to continued investment in their technology platforms, they are looking for an experienced Elastic Stack SME to join the team. THE BENEFITS … secure on-premises estate. You will take technical ownership of Elasticsearch clusters, Logstash, Kibana, Beats, Elastic Agent and Fleet, while delivering solutions across SIEM, observability, threat detection and high-volume data ingestion. You will design highly available and resilient platforms, optimise cluster performance, implement security controls and automation, and provide ...