151 to 167 of 167 Observability Jobs in the East of England

Staff/Lead Python Engineer (FastAPI, Orchestration)

Location
Cambridge, England, United Kingdom
with other apps in aservicearchitecture. Furthering Developer Experience (DevEx) by mentoring others in writing code that is intuitive, clear, and easy to test Developing observability for new and existing ML applications and GenAI/LLM integrations , making use of the Grafana Stack (Prometheus, Loki, Tempo) Develop integrations and services that … owning projects from start to finish, including speccing, architecture, development, testing, deployment, release and monitoring Strong skills in building maintainable tests Strong experience with observability and tracing. Knowledge of best practices for performance optimisation, memory management. Experience mentoring others, especially in good software development practices, patterns, and fundamentals. Drive ...

Bioinformatics Scientist/Engineer - DRAGEN Array

Hiring Organisation
Illumina
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
bioinformatics solutions, integrating established community tools alongside novel methods. You will also contribute to software engineering best practices, including testing automation, CI/CD, observability, and software lifecycle management, helping to ensure DRAGEN Array software is reliable, scalable, and ready for production deployment across local and cloud-based environments. What …/CD workflows to support continuous integration, validation, and reliable delivery of DRAGEN Array software. Implement and evolve telemetry, logging, and monitoring to improve observability, diagnostics, and operational robustness. Execute and improve Software Lifecycle (SLC) practices, including requirements traceability, design documentation, verification, and validation. Use tools such as JAMA (requirements ...

Bioinformatics Scientist/Engineer - DRAGEN Array

Hiring Organisation
Illumina
Location
Saffron Walden, Essex, UK
Employment Type
Full-time
bioinformatics solutions, integrating established community tools alongside novel methods. You will also contribute to software engineering best practices, including testing automation, CI/CD, observability, and software lifecycle management, helping to ensure DRAGEN Array software is reliable, scalable, and ready for production deployment across local and cloud-based environments. What …/CD workflows to support continuous integration, validation, and reliable delivery of DRAGEN Array software. Implement and evolve telemetry, logging, and monitoring to improve observability, diagnostics, and operational robustness. Execute and improve Software Lifecycle (SLC) practices, including requirements traceability, design documentation, verification, and validation. Use tools such as JAMA (requirements ...

Senior SRE - Platform Reliability & Observability Lead

Location
Cambridge, England, United Kingdom
seeking a Senior Site Reliability Engineer to lead reliability, performance and continuous improvement of the Bango Platform. You will own end-to-end reliability, observability and incident response across infrastructure and delivery pipelines, serving as a technical centre of gravity for the SRE function. You will shape the bench across ...

Network Engineer

Hiring Organisation
Third Nexus Group Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Contract
Contract Rate
£375 - £400/annum
network security fundamentals. Automation Practical capability in Python, Ansible and REST APIs. Experience with Terraform, PowerShell and Git/GitHub is desirable. Monitoring & Observability Experience with network monitoring or observability platforms such as netbox, SolarWinds, Auvik, IP Fabric, LogicMonitor, Dynatrace, Azure Monitor or Grafana ...

Software Engineer, Agentic AI Roku, Inc.

Location
Cambridge, England, United Kingdom
product and platform capabilities for Roku TV. You will own the full lifecycle of agent development - from prototyping and architecture through orchestration, evaluation, deployment, observability, and continuous improvement. You will contribute directly to Roku's AI strategy by engineering reusable components, optimizing agent workflows, and ensuring strong real-world performance … systems around them. Create reusable agent templates, modular components, and paved-path patterns that accelerate adoption across teams and use cases. Establish strong evaluation, observability, and monitoring for conversation quality, task success rate, latency, cost, and overall system performance. Build safeguards that improve production readiness and reliability, including testing pipelines ...

Remote DevOps Team Lead

Hiring Organisation
grabjobs
Location
Stowmarket, Suffolk, UK
operation of Runware’s infrastructure and orchestration systems Build automation and tooling to streamline model deployments, scaling, and hardware utilisation across distributed nodes Drive observability, alerting, and reliability practices to detect and resolve issues quickly and proactively Collaborate with engineers to optimise throughput, latency, and platform performance at every layer … similar languages Understand container runtimes like Docker and containerd, and have built or worked with orchestration systems beyond Kubernetes Are fluent in observability and debugging practices across distributed systems, using logs, metrics, traces, and profiling to drive insight and reliability Care deeply about reliability, efficiency, and engineering quality, and know ...

Software Development Engineer II - BDP

Location
Welwyn Garden City, England, United Kingdom
backend event-driven platform using Java and Spring Boot• Pairing with more senior engineers to design, implement, test, and ship code• Learning to use observability tools like New Relic and Splunk to monitor live systems• Participating in planning sessions and team discussions to understand requirements and contribute ideas• Writing automated … feedback, and share what you're learning Nice to have:• Exposure to Spring Boot, NoSQL databases, or cloud services• Curiosity about performance, scalability, and observability in large-scale systems• Familiarity with Git, CI/CD pipelines, and containerisation• Some experience working with platforms like Kafka Whats ...

Senior Director, Data and Information Marketplace

Location
Cambridge, England, United Kingdom
ensuring intuitive experiences for both people and agents across discovery, access, sharing, understanding and use. Advise the development of capabilities for data access, lineage, observability and quality so that data assets are transparent, trusted and usable at scale. Shape enterprise approaches to data, information and knowledge lifecycle management, embedding governance … seamless experiences that are widely adopted by users and machines across multiple enterprise business units. Deep expertise in relevant capability areas, including data quality, observability, lineage, access management, lifecycle management and information governance. Measurable evidence of optimising the data P&L across covering cost/FinOps, value realisation and sustainability. ...

Remote Senior Developer Relations Engineer

Hiring Organisation
grabjobs
Location
Sandy, Bedfordshire, UK
TL;DR: We're seeking a technical, community-first Developer Relations Engineer to become the face of Cloudsmith in the open source world, with impact lasting from today until IPO and beyond. About Cloudsmith Cloudsmith ...

Remote Senior Software Engineer Quality Engineering & Testing

Hiring Organisation
grabjobs
Location
Cambridge, Cambridgeshire, UK
About Us Sophos is a cybersecurity leader defending 600,000 organizations globally with an AI-driven platform and expert-led services. Sophos meets organizations wherever they are in their security maturity and grows with them ...

Staff Engineer

Location
Norwich, England, United Kingdom
Role Overview We are seeking a highly skilled and strategic Staff Engineer to join our growing Engineering department at our Norwich Head Office. In this pivotal role, you will lead the technical strategy and execution ...

IT Infrastructure Solutions Architect

Location
Cambridge, England, United Kingdom
roadmap.**Key responsibilities*** Define and maintain reference architectures and target-state designs for VMware VCF 9.0 platform architecture and lifecycle patterns.* Define Aria Operations observability strategy (telemetry standards, alert philosophy, capacity/performance governance, service reporting) and ensure operational adoption.* Define VCF Automation platform approach (catalog/service design, templates … iSCSI), VSAN, NAS, and software-defined storage concepts.* Experience or exposure to infrastructure-as-code* Proven capability to architect and operationalize enterprise monitoring/observability standards (Logic Monitor and Aria Operations).* Proven capability to architect, govern, and troubleshoot provisioning automation (VCF Automation).* Proven backup/recovery architecture ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Birchanger, Hertfordshire, United Kingdom
Employment Type
Permanent
Salary
GBP 40,000 - 50,000 Annual
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, modern technology estate. Responsibilities … improve speed, accuracy and consistency Supporting major changes, deployments and post-incident reviews with data-driven evidence Qualifications Strong experience with monitoring and observability tools (LogicMonitor, Azure Monitor, App Insights, Log Analytics, Defender for Cloud) Excellent understanding of cloud performance, IaaS/PaaS, networking fundamentals, API performance and capacity modelling ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Stansted, Essex, South East, United Kingdom
Employment Type
Permanent
Salary
£50,000
talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, modern technology estate. Responsibilities … improve speed, accuracy and consistency Supporting major changes, deployments and post-incident reviews with data-driven evidence Qualifications Strong experience with monitoring and observability tools (LogicMonitor, Azure Monitor, App Insights, Log Analytics, Defender for Cloud) Excellent understanding of cloud performance, IaaS/PaaS, networking fundamentals, API performance and capacity modelling ...

Data Platform Architect: Cloud-Native Snowflake & Lakehouse

Location
Basildon, England, United Kingdom
A consulting firm is seeking a Data Platform Solution Architect for onsite work in Basildon, United Kingdom. The role demands strong expertise in Solution Architecture, Cloud-Native Architecture Design, and hands-on experience with Snowflake ...

Agentic AI Engineer: Build Production-Grade TV Agents

Location
Cambridge, England, United Kingdom
Roku TV is seeking a hands-on Agentic AI Engineer to design, build, and maintain intelligent agents and copilots that drive automation and unlock new product capabilities. You will own the full lifecycle from prototyping ...