201 to 225 of 239 Observability Jobs in the South East

Data Platform Engineer

Hiring Organisation
ALTERED RESOURCING LTD
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £75,000 per annum
Data Platform Engineer £70,000 – £75,000 + 10% bonus + excellent benefits We're looking for a hands-on Data Platform Engineer who enjoys getting stuck into infrastructure, automation and platform engineering. This isn ...

Infrastructure Support Engineer

Hiring Organisation
Blues Point Ltd
Location
Shoreham-by-Sea, West Sussex, United Kingdom
Employment Type
Full-Time
Salary
£35,000 - £40,000 per annum
Manage Active Directory, SSL certificates and secure communications Support networking across AWS, including VPCs, IGWs, NAT Gateways and security groups Contribute to monitoring and observability using Grafana Support infrastructure maintenance, migrations, patching, audit and compliance activity Work collaboratively within an Agile/sprint-based environment There will be occasional … Code experience Experience with containers and orchestration PostgreSQL experience Networking and load-balancing knowledge Experience with SSL certificates and endpoint management Grafana/observability experience Good communication skills and the ability to explain technical issues to non-technical audiences A proactive approach and the ability to work independently ...

Data Platform DevOps Analyst

Location
Reading, England, United Kingdom
journey supporting the migration from Teradata to Databricks, validating production readiness, improving monitoring and automation capabilities and helping embed modern DataOps practices across deployment, observability and support processes. What You’ll Bring Here at Primark, we want everyone to feel valued – so please bring your authentic self to work … platforms and DataOps practices with exposure to Azure DevOps, Databricks, dbt Cloud, Azure data services, SQL, Python, ETL/ELT processes, job scheduling, monitoring, observability and deployment automation. Experience supporting production environments and service operations including monitoring, alerting, incident management, ticketing systems, release management, environment management, change governance and enterprise ...

Senior Data Engineer

Location
Maidenhead, England, United Kingdom
data solutions are reliable, scalable, performant, secure, and production‐ready Monitor, troubleshoot, and continuously improve pipeline performance, data quality, and platform stability Drive automation, observability, and supportability across data, analytics, and AI/ML solutions Our Ideal Candidate Strong data engineering experience with hands‐on delivery of scalable data pipelines … productionization, LLM‐based applications, or agentic AI patterns will be an added advantage Experience with DevOps and DataOps practices, including CI/CD, monitoring, observability, and incident support Maersk is committed to a diverse and inclusive workplace, and we embrace different styles of thinking. Maersk is an equal opportunities employer ...

Senior AI Engineer

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
From £700 to £900 per day
predictable failure behaviour. Deploy AI systems across cloud and on-premises environments, with an understanding of the constraints associated with each. Build evaluation and observability capabilities to measure model performance, agent behaviour and system reliability. Take end-to-end ownership of technical workstreams, from architecture and implementation through to deployment. … would be advantageous: Multimodal AI and reasoning. Edge or offline AI deployments. Kubernetes, particularly EKS or OpenShift. MLOps, including model evaluation, monitoring and reproducibility. Observability for agentic AI systems, including model performance, agent behaviour and drift. Agent orchestration and inter-agent communication protocols such as A2A. Model Context Protocol ...

Principal Platform Engineer || Identity Platform (AuthN/AuthZ)

Location
Staines-upon-Thames, England, United Kingdom
each is the right answer Policy‐as‐code exposure: OPA/Rego, Cedar or similar Running the authorisation engine in production on PostgreSQL, with observability and traceability of the decisions it makes Authentication Enterprise‐scale authentication you have architected and operated, not integrated with Hands‐on production Curity and/… approaches and tooling (OPA/Rego, Cedar, or equivalent) Understanding of the operational side: running the authorisation engine in production, backed by PostgreSQL, with observability and traceability of authorisation decisions Authentication (Must Have) Architecting and engineering enterprise‐scale AuthN solutions, demonstrated at production scale Hands‐on production experience with Curity ...

Full Stack Developer

Hiring Organisation
Corus Consultancy
Location
Crawley, West Sussex, United Kingdom
Employment Type
Permanent
Salary
£50000 - £60000/annum
About the Role We are looking for an experienced Full-Stack Developer to join a product engineering team developing and supporting modern digital services. This is a hands-on engineering role where you will work ...

Database Reliability Engineer

Location
Southampton, England, United Kingdom
tune performance across hundreds of instances. Architect Cross‐Cloud Portability: use CNPG and cloud‐native patterns to keep our database layer provider‐agnostic. Evolve Observability & Monitoring: build proactive monitoring and alerting to detect regressions before they affect customers. Support Replication & Mobility: enable data streaming and zero‐downtime migration strategies … provision infrastructure, avoiding manual implementations. Distributed Systems enthusiast: enjoy the challenge of multi‐tenant, multi‐region, multi‐cloud scenarios with rigorous data integrity. Security & Observability mindset: build deep observability (Prometheus/Grafana/OpenTelemetry/Humio) and guardrails for secure operation. Engineering via code: deliver backend services in Java with ...

Platform Engineer – Infrastructure & Cloud Systems

Location
Crawley, England, United Kingdom
deployment and execution across large-scale, distributed environments.You will work across multiple layers of the stack, including infrastructure platforms, container orchestration, authentication systems, and observability tooling. This role focuses on building scalable, reliable systems and ensuring strong integration between infrastructure and higher-level platforms.**About The Team**You will join … relation to underlying infrastructure.**- Security & Access*** Contribute to authentication, authorisation, and service identity systems.* Support secure access patterns including OIDC and access control mechanisms.**- Observability & Reliability*** Implement and maintain observability, monitoring, and logging solutions.* Improve system reliability, performance, and scalability across platform layers.**- Architecture & Performance*** Analyse system performance and optimise ...

Senior AI Engineer

Hiring Organisation
Addition
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£475.00 per day
Developing maintainable, high-quality software using Python. Creating and managing CI/CD pipelines using GitHub Actions and related technologies. Implementing monitoring, evaluation and observability processes for AI systems. Supporting human-in-the-loop workflows and AI model-evaluation frameworks. Troubleshooting complex technical issues across AI applications, data platforms … engineering, particularly involving large language models. Strong proficiency in Python and production-level software development. Experience using LangSmith, Braintrust, Arize or similar AI observability and evaluation platforms. A strong understanding of LLM backend architecture and AI operations. Experience building data platforms and supporting scalable data infrastructure. Strong knowledge of GitHub ...

Software Developer - Data Platform & Distributed Systems

Location
Crawley, England, United Kingdom
Kubernetes deployments.* Support event-driven architectures using messaging systems and caching technologies.**-Architecture & Reliability*** Participate in system design and architecture discussions.* Ensure reliability, observability, and performance of core services.#**Qualifications**## *Required** Proven experience building backend services and distributed systems.* Strong experience with MongoDB and/or PostgreSQL.* Solid understanding … code.* Strong problem-solving skills with the ability to diagnose complex issues.#### *Preferred** Experience working with high-throughput or low-latency systems.* Familiarity with observability tools and performance profiling.* Experience in data-intensive environments.* Experience with Golang or willingness to learn.* Demonstrated technical or project leadership experience.* Competitive salary commensurate ...

AI Engineer

Hiring Organisation
Norton Rose Fulbright LLP
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
full-stack AI engineering role, with applied AI product delivery at its core. You will build AI workflows, tools and integrations, evaluation and observability capabilities, together with the APIs, services and user interfaces needed to ship them reliably. This is not a research-only role. You will apply agreed enterprise … feedback loops. Manage prompts, model configuration, tool schemas and routing as tested product assets, including fallbacks and cost/latency trade-offs. Implement AI observability through traces, logs, metrics and evaluation results, so quality, reliability, failure modes, latency and cost are visible and can be improved. Build clear APIs ...

Platform Engineering Manager (SRE)

Location
Bracknell, England, United Kingdom
accelerate software delivery and operational performance. Build and maintain core platform capabilities: CI/CD pipelines, build and release automation, infrastructure automation, environment management, observability, and developer tooling. Drive standardisation and reduce engineering friction through automation and self-service. Site Reliability Engineering Introduce and embed SRE practices across Engineering. Improve … scale. A background built in cloud product or SaaS companies, where reliability is something customers feel directly. Deep, hands-on SRE expertise: monitoring, observability, alerting, incident management, and operational readiness. You can define what good looks like and stand up foundational SRE practices, from SLOs and error budgets to public ...

Principal Consultant – Agentic AI, Integration & API Architecture

Hiring Organisation
NeosAlpha Technologies
Location
Maidenhead, England, United Kingdom
secure, governed, and production-ready agentic architectures. The role spans Agentic Architecture, Agent Control Planes, AI Gateways, MCP/A2A, Agentic Security & Governance, Observability and Agentic Quality Engineering, alongside enterprise integration and API architecture. As Agentic AI is an emerging field, we do not expect candidates to have many years …/tool discovery and registries o Agent identity and delegated identity o Human-in-the-loop controls o Context engineering and RAG o Agent observability, auditability and cost management o Agentic Security and Governance o Agentic Quality Engineering o Hybrid and multi-cloud agent architectures Assess how platforms such ...

Live Quantum Observability Engineer

Location
Reading, England, United Kingdom
leading quantum computing company, is seeking a Monitoring and Observability Engineer to keep our live cryogenic systems performing at peak reliability. You’ll build dashboards, refine alerts, and automate responses to maximise uptime, collaborating across engineering, operations and reliability teams. Ideal candidates have experience in monitoring, dashboards, Python development … call incident management, with a passion for observability and continuous improvement in complex #J-18808-Ljbffr ...

Senior AI Engineer

Location
Maidenhead, England, United Kingdom
other teams build against. Build evaluation frameworks and developer tooling robust enough for production yet simple enough for non‐specialist developers to adopt. Establish observability standards for AI systems – quality, performance, cost, and regression signals – and build dashboards and reporting that turn those signals into actionable decisions. Drive engineering rigor … machine translation, or content‐generation systems, including metrics such as COMET, chrF++, BLEU, MetricX, and MQM‐style human evaluation. Experience with experimentation and observability tooling, data/test‐set versioning, and rigorous benchmarking workflows. Established practice in AI governance and documentation - model cards, system cards, reproducibility, and responsible‐AI considerations ...

Senior Integration Architect

Hiring Organisation
TXP Technology x People
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£700.00 per day
Senior Integration Architect | Public Sector Transformation Programme Location: Remote with occasional/ad hoc travel to client sites as required Contract: Initial 6-month contract with strong extension potential Rate: Up to £700 per day ...

Data Platform Engineer

Hiring Organisation
Hays
Location
Milton Keynes, Buckinghamshire, UK
Employment Type
Full-time
Reference: 4819090Job ID: 5408590Posted: 2026-08-05Closing date: 2026-11-02Location: Milton Keynes, (Hybrid - 1 day a week on site)Salary: 55 - 65k (DOE) (per annum)Job type: PermanentWorking pattern: Full timeIndustry: Property ...

IT Network Engineer Apprentice

Hiring Organisation
QA
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£26,000 per annum
networking, network security and operational support, while developing awareness of modern data centre networking concepts such as EVPN/VXLAN, software-defined networking, automation, observability and high-performance AI fabrics including NVIDIA InfiniBand. Working under the guidance of senior network, infrastructure and security engineers, the apprentice will contribute to improving … modern data centre connectivity models. Responsibilities: Network support and engineering Firewall, security and access control Monitoring, troubleshooting and operational reliability Modern network automation and observability Documentation and standards Collaboration across teams Required skills: Foundational understanding of networking concepts, including IP addressing, subnetting, routing, switching, DNS, DHCP and common network services ...

Strategic Enterprise Account Executive - Public Sector

Location
Maidenhead, England, United Kingdom
executive leader for named strategic accounts, building trusted relationships with senior business and technology stakeholders and driving adoption of Dynatrace's observability, security and AI-powered platform capabilities. Success in this role comes from understanding how large government organisations operate, navigating complex stakeholder landscapes, influencing major transformation programmes and orchestrating … within assigned accounts. Driving Adoption and Expansion Increase adoption of Dynatrace across departments, programmes and technology domains. Identify opportunities to expand platform usage across observability, application security, infrastructure monitoring, logs, cloud and AI-powered capabilities. Help customers realise greater business value from existing investments while identifying new strategic growth opportunities. ...

Senior Connectivity Engineer

Hiring Organisation
Hackajob Ltd
Location
Wallingford, Oxfordshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
WireGuard/Tailscale or equivalent): access-as-code, policy patterns, posture/health automation, and resilience/disaster recovery planning. Deliver fleet-wide connectivity observability: monitoring, alerting, reporting, and actionable signals that help teams diagnose end-to-end issues quickly. Improve cellular/SIM lifecycle management: provisioning automation, usage/…/PMTUD, conntrack, nftables/iptables) and diagnosing kernel-level networking behaviour. Proficient in Go and/or Python and experienced with modern observability tooling; bonus points for containers/IoT OS, ACL-as-code patterns, and carrier/router API integrations. Benefits Starting from the interview process and continuing ...

Senior Software Engineer, Non-Realtime Controllers

Location
Oxford, England, United Kingdom
common to many controllers rather than specific to one. Work closely with the systems teams and scientists who depend on these services, and strengthen observability, health checking and alerting. Raise the standard of the code around you through review, and bring less experienced engineers on by working alongside them. Requirements … Rust codebase, including async programming. Experience designing and operating networked services such as gRPC or REST, with a real grasp of error handling, observability and failure modes. Experience of hardware and instrument protocols and interfaces such as SCPI, Modbus, serial and Ethernet. Comfortable owning the production behaviour of what ...

Lead ClickHouse Solutions Architect Greenfield Observability

Location
Slough, England, United Kingdom
Colehouse Group is seeking a ClickHouse Solutions Architect to lead a greenfield, enterprise-scale observability platform for a global banking client in Slough. You will shape architecture from first principles to meet high throughput, multi-tenant isolation and storage cost constraints. You'll design data models, ingestion and retention strategies ...

Fleet Connectivity Engineer — Automation & Observability

Location
Wallingford, England, United Kingdom
/or Python, and possess experience in robotics or connected devices. Responsibilities include evolving the connectivity stack, building self-serve workflows, and ensuring observability across the fleet. #J-18808-Ljbffr ...

Technical Delivery Manager - Production / Service Delivery

Hiring Organisation
Michael Page Technology
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£400.00 - £420.00 per day
strategic Production Management & Service Delivery initiatives within a complex financial services environment. Managing around eight concurrent projects, you will drive operational resilience, monitoring and observability, service governance, production support transformation and technology improvements. Client Details An organisation within the financial services industry, located in London. Description Lead … roadmaps and implementation schedules. Oversee project budgets, financial tracking and forecasting activities. Drive delivery of operational resilience initiatives and regulatory commitments. Lead monitoring and observability programmes, including SOC-related enhancements. Support expansion of service management and operational support capabilities. Deliver application improvements, including SSO implementation and production-focused enhancements. Coordinate ...