2,451 to 2,475 of 5,504 Permanent Observability Jobs

Senior Network Engineer

Location
Greater London, England, United Kingdom
access patterns. Support hybrid connectivity models (site‐to‐site VPN, client VPN, ExpressRoute, Direct Connect, SD‐WAN). Monitor network performance and reliability using observability and telemetry tools; proactively address capacity and performance issues. Troubleshoot complex network and cross‐domain infrastructure issues spanning network, compute, and cloud layers. Develop … constructs (VPC/VNet design, routing, security groups, load balancers). Experience with SD‐WAN architectures and implementations. Familiarity with network monitoring, logging, and observability tools (e.g., SNMP, NetFlow, Syslog, modern NPM tools). Working knowledge of compute platforms and operating systems (Windows, Linux, virtualization such as VMware/Hyper ...

Principal Software Engineer - Squad Lead Engineer

Location
Greater London, England, United Kingdom
complete complex bug fixes and performance improvements* Define and uphold Definition of Ready/Done including code quality, automated test coverage, security checks, and observability* Establish/maintain CI/CD pipelines, quality gates, and sensible branching/release strategies* Drive a pragmatic quality strategy: test pyramid balance, contract tests … Windows* Experience with relational and non-relational data stores, performance tuning, and data modelling* Knowledge of CI/CD platforms, containers, cloud technologies, observability, and monitoring practices* Understanding of secure coding, performance optimisation, reliability engineering, and incident response **Work in a Way That Works for You**We promote a healthy ...

AI Platform Engineer

Location
Greater London, England, United Kingdom
patterns that support the safe and scalable adoption of AI-assisted development.Improve developer experience through streamlined workflows, tooling integration and self-service capabilities. Develop observability and measurement capabilities to provide insights into engineering productivity, quality and platform adoption. Collaborate with Technical Leads and AI Software Engineers to identify recurring engineering … cost optimisation, including token monitoring, caching strategies and model selection considerations. Knowledge of approaches for managing and reducing token consumption costs.Deep understanding of observability, automated testing and software delivery tooling. Knowledge of platform security, governance and operational controls. Strong programming, automation and problem-solving capabilities. Passionate about improving developer experience ...

Senior Lead Software Engineer - Mobile Engineering

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
drive measurable improvements in stability and release confidence. Own mobile build/release and operational maturity: CI/CD pipelines, distribution, feature flags, observability, crash/performance monitoring, and incident response. Mentor and coach engineers; support team growth through feedback, technical guidance, and strong engineering culture. Communicate clearly with senior … patterns, accessibility standards, component libraries).Experience with CI/CD for mobile (e.g., build automation, signing, distribution, feature flags, release trains).Experience with observability and production support practices: crash analytics, performance monitoring, logging, alerting, and operational readiness. Experience leading multiple engineers/teams (people leadership or strong matrix leadership), including ...

Infrastructure Lead

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
other teams, with documentation and examples. In addition, you should have experience working with Kubernetes in production environments at scale, and be familiar with observability tools such as Prometheus and Grafana. Strong Linux Server Administration and Configuration Management skills, as well as some networking experience, are also required. The ideal ...

Senior Platform Manager (Server Infrastructure)

Location
Gloucester, England, United Kingdom
current and future organisational needs. This is an exciting opportunity to play a key role in modernising infrastructure services and driving adoption of automation, observability, and platform reliability best practices. What you would be doing You will be responsible for the operational management, maintenance, and continual improvement of enterprise server … looking for Proven experience in managing enterprise server infrastructure in a complex environment. Experience in Leading Technical Teams. Knowledge of infrastructure monitoring, alerting, and observability tooling. Strong troubleshooting and problem-solving skills with the ability to manage competing priorities effectively. Excellent communication and stakeholder engagement skills, with the ability ...

Principal Engineer - Integration Services

Location
Greater London, England, United Kingdom
duplication and improve interoperability. Providing technical leadership on significant integration decisions spanning multiple platforms, teams and domains. Ensuring integration approaches consider security, resilience, scalability, observability, operability and long‐term maintainability. Working with Architecture and senior engineering leaders to align integration direction with wider FT technology strategy. Establishing a clear view … progress, risks, trade‐offs and technical constraints. Operational Excellence & Reliability Owning the operational performance of shared integration services. Establishing appropriate approaches to monitoring, observability, incident management, problem management and operational readiness. Ensuring critical integrations have clear ownership, appropriate resilience and effective recovery mechanisms. Leading the response to significant incidents affecting ...

Applied AI Engineer

Hiring Organisation
Nufin
Location
London, UK
Employment Type
Full-time
business impact. Create representative test datasets, evaluation criteria, regression tests, and human-review processes. Measure and improve accuracy, latency, cost, and user experience. Establish observability and feedback loops that make agent behavior understandable and continuously improvable. Develop the AI Application ArchitectureApply context-engineering techniques such as RAG, MCP and knowledge … LlamaIndex, or comparable frameworksContext engineering: RAG, MCP, knowledge graphs, tool use, memory, and retrieval systemsAI evaluation: offline and online evaluations, test datasets, regression testing, observability, and human reviewLanguage models: Gemini, OpenAI, Anthropic, Llama, Mistral, or similarBackend engineering: Python or Java, REST APIs, Kafka, microservices, and distributed systemsData systems: SQL, PostgreSQL ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Hiring Organisation
First Up
Location
Dungannon, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Cardiff, Glamorgan, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Chorley, Lancashire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Horsham, Sussex, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Norwich, Norfolk, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Clydebank, Dunbartonshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Okehampton, Devon, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Peterlee, Durham, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Wrexham, Denbighshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Southwell, Nottinghamshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Auchterarder, Perthshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Tamworth, Staffordshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Ventnor, Hampshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Leven, Fife, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Remote Sr. Software Engineer, Fullstack (UK)

Location
Leeds, West Yorkshire, United Kingdom
post-incident reviews in a "you build it, you run it" environment. Identify, analyse, and resolve system availability, reliability, and performance issues, contributing to observability and resiliency improvements. Partner with Product Management and Design to translate business requirements into scalable technical solutions. Minimum Qualifications Bachelor's degree in Computer Science … HRIS platforms such as Workday, SAP SuccessFactors, Dayforce, or similar enterprise HR systems. Experience with Kubernetes, Docker, and Helm. Experience with Datadog or similar observability and monitoring platforms. Demonstrated use of Generative AI tools or coding agents in development workflows. Experience in enterprise SaaS organisations, particularly HR Tech or regulated ...

Principal Software Engineer - Customer Platforms

Location
Greater London, England, United Kingdom
customers to a desired outcome, without prescribing it Authoritative skills at cloud computing (network, security, serverless, Kubernetes etc) and automation Experience with implementation of Observability and Reliability using market technologies (e.g.: New Relic) Good experience with Performance Engineering (load testing, derivations, tuning, core web vitals, page speed etc.) Expertise … organisation(s) Tech Stack M&S uses a variety of technologies including; Java, Spring, SpringBOOT, Micronaut React, Next.js, Typescript, Angular Azure Cloud, Kubernetes, Dynatrace (observability) SQL Server, MongoDB Ignite, Redis What’s In It For You Working at M&S means being part of something bigger - helping to deliver quality ...