1,901 to 1,925 of 5,501 Permanent Observability Jobs

Principal Software Engineer - Squad Lead Engineer

Location
City Of London, England, United Kingdom
complete complex bug fixes and performance improvements Define and uphold Definition of Ready/Done including code quality, automated test coverage, security checks, and observability Establish/maintain CI/CD pipelines, quality gates, and sensible branching/release strategies Drive a pragmatic quality strategy: test pyramid balance, contract tests … Windows Experience with relational and non‐relational data stores, performance tuning, and data modelling Knowledge of CI/CD platforms, containers, cloud technologies, observability, and monitoring practices Understanding of secure coding, performance optimisation, reliability engineering, and incident response Work in a Way That Works for You We promote a healthy ...

Software Engineering Manager

Hiring Organisation
WTW
Location
Surrey, United Kingdom
Employment Type
Full Time
maintenance and operational support. Strong practical understanding of modern software engineering principles, including maintainable architecture, automated testing, CI/CD, secure-by-design practices, observability and reliability. Successful track record of growing teams and leading transformation initiatives across engineering effectiveness, Agile or flow-based delivery, people development, cloud/SaaS … cloud-native architectures, SaaS technologies, platform engineering or internal developer platforms. Appreciation of current and emerging technologies, including AI-assisted software delivery, automation, DevSecOps, observability and developer-experience tooling, along with their benefits, limitations and risks. General knowledge of the insurance industry, actuarial or reserving domains, or other regulated software ...

Senior Software Engineer, Data

Location
Greater London, England, United Kingdom
/day scale Contribute to the technical design of the Data Transfer Hub, making pragmatic trade-offs across UX, careliability, throughput, cost, partner constraints, observability and operational support Build parallelised, distributed data transfer pipelines using Flyte for workflow orchestration, Kafka/event-driven patterns for lifecycle tracing, and Azure … Background in autonomous vehicles, robotics, geospatial data, video processing or other data-intensive domains Experience with Kafka or other event-driven systems for transfer observability and lifecycle tracking Experience building dashboards to communicate transfer job status Multi-cloud experience across Azure, AWS and/or GCP Experience improving cost efficiency ...

Staff Network Engineer

Location
Greater London, England, United Kingdom
root-cause analysis for performance and stability issues, and systematically reducing reactive toil through runbooks, automation, and measurable SLOs. Set the direction for network observability, telemetry, monitoring, and alerting to provide clear visibility into fabric health, performance, and traffic patterns. Ensure the accuracy and reliability of source-of-truth network … Juniper SRX and/or Palo Alto, including security policy architecture, high-availability design, and multi‐tenant segmentation. Experience designing network telemetry and observability for high-throughput, performance-sensitive environments. Proven ability to lead complex technical decisions and incidents across networking, systems, storage, and HPC/AI workload teams, with ...

Principal Product Engineer

Location
Greater London, England, United Kingdom
home in NestJS (or a similar Node.js framework) and React. Cloud and DevOps minded. Comfortable across GCP or a similar cloud, CI/CD, observability, and modern infrastructure tooling. Use AI daily. AI assistants are part of how you build. You bring back patterns that help others get more … code. You measure your work by what changed for the customer, not the lines of code shipped. Production minded. Solid grasp of API design, observability, scaling, reliability, and security. Care about craft. Clean APIs, attention to detail, and how the system feels to work in. Comfortable in uncertainty. You move ...

Lead AI Engineer

Location
Manchester, England, United Kingdom
traffic, within real latency budgets and reliability realities. Hold a high engineering bar on AWS and TypeScript — clean CI/CD, infrastructure as code, observability, testing, and LLMOps for running model‐backed systems in production. Shape delivery against the roadmap with the VP, product managers, data analysts and Staff Engineers … powered applications. Extensive experience building and operating cloud‐native applications on AWS, with strong knowledge of CI/CD, infrastructure as code, observability and modern engineering practices. Experience delivering high‐performance, real‐time systems that operate reliably at scale with demanding latency requirements. Demonstrated experience leading the successful transition ...

Principal Product Engineer

Hiring Organisation
Zapp
Location
London, UK
Employment Type
Full-time
home in NestJS (or a similar Node.js framework) and React. Cloud and DevOps minded. Comfortable across GCP or a similar cloud, CI/CD, observability, and modern infrastructure tooling. Use AI daily. AI assistants are part of how you build. You bring back patterns that help others get more … code. You measure your work by what changed for the customer, not the lines of code shipped. Production minded. Solid grasp of API design, observability, scaling, reliability, and security. Care about craft. Clean APIs, attention to detail, and how the system feels to work in. Comfortable in uncertainty. You move ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
data retrieval layers (pgVector, Pinecone) and ensure efficient embedding pipelines for AI contexts. Partner with AI teams to monitor data latency, cost efficiency, and observability metrics. Collaboration & Governance Partner with Platform Operations and Security to enforce privacy, compliance, and access control frameworks (GDPR, SOC2). Work cross‐functionally with analysts … platform data (Google Ads, Meta, TikTok, DV360, Amazon Ads). Demonstrated expertise in code versioning (GitHub), CI/CD integration, and data observability practices. Ability to write clean, modular, testable code and review peers’ contributions for maintainability and performance. Additional Information Publicis Groupe has fantastic benefits on offer ...

Cloud Advisory Architect

Location
Greater London, England, United Kingdom
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Manage & Mentor: Lead teams of architects and engineers, providing technical coaching, career counselling, performance management, and coaching ...

Infrastructure/DevOps Engineer

Location
Greater London, England, United Kingdom
enjoys building strong, cost-efficient infrastructure. Your work will keep systems audit-ready and high-performing. In this senior role, you’ll lead DevOps, observability, and compliance. Tasks include architecting cloud infrastructure for growth, prepping for SOC 2/ISO 27001 audits, and setting up CI/CD pipelines ...

Agentic Commerce Architect

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
intelligence — including embedding pipelines, vector retrieval, semantic search, and structured product data — ensuring agent outputs are accurate, grounded, and commercially reliable Embedding AI governance, observability, and responsible AI principles into architecture design — including audit logging, human-in-the-loop escalation points, and performance monitoring — as first-class concerns rather than … deployment, including familiarity with cloud-native services and CI/CD-based delivery practices Understanding of AI governance and responsible AI principles — including observability, tracing, auditability, and how to build guardrails into production AI systems Preferred Experience Experience contributing to technical proposals, reference architectures, or delivery accelerators in a consulting ...

Senior Software Engineer (Java)

Location
City Of London, England, United Kingdom
market connectivity workflows Knowledge of Linux engineering, troubleshooting, and performance optimisation Experience with Spring Boot or Google Guice dependency injection frameworks Experience with observability stacks (Open Telemetry, Grafana) Experience with distributed caching solutions such as Hazelcast Experience with BDD and automation frameworks (Cucumber) #J-18808-Ljbffr ...

Consultant - Service Management Transformation

Location
Greater London, England, United Kingdom
service management, including intelligent automation, virtual agents, knowledge management, workflow optimisation and service analytics.AIOps & Intelligent Operations – Develop an understanding of modern operational practices including observability, monitoring, intelligent alerting, event management and AIOps capabilities, supporting clients in improving operational performance through automation and data-driven insights.Cloud & Digital Operations – Support the design … automation can be applied within service management to improve operational efficiency, employee experience and service outcomes.Awareness of modern AIOps and observability capabilities, including monitoring, event correlation, anomaly detection, intelligent alerting, operational analytics and automation.Familiarity with enterprise service management platforms such as ServiceNow, Jira Service Management, Freshworks or similar technologies.Experience contributing ...

Remote Engineering Manager, Cloud Platform (UK)

Location
United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Data Platform Engineer- BPL-CIO

Location
Greater London, England, United Kingdom
automation, GitOps operating models AWS Platform Engineering Hands‐on AWS engineering with a focus on: IAM and security patterns, Networking integration, Storage and encryption, Observability, Resilience and operational readiness, Automation and supportability Databricks on AWS Experience with: Workspace deployment, Databricks Terraform provider, Identity integration, Storage integration, Network connectivity patterns Platform … Engineering/Internal Developer Platforms Experience building: Self‐service capabilities, Golden paths, Paved roads, Reusable engineering patterns, Internal platform products Observability, Reliability & Operational Engineering Experience with: Monitoring and alerting, Operational telemetry, Incident management, Resilience testing, Recovery automation, Production support You may be assessed on the key critical skills relevant ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Kent, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Warwickshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Larbert, Stirlingshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Pwllheli, Caernarfonshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Airdrie, Dunbartonshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Brecon, Radnorshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Chesterfield, Derbyshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Newport, Monmouthshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Doune, Perthshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...

Remote Engineering Manager, Cloud Platform (UK)

Location
Wrexham, Denbighshire, United Kingdom
experience. 8+ years of software engineering experience, with 3+ years leading engineering teams Deep expertise architecting distributed systems on AWS (compute, networking, data, IAM, observability) js background and a clear point of view on building maintainable services at scale Track record of improving performance, scalability, and deployment velocity in production ...