1 to 25 of 165 Observability Jobs in Central London

Senior Software Engineer

Location
City of Westminster, England, United Kingdom
fintech, payments, or enterprise SaaS platforms Exposure to event-driven architecture (Kafka, RabbitMQ) Familiarity with infrastructure-as-code tools (Terraform, CloudFormation) Understanding of observability tools (Prometheus, Grafana, ELK stack) L’état d’esprit Edenred - Nous sommes une entreprise unique. Nous recherchons de nouveaux collaborateurs prêts à prendre part ...

Production Engineer

Location
City Of London, England, United Kingdom
identify patterns, and drive intelligent automation solutions Hands-on experience with containerization technologies such as Docker and Kubernetes, including cluster management, deployment, scaling, and observability Deep practical knowledge of algorithmic trading workflows, including the behaviour, lifecycle, and risk controls of execution algos used across the EMEA markets Experience designing ...

Lead Data Engineer

Hiring Organisation
Hackajob Ltd
Location
Westminster, Greater London, UK
processing and data quality frameworks using Python, PySpark, and dbt Build and optimize batch and streaming data pipelines with strong performance, fault tolerance, and observability Develop and operate workflow orchestration (e.g., Apache Airflow) to schedule, monitor, and manage data movement and transformations Model and transform data for analytics using ...

Engineer C# (Full Stack)

Location
City Of London, England, United Kingdom
collaboration skills. Desired: Experience with microservices and event‐driven architectures. Experience with AI‐assisted design and coding. Knowledge of GraphQL and WebSockets. Familiarity with observability tools such as OpenTelemetry and Grafana. Experience with Infrastructure as Code (Terraform). Understanding of financial markets or trading systems. Contribution to open‐source projects. ...

Senior Full-Stack Engineer (Java, Typescript & Azure)

Location
City Of London, England, United Kingdom
stakeholders to understand business needs and shape technical solutions Designing, developing and deploying scalable applications using modern engineering practices Driving quality through automated testing, observability and continuous improvement initiatives Leading technical problem-solving and supporting incident investigation and resolution Championing secure software development practices and promoting engineering best practice Mentoring ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
Hackajob Ltd
Location
Westminster, Greater London, UK
Preferred qualifications, capabilities, and skills Advanced knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh ...

Associate Vice President, Credit Technology Engineering

Location
City of Westminster, England, United Kingdom
deliverables, estimate work, and execute delivery plans. Cloud & Platform Engineering Build and deploy applicationsutilizingcloud-native technologies. Implement containerizedsolutionsleveragingKubernetes and cloud platform services. Develop platform observability, logging, monitoring, and operational support capabilities. Support modernization initiatives focused on automation, scalability, resiliency, and operational efficiency. Integration & Event-Driven Architecture Design and develop APIs ...

Oracle OCI Multi Cloud Engineer

Location
City Of London, England, United Kingdom
Ansible, OCI Resource Manager — to automate cloud provisioning Build and maintain CI/CD pipelines for infrastructure and database change deployment Implement monitoring and observability solutions (OCI Monitoring, Azure Monitor, CloudWatch, Prometheus/Grafana) Automate routine DBA and EBS administration tasks Required Skills & Experience Oracle & Database Solid experience in Oracle ...

Python Developer - AI

Hiring Organisation
83zero Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£55,000
Experience with CI/CD using GitHub, GitLab or Jenkins Agile engineering experience Experience with AI agents, tool calling, embeddings, prompt engineering or LLM observability is beneficial React/TypeScript and Terraform/IaC experience is beneficial Experience taking GenAI POCs into production is highly beneficial Why this role ...

AI Platform Engineer

Hiring Organisation
The Portfolio Group
Location
City of London, London, Castle Baynard, United Kingdom
Employment Type
Permanent
Salary
£80000 - £100000/annum
platform components. Deploying infrastructure using Terraform and supporting containerised applications. Building and maintaining CI/CD pipelines using GitHub and Azure DevOps. Improving observability, monitoring and platform resilience. Supporting vector search, embedding pipelines and knowledge ingestion. Applying security and governance best practice across the AI platform. What we're looking ...

Senior Machine Learning Scientist

Location
Westminster, West End, United Kingdom
enable repeatable releases Develop reusable engineering assets (libraries, templates, reference architectures, infrastructure-as-code patterns) to reduce technical debt and accelerate delivery Implement observability for AI services (logging/metrics/tracing), model performance monitoring, and quality/drift checks with actionable alerting Partner with data scientists, data engineers, platform ...

Senior Software Engineer (Full-Stack - TypeScript/Node)

Location
City Of London, England, United Kingdom
like code review, TDD, CI/CD and pairing using tools like Git and GitHub. Experience of operationally managing software components once live, including; observability, logging, metrics, error reporting, debugging and live incident management. Experience of working with sensitive personal data. Competencies Experience working in/with cross-functional teams ...

Finance Data Engineer (Senior Manager)

Location
City Of London, England, United Kingdom
evidence sufficient for audit, so that agent output can be traced to source, with automated data-quality checks, finance reconciliations, completeness and freshness measures, observability, alerts and recoverable operations. Define data contracts between delivery pods, and between Accenture delivery and client platform teams, including schemas, quality thresholds, access rules, service ...

SC-Cleared DevOps Engineer - Secure Infra, Hybrid

Location
City Of London, England, United Kingdom
problem solving within an agile setup. Ideal candidates will have on-prem Kubernetes/OpenShift exposure, Terraform IAC, Jenkins CI/CD, and strong observability with ELK, Splunk, and Grafana across Azure or AWS cloud platforms. #J-18808-Ljbffr ...

DevOps Engineer - SC Cleared - Hybrid - Inside IR35

Location
City Of London, England, United Kingdom
Experience of working within a production environment, with strong on-prem Kubernetes/OpenShift deployments IAC using Terraform with CI/CD pipelines (Jenkins) Observability tools to include design and operate end to end logging, metrics, tracing, dashboards to include alerting systems using ELK, Splunk, Grafana Cloud platforms to include ...

Senior Data Engineer

Location
City of Westminster, England, United Kingdom
using Databricks & Spark Convert requirements into clear technical designs and conceptual data models Ingest large, complex datasets and automate manual processes Improve data quality, observability, reliability, and delivery performance Champion a build‐once‐consume‐many culture across the platform Collaborate with stakeholders to solve data‐related engineering challenges Contribute ...

Software Engineer

Hiring Organisation
McGregor Boyall Associates Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Desirable Experience Adobe Experience Manager (AEM) Customer onboarding or digital acquisition platforms KYC, financial services onboarding or regulated industry experience AWS Certifications Monitoring and observability tooling experience What's on Offer? Up to £100,000 total compensation package Opportunity to work on high-profile digital transformation initiatives Modern technology stack ...

Sr. Full Stack Developer (Java, React & Python)

Location
City Of London, England, United Kingdom
developing IT solutionsProficient in Python & ReactJS and related technologies is required.Experience with GenAI application development including vector stores, RAG solutions, context engineering, prompt engineering, observability is essentialExperience with Agentic AI development & Java Spring Framework would be advantageousUnderstanding of full stack development and architecture patterns.Experience with RESTful API development and integration.Proven ...

Senior Quality Engineer

Location
City Of London, England, United Kingdom
ensure we are building the right software for our customers and embracing a shift-left approach to software quality Enhancing and utilising our observability stack and embracing a shift-right approach to software quality Implement risk based testing strategies to ensure we are approaching quality at the correct level Actively ...

Engineering Team Lead | Tech | London, UK

Location
City Of London, England, United Kingdom
other multi-agent/LLM orchestration frameworks Experience configuring Azure services (App Service, Container Registry, Blob Storage) or AWS equivalents Familiarity with LLM observability/evaluation tooling (e.g. Langfuse) and resilience patterns for third-party AI providers (e.g. circuit breakers, provider fallback) Awareness of data protection and AI transparency considerations ...

Site Reliability Engineering Lead

Location
City Of London, England, United Kingdom
Risk at https://risk.lexisnexis.com/insurance About our Team The IC (Insurance Core Services) team is responsible for establishing and driving reliability, observability, automation, and operational excellence standards across Insurance technology platforms. The team partners closely with application, infrastructure, database, and cloud engineering teams to improve platform availability … scalability, performance, and resilience. ICS leads strategic initiatives including SLO/SLI implementation, observability platform adoption, cloud modernization, operational readiness reviews, performance engineering, and reliability automation. The team also develops reusable engineering frameworks, standards, and best practices that enable product teams to build and operate highly reliable cloud-native services ...

Senior Software Engineer (Java)

Location
City Of London, England, United Kingdom
market connectivity workflows Knowledge of Linux engineering, troubleshooting, and performance optimisation Experience with Spring Boot or Google Guice dependency injection frameworks Experience with observability stacks (Open Telemetry, Grafana) Experience with distributed caching solutions such as Hazelcast Experience with BDD and automation frameworks (Cucumber) #J-18808-Ljbffr ...

Senior Manager - AWS Solution Architect

Hiring Organisation
Anson Mccade
Location
Central London, London, United Kingdom
Employment Type
Permanent
data, AI, security and integration • Work with senior stakeholders and present architecture and strategy at executive level • Drive cloud cost optimisation, resilience, performance and observability • Mentor architects, managers and consultants across the wider team • Help clients adopt modern cloud-native technologies and engineering practices What they're looking for: • Significant ...

Lead Agentic AI Engineer

Location
City Of London, England, United Kingdom
implement evaluation frameworks for quality, grounding, task success, safety, latency, and cost Experience deploying AI solutions into production with focus on cost optimisation, scalability, observability, security, and governance Experienced in using Microsoft GenAI ecosystem, including M365, Copilot Studio, Azure AI Foundry, and Microsoft Agent Framework Proficient in Python and production ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
user experiences. ThousandEyes is deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform ...