51 to 75 of 1,783 Observability Jobs in London

Cloud/ DevOps Engineer

Location
Greater London, England, United Kingdom
Microsoft Azure. Based in Richmond, London, you will work closely with software engineering and cloud teams to improve cloud infrastructure, deployment automation, observability and platform reliability. ## Responsibilities* Build, maintain and support AWS cloud infrastructure using Infrastructure as Code principles.* Support AWS services including EKS, EC2, ECS/Fargate ...

Engineer C# (Full Stack)

Location
Greater London, England, United Kingdom
solving, communication, and collaboration skills.Desired* Experience with microservices and event-driven architectures.* Experience with AI assisted design and coding.* GraphQL, and WebSockets.* Knowledge of observability tools (e.g., OpenTelemetry, Grafana).* Familiarity with Infrastructure as Code (Terraform).* Understanding of financial markets or trading systems.* Contribution to open-source projects.* Awareness ...

Junior Data Engineer

Location
Greater London, England, United Kingdom
primary), some GCP Warehouse & Storage: Snowflake, S3/Parquet Data & ETL: dbt, Fivetran Platform & Infra: Kubernetes, Kafka, RabbitMQ, Argo, GitHub Actions, HashiCorp Vault Observability: Datadog, Grafana Dashboarding: Preset Other : Claude What we’re looking for Strong fundamentals and the ability to apply them pragmatically: Solid programming ability (Python or similar ...

Senior Engineer (Java, Typescript & Azure)

Location
City Of London, England, United Kingdom
stakeholders to understand business needs and shape technical solutions Designing, developing and deploying scalable applications using modern engineering practices Driving quality through automated testing, observability and continuous improvement initiatives Leading technical problem-solving and supporting incident investigation and resolution Championing secure software development practices and promoting engineering best practice Mentoring ...

Senior System Engineer London

Location
Greater London, England, United Kingdom
code tools such as Ansible, Terraform, Chef, Puppet, or Salt. Proven experience deploying, managing, and troubleshooting Kubernetes clusters in production environments. Strong understanding of observability and monitoring practices, including experience with tools such as Prometheus, Grafana, or similar platforms. Demonstrated ability to work confidently across both Unix based systems (Ubuntu ...

Principal Network Platform Lead

Location
Greater London, England, United Kingdom
documentation, and operational support. Partner with Platform Engineering to ensure reliability, resilience, performance, and rapid issue resolution in production. Drive continuous improvement across architecture, observability, engineering practices, and operational efficiency. Technical Leadership Provide technical leadership and mentorship to engineers working across network automation and APIs. Foster an automation-first software ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Location
London, United Kingdom
Preferred qualifications, capabilities, and skills Advanced knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Preferred qualifications, capabilities, and skills Advanced knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh ...

Senior Fullstack Developer — Java or Kotlin, React and TypeScript

Location
Greater London, England, United Kingdom
processes using BPMN or comparable technologies.* Experience building shared platforms, common services or greenfield technical foundations.* API governance in a multi-team engineering environment.* Observability practices covering logs, metrics and distributed tracing.* Experience working in regulated, financial-services or high-throughput environments.* Exposure to payments, transaction processing or other complex ...

Platform Engineer

Location
Greater London, England, United Kingdom
container orchestration platforms; Experience in using a scripting language (such as Bash or Python), for CI/CD and general problem‐solving; Knowledge of observability concepts and ability to utilise monitoring tools (logs, metrics, dashboards, alerts – we use Grafana) to identify problems and find solutions. Nice‐to‐haves Hands ...

Principal Engineer (Gen AI and MACH Architecture)

Location
Greater London, England, United Kingdom
existing models for production use cases; including data preparation, evaluation, versioning and deployment within enterprise governance, cost and reliability constraints. AI cost‐value analysis, observability, governance, testing and evaluation frameworks for production systems. Practical application of Generative AI to marketing and experience challenges; e.g. personalisation, content generation, campaign optimisation; with ...

Senior Engineering Manager Software engineering London

Location
Greater London, England, United Kingdom
production. Data & platform: SQL, plus experience designing batch and stream data-processing systems. Cloud & tooling: AWS and/or GCP, Docker, Terraform, and observability tooling such as Datadog. Knowledge of Kubernetes, Kafka, and CI/CD pipelines is highly beneficial. Additional Information Bring all of you to work We create ...

Cloud Platform Product Manager - UK Security Clearance eligibility required

Location
Greater London, England, United Kingdom
Support the definition and rollout of outcome focussed SOWs, ensuring clear lines of responsibility between different platform focussed product teams. (eg CI/CD, Observability etc) Ensure delivery aligns with DevSecOps best practices (automation, IaC, continuous assurance). Governance & Reporting Define and monitor KPIs/OKRs for product success, including ...

Senior Architect, ADC/Quant

Location
Greater London, England, United Kingdom
compliance requirements typical of the London financial services sector.* Resilience & SRE: Advanced knowledge of building fault-tolerant architectures leveraging modern SRE principles, robust observability, and cloud-native resilience patterns.* Matrix Leadership: Exceptional stakeholder and client communication skills, with a track record of influencing cross-functional engineering pods, product managers ...

DevOps Engineer - UK

Location
Greater London, England, United Kingdom
/CD pipelines Strong AWS architecture knowledge Experience with large-scale distributed systems Hands-on experience with containerization & orchestration Experience with monitoring and observability tools Solid understanding of infrastructure security British Citizen or right to work in UK (no visa sponsorship available) Strong communication & problem-solving skills Experience working with ...

Technology Integration Specialist

Hiring Organisation
Ncounter
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£130,000 - £150,000 per annum
/CD knowledge across tooling such as Jenkins, GitLab CI or ArgoCD Strong Linux knowledge and experience working across AWS, GCP or Azure Observability experience with technologies such as Prometheus or Grafana Scripting or software development capability, ideally Python or Go Experience leading complex technical projects involving multiple engineering teams ...

cloud engineer in cloud platforms

Location
Greater London, England, United Kingdom
cloud platforms and services with Architects and engineering teams Support Azure networking, identity, security, compute, storage and platform services Implement monitoring, logging, alerting and observability across cloud environments Troubleshoot complex platform, deployment and infrastructure issues Improve reliability, scalability and operational performance through automation and engineering best practice Support containerised workloads ...

Junior Data Engineer

Location
Greater London, England, United Kingdom
primary), some GCP* **Warehouse & Storage:** Snowflake, S3/Parquet* **Data & ETL:** dbt, Fivetran* **Platform & Infra:** Kubernetes, Kafka, RabbitMQ, Argo, GitHub Actions, HashiCorp Vault* **Observability:** Datadog, Grafana* **Dashboarding:** Preset* **Other**: Claude### **What we’re looking for**Strong fundamentals and the ability to apply them pragmatically:* Solid programming ability (Python or similar ...

Senior Software Engineer - Backend & Distributed Systems Engineer

Location
Greater London, England, United Kingdom
operations team. Automate Delivery and Infrastructure: Manage infrastructure using Pulumi or Terraform and improve automated testing and deployment through GitHub Actions. Improve Reliability and Observability: Build effective monitoring, logging, tracing and alerting across our data, service, model and agent infrastructure. What We’re Looking For 7+ years of professional experience ...

Software Engineering Manager (Test & Devops)

Location
Greater London, England, United Kingdom
cybersecurity guidance) and SOUP management Experience managing complex build environments (CMake, Conan, Ceedling) or cloud-native infrastructure-as-code (Terraform, CDK) Familiarity with observability tooling such as Grafana and Prometheus Benefits: Company equity plan so all employees share in the success of the company Salary-sacrifice pension scheme Private medical ...

DevOps Engineer

Location
Greater London, England, United Kingdom
Infrastructure as Code (Terraform, Pulumi, etc.) Solid understanding of CI/CD practices and tools (GitHub Actions, CircleCI, etc.) Experience with monitoring and observability tools (Datadog, CloudWatch, etc.) Experience with containerisation and container orchestration (Docker, ECS, etc.) Proficiency in at least one programming/scripting language (Python, Go, etc.) Strong ...

Senior Software Engineer – Data

Location
Greater London, England, United Kingdom
support. Perform thoughtful peer reviews that raise code quality and share best practices across the team. Contribute to platform‐wide initiatives that improve reliability, observability, and cost efficiency. Help shape the technical roadmap for data engineering across Elliptic. Leverage and deploy AI and agentic systems to operate more effectively. Tech ...

Principal Data Scientist - Optimisation Engineering

Location
Greater London, England, United Kingdom
correctness, explainability, and maintainability. Establish strong engineering practices: code review, automated testing, CI/CD, release management, incident response, and post‐incident learning. Build observability into optimisation services (KPIs, logs, traces) and manage performance tuning (latency, throughput, cost) across environments. Contribute hands‐on when needed (prototyping, critical‐path coding, reviews ...

Principal AI Quality Engineer

Location
Greater London, England, United Kingdom
keep Fourth’s testing practices at the leading edge. Experience with CI/CD tooling, e.g. Jenkins, Azure DevOps, and Octopus Deploy. Experience with observability tooling such as Prometheus, Grafana, and Sumo Logic. Benefits Holidays. We all need to rest so you get 25 basic holidays with the option ...

Vice President, Software Engineer (Private Markets Data)

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
implement reproducible, testable data workflows that enable advanced analytics and Artificial Intelligence and Machine Learning (AI/ML) use cases, incorporating validation, monitoring, observability, experimentation, versioning, and production deployment capabilities. Lead architectural decisions and influence technical prioritisation, partnering closely with product and delivery teams to align engineering effort with business ...