1,276 to 1,300 of 1,896 Observability Jobs

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
considering sustainability and cost. Bonus points if: You are familiar with search engine technology such as OpenSearch, ElasticSearch, or Vespa. You are familiar with observability, tracking, and data pipeline tools and methodologies. Additional Information Health & Mental Wellbeing: PMI and cash plan healthcare access with Bupa, subsidised counselling and coaching with ...

Software Engineer, Privacy Engineering (Lawful Access)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
intuitive operator experiences. Identify responsible automation opportunities that reduce repetitive work while preserving human review, judgment, and accountability. Own production systems through testing, observability, incident response, documentation, and continuous reliability improvements. Help define the architecture and roadmap for reusable privacy and legal-infrastructure foundations as the company's products ...

Founding Forward Deployed Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
next iteration. Make speed/quality/cost trade‐offs in real time. Choose build vs. buy vs. configure existing platform capability. Own observability and evals once live. Iterate until the system is reliable in the customer's real conditions. Earn trust with the client's own engineering and security ...

Staff Software Engineer (Web & Mobile)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
offs not just in code but in product impact, time-to-ship, and operational cost Drive operational maturity wherever it’s weakest - release management, observability, incident response, performance monitoring - including in the mobile apps Partner with PMs, designers, and engineering leaders to shape what we build, why, and in what ...

Data Engineer

Hiring Organisation
Aristocrat
Location
Austin, Texas, United States
Employment Type
Permanent
Salary
USD 81 Hourly
Optimize queries and data processing logic to handle large volumes of transactional data efficiently. Ensure data integrity across the pipeline through validation, testing, and observability practices. Investigate and resolve data quality issues, pipeline failures, and performance bottlenecks. Collaborate with data scientists, analysts, and product teams to understand data requirements ...

Architect/Staff Embedded Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
shipped across multiple generations in large and fast‐moving organizations. Track record driving technical outcomes in organizations with high reliability expectations, including robust observability, incident management, and close collaboration with hardware and silicon teams on field issues. Outstanding technical communicator. You can articulate architectural decisions and their consequences clearly ...

Senior Engineering Manager, Global Bank

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
defining scope, aligning with product leadership, and driving delivery across squads and tribes Establish and uplift tribe-wide engineering practices across areas such as observability, incident response, security, or AI workflows, setting standards that go beyond a single squad Act as a senior escalation point for production incidents and complex ...

Mid-Market Sales Engineer UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
customer demonstrations, POCs, and overseeing day-to-day account-level activities. You will be responsible for evangelising, positioning, and architecting Rubrik’s Data Resilience, Observability, and Remediation tools to a targeted list of new & existing customers throughout the UK/I Region. What you’ll do: Provides technical leadership ...

AWS Alliance Director — EMEA

Hiring Organisation
Jobleads-UK
Location
United Kingdom
advantage of all structured and unstructured data — securing and protecting private information more effectively — Elastic’s complete, cloud-based solutions for search, security, and observability help organizations deliver on the promise of AI. What is The Role: As a 3X global award-winning partner of AWS with deep integrations across ...

Engineering Manager, Search

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
cost in mind through thinking thrift.Bonus points if:You are familiar with search engine technology such as OpenSearch, ElasticSearch, Vespa..You are familiar with observability, tracking and data pipeline tools and methodologies.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with BupaSubsidised counselling and coaching with Self SpaceCycle to Work ...

Senior Forward Deployed ML Engineer, Agents

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
agent development, MLOps pipeline implementation, and production optimization. You understand what makes agents perform well in production and how to systematically improve quality through observability and evaluation. Experience with voice AI platforms, RAG systems, and LLM orchestration frameworks is highly desirable. You bring exceptional communication skills, customer empathy … validate datasets for fine-tuning, evaluation, and synthetic data generation Work with other MLEs, MLOps, SREs to carry out model deployment and productionization Observability, Evaluation & Production Operations Implement LLM and agents observability and monitoring tracking token usage, latency, costs, and quality metrics across deployments on aion's infrastructure Instrument applications ...

Principal Software Engineer - Full Stack - AI

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
full stack, guiding the development of responsive frontend applications (React/TypeScript) and robust, scalable backend services (Python, Java, Kotlin, or Node.js). LLM Observability & Reliability: Establish robust LLM observability, evaluations, and caching, implementing latency optimisations and comprehensive monitoring (logging, usage tracking, agent behaviour). Operational Excellence: Champion high availability … performance optimisation, and observability across frontends and backend microservices, focusing on practices that maintain platform reliability and optimise MTTD and MTTR. Mentorship & Collaboration: Elevate the engineering organisation by mentoring senior and junior engineers, conducting rigorous code and system design reviews, and partnering with product managers to translate product visions into ...

Senior or Staff Software Engineer, SRE/ Platform Team

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
code with Kubernetes and Terraform. You'll be at the forefront of shaping our foundational architecture, ensuring it’s both resilient and scalable. Drive Observability and Monitoring: Establish and maintain a state‐of‐the‐art observability and monitoring stack. Your insights will enable us to stay ahead of potential issues ...

Senior AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
developer tooling that enable self‐service AI development across engineering teams. Design secure, scalable deployment pipelines for AI models and applications. Build AI observability capabilities including monitoring, tracing, evaluation, cost optimisation, and production quality measurement. Collaborate closely with AI Engineers, Backend Engineers and Engineering Leadership to define platform architecture … containerised deployments. Experience with solutions such as AWS Bedrock and AgentCore. Understand how to deploy, monitor, and operate AI services in production. AI Operations & Observability Experience implementing monitoring, tracing, evaluation, and cost optimisation for AI systems. Experience with observability solutions such as Arize Phoenix, Langfuse, or Langsmith. Understand the operational ...

AI Software Engineering Associate Director

Hiring Organisation
Jobleads-UK
Location
Newcastle upon Tyne, England, United Kingdom
production-grade agentic systems at enterprise scale: multi-agent orchestration across complex environments, RAG pipelines, policy-based routing, memory management, and programme-level lifecycle observability Define RAG pipeline standards across engagements: establish chunking and embedding strategies, set quality benchmarks, and ensure metric-backed tradeoff decisions are documented and transferable … standard design practice across providers including OpenAI, Anthropic, Vertex AI, and open-source models Own LLMOps at programme scale: eval strategy, prompt governance, observability tooling standards, safety monitoring and cost controls across multiple concurrent systems Lead client engineering engagements at senior level - facilitate architecture design sessions, lead proof-of-concept ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, UK
embedded within the company’s engineering teams to co-deliver reusable, production-ready platform components, Infrastructure as Code (IaC), CI/CD automation, observability, resilience, and FinOps practices aligned to the client's backlog and sprint cadence. The role will support secure, scalable, and cost-effective cloud services while enabling …/CD & Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog ...

Platform Engineer

Hiring Organisation
Morgan McKinley
Location
Newbury, England, United Kingdom
embedded within the company’s engineering teams to co-deliver reusable, production-ready platform components, Infrastructure as Code (IaC), CI/CD automation, observability, resilience, and FinOps practices aligned to the client's backlog and sprint cadence. The role will support secure, scalable, and cost-effective cloud services while enabling …/CD & Automation: Implement and improve CI/CD pipelines, deployment automation, environment management, and release practices for key platform services. Operational Excellence: Embed observability, monitoring, resilience, security, and FinOps practices into platform components to improve reliability, cost transparency, and operational readiness. Agile Contribution: Actively contribute to the engineering backlog ...

Principal Platform Engineer

Hiring Organisation
SF Partners Admin
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
capabilities. Design and operate production-grade Kubernetes platforms, including EKS, AKS or OpenShift. Define engineering standards, golden paths, reusable modules and platform patterns. Build observability strategies using Prometheus, Grafana, OpenTelemetry and modern APM tooling. Improve reliability through SLOs, incident reviews and Site Reliability Engineering (SRE) practises. Embed DevSecOps, supply-chain … Infrastructure as Code (IaC). CI/CD automation. GitOps tools such as ArgoCD or Flux. Internal Developer Platforms or self-service engineering. Observability tools including Prometheus, Grafana, OpenTelemetry, ELK, Datadog, Dynatrace or New Relic. DevSecOps and supply-chain security. SRE practises, SLOs, SLIs and incident management. Platform governance, cloud ...

Senior Infrastructure Engineer, Public Cloud

Hiring Organisation
Jobleads-UK
Location
Halifax, England, United Kingdom
Develop self-service platform capabilities to improve developer experience Apply Site Reliability Engineering (SRE) practices to platform operations Support incident response, monitoring, observability and continuous improvement Diagnose issues across performance, scaling, storage and automation Contribute to a 24x7 on-call rotation Implement policy-as-code controls (e.g. OPA Gatekeeper, RBAC …/egress patterns (e.g. Istio, Anthos, Cloud Service Mesh) Support cloud networking (VPCs, DNS, NAT, VPN, routing, connectivity) Integrate shared platform services (cert-manager, observability, cost tooling) Requirements Strong experience in Platform Engineering, DevOps or SRE Proven delivery of production Kubernetes platforms, ideally GKE Experience with multi-tenant platform environments ...

Staff Software Engineer – Payments Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
based platforms and high-traffic applications, with experience across Azure (preferred), AWS or GCP. Strong understanding of platform engineering, Site Reliability Engineering (SRE) principles, observability and monitoring tools such as Dynatrace or equivalent. Experience implementing CI/CD practices and modern DevOps tooling, ideally including GitHub Actions and automated delivery …/CD Implementation DevOps Practices ATS Optimization Keywords Hard Skills Java Platform Engineering Cloud Computing DevOps CI/CD Automation Incident Management Observability Monitoring Tools Testing Soft Skills Mentoring Coaching Knowledge Sharing Industry Keywords Payment Systems Engineering Excellence Operational Improvements Service Performance High-Traffic Applications Tools & Technologies Azure ...

Principal Infrastructure Engineer

Hiring Organisation
Sidram tech
Location
San Francisco, California, United States
Employment Type
Permanent
Salary
USD Annual
secure, highly available Kubernetes platforms across multi-cloud environments. Collaborate with customers, product teams, and engineering to deliver scalable infrastructure solutions. Drive platform reliability, observability, security, and performance optimization. Conduct infrastructure code reviews, security reviews, and architecture improvements. Support production environments and ensure operational excellence. Required Qualifications Bachelor's Degree … platforms. Knowledge of data and ML platforms including Snowflake and Databricks. Experience with cloud networking (AWS VPC, Azure VNet, GCP VPC). Experience with observability tools (Prometheus, ELK Stack, Grafana, or similar). Experience with service mesh technologies such as Istio or Linkerd. Strong understanding of cloud security, scalability, reliability ...

DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Product Engineering teams to align on strategy, shape technical direction, and embed a true DevOps culture across the organisation. Championing infrastructure as code, observability, and a bias toward automation at every layer of the stack, you'll influence design decisions and cross‐team ways of working that raise … engineers to deeply understand their development workflows and release processes, identifying and removing friction points that slow delivery Own and continuously improve the observability stack — from instrumenting services with logging, metrics, and tracing, to building dashboards, alerts, and runbooks that enable proactive detection and faster resolution Measure and improve ...

Performance Test Engineer - SC Cleared

Hiring Organisation
Lorien
Location
London, UK
Employment Type
Full-time
performance testing strategies and non-functional requirements. \n Identify performance bottlenecks and provide actionable recommendations for improvement. \n Monitor application and infrastructure performance using observability and diagnostic tools. \n Collaborate with engineering teams to improve system scalability and resilience. \n Integrate performance testing into CI/CD pipelines and automated … analysing performance, load, stress, and volume tests. \n Strong knowledge of cloud platforms including Azure, AWS, or GCP. \n Experience with monitoring and observability tools such as Grafana, Splunk, New Relic, or Dynatrace. \n Scripting and automation skills using Python, JavaScript, or similar technologies. \n Understanding of Docker, Kubernetes ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Basingstoke, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Drive platform improvements and DevOps best practices. Design and implement self-service infrastructure and tooling. Deliver scalable, secure, and highly available systems. Enhance monitoring, observability, and operational performance. Support engineering teams with technical expertise and guidance. Skills & Experience Experience designing and implementing CI/CD pipelines and software delivery processes. … Infrastructure as Code experience using tools such as Terraform or Ansible. Experience with monitoring and observability tools. Strong knowledge of Docker, Kubernetes, AWS, and cloud technologies. Excellent communication skills and ability to collaborate across teams. A passion for automation, platform engineering, and continuous improvement. This is a full-time, permanent ...

Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
leading high-performance, distributed computing products to run reliably, securely and efficiently at scale. This includes infrastructure automation, runtime orchestration, CI/CD enablement, observability and performance optimisation. You will work closely with software engineers, data scientists, product owners, delivery leads, and central IT to define and deliver the platform … scheduling. Implement Infrastructure as Code (IaC) and automated environment provisioning. Build and maintain CI/CD pipelines supporting distributed systems and shared components. Implement observability through logging, metrics and alerting, improving platform reliability and debuggability. Monitor and optimise system performance, throughput and resource utilisation. Ensure platform security, access control ...