2,126 to 2,150 of 2,912 Remote Observability Jobs

Senior AI Product Engineer

Hiring Organisation
Elliptic
Location
London, UK
Employment Type
Full-time
more junior engineers through pair programming, code review, and design feedback. Raise the engineering bar across the team by promoting good practices in testing, observability, and AI system reliability. Influence cross-team decisions on how AI capabilities integrate with the rest of the Elliptic platform. What you will achieve … technical direction of an AI workstream, including architecture, evaluation, and rollout. Established or improved at least one team practice for building AI systems (evals, observability patterns, prompt management, rollout safety).Mentored junior engineers on AI engineering practices and contributed to their growth. Built strong working relationships across product, web engineering ...

Platform Engineer

Location
United Kingdom
Terraform, CloudFormation, or CDK), CI/CD pipelines, API Gateway, Lambda, and Aurora PostgreSQL — and confident applying fundamentals such as high availability, fault tolerance, observability, and cost control in a live environment. Payments or fintech exposure is an advantage but not required; what matters more is a practical, first‐principles … rollback processes Support deployment of AI generated applications and tooling Support change control and release management alongside Engineering and the Information Security Officer Reliability, Observability & Data Infrastructure Design and maintain systems for high availability, fault tolerance, and resilience by default Implement logging, monitoring, and alerting (for example CloudWatch) across services ...

Enterprise Network Architect

Location
Merley, England, United Kingdom
security frameworks, firewalls, endpoint protection, and SIEM tools. Strong knowledge of data management platforms, databases, data lakes, Fabric and ETL processes. Experience with observability tools and practices, including monitoring, logging, tracing, and metrics collection using platforms such as ELK stack, Grafana, Solarwinds & Azure Monitor. Ability to design and implement observability ...

Senior DevOps Engineer

Hiring Organisation
Experis
Location
Derby, Derbyshire, East Midlands, United Kingdom
Employment Type
Contract
/CD pipelines and deployment orchestration. Support Kubernetes and OpenShift platform troubleshooting and optimisation. Deliver secure and compliant infrastructure solutions. Implement monitoring, logging, and observability tooling across environments. Collaborate with engineering, architecture, and delivery teams to improve deployment efficiency and platform reliability. Champion automation-first approaches to infrastructure and application … designing and maintaining enterprise-scale CI/CD pipelines . Strong understanding of cloud security and secure delivery practices. Experience implementing monitoring, logging, and observability solutions. Ability to define technical standards, governance, and reusable deployment frameworks. Experience working within large-scale enterprise transformation programmes. Desirable Skills Experience within highly regulated ...

Lead Software Engineer

Hiring Organisation
Mondo
Location
Phoenix, Arizona, United States
Employment Type
Permanent
Salary
USD Annual
engineers and guide collaborative technical problem-solving efforts Participate in code reviews, CI/CD improvements, and engineering best practices discussions Support production operations, observability, troubleshooting, and continuous improvement efforts Improve engineering quality, delivery execution, and overall team effectiveness Balance approximately 80% hands-on engineering work with 20% leadership responsibilities … functional teams Preferred Qualifications: Angular experience Ruby on Rails exposure PostgreSQL experience Docker experience CI/CD pipeline experience, preferably with CircleCI Experience with observability and monitoring tools such as DataDog and Sentry Financial services or fintech industry experience Experience within enterprise SaaS or customer-facing product organizations ...

Lead Software Engineer

Hiring Organisation
Mondo
Location
Charlotte, North Carolina, United States
Employment Type
Permanent
Salary
USD Annual
engineers and guide collaborative technical problem-solving efforts Participate in code reviews, CI/CD improvements, and engineering best practices discussions Support production operations, observability, troubleshooting, and continuous improvement efforts Improve engineering quality, delivery execution, and overall team effectiveness Balance approximately 80% hands-on engineering work with 20% leadership responsibilities … functional teams Preferred Qualifications: Angular experience Ruby on Rails exposure PostgreSQL experience Docker experience CI/CD pipeline experience, preferably with CircleCI Experience with observability and monitoring tools such as DataDog and Sentry Financial services or fintech industry experience Experience within enterprise SaaS or customer-facing product organizations ...

Lead Software Engineer

Hiring Organisation
Mondo
Location
Baltimore, Maryland, United States
Employment Type
Permanent
Salary
USD Annual
guide technical problem-solving efforts collaboratively Participate in code reviews, CI/CD improvements, and engineering best practices discussions Support production operations, monitoring, observability, troubleshooting, and continuous improvement initiatives Help improve engineering quality, delivery execution, and overall team effectiveness Minimum Requirements: 10+ years of professional software engineering experience Strong frontend … experience Ruby on Rails exposure PostgreSQL experience Docker and containerization experience CI/CD pipeline experience (CircleCI preferred) Experience with DataDog, Sentry, or other observability tools Financial services, fintech, or enterprise SaaS experience ...

Lead Software Engineer

Hiring Organisation
Mondo
Location
Dallas, Texas, United States
Employment Type
Permanent
Salary
USD Annual
engineers and guide collaborative technical problem-solving efforts Participate in code reviews, CI/CD improvements, and engineering best practices discussions Support production operations, observability, troubleshooting, and continuous improvement efforts Improve engineering quality, delivery execution, and overall team effectiveness Balance approximately 80% hands-on engineering work with 20% leadership responsibilities … authorization required Preferred Qualifications: Angular experience Ruby on Rails exposure PostgreSQL experience Docker experience CI/CD pipeline experience, preferably with CircleCI Experience with observability and monitoring tools such as DataDog and Sentry Financial services or fintech industry experience Experience within enterprise SaaS or customer-facing product organizations ...

Lead Software Engineer

Hiring Organisation
Mondo
Location
Fort Worth, Texas, United States
Employment Type
Permanent
Salary
USD Annual
engineers and guide collaborative technical problem-solving efforts Participate in code reviews, CI/CD improvements, and engineering best practices discussions Support production operations, observability, troubleshooting, and continuous improvement efforts Improve engineering quality, delivery execution, and overall team effectiveness Balance approximately 80% hands-on engineering work with 20% leadership responsibilities … authorization required Preferred Qualifications: Angular experience Ruby on Rails exposure PostgreSQL experience Docker experience CI/CD pipeline experience, preferably with CircleCI Experience with observability and monitoring tools such as DataDog and Sentry Financial services or fintech industry experience Experience within enterprise SaaS or customer-facing product organizations ...

GCP Lead Engineer - Google Cloud Platform

Location
Greater London, England, United Kingdom
engineering assurance Improve Terraform/Infrastructure as Code standards and reusable patterns Develop and enhance CI/CD and deployment automation Improve monitoring, observability, alerting and operational readiness Support cloud‐based data integration and analytical workloads Ensure solutions meet demanding security, governance and audit requirements Work with technical and business … Terraform/Infrastructure as Code Cloud and Platform Engineering DevOps and CI/CD Production‐grade cloud environments Cloud security and governance Monitoring and observability Data platforms and data‐intensive applications BigQuery Python and SQL APIs and data integration Authentication and access management, including OAuth/OIDC Most importantly ...

SRE | Permanent | London, Hybrid, AWS

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
scalability. Key responsibilities Partner with engineering teams to define, measure, and manage SLOs/SLIs, using error budgets to guide delivery decisions. Enhance observability across services (metrics, logs, traces) to detect and resolve issues proactively. Lead cost optimisation: monitor spend, right-size workloads, tune autoscaling, and improve infrastructure efficiency. Improve … Kubernetes operational experience (on-prem and AWS EKS).Hands-on experience defining and operating SLOs/SLIs, alerting, and incident workflows. Deep understanding of observability and telemetry (monitoring, logging, tracing).Infrastructure as Code with Terraform; experience with GitOps workflows and CI/CD.Scripting proficiency in Python, Bash, or Go. Proven ...

Forward Deployed AI Engineer

Location
Greater London, England, United Kingdom
enabled systems. You’ll bring deep expertise across modern full-stack technologies (.NET, Azure, SQL, React/Angular), along with experience in distributed systems, observability, and AI tooling such as LLMs, retrieval pipelines, agentic workflows, and platforms such as Anthropic Claude. Experience designing and deploying AI agents, leveraging Model Context … orchestration, evaluation loops, and human-in-the-loop controls. Enterprise integration: Integrate AI solutions with enterprise systems, APIs, data platforms, document repositories, workflow tools, observability platforms, and identity and access management services. Production engineering: Ensure AI solutions meet enterprise standards for reliability, scalability, latency, maintainability, cost control, logging, monitoring ...

DevOps Engineer

Location
Greater London, England, United Kingdom
build and operate the infrastructure that every engineering team at Invisible depends on — Kubernetes, CI/CD, identity, secrets management, networking, and observability — along with the internal tooling that allows engineers and AI coding agents to ship safely and quickly. The Platform team is small relative to the surface area … entire engineering organisation. We are looking for engineers who don’t just apply best practices in IaC, Kubernetes, CI/CD and observability, but understand the problems those practices were designed to solve. Most were built around the pace and failure modes of human engineers. Those assumptions are changing ...

Platform Engineer London, UK · Full time · Hybrid

Location
Greater London, England, United Kingdom
engineers. The role combines infrastructure and software development, with an approximate 65/35 split. You’ll work on cloud infrastructure, deployment automation, observability, CI/CD, and internal tooling. The goal is to make development and releases more reliable, efficient, and straightforward. What you’ll do Improve developer experience … Find the causes of slow builds, failed pipelines, and flaky tests. Develop internal tools, improve test infrastructure, and occasionally work on product features. Improve observability across our infrastructure and delivery workflows using OpenTelemetry, ELK. Leverage AI to improve engineering workflows by building cloud-based agent tools and helping teams adopt ...

Developer - Scala

Hiring Organisation
Hackajob Ltd
Location
Newcastle Upon Tyne, Tyne and Wear, North East, United Kingdom
Employment Type
Permanent, Work From Home
frontend engineers, QA, product owners, solution designers, and other backend developers to deliver high-quality product increments. Support production stability by investigating issues, improving observability, and continuously reducing technical debt. Required Skills and Experience: Professional backend development experience, ideally in enterprise, SaaS, or cloud-based product environments. Strong hands … problems. Nice to Have: Experience with Squeryl, Doobie, or similar Scala data access libraries. Familiarity with Grafana or Kibana dashboards, alerting, logging, and production observability practices. Experience with large distributed systems, horizontal scaling, resilient service design, or high-throughput Play/Pekko applications. Previous experience in Payroll ...

Senior Backend Engineer (London, Barcelona, Madrid)

Location
Greater London, England, United Kingdom
product and stakeholders to stay aligned on direction. Contribute to design reviews with a clear view on the trade-offs. Keep your services healthy - observability, feedback loops, incident response. Raise the bar around you through code review, pairing and knowledge sharing. ABOUT YOU: 4+ years as a software engineer, with … equivalent is fine. A clear communicator and a pragmatic problem-solver. NICE TO HAVE EXPERIENCE: Data-intensive applications at scale. Building and interpreting observability - metrics, logging, tracing. Working within or alongside DevOps, SRE or infrastructure teams. Complex distributed environments - high throughput, low latency, large datasets. Any exposure to commodities, energy ...

Agentic Platform Engineer · Manchester, UK ·

Location
Manchester, England, United Kingdom
Protocol (MCP). Build production-grade agent services using Python, cloud-native architectures, event-driven design, automation and Infrastructure as Code. Implement robust evaluation, observability and continuous improvement capabilities, including testing, tracing, telemetry and performance optimisation. Embed security, governance and responsible AI principles through least-privilege access, policy enforcement, auditability … workflow state, retrieval-augmented generation and human-in-the-loop patterns. Experience building secure integrations with enterprise APIs, repositories, cloud services, CI/CD, observability or ITSM platforms. Experience creating evaluation frameworks for agent quality, task completion, safety, reliability, latency and cost. Strong understanding of agent security, including workload identity ...

Software Engineering Manager

Location
Manchester, England, United Kingdom
appropriate technical documentation. Champion strong engineering practices, including collaborative programming, testing approaches such as TDD, CI/CD and production ownership. Ensure strong observability and operational health, with useful code‐quality and Production metrics, appropriate KPIs, well‐calibrated alarms and clear ownership of actions following incidents and PIRs. Build relationships … Data/Data Science, CRM, Security, Legal, Privacy and other engineering teams. Experience establishing strong operational ownership of production software, including CI/CD, observability, support practices and continuous improvement. Commitment to inclusive leadership, frequent feedback and creating an environment where engineers can grow, challenge ideas and take meaningful ownership. ...

Software Engineering Manager

Location
Greater London, England, United Kingdom
appropriate technical documentation. Champion strong engineering practices, including collaborative programming, testing approaches such as TDD, CI/CD and production ownership. Ensure strong observability and operational health, with useful code‐quality and Production metrics, appropriate KPIs, well‐calibrated alarms and clear ownership of actions following incidents and PIRs. Build relationships … Data/Data Science, CRM, Security, Legal, Privacy and other engineering teams. Experience establishing strong operational ownership of production software, including CI/CD, observability, support practices and continuous improvement. Commitment to inclusive leadership, frequent feedback and creating an environment where engineers can grow, challenge ideas and take meaningful ownership. ...

Senior Engineer

Location
Greater London, England, United Kingdom
turn ideas into working product quickly Contribute to product decisions and pragmatic technical trade-offs Diagnose and fix bugs quickly Improve testing, monitoring, and observability Maintain data integrity, system stability, and security Work closely with technical leadership Collaborate with our senior technical advisor on architecture and technical direction Implement technical … processes Design and maintain data-compliant systems with privacy, security, and regulatory standards embedded by default (e.g. UK GDPR, NHS DSPT, MHRA guidance) Implement observability and metrics to understand product performance, user behaviour, and system health Ensure our systems and features are audit-ready and aligned with healthcare regulatory requirements ...

Senior Software Engineer

Location
City Of London, England, United Kingdom
Make the architectural calls on your domain, write them down, and defend them in front of the tribe. Build in security, availability, reliability and observability from the start, and instrument your services so the answer to what's happening is already in a dashboard. Use AI tooling well. … awkward parts that come with them: idempotency, ordering, retries, poison messages, exactly-once as a promise nobody can keep. Security, availability, reliability and observability built in from the start, not bolted on at the end. You think about who can reach what, you know how your service behaves when ...

Software Engineering Manager

Location
Bristol, England, United Kingdom
risks early and transparently. Establish engineering guardrails across scope, quality, and non-functional requirements, enabling teams to design optimal solutions within them. Champion observability and operational excellence, ensuring system health, SLOs, and alerting are visible and actively managed. Partner with Tech Leads and Architects on system design and evolution, bringing … architecture, including API design and integration, performance optimisation, security, and microservice or event‐driven patterns. Experience with CI/CD, modern development workflows, and observability practices. Proven track record of leading high‐performing teams in fast‐paced, complex or regulated environments. Passion for mentoring and developing engineers through coaching, feedback ...

Senior Software Engineer - Data

Hiring Organisation
ApartmentIQ
Location
Portland, Oregon, United States
Employment Type
Permanent
Salary
USD Annual
Teams: Work closely with Product and Application Engineering to ensure new datasets unlock measurable features and actionable customer insights. Ensure Operational Excellence: Build strong observability into our systems, including daily monitoring, anomaly detection, and automated alerts to maintain ingest health. Manage Roadmap Execution: Lead the delivery of data milestones … style data workflows using tools like dbt, Airflow, or similar equivalents. Cloud Native: Deep expertise in AWS infrastructure (RDS, ECS, S3, Lambda, OpenSearch). Observability Minded: Familiar with modern stacks for performance tracking, error monitoring, and distributed tracing. Strategic Communicator: Strong ability to translate complex technical details into clear, actionable ...

Senior Software Engineer - Data

Hiring Organisation
ApartmentIQ
Location
Denver, Colorado, United States
Employment Type
Permanent
Salary
USD Annual
Teams: Work closely with Product and Application Engineering to ensure new datasets unlock measurable features and actionable customer insights. Ensure Operational Excellence: Build strong observability into our systems, including daily monitoring, anomaly detection, and automated alerts to maintain ingest health. Manage Roadmap Execution: Lead the delivery of data milestones … style data workflows using tools like dbt, Airflow, or similar equivalents. Cloud Native: Deep expertise in AWS infrastructure (RDS, ECS, S3, Lambda, OpenSearch). Observability Minded: Familiar with modern stacks for performance tracking, error monitoring, and distributed tracing. Strategic Communicator: Strong ability to translate complex technical details into clear, actionable ...

Senior Software Engineer - Data

Hiring Organisation
ApartmentIQ
Location
Minneapolis, Minnesota, United States
Employment Type
Permanent
Salary
USD Annual
Teams: Work closely with Product and Application Engineering to ensure new datasets unlock measurable features and actionable customer insights. Ensure Operational Excellence: Build strong observability into our systems, including daily monitoring, anomaly detection, and automated alerts to maintain ingest health. Manage Roadmap Execution: Lead the delivery of data milestones … style data workflows using tools like dbt, Airflow, or similar equivalents. Cloud Native: Deep expertise in AWS infrastructure (RDS, ECS, S3, Lambda, OpenSearch). Observability Minded: Familiar with modern stacks for performance tracking, error monitoring, and distributed tracing. Strategic Communicator: Strong ability to translate complex technical details into clear, actionable ...