1,651 to 1,675 of 2,934 Remote/Hybrid Observability Jobs

Senior Staff Software Engineer, Motion Planning

Hiring Organisation
Agility Robotics
Location
Kansas City, Missouri, United States
Employment Type
Permanent
Salary
USD Annual
grasping, and loco-manipulation Design, implement, test, and deploy motion planning and trajectory optimization algorithms for humanoid robots Architect motion planning systems for modularity, observability, and clean integration with perception, state estimation, and control Develop algorithms robust to environmental uncertainty and imperfect state estimation Lead cross-team architectural decisions ...

Senior Staff Software Engineer, Motion Planning

Hiring Organisation
Agility Robotics
Location
Providence, Rhode Island, United States
Employment Type
Permanent
Salary
USD Annual
grasping, and loco-manipulation Design, implement, test, and deploy motion planning and trajectory optimization algorithms for humanoid robots Architect motion planning systems for modularity, observability, and clean integration with perception, state estimation, and control Develop algorithms robust to environmental uncertainty and imperfect state estimation Lead cross-team architectural decisions ...

Senior Staff Software Engineer, Motion Planning

Hiring Organisation
Agility Robotics
Location
Salt Lake City, Utah, United States
Employment Type
Permanent
Salary
USD Annual
grasping, and loco-manipulation Design, implement, test, and deploy motion planning and trajectory optimization algorithms for humanoid robots Architect motion planning systems for modularity, observability, and clean integration with perception, state estimation, and control Develop algorithms robust to environmental uncertainty and imperfect state estimation Lead cross-team architectural decisions ...

Senior Staff Software Engineer, Motion Planning

Hiring Organisation
Agility Robotics
Location
Rapid City, South Dakota, United States
Employment Type
Permanent
Salary
USD Annual
grasping, and loco-manipulation Design, implement, test, and deploy motion planning and trajectory optimization algorithms for humanoid robots Architect motion planning systems for modularity, observability, and clean integration with perception, state estimation, and control Develop algorithms robust to environmental uncertainty and imperfect state estimation Lead cross-team architectural decisions ...

Senior Staff Software Engineer, Motion Planning

Hiring Organisation
Agility Robotics
Location
Sioux Falls, South Dakota, United States
Employment Type
Permanent
Salary
USD Annual
grasping, and loco-manipulation Design, implement, test, and deploy motion planning and trajectory optimization algorithms for humanoid robots Architect motion planning systems for modularity, observability, and clean integration with perception, state estimation, and control Develop algorithms robust to environmental uncertainty and imperfect state estimation Lead cross-team architectural decisions ...

Product Owner – Knowledge Repository & Management Platform

Location
Greater London, England, United Kingdom
consumable Develop, maintain, and prioritize the product backlog across functional and non-functional requirements — covering the Integration Fabric, Knowledge Graph, Search, Permissions, Governance, Observability, APIs, and MCP layers Author clear, well-groomed user stories with defined acceptance criteria; lead sprint planning, estimation, retrospectives, and scrum-of-scrums ceremonies Work closely ...

Senior / Staff Software Engineer, Mapping

Hiring Organisation
Waabi
Location
Pittsburgh, Pennsylvania, United States
Employment Type
Permanent
Salary
USD Annual
long-term mapping infrastructure. - Build robust, scalable pipelines to automatically create, validate, maintain, and visualize massive amounts of data. - Design and develop comprehensive metrics, observability, and anomaly detection across various systems. - Partner directly with downstream customers (e.g. Perception, Motion Planning, Simulation) to capture complex requirements and implement intuitive APIs. - Engineer ...

Engineering Manager - Customer London, UK

Location
Greater London, England, United Kingdom
team owns. Collaborate with tech leads, architects, and other EMs to share context, resolve dependencies, and improve system design. Encourage good practices around testing, observability, performance, and resilience – ensuring your team owns their systems in production. Hiring and Talent Development Take a leading role in hiring, onboarding, and growing diverse ...

Senior Software Engineer - UI Developer Tooling (Hybrid, London)

Location
Greater London, England, United Kingdom
push hooks so engineers catch issues in seconds, not in a failed pipeline twenty minutes later. Work with UI Quality, UI Release and Observability so the platform feels like one thing through a single CLI. Pair with Product Group leads to migrate legacy build and tooling patterns onto the paved ...

Lead AI Experience Engineer (Remote, United Kingdom)

Location
United Kingdom
Experience defining AI evaluation frameworks and using interaction data, testing, and analytics to improve performance. Knowledge of RAG, semantic search, enterprise knowledge systems, grounding, observability, and regression testing. Working knowledge of APIs, cloud services, software architecture, integration patterns, and modern engineering practices. Ability to lead complex initiatives, navigate ambiguity ...

Senior DevOps & Infrastructure Engineer

Location
United Kingdom
monitoring, release management, and platform operations. Help define practical, secure, and scalable ways of using AI in DevOps processes, whilemaintainingstrong engineering standards and governance. Observability, Reliability & Continuous Improvement Improve monitoring, alerting, logging, and observability practices to help teams detect issues earlier and resolve incidents faster. Analyze platform performance, deployment quality ...

Principal Technology Architect Hybrid Cloud Platforms -Germany, UK, Netherlands

Hiring Organisation
Infosys Technologies
Location
London, UK
Employment Type
Full-time
across cloud and on‐prem: IAM, PAM, KMS, certificates, secrets management. Partner with security architects to ensure platform designs meet enterprise security requirements.5. Resilience, Observability & Performance EngineeringArchitect high‐availability and disaster recovery models across hybrid environments. Define availability zones, failover strategies, multi‐region patterns, and RTO/RPO targets. Implement … full‐stack observability: logging, metrics, tracing, synthetic testing, SLO/SLI models. Ensure platform performance aligns with business workloads, including real-time and latency-sensitive applications.6. Governance, Risk & Compliance (GRC)Establish enterprise policies for segmentation, tagging, lifecycle management, patching, and resource standards. Govern platform consistency through enterprise architecture boards ...

Principal Java Engineer

Location
Wallingford, England, United Kingdom
continuous improvement. Production systems are reliable, observable and operationally excellent Lead root cause analysis and resolution of complex production issues. Drive improvements in system observability, monitoring and operational performance. Ensure applications are designed and operated to meet reliability, availability and performance targets. Partner with Operations, DevOps and QA teams … Claude, Codex, Gitlab Duo, etc) REST APIs, OpenAPI, Microservices, Event-driven architecture (RabbitMQ) Containers, Docker, AWS, Linux CI/CD with GitLab Pipelines & Jenkins Observability: logging, metrics and monitoring MySQL, Apache Solr Front-end UI (e.g. Angular) Person Specification Strategic and systems-thinking mindset Excellent communication and stakeholder management skills ...

Senior Reliability Engineer

Hiring Organisation
Fitch Ratings
Location
London, UK
Employment Type
Full-time
someone who is curious about the evolving role of AI in infrastructure engineering, someone who actively explores how AI-assisted tooling, automation, and intelligent observability can raise the bar for reliability and developer experience. You will collaborate closely with global development and engineering teams to deliver reliable, resilient, and high … reliability, security, and efficiencyIdentify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applicationsImplement proactive monitoring and observability practices to detect and prevent issues before they impact usersDevelop and maintain automation and tooling solutions, including AI-assisted approaches to reduce toil and accelerate ...

Senior Cloud SRE - AI/ML Platform & GPU Compute

Location
Greater London, England, United Kingdom
escalation, communications, and root cause analysis. Translate post-incident learning into durable architectural or automation improvements. Continuously reduce alert noise and recurring operational burden. Observability & Operational Excellence Design and operate monitoring, logging, tracing, and alerting systems that enable rapid detection and recovery. Build dashboards that reflect real user-centric platform … Python, Go, C++) with a bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g., Datadog, Prometheus, Grafana, OpenTelemetry). Clear communication skills, including leading incidents, writing post‐mortems, and influencing teams to prioritise reliability improvements. Desirable skills ...

Senior Software Engineer (Remote Sensing)

Hiring Organisation
Umbra
Location
Arlington, Virginia, United States
Employment Type
Permanent
Salary
USD Annual
ensure system uptime, performance, and operational excellence. Develop and maintain APIs, backend services, and data workflows that support autonomous satellite operations. Continuously improve observability, testing, deployment, and operational processes. Requirements Required Qualifications Bachelor of Science in Computer Science, Software Engineering, or a related field. 5-8+ years of professional … Experience building software to automate space operations. Experience designing and implementing scheduling systems, optimization algorithms, or automated planning systems. Strong understanding of infrastructure monitoring, observability, and operational best practices. Experience designing and documenting APIs using Swagger/OpenAPI. Track record of improving team effectiveness through mentorship, documentation, or knowledge sharing. ...

Software Developer

Hiring Organisation
TaxSlayer
Location
Evans, Georgia, United States
Employment Type
Permanent
Salary
USD Annual
Vault, Container Registry, and Blob Storage using existing platform standards. Maintain CI/CD pipelines in Azure DevOps to support reliable application delivery. Use observability tooling (for example Grafana) to monitor production behavior and proactively identify issues. Write automated tests across service, integration, and UI layers using tools such … Azure DevOps. Experience deploying or supporting applications in Kubernetes-based environments. Advanced SQL Server performance tuning experience in high-volume production systems. Experience with observability and production diagnostics in complex systems. Education & Certifications Bachelor's degree in Computer Science, Software Engineering, or related field is preferred. Equivalent professional experience will ...

Software Developer

Hiring Organisation
TaxSlayer
Location
North Augusta, South Carolina, United States
Employment Type
Permanent
Salary
USD Annual
Vault, Container Registry, and Blob Storage using existing platform standards. Maintain CI/CD pipelines in Azure DevOps to support reliable application delivery. Use observability tooling (for example Grafana) to monitor production behavior and proactively identify issues. Write automated tests across service, integration, and UI layers using tools such … Azure DevOps. Experience deploying or supporting applications in Kubernetes-based environments. Advanced SQL Server performance tuning experience in high-volume production systems. Experience with observability and production diagnostics in complex systems. Education & Certifications Bachelor's degree in Computer Science, Software Engineering, or related field is preferred. Equivalent professional experience will ...

AI Engineer – LLMs, NLP & Market Intelligence

Location
Greater London, England, United Kingdom
automated research workflows Retrieval, embeddings and semantic search Narrative detection and clustering Multilingual text intelligence Large-scale data and model pipelines Model evaluation, observability and monitoring Production ML infrastructure on AWS The problems are often open-ended. You might be evaluating how reliably different models identify changes in a market … models RAG and vector databases Apache Airflow AWS S3 ECS/EKS Redshift Docker Kubernetes GitHub Actions Pulumi or Terraform SQL Model monitoring and observability Distributed processing Financial markets, economics or commodities Financial-market experience is not required. Curiosity about how markets, economics and global events interact is more important. ...

Data Engineer, Vice President

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
datasets, and extensible pipelines that support multiple Company Intelligence products and advanced analytics use cases. Ensure high standards of data quality, governance, lineage, and observability, proactively managing operational and compliance risks across enterprise-grade data products. Partner effectively with product, analytics, and business leaders, developing a deep understanding of strategic … control, testing, CI/CD). Experience building and operating data pipelines using workflow orchestration frameworks (e.g. Apache Airflow), with a focus on reliability, observability, dependency management, and operational resilience. Experience designing and operating cloud-native data platforms (AWS or Azure preferred) and enterprise data warehouses (Snowflake preferred), including performance ...

Data Engineer, Vice President

Hiring Organisation
Hackajob Ltd
Location
Slough, Berkshire, UK
Employment Type
Full-time
datasets, and extensible pipelines that support multiple Company Intelligence products and advanced analytics use cases. Ensure high standards of data quality, governance, lineage, and observability, proactively managing operational and compliance risks across enterprise-grade data products. Partner effectively with product, analytics, and business leaders, developing a deep understanding of strategic … control, testing, CI/CD). Experience building and operating data pipelines using workflow orchestration frameworks (e.g. Apache Airflow), with a focus on reliability, observability, dependency management, and operational resilience. Experience designing and operating cloud-native data platforms (AWS or Azure preferred) and enterprise data warehouses (Snowflake preferred), including performance ...

Production AI Engineer - Vice President

Location
Greater London, England, United Kingdom
engineering techniques to integrate large language models (LLMs) into operational tooling, incident response pipelines, and developer productivity platforms. Leads the development of AI-native observability solutions — leveraging intelligent agents to detect anomalies, predict failures, and automate remediation before issues impact end users. Writes clean, well-tested, and well-documented code … . Operational experience of using middleware technologies (MQ, Apache Kafka, etc.) to run services at scale is desirable. Strong experience with end-to-end observability stacks (Datadog, AppDynamics, Dynatrace, etc.) is desirable. Degree in Computer Science, Mathematics, Physics, or a related technical subject is desirable. Experience of senior stakeholder management. ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that ensures the reliability of systems critical to the operation of satellite megaconstellations and missions to deep space. This is a greenfield/brownfield … expert, helping to define and implement the strategy and building the tools that empower our engineers. You will support the roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). ...

Senior Backend Engineer — Commercial Planning (Hybrid)

Location
City of Westminster, England, United Kingdom
thousands of colleagues. You’ll work on Commercial Planning and Fashion, Home & Beauty transformation initiatives, collaborating with architecture, product managers and partners to raise observability, security and performance on cloud-native platforms. #J-18808-Ljbffr ...

Senior Backend Engineer: Chaos & Reliability (Remote)

Location
Greater London, England, United Kingdom
guide product direction, and collaborate with friendly colleagues who live our FAITH values. You’ll design and run chaos experiments, improve load testing and observability, and introduce new tooling to boost reliability. Global teams collaborate on scalable solutions. #J-18808-Ljbffr ...