1,526 to 1,550 of 1,866 Observability Jobs

Backend Engineer - Platform - Stacks | UK | Remote

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand ...

Team Lead, Integrations

Hiring Organisation
Genesis10
Location
Irving, Texas, United States
Employment Type
Permanent
Salary
USD Annual
Code & DevOps: IaC with Terraform; able to review and approve infrastructure changes Azure DevOps pipeline design, branch policies, and mandatory-review gate configuration Observability & Monitoring: Application Monitoring using Application Insights Log Management and Analytics Platform Monitoring using Azure Monitor Query & Analysis using KQL (Kusto Query Language) Alerting, dashboards, and validating … solution observability Application Development: Minimum 8 years of solid .NET coding experience (C#, ASP.NET Core, background/worker services) Expert knowledge of common design and patterns including .NET patterns, libraries and Azure services Documentation & Communication: Strong technical writing skills Visual Modeling with Lucid chart/Azure architecture diagramming Comfortable delivering ...

Infrastructure Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
tenant environments for enterprise customers. Developer Velocity: Architect CI/CD pipelines (GitHub Actions, Docker) that allow our team to ship safely and quickly. Observability & Reliability: Instrument the stack with OpenTelemetry and Datadog to ensure we detect issues before our users do. AI Performance Tuning: Tune autoscaling and network routing … Haves: Experience with ECS, container orchestration, and distributed task queues (Celery/SQS). Strong Python skills. Familiarity with the Datadog/Grafana observability stack. A background in offensive security, CTFs, or AI/ML infrastructure. What We Offer Competitive Salary + Significant Equity: We want you to have true ...

Software Engineer / AI Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
within a defined problem, building and testing tool use, retrieval pipelines and agent workflows, integrating AI capabilities into enterprise systems, and contributing to evaluation, observability and guardrails. You will hold a high bar on code quality, flag risks and blockers early, and work alongside host‐function stakeholders to make sure … agentic AI solutions to production standard within a defined technical approach. Implement and test tool use, retrieval pipelines, and agent workflows. Contribute to evaluation, observability and guardrails for agentic systems. Integrate AI capabilities into existing enterprise workflows and systems. Maintain high code quality and documentation so patterns can be reused. ...

Lead Ai Engineer (AI/ML/R&D)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
powered applications. Extensive experience building and operating cloud-native applications on AWS, with strong knowledge of CI/CD, infrastructure as code, observability and modern engineering practices. Experience delivering high-performance, real-time systems that operate reliably at scale with demanding latency requirements. Demonstrated experience leading the successful transition … powered applications. Extensive experience building and operating cloud-native applications on AWS, with strong knowledge of CI/CD, infrastructure as code, observability and modern engineering practices. Experience delivering high-performance, real-time systems that operate reliably at scale with demanding latency requirements. Demonstrated experience leading the successful transition ...

Engineering Manager (Remote - UK)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
evolving AWS-native and cloud-based platforms, alongside legacy systems Set and uphold strong engineering standards across code quality, testing, CI/CD, observability and documentation Stay close to technical decisions through design reviews, architecture discussions and hands‐on coaching Balance new feature delivery with technical debt, reliability, security … least one modern programming language Strong grasp of system design and software engineering fundamentals Experience with Infrastructure as Code, CI/CD, observability and secure production systems Able to communicate technical ideas clearly to both technical and non‐technical stakeholders What Altus Group Offers Rewarding performance: competitive compensation, incentive ...

Senior Machine Learning Engineer - News

Hiring Organisation
Disney Entertainment and ESPN Product & Technology
Location
New York, United States
Employment Type
Permanent
Salary
USD Annual
sensitive outcomes, while proactively identifying, communicating, and mitigating risks to ensure successful execution Champion engineering best practices across code quality, testing, CI/CD, observability, and incident response Mentor and coach engineers, fostering a culture of ownership, collaboration, and continuous improvement Contribute to technical documentation and promote knowledge sharing across … Databricks, Kinesis, Kafka Proven leadership, coaching, and mentoring skills, with the ability to inspire and empower a team towards achieving business goals Experience with observability tools for metrics, logging, and monitoring such as Datadog Experience working in Agile/Scrum development environments Excellent communication skills and a commitment to collaboration ...

Senior Machine Learning Engineer - News

Hiring Organisation
Disney Entertainment and ESPN Product & Technology
Location
Glendale, California, United States
Employment Type
Permanent
Salary
USD Annual
sensitive outcomes, while proactively identifying, communicating, and mitigating risks to ensure successful execution Champion engineering best practices across code quality, testing, CI/CD, observability, and incident response Mentor and coach engineers, fostering a culture of ownership, collaboration, and continuous improvement Contribute to technical documentation and promote knowledge sharing across … Databricks, Kinesis, Kafka Proven leadership, coaching, and mentoring skills, with the ability to inspire and empower a team towards achieving business goals Experience with observability tools for metrics, logging, and monitoring such as Datadog Experience working in Agile/Scrum development environments Excellent communication skills and a commitment to collaboration ...

Software Engineer / AI Engineer

Hiring Organisation
Elsevier
Location
Greater London, United Kingdom
Employment Type
Full Time
within a defined problem, building and testing tool use, retrieval pipelines and agent workflows, integrating AI capabilities into enterprise systems, and contributing to evaluation, observability and guardrails. You will hold a high bar on code quality, flag risks and blockers early, and work alongside host-function stakeholders to make sure … agentic AI solutions to production standard within a defined technical approach. Implement and test tool use, retrieval pipelines, and agent workflows. Contribute to evaluation, observability and guardrails for agentic systems. Integrate AI capabilities into existing enterprise workflows and systems. Maintain high code quality and documentation so patterns can be reused. ...

Principal Software Engineer SC&L

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
source technology Support recruitment, onboarding and internal and external brand outreach activities Tech Stack Java, Micronaut, PL/SQL ReactJS, Next.js Azure Cloud, Dynatrace (observability) Mule, Kafka, MQ Blue Yonder Dispatcher Essential Experience Significant track record of strategic and innovative thinking, as well as execution and implementation Specialist in clean … customers to a desired outcome, without prescribing it Authoritative skills at cloud computing (network, security, serverless, Kubernetes etc) and automation Experience with implementation of Observability and Reliability using market technologies (e.g.: Dynatrace) Advocate and experience of Continuous Integration and Continuous Delivery Advanced experience of DevOps: you build ...

Lead AI Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
solutions Enhance CI/CD pipelines for AI systems, introducing automation, traceability and controlled release processes to ensure safe and repeatable deployments. Define monitoring, observability and operational strategies, improving visibility across quality, performance, safety, cost and reliability to support effective issue diagnosis and resolution Set standards for prompt engineering, evaluation … using LLM APIs/Bedrock, including RAG architectures, vector databases and orchestration frameworks DevOps skills including CI/CD pipelines, containerisation and monitoring/observability, with experience defining release and operating practices for AI services Experience Leading Engineering teams: providing technical guidance, aligning on standards/patterns, and adapting plans ...

Software Engineer

Hiring Organisation
RWS
Location
Sheffield, United Kingdom
Employment Type
Full Time
security, reliability, and operation of the services you build, with a DevSecOps approach throughout Improving engineering practices including CI/CD pipelines, automated testing, observability, security scanning, and deployment workflows Collaborating closely with product managers, designers, domain experts, and other engineers to deliver meaningful outcomes Mentoring and supporting engineers across … DevSecOps mindset with experience owning the security and operation of services you build Understanding of modern delivery practices: CI/CD, automated testing, observability, and production ownership Ability to work across the full development lifecycle, from early design through to deployment and operations Clear communication skills and a collaborative working ...

Senior Software Engineer - Customer Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams, including … build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such ...

SRE Technical Lead

Hiring Organisation
Capgemini
Location
Surrey, United Kingdom
Employment Type
Full Time
point for major incidents and high risk releases, protecting service stability and ensuring blameless post incident reviews lead to measurable improvement. • Define and govern observability and capacity practices so reliability risks are visible, actionable, and proactively managed. • Ensure SRE practices align with service governance, security, and compliance requirements, and contribute … including: • Strong expertise in Kubernetes and OpenShift. • Experience with multi cloud and hybrid architectures, including service mesh (e.g. Istio). • Hands on experience with observability platforms such as Prometheus, Grafana, Loki, Tempo, and OpenTelemetry. • Strong Infrastructure as Code and GitOps experience (Helm, Kustomize, ArgoCD, Tekton). • Experience with CI/ ...

Senior Software Engineer - Customer Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams, including … build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such ...

Senior Engineering and Operations Manager

Hiring Organisation
Jobleads-UK
Location
East Midlands, England, United Kingdom
proof‐of‐concepts, and helping teams resolve complex engineering issues. Champion scalable, maintainable and secure engineering practices, including appropriate use of AI, testing, monitoring, observability and DevOps ways of working. Maintain visibility of technical debt, platform health, system resilience and engineering investment priorities. Provide hands‐on technical support and challenge … Microsoft technology stacks, .NET, SQL Server, Azure, Microsoft 365 or similar enterprise platforms. Experience with cloud platforms, CI/CD pipelines, automation, monitoring, observability and DevOps‐aligned operating models. Experience improving engineering maturity, delivery governance or operational effectiveness. Experience contributing to platform modernisation, technology transformation or business‐critical system improvement ...

Software Engineer - Senior

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
regressions are caught before production Take end-to-end ownership from discovery and design through build, rollout, and operational excellence Instrument systems with the observability, cost tracking, and audit trails needed to know when they degrade Apply FinOps and cost-optimization practices to AI workloads, tracking and managing token, inference … example LangChain, LlamaIndex, or the Model Context Protocol) Experience with MLOps/LLMOps tooling such as experiment tracking, model versioning, monitoring, evaluation/observability platforms, or CI/CD for ML and LLM systems Experience implementing responsible-AI or AI-governance controls such as guardrails, human-in-the-loop oversight ...

Software Engineer - Senior

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
regressions are caught before production.* Take end-to-end ownership from discovery and design through build, rollout, and operational excellence. Instrument systems with the observability, cost tracking, and audit trails needed to know when they degrade.* Apply FinOps and cost-optimization practices to AI workloads, tracking and managing token, inference … example LangChain, LlamaIndex, or the Model Context Protocol).* Experience with MLOps/LLMOps tooling such as experiment tracking, model versioning, monitoring, evaluation/observability platforms, or CI/CD for ML and LLM systems.* Experience implementing responsible-AI or AI-governance controls such as guardrails, human-in-the-loop ...

Senior C++ / Java Developer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
## Senior C++/Java DeveloperApplylocations: London, United Kingdomtime type: Full timeposted on: Posted 3 Days Agojob requisition id: R0119302**ABOUT US:**LSEG (London Stock Exchange Group) is more than a diversified global financial markets ...

Data Architect

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
We believe in the power of ingenuity to build a positive human future.We challenge where it matters and own the outcome.As strategies, technologies, and innovation collide, we create opportunity from complexity. Our teams of interdisciplinary ...

Hybrid SRE Engineer — Observability & Cloud (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
help transform current workloads toward an SRE model while working in a hybrid setup, visiting the London office twice weekly. The role focuses on observability, high availability and incident management, with collaboration across Product Engineering and Infrastructure teams. Strong AWS, Terraform, Python and Kubernetes skills are valued. #J-18808-Ljbffr ...

Senior SRE Technical Lead — Reliability & Observability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
seeking a Technical Lead SRE in Greater London. In this role, you will enhance the reliability engineering capabilities, collaborating with various teams to establish observability standards and ensure operational excellence. The ideal candidate will have over 10 years of experience in SRE or related fields, strong AWS and Kubernetes skills ...

Senior Observability Engineer: Honeycomb & OpenTelemetry

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Group is seeking a Senior Observability Engineer in London to own and evolve the observability platform, primarily centered around Honeycomb and OpenTelemetry. You will instrument services in Java, Python, and C++, collaborate with SRE and development teams, and drive reliability through SLOs and incident response. This role requires … years in observability or related field, with UK working hours considered. #J-18808-Ljbffr ...

Senior Data Architect

Hiring Organisation
Hilti
Location
Allen, Texas, United States
Employment Type
Permanent
Salary
USD Annual
streaming patterns (e.g.,MSK/Kinesis,DMS,Glue). Partner with Solution/Platform Architects to ensuremicroservicesandevent drivendesigns are data efficient and observability ready. Data Governance, Quality & Observability Implementdata governance(policies, stewardship, data classification),cataloging(Glue Data Catalog),lineage(Open Lineage compatible), andquality(rules, thresholds, SLAs/SLOs). Embedmonitoring …/tokenization, RBAC/ABAC. AI/LLM enablement:Amazon Bedrock(Knowledge Bases, Guardrails, Agents), embeddings, chunking, retrieval,prompt design,token optimization, evaluation loops. Observability & FinOps: Cloud logs/metrics, lineage/quality SLAs,cost controls(storage/compute), workload rightsizing. DevOps/MLOps/IaC:Git,CI/CDfor ...

Observability Engineer (Dynatrace) — Telemetry & Performance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Computacenter is seeking a Monitoring & Observability Engineer (Dynatrace) to design, implement and manage observability across customer IT estates in the UK. You will collect telemetry, diagnose issues and drive proactive improvements across teams. The role requires strong Dynatrace/Grafana/Splunk experience, scripting skills, cloud familiarity (Azure/ ...