1,251 to 1,275 of 1,791 Permanent Observability Jobs

Software Engineer III - Python

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
infrastructure-as-code using Terraform within established team patterns across modules, environments, and state management Improve operability of services by adding and using observability tooling including logs, metrics, traces, dashboards, and alerts, and participate in incident response and root-cause analysis Leverage enterprise-authorized AI coding assist tools within … implementing application logic and APIs on top of relational data Experience building APIs and microservices using REST or gRPC, including contracts, security basics, and observability Practical experience delivering LLM-based features as part of software systems, with familiarity with agentic patterns Working knowledge of delivery and operations including CI/ ...

Lead Java Developer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
adoption and ensure successful rollout of new capabilities.* Lead root cause analysis on production issues, drive long‐term stability improvements, and strengthen monitoring and observability across the platform.**Recommended Experience:*** Strong experience in Core Java, J2EE, Spring Framework* Exposure to Python scripting and data analysis* Experience in fast moving Capital … such as Kafka, JMS, gRPC etc* Proficient in latency measurement and performance optimization of Java based platforms with focus on JVM tuning* Experience with observability stacks like ELK, Prometheus, Grafana, Kiali, Jaeger etc.* Sound knowledge for persistence technologies such as relational databases, NoSQL databases, off heap storages and distributed caches ...

Staff or Principal Data Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
create unnecessary complexity, risk, duplicated capability or long‐term support burden. Raise the quality bar for data products through clear ownership, robust testing, reconciliation, observability, lineage, documentation, performance and supportability. Collaborate with cross‐functional teams to address security, GDPR, PII handling, role‐based access, auditability and data governance are designed … services across batch, streaming and event‐driven patterns. Deep understanding of engineering practice: clean design, testing strategy, CI/CD, infrastructure as code, observability, performance, security, incident response and DevSecOps. Experience with cloud data services and modern data stacks. Relevant technologies may include Snowflake, Azure/AWS/GCP data ...

Engineering Manager, Corp Tech

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Oversight of development teams and ensure timely delivery of solutions Manage incident resolution and performance improvements Establish API‐first development, data quality, and platform observability Optimize and orchestrate workflows using BPM tools, decision engines, and custom microservices Pilot and scale emerging technologies such as Generative AI where appropriate. Implementing … platform providers, own technical due diligence, integration planning, and vendor alignment with enterprise architecture Ensure third‐party platforms meet enterprise standards around security, integration, observability, and support Present strategy, solution design, and progress updates to steering committees and senior technology forums Ensure platforms are audit‐ready and meet data residency ...

Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
across the Platform team. Eliminate Toil: Automate manual processes and remove friction from the developer day‐to‐day. Enable Ownership: Coach squads on SLOs, observability, and AI tooling, ensuring they are capable of leading their production behaviour without relying on Platform. Strategic Impact: Proactively identify, scope, and deliver high‐leverage … balance performance, cost, and reliability. Our stack includes ECS Fargate, RDS, ElastiCache, DynamoDB, Lambda, and EventBridge. Have good judgement across CI/CD and observability, comfortable with safe and reliable deployment practices (we use GitHub Actions and Terraform Cloud), and comfortable using metrics, logs, and traces to understand how applications ...

Head of Technology Operations

Hiring Organisation
Jobleads-UK
Location
Halifax, England, United Kingdom
adoption of infrastructure as code (IaC), CI/CD pipelines, and automated testing within platform operations. Champion site reliability engineering (SRE) practices, embedding monitoring, observability, and incident response, and continuously improving performance metrics including application load times, throughput, and error rates. Partner with the Head of Development and the Director … technologies. Deep expertise in Kubernetes, cloud networking, CI/CD pipelines, infrastructure as code, and platform security. Experience leading site reliability engineering (SRE), monitoring, observability, and performance optimisation (load times, application speed, latency management). Experience providing senior‐level escalation support for complex infrastructure, firewall, networking, and systems issues. ITIL ...

Backend Engineer

Hiring Organisation
Page Group
Location
Whippany, New Jersey, United States
Employment Type
Permanent
Salary
USD Annual
The Backend Engineer will focus on designing, developing, and maintaining strong server-side solutions to support the company's financial services operations. This role involves collaborating closely with cross-functional teams to ensure seamless integration ...

Senior Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
Job ContextEngineering • Leeds • Full-Time • On-SiteAre you a people-first engineering leader who thrives on turning complex technical challenges into elegant, user-centric software As our Senior Engineering Manager, you will sit at the ...

AI & ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About the role The AI & ML Engineering team accelerates the adoption of AI across the business, championing innovation while ensuring our machine learning solutions are robust, scalable, and cost-efficient. We enable teams to solve ...

AI & ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About Charlotte Tilbury Beauty Founded by British makeup artist and beauty entrepreneur Charlotte Tilbury MBE in 2013, Charlotte Tilbury Beauty has revolutionised the face of the global beauty industry by de-coding makeup applications for ...

Devops Engineer

Hiring Organisation
Jackson Hogg Ltd
Location
Newcastle upon Tyne, Tyne & Wear, United Kingdom
Employment Type
Permanent
Salary
£50000 - £60000/annum
We're looking for a DevOps Engineer to join a growing technology team responsible for building, supporting and evolving critical cloud platforms. This role sits at the heart of a modern AWS environment, helping to ...

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Peak jackpot proactive monitoring Own end‐to‐end incident lifecycle: Detection -> triage -> resolution -> post‐incident review Ensure blameless post‐mortems with clear remediation ownership Observability & service insight Define and evolve observability strategy using: Splunk (log analytics) CloudWatch (AWS telemetry) Grafana (metrics visualisation) Quantum Metric (user behaviour insight) Standardise: Alerting quality … safety (CI/CD, progressive delivery patterns) Key contributor to transition strategy for ECS → EKS (Kubernetes adoption) Reduce operational toil through tooling, self‐healing, observability and platform improvements Empower Level‐1 operational teams with safe, controlled access to the tools they need to operate autonomously Capacity & performance engineering Own capacity ...

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Peak jackpot proactive monitoring Own end‐to‐end incident lifecycle: Detection triage resolution post‐incident review Ensure blameless post‐mortems with clear remediation ownership Observability & service insight Define and evolve observability strategy using: Splunk (log analytics) CloudWatch (AWS telemetry) Grafana (metrics visualisation) Quantum Metric (user behaviour insight) Standardise: Alerting quality … safety (CI/CD, progressive delivery patterns) Key contributor to transition strategy for ECS EKS (Kubernetes adoption) Reduce operational toil through tooling, self‐healing, observability and platform improvements Empower Level‐1 operational teams with safe, controlled access to the tools they need to operate autonomously Capacity & performance engineering Own capacity ...

Observability Architect - 12 Month FTC

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You’re comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with … grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate ...

Observability Architect - 12 Month FTC

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You’re comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with … grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Operate Service Manager

Hiring Organisation
17918
Location
Bristol, Gloucestershire, United Kingdom
Summary Apto Operate is our managed service for telemetry, security, and observability platform management and it is growing. The Operate Service Manager owns the service itself: the quality and consistency of delivery across every customer account, the performance and development of our engineering team, and the continuous improvement ...

Federal Deployment Engineer - Data Security Platform

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Virtru is seeking a Forward Deployed Engineer to enhance observability, performance, and reliability across our platform infrastructure. You will drive secure data protection deployments for government clients and lead integration projects with federal agencies. The role requires active security clearance (DV preferred) and 5+ years in deployment or solutions engineering ...

AI/ML Enterprise Architect - Governance & Security

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will collaborate with Cyber, Risk, Legal, Privacy and Compliance teams, translate complex AI/ML designs into clear docs, and ensure security, resilience and observability are baked into every solution. #J-18808-Ljbffr ...

Global Head of SRE & Reliability – Hybrid Role

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
office and collaboration across Engineering, Infrastructure Operations and Security to boost reliability and performance of critical platforms. You will define reliability strategy, drive automation, observability, and incident maturity, and scale the organization while embedding reliability into the #J-18808-Ljbffr ...

Site Reliability Engineer - Scale, Observe, Automate

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
seeking a Site Reliability Engineer to support the reliability and performance of our digital services. You will work with production systems, automation, and observability to keep services stable, scalable and well-instrumented across peak events. You will collaborate with senior SREs and engineering teams to drive automation, incident response ...

Cloud SRE Delivery Lead — Global Platform Transformation

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
will drive governance, manage investments in reliability, and ensure end-to-end SLIs/SLOs with robust reporting and risk visibility, while advancing observability and automation to reduce toil and improve throughput. #J-18808-Ljbffr ...