26 to 34 of 34 Observability Jobs in Glasgow

Context Plane Python Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
data sources and services across the firm, including enterprise AI and large language model gateways Own quality across your components: automated testing, code reviews, observability, and resilient, secure service design Partner with Corporate Technology AI, product, and data science colleagues to translate concrete use cases into working, measurable capabilities Contribute … working with cloud infrastructure (AWS) and containerized services (Docker/ECS) Ability to own technical components end-to-end — from design through deployment and observability Strong collaboration skills with the ability to work across engineering, product, and data science disciplines Hands-on experience using enterprise-authorized AI-assisted software development ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You’ll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. If you enjoy solving hard production problems and making platforms measurably better, you’ll find meaningful impact and growth here. … large language models on cloud-based container orchestration platforms and on-premises GPU clusters using reproducible infrastructure as code and continuous delivery pipelines Implement observability across logs, metrics, and traces with dashboards and actionable alerting for large language model and GPU workloads Tune GPU and accelerator capacity, autoscaling, and cost ...

Lead Architect- Data & Database Systems

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
optimize queries for performance and correctness. Monitor and tune database performance, including query tuning, indexing, partitioning, and caching. Build and maintain monitoring, alerting, and observability for database systems. Support production incidents and perform root cause analysis; participate in on-call rotations. Assist application teams with data access patterns, migrations … MySQL/MariaDB, Microsoft SQL Server, or Oracle. Performance tuning experience, including indexing strategies, query profiling, and execution plan interpretation. Experience with monitoring and observability tooling such as Prometheus, Grafana, Datadog, Dynatrace or New Relic. Familiarity with versioned schema migrations using tools such as Flyway, Liquibase, Alembic, or sqitch. Good ...

Unix Engineer

Hiring Organisation
Wolviston Management Services
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
server estates. Automation of infrastructure operations and maintenance activities. Development of orchestration workflows using Apache Airflow. Building automation tooling using Python and Ansible. Enhancing observability, monitoring and operational controls. Supporting the introduction of live-patching technologies. The role requires a strong infrastructure engineering mindset, with particular emphasis on UNIX/… orchestration workflows within Apache Airflow. Create automated pre-change validation and post-change verification processes. Improve efficiency and reduce manual intervention through automation. Platform Observability Implement monitoring and observability capabilities across automated infrastructure workflows. Develop dashboards, alerting and operational reporting. Work with tooling including: Prometheus Grafana Loki Ensure automation ...

Lead Data Architect: Cloud DB Solutions & Observability

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
native database solutions and partner with engineering teams and business stakeholders to define data strategies. You will tune complex queries, manage indexing strategies, build observability frameworks, and support production incidents across multiple business functions, shaping future data architectures. #J-18808-Ljbffr ...

Lead Site Reliability Engineer: Drive Stability & Observability

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
will guide incident response, drive service level objectives, and mentor a team of engineers across multiple domains. The role emphasizes deep expertise in observability, container orchestration, and OpenTelemetry, with a focus on performance, security, and scalability. Collaboration with cross‐functional teams is essential. #J-18808-Ljbffr ...

Product Delivery Manager - Public Cloud SRE

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
bottlenecks. Ensure operational readiness and change safety requirements are executed in delivery workflows, improving release outcomes and reducing change-related incidents. Drive delivery of observability improvements (metrics/logs/traces coverage, alert quality, dashboards) that improve signal quality and operational decision-making. Deliver measurable automation and toil reduction, increasing … design, and data analytics Deep experience in multi-cloud platforms, infrastructure services, automation, and operational tooling. SRE domain expertise: SLIs/SLOs, error budgets, observability, incident/problem management, operational readiness, and reliability analytics. Proven transformation leadership across matrixed, global organizations; strong executive communication and stakeholder influence. ABOUT US J.P. ...

AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
Responsibilities Design, develop and improve software, utilizing various engineering methodologies. Develop and deliver high-quality software solutions using industry-aligned programming languages, frameworks, and tools. Ensure that code is scalable, maintainable, and optimized for performance. ...

Product Delivery Manager - Public Cloud SRE

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
bottlenecks. Ensure operational readiness and change safety requirements are executed in delivery workflows, improving release outcomes and reducing change-related incidents. Drive delivery of observability improvements (metrics/logs/traces coverage, alert quality, dashboards) that improve signal quality and operational decision-making. Deliver measurable automation and toil reduction, increasing … design, and data analytics Deep experience in multi-cloud platforms, infrastructure services, automation, and operational tooling. SRE domain expertise: SLIs/SLOs, error budgets, observability, incident/problem management, operational readiness, and reliability analytics. Proven transformation leadership across matrixed, global organizations; strong executive communication and stakeholder influence. #J-18808-Ljbffr ...