5 of 5 Remote/Hybrid ONNX Jobs in the UK excluding London

Staff Applications Engineer

Location
Cambridge, England, United Kingdom
teams. Excellent communications skills both written and verbal “Nice To Have” Skills and Experience: Experience with AI/ML frameworks such as TensorFlow, PyTorch, ONNX, or inference runtimes. System bring-up and JTAG debugging expertise Good background of system performance analysis Experience with RTL simulation tools and software development tools ...

Principal Product Manager - AI Tooling

Location
Cambridge, England, United Kingdom
demonstrate: Proven product management in development tools and software including CLIs, SDKs and APIs. Knowledge of machine learning frameworks, runtimes, infrastructure such as PyTorch, ONNX, ExecuTorch, Llama.cpp, vLLM and LiteRT. Understanding of model deployment to edge, embedded or heterogeneous computing, with an understanding of trade-offs between accuracy, performance ...

AI Compiler Optimization Engineer (Hybrid CPU/XPU)

Location
City of Edinburgh, Scotland, United Kingdom
model inference performance on CPU and CPU/XPU hybrid systems, using advanced compiler techniques. You will profile frameworks such as TensorFlow, PyTorch and ONNX, optimize graph execution, and contribute to open research with practical insights and publications. #J-18808-Ljbffr ...

Staff ML Engineer - Developer Tools

Location
Cambridge, England, United Kingdom
model optimisation techniques such as quantisation, graph optimisation, operator fusionandprecision reduction Experience across multiple ML frameworks, model formats and inference runtimes, such as PyTorch, ONNX/ONNX Runtime, ExecuTorch, TensorFlow/LiteRT and OpenVINO. Experience analysing, profiling and debugging ML workloads, including model compatibility, performance and the trade-offs between ...

Senior Machine Learning Engineer

Hiring Organisation
Hackajob Ltd
Location
Slough, England, United Kingdom
optimize open-source SLMs (e.g., Gemma 3, Llama 3) and vision-language models for execution on low-power edge runtimes (LiteRT/TensorFlow Lite, ONNX Runtime, ExecuTorch). Knowledge Analytics & Graph Processing: Design, implement, and maintain lightweight on-device graph databases and relationship extraction pipelines (Python, Rust, or C++ … V1.1/Dependable AI ). ML & Edge Inference Mastery 3+ years of production experience deploying ML models to edge runtime environments (LiteRT/TFLite, ONNX, C++ bindings). Experience in model quantization techniques (INT8, INT4, AWQ) and execution acceleration across NPU/GPU hardware. Proficiency in Python and PyTorch/ ...