14 of 14 ONNX Jobs in the UK excluding London

Embedded Machine Learning Engineer

Hiring Organisation
KO2 Embedded Recruitment Solutions LTD
Location
Dunfermline, Fife, UK
Employment Type
Full-time
world experience working with sensor data, time-series data, or IoT data streams Familiarity with embedded ML tools and approaches: TensorFlow Lite, Edge Impulse, ONNX Runtime, or equivalent Hands-on mindset; comfortable getting close to hardware, firmware code, and the real-world constraints of device deployment Clear communication; ability ...

Embedded Machine Learning Engineer

Hiring Organisation
KO2 Embedded Recruitment Solutions LTD
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent
Salary
£70,000
world experience working with sensor data, time-series data, or IoT data streams Familiarity with embedded ML tools and approaches: TensorFlow Lite, Edge Impulse, ONNX Runtime, or equivalent Hands-on mindset; comfortable getting close to hardware, firmware code, and the real-world constraints of device deployment Clear communication; ability ...

Embedded Machine Learning Engineer

Location
Edinburgh, Midlothian, United Kingdom
world experience working with sensor data, time-series data, or IoT data streams Familiarity with embedded ML tools and approaches: TensorFlow Lite, Edge Impulse, ONNX Runtime, or equivalent Hands-on mindset comfortable getting close to hardware, firmware code, and the real-world constraints of device deployment Clear communication ability ...

Staff Applications Engineer

Location
Cambridge, England, United Kingdom
teams. Excellent communications skills both written and verbal “Nice To Have” Skills and Experience: Experience with AI/ML frameworks such as TensorFlow, PyTorch, ONNX, or inference runtimes. System bring-up and JTAG debugging expertise Good background of system performance analysis Experience with RTL simulation tools and software development tools ...

Principal Product Manager - AI Tooling

Location
Cambridge, England, United Kingdom
demonstrate: Proven product management in development tools and software including CLIs, SDKs and APIs. Knowledge of machine learning frameworks, runtimes, infrastructure such as PyTorch, ONNX, ExecuTorch, Llama.cpp, vLLM and LiteRT. Understanding of model deployment to edge, embedded or heterogeneous computing, with an understanding of trade-offs between accuracy, performance ...

AI Compiler Optimization Engineer

Location
City of Edinburgh, Scotland, United Kingdom
improvements Preferred: Experience with LLVM/MLIR development AI Model Profiling & Framework Optimization: Profile end-to-end inference workflows on frameworks like TensorFlow, PyTorch, ONNX, and llama.cpp to identify hotspots and bottlenecks Propose and implement optimization strategies (e.g., kernel tuning, graph-level optimizations) Preferred: Experience optimizing models on multiple ...

AI Compiler Optimization Engineer (Hybrid CPU/XPU)

Location
City of Edinburgh, Scotland, United Kingdom
model inference performance on CPU and CPU/XPU hybrid systems, using advanced compiler techniques. You will profile frameworks such as TensorFlow, PyTorch and ONNX, optimize graph execution, and contribute to open research with practical insights and publications. #J-18808-Ljbffr ...

AI Compiler Optimization Engineer - Edinburgh

Hiring Organisation
Microtech Global Ltd
Location
Livingston, West Lothian, UK
Employment Type
Full-time
improvements Preferred: Experience with LLVM/MLIR development AI Model Profiling & Framework Optimization: Profile end-to-end inference workflows on frameworks like TensorFlow, PyTorch, ONNX, and llama.cpp to identify hotspots and bottlenecks Propose and implement optimization strategies (e.g., kernel tuning, graph-level optimizations) Preferred: Experience optimizing models on multiple ...

AI Compiler Optimization Engineer - Edinburgh

Hiring Organisation
Microtech Global Ltd
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent
improvements Preferred: Experience with LLVM/MLIR development AI Model Profiling & Framework Optimization: Profile end-to-end inference workflows on frameworks like TensorFlow, PyTorch, ONNX, and llama.cpp to identify hotspots and bottlenecks Propose and implement optimization strategies (e.g., kernel tuning, graph-level optimizations) Preferred: Experience optimizing models on multiple ...

AI Compiler Optimization Engineer - Edinburgh

Location
Bonnyrigg, Midlothian, United Kingdom
improvements Preferred: Experience with LLVM/MLIR development AI Model Profiling & Framework Optimization: Profile end-to-end inference workflows on frameworks like TensorFlow, PyTorch, ONNX, and llama.cpp to identify hotspots and bottlenecks Propose and implement optimization strategies (e.g., kernel tuning, graph-level optimizations) Preferred: Experience optimizing models on multiple ...

Staff ML Engineer - Developer Tools

Location
Cambridge, England, United Kingdom
model optimisation techniques such as quantisation, graph optimisation, operator fusionandprecision reduction Experience across multiple ML frameworks, model formats and inference runtimes, such as PyTorch, ONNX/ONNX Runtime, ExecuTorch, TensorFlow/LiteRT and OpenVINO. Experience analysing, profiling and debugging ML workloads, including model compatibility, performance and the trade-offs between ...

Senior Machine Learning Engineer

Hiring Organisation
Hackajob Ltd
Location
Slough, England, United Kingdom
optimize open-source SLMs (e.g., Gemma 3, Llama 3) and vision-language models for execution on low-power edge runtimes (LiteRT/TensorFlow Lite, ONNX Runtime, ExecuTorch). Knowledge Analytics & Graph Processing: Design, implement, and maintain lightweight on-device graph databases and relationship extraction pipelines (Python, Rust, or C++ … V1.1/Dependable AI ). ML & Edge Inference Mastery 3+ years of production experience deploying ML models to edge runtime environments (LiteRT/TFLite, ONNX, C++ bindings). Experience in model quantization techniques (INT8, INT4, AWQ) and execution acceleration across NPU/GPU hardware. Proficiency in Python and PyTorch/ ...

Senior Machine Learning Engineer

Location
Cambridge, England, United Kingdom
robust, scalable and reproducible training and inference workflows. What you will do Develop robust, scalable and reproducible inference pipelines. Deploy models into production using ONNX, TensorRT or similar frameworks. Build and optimise scalable machine learning training workflows. Optimise data loading, logging, checkpointing and resource utilisation for large-scale model training. … optimising and maintaining ML training and inference pipelines. Experience with Docker and containerised ML workflows. Experience with model deployment and optimisation frameworks such as ONNX, TensorRT or similar tools. Experience with GPU-based training, model serving and compute optimisation. Experience building and maintaining CI pipelines. Experience with cloud environments. Strong ...

Senior Machine Learning Applications and Compiler Engineer, LPX

Hiring Organisation
Nvidia
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep learning, crafting how neural network workloads ...