22 of 22 Remote/Hybrid ONNX Jobs

Senior MLOps Engineer

Hiring Organisation
Appcast
Location
Cambridge, Cambridgeshire, UK
similar platform.Experience with distributed training frameworks such as Ray, DeepSpeed, Horovod or PyTorch FSDP.Hands-on experience optimising models for inference using TensorRT, ONNX Runtime, OpenVINO, quantisation or pruning.Strong Docker and containerisation skills, including multi-stage or multi-architecture builds.Experience building CI/CD pipelines for Machine Learning systems.Experience with monitoring ...

Senior MLOps Engineer

Hiring Organisation
MFK Recruitment
Location
London, United Kingdom
Salary
£ 80 K
similar platform.Experience with distributed training frameworks such as Ray, DeepSpeed, Horovod or PyTorch FSDP.Hands-on experience optimising models for inference using TensorRT, ONNX Runtime, OpenVINO, quantisation or pruning.Strong Docker and containerisation skills, including multi-stage or multi-architecture builds.Experience building CI/CD pipelines for Machine Learning systems.Experience with monitoring ...

Sr. Machine Learning Engineer - Hybrid schedule (10-12 onsite a month)

Hiring Organisation
Calance US
Location
Los Angeles, California, United States
Employment Type
Permanent
Salary
USD Annual
frameworks relevant to GenAI application development (e.g., LangChain) Exhibit proficiency with ML frameworks (e.g., PyTorch, TensorFlow, Scikit-learn), and serving tools (e.g., TorchServe, ONNX, Triton) Display proficiency in training or fine-tuning language models (e.g., BERT, Llama2, GPT), and their optimization (LoRA, knowledge distillation, pruning, and quantization) And have ...

Senior Audio AI Research Engineer

Hiring Organisation
Logitech
Location
London, United Kingdom
Salary
£ 70 K
collaborative coding practices (e.g., open-source contributions).Expertise in performance analysis and optimization of ML systems.Hands-on deployment experience with systems like TensorFlow Lite, ONNX, TVM, and Glow.Strong familiarity with cloud compute environments, ideally AWS, and data ingestion (e.g., TF Data pipelines).Knowledge of source control and project tracking systems ...

Machine Learning Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
following would be beneficial: LangChain and/or LangGraph Python memory management and performance optimisation GPU optimisation and acceleration PyTorch ONNX/ONNX Runtime Model inference and acceleration Responsibilities for Machine Learning Engineer: Develop and maintain high-quality Python software within advanced commercial AI products Deploy and integrate new machine …/Machine Learning Software Engineer/AI Integration Engineer/Python/Large Language Models/LLM/LangChain/LangGraph/PyTorch/ONNX/ONNX Runtime/NumPy/pandas/Machine Learning/Artificial Intelligence/GPU Optimisation/Model Inference/Model Acceleration/Deep Learning ...

Machine Learning Framework/Runtime Software Engineer

Location
United Kingdom
building machine learning inference engines, runtime systems, or backend integration frameworks. Experience working with a machine learning inference framework such as LiteRT, TensorFlow Lite, ONNX Runtime, or a similar technology is also important. A good understanding of how AI models execute in practice is essential, including graph processing, operator execution ...

Experienced Machine Learning Framework/Runtime Software Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
building machine learning inference engines, runtime systems, or backend integration frameworks. Experience working with a machine learning inference framework such as LiteRT, TensorFlow Lite, ONNX Runtime, or a similar technology is also important.A good understanding of how AI models execute in practice is essential, including graph processing, operator execution, memory ...

ML Embedded Software Engineer

Location
United Kingdom
mobile applications, portable devices and home automation. Required skills and experience: Good programming skills - preferably C++ and OOP Experience of ML frameworks such as ONNX Runtime or PyTorch/ExecuTorch A desire to help developers deploy ML workloads to constrained edge devices A few years' relevant engineering experience 'Nice ...

ML Embedded Software Engineer

Location
Cambridge, England, United Kingdom
application - and potentially welcoming you to Arm. Required Skills And Experience Good programming skills - preferably C++ and OOP Experience of ML frameworks such as ONNX Runtime or PyTorch/ExecuTorch A desire to help developers deploy ML workloads to constrained edge devices A few years' relevant engineering experience Nice ...

Technical Architect

Hiring Organisation
Venturi
Location
City of London, London, United Kingdom
registry, versioning, CI/CD, drift detection, and retraining logic Working out how to run models in constrained or edge settings: quantisation, pruning, distillation, ONNX/TensorRT, and picking the right accelerators for the job Setting the evaluation approach - where precision and recall trade off, where thresholds ...

Staff Applications Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
functional teams.Excellent communications skills both written and verbal“Nice To Have” Skills and Experience:Experience with AI/ML frameworks such as TensorFlow, PyTorch, ONNX, or inference runtimes.System bring-up and JTAG debugging expertiseGood background of system performance analysisExperience with RTL simulation tools and software development toolsIn Return ...

Staff Applications Engineer

Location
Cambridge, England, United Kingdom
teams. Excellent communications skills both written and verbal “Nice To Have” Skills and Experience: Experience with AI/ML frameworks such as TensorFlow, PyTorch, ONNX, or inference runtimes. System bring-up and JTAG debugging expertise Good background of system performance analysis Experience with RTL simulation tools and software development tools ...

ML Embedded Software Engineer

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
receiving your application - and potentially welcoming you to Arm.Required skills and experience:Good programming skills - preferably C++ and OOPExperience of ML frameworks such as ONNX Runtime or PyTorch/ExecuTorchA desire to help developers deploy ML workloads to constrained edge devicesA few years' relevant engineering experience'Nice to have' abilities ...

Principal Product Manager - AI Tooling

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 70 K
demonstrate:Proven product management in development tools and software including CLIs, SDKs and APIs.Knowledge of machine learning frameworks, runtimes, infrastructure such as PyTorch, ONNX, ExecuTorch, Llama.cpp, vLLM and LiteRT.Understanding of model deployment to edge, embedded or heterogeneous computing, with an understanding of trade-offs between accuracy, performance and power consumption.You ...

Principal Product Manager - AI Tooling

Location
Cambridge, England, United Kingdom
demonstrate: Proven product management in development tools and software including CLIs, SDKs and APIs. Knowledge of machine learning frameworks, runtimes, infrastructure such as PyTorch, ONNX, ExecuTorch, Llama.cpp, vLLM and LiteRT. Understanding of model deployment to edge, embedded or heterogeneous computing, with an understanding of trade-offs between accuracy, performance ...

AI Compiler Optimization Engineer (Hybrid CPU/XPU)

Location
City of Edinburgh, Scotland, United Kingdom
model inference performance on CPU and CPU/XPU hybrid systems, using advanced compiler techniques. You will profile frameworks such as TensorFlow, PyTorch and ONNX, optimize graph execution, and contribute to open research with practical insights and publications. #J-18808-Ljbffr ...

Compiler Engineer

Hiring Organisation
Intellectual Capital Resources
Location
Bristol, Gloucestershire, United Kingdom
Salary
£ 100 K
compiler stack Developing intermediate representations, lowering pipelines and transformation passes Building hardware-aware compiler optimisations Integrating with technologies such as Google HEIR, PyTorch and ONNX Runtime Translating high-level workloads into efficient code for a novel accelerator Working closely with hardware, simulation, cryptography and algorithms engineers Benchmarking, profiling and debugging ...

Staff ML Engineer - Developer Tools

Location
Cambridge, England, United Kingdom
model optimisation techniques such as quantisation, graph optimisation, operator fusionandprecision reduction Experience across multiple ML frameworks, model formats and inference runtimes, such as PyTorch, ONNX/ONNX Runtime, ExecuTorch, TensorFlow/LiteRT and OpenVINO. Experience analysing, profiling and debugging ML workloads, including model compatibility, performance and the trade-offs between ...

Staff ML Engineer - Developer Tools

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 70 K
aware model optimisation techniques such as quantisation, graph optimisation, operator fusionandprecision reductionExperience across multiple ML frameworks, model formats and inference runtimes, such as PyTorch, ONNX/ONNX Runtime, ExecuTorch, TensorFlow/LiteRT and OpenVINO.Experience analysing, profiling and debugging ML workloads, including model compatibility, performance and the trade-offs between latency ...

Senior Machine Learning Engineer

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
optimize open-source SLMs (e.g., Gemma 3, Llama 3) and vision-language models for execution on low-power edge runtimes (LiteRT/TensorFlow Lite, ONNX Runtime, ExecuTorch). Knowledge Analytics & Graph Processing: Design, implement, and maintain lightweight on-device graph databases and relationship extraction pipelines (Python, Rust, or C++ … V1.1/Dependable AI ). ML & Edge Inference Mastery 3+ years of production experience deploying ML models to edge runtime environments (LiteRT/TFLite, ONNX, C++ bindings). Experience in model quantization techniques (INT8, INT4, AWQ) and execution acceleration across NPU/GPU hardware. Proficiency in Python and PyTorch/ ...

Senior Software Engineer, Machine Learning Services

Location
Greater London, England, United Kingdom
specific needs. Dive deep into the entire stack, from Kubernetes and container orchestration, through gRPC‐based service communication, to the performance tuning of ONNX‐based inference on GPU‐accelerated hardware. Write clean, efficient, and rigorously tested code. We value simplicity, correctness, and peer review. What you'll bring … challenges of managing the lifecycle of models in a multi‐tenant, high‐availability system. Familiarity with building ML inference services, model serialization (e.g., ONNX), and GPU programming (CUDA). You've built or worked on custom storage or job‐queueing systems before and have the scars to prove it. Maybe ...

Edge ML Embedded Engineer: C++/ONNX on Cortex-M

Location
United Kingdom
help demonstrate AI capabilities on edge devices. The role requires strong C++ and object-oriented programming skills, plus experience with ML frameworks such as ONNX Runtime or PyTorch/ExecuTorch. You will collaborate with diverse teams across regions on projects that reach developers globally. The position offers hybrid working arrangements ...