4 of 4 Permanent Qwen Jobs in London

Software Engineer, Machine Learning Infrastructure

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves - real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs - delivering large cost and latency wins (for example ...

Senior Software Engineer, GenAI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves - real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs - delivering large cost and latency wins (for example ...

LLM & Generative AI Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
private models. You will be responsible for tailoring deep learning models to specialized domain tasks. Key Responsibilities Fine-tune open-source models (Llama, Mistral, Qwen) for specific domain functions Optimize model deployment pipelines for low latency and high throughput Build advanced context management and semantic search solutions Implement prompt evaluation ...

Senior Software Engineer - Applied AI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
constrained hardware such as Jetson, Android, iPad, Raspberry Pi or microcontrollers. Hands-on with llama.cpp or similar, and evaluating small on-device models (e.g. Qwen, Gemma, MedGemma, custom). Some hardware exposure: anyone doing genuine edge AI tends to have this. Cloud experience (AWS), particularly running LLMs in a cloud ...