LLM & Generative AI Engineer
- Location
- Greater London, England, United Kingdom
About the Role Join our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks. Key Responsibilities Fine-tune open-source models (Llama, Mistral, Qwen … semantic search solutions Implement prompt evaluation frameworks and guardrail architectures Requirements 3+ years of experience focusing on Natural Language Processing and Generative AI Hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA) Experience deploying models with vLLM, Ollama, or Triton Inference ...