MLOps Engineer (LLM/GenAI)
- Location
- Sheffield, England, United Kingdom
quantisation (INT4/FP8/GPTQ/AWQ), operator optimisation, framework integration (vLLM/TensorRT-LLM/SGLang) Production hosting experience with Docker/Kubernetes and cloud platforms (AWS/GCP/Azure) End-to-end fine-tuning expertise: data preparation, distributed training, hyperparameter tuning, HF/Accelerate/LoRA ...