Enterprise Architect - AI
- Hiring Organisation
- World Wide Technology
- Location
- London, UK
- Employment Type
- Full-time
frameworks.Inference & serving: NVIDIA Triton, vLLM, TensorRT-LLM, or equivalent high-throughput serving platforms.MLOps/LLMOps: Kubeflow, MLflow, and at least one hyperscaler ML platform (SageMaker, Azure ML, or Vertex AI).Generative AI: LLM fine-tuning (LoRA/QLoRA), RAG architecture design, vector databases (Pinecone, Milvus, Weaviate), and agentic frameworks ...