Large Language Models. You will help develop and implement the organisation’s LLM evaluation framework, define best practice validation standards, and drive the automation of testing and assurance processes.Alongside the technical challenge, you will work closely with senior stakeholders across risk, technology, data science and business leadership to shape … learning and Generative AI modelsDesign and implement robust LLM evaluation and testing frameworksDevelop approaches for assessing model performance, reliability, explainability, bias and operational riskDrive automation and innovation within the model validation processPartner with AI development teams to improve model quality and governance standardsPresent findings and recommendations to senior stakeholders ...