View, CA, USA . PhD in Computer Science, Artificial Intelligence, or a related quantitative field, or equivalent practical experience. 4 years of experience in LLM fine-tuning/Reinforcement Learning, autonomous and human-on-loop agent development, and designing evaluation frameworks for real-world applications. Experience providing technical leadership … Experience with training agents for long-horizon tasks, computer control, domain safety, uncertainty calibration, and human-on-the-loop systems. Experience in full-lifecycle LLM development, deploying autonomous/human-on-loop agents in safety-critical real-world environments. Track record of publishing in AI venues (e.g., NeurIPS, ICML, ICLR ...