[Expression of Interest] Research Engineer / Scientist, Alignment - London
- Location
- Greater London, England, United Kingdom
prompts to efficiently produce evaluation questions to test models’ reasoning abilities in safety‐relevant contexts. Contribute ideas, figures, and writing to research papers, blog posts, and talks. Run experiments that feed into key AI safety efforts at Anthropic, such as the design and implementation of our Responsible Scaling Policy. ...