Anthropic is a public benefit corporation focused on creating reliable and interpretable AI systems. The Research Lead on the Training Insights team will develop strategies and lead research efforts to measure and characterize model capabilities across training and deployment, while mentoring a small team of researchers and engineers.
Responsibilities:
- Build new novel and long-horizon evaluations
- Develop novel measurement approaches for understanding how model capabilities emerge and evolve during RL training
- Lead strategic evaluation coverage across the company
- Shape the evaluation narrative for model releases
- Lead and mentor a small team of researchers and research engineers, setting research direction and fostering a culture of rigorous, creative research
- Design evaluation frameworks that balance scientific rigor with the practical demands of production training schedules
- Build and maintain relationships across Anthropic's research organization to ensure evaluation insights inform training and deployment decisions
- Contribute to the broader research community through publications, open-source contributions, or external engagement on evaluation best practices