Peter Henderson’s stream focuses on developing safe, aligned AI agents, with projects on scalable oversight rules informed by law and game theory, safe long-horizon exploration, and measuring “jagged” capability/safety frontiers. Scholars will join an independently driven, engineering-heavy research environment, collaborating with other MATS scholars and PhD students, with weekly 1:1s and active async mentorship.
I'd be interested in a variety of potential projects, but three directions that I'm excited by are:
New York City
Peter is an assistant professor at Princeton University, where he works on reinforcement learning, alignment, and law. He received a J.D. and Ph.D. in computer science from Stanford University.
Essential:
Nice to have, but not necessary:
Not a good fit:
Mentors in the group will pitch projects, and scholars will try ones they find interesting for a week. We'll iterate together at the end of week 1 and pick final assignments in week 2.