Alignment-science style work eg:
Tech-gov style work:
Other:
Rhys is the research director of Arrow: a new AI safety non-profit based in London.
He is currently most excited about working on alignment science (e.g., model organisms research) and technical governance style work which reduces public uncertainty regarding important questions (e.g., his recent work on no-CoT time-horizons).
Rhys started working on AGI risk in 2019. He has previously worked at: Redwood Research, LawZero, UK AISI, GovAI, CLR, and the Centre for Assuring Autonomy. He did his PhD in AI Deception at Imperial College London.
Technical skills, experience building evals with inspect, or working on agents or fine-tuning models with tinker.
It's good if you have at least one completed AI safety project.
The Winter 2027 cohort offers a wide range of research streams led by experts across AI alignment, interpretability, governance, and safety. Each stream provides its own research agenda, methodology, and mentorship focus.