I'm interested a variety of projects, especially in multi-agent systems, AI control red-teaming, eval design, and propensity studies. Here are two multi-agent flavored research directions I would be excited about:
High-touch (+2 hours of weekly 1:1s)
London
No preference for this location
No preference for this location
I work at the UK AI Security Institute. In the past, I’ve done research in high-performance computing, language model pretraining, interpretability, and hardware enabled governance.
I would be happy to suggest concrete project ideas and help with brainstorming topic choices, or help guide an existing project that the scholar is interested in. My preference is that the scholar picks a category that overlaps with an area I actively work on so that I can give effective high-level advice.