A few project ideas:
Standard (1-2 hours of weekly 1:1s)
London
No preference for this location
No preference for this location
I currently lead the science of evaluation team at the AI Security Institute in London. I joined AISI early in its life, and have worked in several roles, including co-leading the team responsible for our pre-deployment testing programme.
I'm generally interested in topics around dangerous capability evals, and understanding agent behaviours and their implications for policy. In particular, I'm interested in:
Before AISI I was chief scientist at a startup in Cambridge, where I led a team of 25 researchers with a mission to optimize decision making in electricity grids, and improve economic efficiency and reduce emissions.
I have a PhD in physics from the University of Waterloo and Perimeter Institute for Theoretical Physics. My focus was on reconstructing quantum theory from simple first principles, so we can all stop worrying about the reality of the wave-function.
1. Epistemic rigour, strong conceptual thinking and creativity: Able to dissect flawed assumptions in evaluation designs, and spot hidden pitfalls in claims.
2. Hands‑on prompting and agent scaffolding experience; engagement with experimenting and analysing model/agent behaviours.
3. Comfort with experimental design/analysis, uncertainty quantification, statistical modelling.