MATS mentors are advancing the frontiers of AI alignment, transparency, and security

I previously worked on the alignment team at DeepMind, and on the governance team at OpenAI. I'm currently an independent researcher focusing on multi-agent intelligence. My research is in the tradition of natural philosophy; I'm trying to develop vague intuitive concepts (like trust, identity, and integrity) to the point where they can serve as seeds for new scientific paradigms.

Focus:
Theory
Agent Foundations
Alek Westover
Redwood Research
,
Member of Technical Staff

Alek is working on AI safety at Redwood Research.  He recently graduated from MIT where he studied Math, CS and AI.  Before working on AI safety he did theoretical computer science research (data structures, online algorithms, and algorithmic graph theory).

Focus:
Empirical
AI Control and Monitoring, Misalignment Science, Alignment Training Methods, Forecasting and Strategy

Ryan is Chief scientist at Redwood Research, focused on technical AI safety research to reduce risks from rogue AIs.

Focus:
Empirical
AI Control and Monitoring, Misalignment Science, Forecasting and Strategy

I am an Assistant Professor of Statistics and EECS at UC Berkeley, where I’m also part of BAIR and CLIMB. I am also Founder & CEO of Transluce, a non-profit research lab building open, scalable technology for understanding frontier AI systems.

Focus:
Empirical
AI Control and Monitoring, Capability and Propensity Evaluations, Adversarial Robustness and Safeguards, Misalignment Science, Forecasting and Strategy, Interpretability

I'm a research scientist at the UK AI Security Institute, working on AI control red teaming and model organisms of misalignment. I was previously a postdoc with Sam Bowman at NYU, did MATS with Owain Evans, and mentored for the MATS, SPAR and Pivotal fellowships. I got my PhD at the University of Edinburgh, supervised by Iain Murray.

Focus:
Empirical
AI Control and Monitoring, Capability and Propensity Evaluations, Adversarial Robustness and Safeguards, Misalignment Science
Adam Kaufman
Redwood Research
,
Member of Technical Staff

Adam is an AI Safety researcher and member of technical staff at Redwood Research.

Focus:
Empirical
AI Control and Monitoring
Stephen McAleer
Anthropic
,
Member of Technical Staff

Stephen is currently a researcher at Anthropic where he researches how to align and control superintelligence. He was previously a researcher at OpenAI and, before that, a postdoc at CMU working with Tuomas Sandholm. Stephen received his PhD in computer science from the University of California, Irvine working with Pierre Baldi. During his PhD, he did research scientist internships at Intel Labs and DeepMind. Before that, Stephen received his bachelor's degree in mathematics and economics from Arizona State University in 2017. Projects he is interested in include:

  • Anything related to control/monitoring for coding agents
  • Scalable oversight for agent alignment
  • Scheming evaluations and mitigations
  • Adversarial training for robust monitors / reward models
  • Reward hacking / deception in agents
Focus:
Empirical
AI Control and Monitoring, Adversarial Robustness and Safeguards, Misalignment Science

Gabriel runs ISL, which focuses on how to secure the most sensitive AI data centers against the most sophisticated current and future threats. ISL is a nonprofit R&D org focused on implementation-driven R&D for high-security AI systems. Previously, Gabriel was a fellow at RAND on hardware-enabled governance mechanisms and international verification of agreements. He holds a master's degree in computer science.

Focus:
Systems Security
Technical AI Governance, AI Systems Security
Kyle Fish
Anthropic
,
Model Welfare Lead

Kyle works on model welfare at Anthropic. He previously co-founded Eleos AI Research, Telis Bioscience, and Alvea.

Focus:
Empirical
Misalignment Science, AI Welfare
Julian Stastny
Redwood Research
,
Associate Member of Technical Staff

Julian leads the diffuse control team at Redwood Research.

Focus:
Empirical
AI Control and Monitoring, Misalignment Science, Forecasting and Strategy

Frequently asked questions

What is the MATS Program?
Who are the MATS Mentors?
What are the key dates of the MATS Program?
Who is eligible to apply?
How does the application and mentor selection process work?