
Anthropic
—
Member of Technical Staff
Links
Focus
AI Control and Monitoring, Capability and Propensity Evaluations, Misalignment Science, Alignment Training Methods
Stream
Anthropic
Fabien Roger is an AI safety researcher at Anthropic and previously worked at Redwood Research. Fabien’s research focuses on AI control and dealing with alignment faking.