Bloom: an open source tool for automated behavioral evaluations

MATS Fellow:

Isha Gupta

Authors:

Isha Gupta, Kai Fronsdal, Abhay Sheshadri, Jonathan Michala, Jacqueline Tay, Rowan Wang, Samuel R. Bowman, Sara Price

Citations

0 Citations

Abstract:

We are releasing Bloom, an agentic framework for developing behavioral evaluations. Bloom's evaluations are reproducible and targeted: unlike open-ended auditing, Bloom takes a researcher-specified behavior and quantifies its frequency and severity across automatically generated scenarios. Bloom's evaluations correlate strongly with our hand-labelled judgments and reliably separate baseline models from intentionally misaligned ones. As examples, we also release benchmark results for four alignment relevant behaviors on 16 models. Bloom is available at github.com/safety-research/bloom.

Recent research

Synthetic Persona Pretraining: Alignment from Token Zero

Authors:

Julian Minder

Date:

August 13, 2026

Citations:

Capability Provenance in Language Models: A Case Study in Social Reasoning

Authors:

Glenn Matlin, Taywon Min

Date:

August 10, 2026

Citations:

Frequently asked questions

What is the MATS Program?
How long does the program last?