Agentic Property-Based Testing: Finding Bugs Across the Python Ecosystem

MATS Fellow:

Muhammad Maaz

Authors:

Muhammad Maaz, Liam DeVoe, Zac Hatfield-Dodds, Nicholas Carlini

Citations

3 Citations

Abstract:

We developed an agent that can efficiently identify bugs in large software projects. To do this, our agent infers general properties of code that should be true, and then by applying property-based testing—a technique similar to fuzz testing—we are able to discover bugs in top Python packages like NumPy, SciPy, and Pandas. After extensive manual validation, we are in the process of reporting these bugs to the developers, several of which have already been patched.

For more information, read the full paper, take a look at the GitHub repository, or browse the bugs we found at our site.

Recent research

Underwriting the Agent Economy: The Blueprint for an AI Insurance Stack

Authors:

Anita Srinivasan

Date:

July 14, 2026

Citations:

When Role-playing, Do Models Believe What They Say?

Authors:

Benjamin Sturgeon

Date:

June 25, 2026

Citations:

Agentic Property-Based Testing: Finding Bugs Across the Python Ecosystem

Recent research

Frequently asked questions