This stream will work on projects that empirically assess national security threats of AI misuse (CBRN terrorism and cyberattacks) and improve dangerous capability evaluations. Threat modeling applicants should have a skeptical mindset, enjoy case study work, and be strong written communicators. Eval applicants should be able and excited to help demonstrate concepts like sandbagging elicitation gaps in an AI misuse context.
This stream is primarily interested in mentoring projects in biosecurity that either (1) create rigorous threat models of AI biological misuse or (2) create benchmarks and tools that allow us to evaluate and mitigate these risks, as well as verifying that companies are taking suitable precautions.
Potential example projects include:
Nelly Mak is a research scientist on SecureBio's AI team, working on AIxBio evaluations and safeguards against AI-enabled biorisks. Mak previously completed a postdoc in the Jolly lab on HIV induction of tissue residency, and holds a PhD on the natural hosts of zoonotic viruses.
Peter Peneder is a research scientist on SecureBio's AI & Biotechnology Risks team, where his work focuses on building evaluations that assess the biorisk posed by frontier AI models. Peneder previously developed multimodal deep learning approaches for biomedical data during a PhD in bioinformatics.
Coleman Breen works at SecureBio on technical evaluations of coding agents, EU AI Act implementation, and policymaker engagement. Before SecureBio, Breen was a fellow at the Johns Hopkins Center for Health Security working on AIxBio policy.
Typically, this would include weekly meetings, detailed comments on drafts, and asynchronous messaging.
For threat modeling work:
For evaluations, mitigations, and verification work:
Mentor(s) will talk through project ideas with scholar
The Winter 2027 cohort offers a wide range of research streams led by experts across AI alignment, interpretability, governance, and safety. Each stream provides its own research agenda, methodology, and mentorship focus.