政策与治理方向研究先进 AI 目前如何治理,以及应该如何治理。随着 AI 能力快速提升,许多最棘手的问题已不再只是技术问题,还涉及国际协调、不确定性下的政策制定、国家能力、监管设计,以及如何将安全目标转化为现实政策与制度。未来 6 至 12 个月作出的决定,将影响实验室开发什么、政府提出哪些要求,以及未来数年的监督方式。
本方向涵盖多类研究。有些研究流关注具体治理机制,包括评估、标准、安全措施、监控系统和执行机制;另一些则开展政策与制度分析,例如比较不同治理制度、监管框架和国际协调难题。还有一些研究流探讨更广泛的问题,例如先进 AI 如何改变全球权力格局,以及治理方式能否切实降低灾难性风险。
我们希望研究员能够清晰地分析和撰写相关议题。以往届别中的优秀申请者拥有多种背景,包括政策、经济、法律、政治学、公共管理、安全研究、哲学、计算机科学、预测研究、社会学、历史、新闻,以及科学与技术研究。
我们会根据契合度为研究员匹配导师,并规划项目,使其在项目结束前产出具体成果,例如政策备忘录、监管意见、技术规范、比较分析和同行评审研究。本方向的成果面向 AISI 工作人员、实验室治理团队、监管机构、标准组织,以及参与前沿 AI 治理的研究与政策社区。
TBD. Please check back later for details. Broadly speaking I'm open to a wide range of projects across AI policy and technical AI governance, and I'm open to pitches. Right now I'm sorting out some career plans that will affect the kinds of projects that I'd support.
Strong understanding of the AI x-risk worldview, including detailed and well-informed views on AI timelines, the most important threats, the key technical and policy solutions, etc.
Strong writing ability, autonomy, and a track record of having led large impressive projects
TBD
My stream focuses on preserving checks and balances as governments adopt increasingly powerful AI. Fellows will work on questions like how Congress can maintain oversight of an AI-accelerated executive branch (including via privacy-preserving AI auditors) and what a positive vision for government AI adoption looks like. Projects will typically produce a public report and sometimes involve engaging directly with policymakers and other stakeholders.
We work arms-length with policymakers in Congress and the executive branch on timely, practical AI policy issues, while also conducting original research into future-oriented topics that may arise under plausible AI trajectories that have yet to enter the mainstream. High level themes include frontier AI governance, robotics / reindustrialization, and institutional disruption.
- Strong writing and communication ability
- Basic familiarity with core concepts in public policy and economics
This stream focuses on identifying tractable policy and technical interventions to gradual disempowerment, focusing on economic disempowerment and the intelligence curse. Possible project areas include:
We’ll meet 1:1 for 30 minute slots twice a week, once with each mentor. We’ll be active on Slack (default to over-slacking us), and can do quick ad-hoc calls as well. Once a week, we expect you to have some artifact that we will give feedback on.
We're excited about applications from a variety of backgrounds. Use the list below as general guidance, not as an exhaustive list.
Essential:
We provide three projects as options we are excited about, but they are not inclusive of all ideas. During the application process, we will ask potential mentees to either a) sharpen these proposals into a more specific question incorporating their own interests, or b) propose their own projects.
We expect fellows to come in with inner conviction towards a starting point that fits within the above themes, and expect that the best work in this stream will come from self-directed fellows pursuing their own research taste. However, we will require sign off to pursue a project and may require fellows to shift scope if they move outside the target area.
Escalation risks from state perceptions of AI capability, AI-enabled targeting, AI-enabled decision manipulation, and the impact of AI integration into nuclear command and control.
Mentorship will mostly consist of calls, sorting through research ideas and providing feedback. I'll be up to review papers, and potentially to meet in person depending on timing.
Looking for intellectually curious and honest scholars, with some background on topics related to national security, game theory, or AI-enabled military and influence capabilities.
I'll talk through project ideas with scholar, or the scholar can pick from a list of projects
Our stream focuses on AI verification, as in how actors can check that the use of AI compute is compliant with policy, especially for enabling international agreements on AI. This sense of verification is much broader than formal verification.
We'll meet once or twice a week (~1 hr/wk total, as a team if it's a team project). I'm based in DC, so we'll meet remotely. I (Mauricio) will also be available for async discussion, career advising, and detailed feedback on research plans and drafts.
Strong analytical and writing skills, research pragmatism and judgment, fast learner, proactive, and AI landscape context.
I'll talk through project ideas with scholar
This stream will focus on preparing AI governance policies for future policy windows through scenario mapping and policy architecture.
Research papers (technical governance or ML) related to evaluating and mitigating dangerous AI capabilities, with a focus on what's actionable and relevant for AGI companies
I like to get daily standup messages about progress that has been made on the project, and I'm happy to provide some quick async feedback on new outputs. I'll also have weekly meetings. I'm based in Constellation in Berkeley.
Good writers/researchers who can work independently and autonomously! I'm looking for scholars who can ship a meaningful research output end-to-end and ideally have prior experience in writing relevant papers.
I may assign a project, have you pick from a list of projects, or talk through project ideas with you.
MATS 项目是一项为期 10 周的研究奖学金计划,旨在培养和支持从事人工智能对齐、透明度和安全领域工作的新兴研究人员。研究员将与世界一流的导师合作,获得专门的研究管理支持,并加入位于伯克利、致力于推动人工智能安全与可靠发展的活跃社区。该项目提供开展高影响力研究并开启人工智能安全领域长期职业生涯所需的架构、资源和指导。
MATS 导师均为来自人工智能安全、对齐、治理、领域建设及安全等广泛领域的顶尖研究人员。他们包括学术界人士、行业研究员以及独立专家,负责指导学者开展研究项目、提供反馈,并助力每位学者的研究成长。导师们的专业领域涵盖:
查看 往届及现任导师
关键日期
申请:
主项目将于 9 月 28 日至 12 月 4 日进行,获选研究员的延展阶段将于 12 月开始。
MATS 欢迎来自不同学术和专业背景的申请者——从机器学习、数学和计算机科学,到政策、经济学、物理学、认知科学、生物学和公共卫生,同时也欢迎没有传统研究背景的创业者、运营人员和领域建设者。主要要求是具备为人工智能安全做出贡献的强烈动机,并展现出技术能力、研究潜力或相关的运营经验。具备人工智能安全相关经验会有所帮助,但并非必要条件。