Applications are now closedRegister for future cohorts →

Project bank

Projects and contributors

Real problems sent in by people working in AI safety, plus open problems with no set contributor yet. Everyone in the cohort gets one project, cut down until it can be finished in five days and used afterwards.

Harry Waterman

USG Pastcasting

Harry Waterman · Generator Residency · Fellow

Build a run-ready workshop that helps AI safety people improve their models of how the US government responds to crises. Participants get a historical crisis as it looked on day 0, predict which agencies will act, under what authorities, and on what timeline, then compare…

PolicyWorkshopsField-building
Matt Handzel

Personal fit infrastructure for AI safety roles

Matt Handzel · Kairos / Constellation · Generator Resident

Create a methodology (and the material behind it) that lets applicants work out whether they'd actually enjoy a given AI safety role before sinking effort into applications and work trials.

Career pipelineField-buildingResearch
Sohan Venkatesh

Deception Probe Robustness Harness

Sohan Venkatesh · LASR Labs · Research Fellow

A reusable toolkit that stress-tests any linear deception probe against the failure modes recent research has identified, but never packaged. Given a trained probe and a target model, it would run a standardized battery: an eight-style distributional shift test, cross-domain…

TechnicalInterpretability
Alejandro Acelas

AI Uplift Guide

Alejandro Acelas · Impact Ops / 80,000 Hours · AI Uplift Person

Make a written guide that takes an extremely smart person (e.g., the median 80k employee) from "I use Claude Chat to assist me with some things" to "man, I can't imagine not trying AI for absolutely everything I do" I suspect there's generally-applicable intuitions on getting…

AI upliftWriting
Francesca Gomez

Evaluating paths to impact for cooperative AI cyber defenses

Francesca Gomez · Wiser Human · AI safety researcher

Evaluate which intervention approaches would get cooperative AI cyber defenses adopted: who is most exposed, which deployment routes are most tractable, which funding models move fastest, and how we'd measure whether it's working.

SecurityStrategyGovernance
Mark Aiken

Hiring senior staff: how scaling AI safety orgs can scope and assess senior roles

Mark Aiken · Manager

This project will help develop a coherent framework for recruitment in AI safety, by modelling competency based approaches to defining positions and recruitment processes based on best practices. In one week a participant would ship a usable toolkit: (1) a job analysis and…

HiringOperations
Garrett Xu

Building Brown AI Safety's Policy Capacity Across Curriculum, Legislative Drafting, and Lab Transparency Research

Garrett Xu · Brown AI Safety · Executive — Policy Lead

Curriculum and workshops (~2 people, ~2 days): Help design Brown AI Safety's Policy curriculum and workshop series (e.g. memo writing workshop and one-pagers workshop) for Brown AI Safety, with materials other university groups could adapt. Brown AI Safety Practical Policy (~2-3…

PolicyField-building
Srijit Paul

Wales AI Safety Hub (WAISH)

Srijit Paul · AI Safety Cardiff University · President

Ship the operating system for a new AI safety hub in Wales: an intake process for its free applied research service, a one-page explainer for prospective partners, and a first draft of the protective assessment framework researchers would use.

Field-buildingGovernanceOperations
Noah Erlwein

Authorised actor does not equal authorised action: a benchmark for agent action-level authorisation

Noah Erlwein · Invariant Systems, Inc. · Founder

Build a scenario corpus and minimal evaluator that test whether current agent frameworks catch actions taken outside the intent a user actually granted, and report which categories they miss.

EvalsAgentsSecurity
Peter Kazys

Measuring the AI uplift in re-identifying "anonymized" data-broker records

Peter Kazys · SAP Concur · Client Success Manager

Run a controlled, synthetic-data experiment measuring how much modern AI lowers the cost, time and skill needed to re-identify "anonymized" broker records, and turn the number into a threat model and policy brief.

PrivacyPolicyEvals
Andrew Nelson

Navigating transformative Physical AI (robotics): mapping the opportunity space

Andrew Nelson · Safe Physical AI (working name) · Cofounder, general leadership and product

Build a comprehensive field map of Physical AI: robotics manufacturers, model developers, evals orgs and safety researchers, so a new initiative can judge neglectedness and target the right partners.

Field-buildingRoboticsStrategy
Amir Mohammad Farhang

Onboarding runbook + participant-ops system for an AI governance benchmark's first public competition

Amir Mohammad Farhang · Compex · Founder & CEO, Compex; organizer, FinBoundBench (ICAIF 2026 competition track)

Build the human-facing operations layer for FinBoundBench, an open benchmark competition launching Sept 1: participant onboarding flow, a support runbook for the six-week development phase, and an outreach tracker for labs, fintech teams, and governance research groups.

OperationsGovernanceEvals
Yaseen Ismail

Gamified skill-tree to AI safety upskilling and contribution

Yaseen Ismail · Independent technical research · Ex-Yale researcher

Build a linear, gamified skill-tree style pick-your-path system: an organised route from first hearing about AI safety to the different goals and paths you may reach for, and what you need on the way.

Field-buildingCareer pipelineProduct
Helen Adepoju

AI Safety Navigator

Helen Adepoju · Founder, AI Safety Navigator

Scale a filterable directory of AI safety opportunities so people can find fellowships, training and roles that fit their background, time and timeline, and help find an org to own it long term.

Career pipelineField-buildingProduct