Abeer Sharma
Mekaoui Helmy
When an AI agent spawns sub-agents, its safety limits do not follow. I build and deploy the layer that makes them inherited and non-strippable.
Constance Li
Funding compute/API costs for Incubator projects that build nonhuman welfare consideration into AI safety work
Phil Palmer
Deploying customer screening software at DNA synthesis providers to reduce AI-enabled biothreats
Anthony Ozerov
Help me evaluate the safety, ethics, and values of the quantized and fine-tuned open-weight LLMs that individuals and enterprises are actually using.
Sofia Yablonskaya
Monthly analysis of China's algorithm-filing registry and binding AI security standards, read in Chinese, for the people calibrating AI rules in the West.
Gabriel Sherman
A playbook to help AI safety policy advocates communicate with the U.S. government during the window of opportunity during an AI-related crisis.
Bobcat
It's an arthouse horror film. It's got a girl and a gun and global catastrophic risk.
Jai Dhyani
Creating conditions for cooperative strategies to dominate adversarial ones among near-future AIs while we still can
Eitan Sprejer
The Argentinian AI Safety community (BAISH, baish.com.ar) is the largest in Latin-America. Support BAISH's growth, by providing funding for paying salaries.
Pip Foweraker
A nightmarishly hard AI safety strategy game about holding p(Doom) down. You can't win; you can only buy time.
David Yu
Genevieve Shea
A practical evaluation framework to identify governance failures in frontier AI systems during elections.
Victor Porton
Jordan Arel
A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures
Paul Wang
Low-overhead zero-knowledge proofs of properties of training
explore26
A research bridge to safer deployment
Lawrence Wagner
Providing GPU credits and instructional support for 40 participants completing the ARENA AI Safety curriculum through Black in AI Safety and Ethics (BASE)
David Franklin
Yves
Reinforcement learning for cooperative LLM agents to address multi-agent risks.