I'm excited about work on CompassionBench and similar evals, to help better understand how well models are doing on animal welfare -relevant axes
@ethanjperez
I lead the adversarial robustness team at Anthropic, where I’m hoping to reduce existential risks from AI systems. I helped to develop Retrieval-Augmented Generation (RAG), a widely used approach for augmenting large language models with other sources of information. I also helped to demonstrate that state-of-the-art AI safety training techniques do not ensure safety against sleeper agents. I received a best paper award at ICML 2024 for my work showing that debating with more persuasive LLMs leads to more truthful answers. I received my PhD from NYU under the supervision of Kyunghyun Cho and Douwe Kiela and funded by NSF and Open Philanthropy. Previously, I’ve spent time at DeepMind, Facebook AI Research, Montreal Institute for Learning Algorithms, and Google. I was also named one of Forbes’s 30 Under 30 in AI.
ethanperez.netThis is a donation to this user's regranting budget, which is not withdrawable.
$13,315 in pending offers
I'm interested in funding anything related to making AI go well, including work on AI safety, policy, welfare, and more.
Ethan Josean Perez
11 days ago
I'm excited about work on CompassionBench and similar evals, to help better understand how well models are doing on animal welfare -relevant axes
Ethan Josean Perez
8 months ago
I'm a huge fan of Transluce's work on investigator agents, automated red teaming and jailbreaking, and introducing fresh angles in interpretability/explainability. Their work has consistently been thought-provoking and among the best safety research out there. They also seem to have quite a talented and productive team as well.
Ethan Josean Perez
8 months ago
Forethought’s work has been some of the biggest influence on my safety research work, raising several new areas of research I and some of my collaborators hadn’t previously considered or prioritized. Keep up the amazing work!
| For | Date | Type | Amount |
|---|---|---|---|
| CaML - AGI alignment to nonhumans | 11 days ago | project donation | 26685 |
| Manifund Bank | 3 months ago | deposit | +25000 |
| Transluce: Fund Scalable Democratic Oversight of AI | 5 months ago | project donation | 10000 |
| Forethought | 6 months ago | project donation | 100000 |
| Manifund Bank | 8 months ago | deposit | +25000 |
| Manifund Bank | 12 months ago | deposit | +100000 |
| Manifund Bank | almost 3 years ago | withdraw | 200100 |
| Compute and other expenses for LLM alignment research | almost 3 years ago | project donation | +200000 |
| Compute and other expenses for LLM alignment research | almost 3 years ago | project donation | +100 |
| Compute and other expenses for LLM alignment research | almost 3 years ago | project donation | +200000 |
| Manifund Bank | almost 3 years ago | withdraw | 200000 |