@evhub
AGI safety Research Scientist at Anthropic. Previously Research Fellow at Machine Intelligence Research Institute.
https://www.alignmentforum.org/users/evhubThis is a donation to this user's regranting budget, which is not withdrawable.
$0 in pending offers
All of my grant-making will go towards reducing existential risks from artificial intelligence. Some reasons to think I'll do a good job at that:
I have a good deal of grant-making experience, having previously served as a Fund Manager for the EA Long-Term Future Fund (example grant write-ups: https://forum.effectivealtruism.org/posts/HYKDh2mLjapsgj9nB/long-term-future-fund-july-2021-grant-recommendations#Grant_reports_by_Evan_Hubinger and https://forum.effectivealtruism.org/posts/ddBLtdQjcjvZH5JvF/long-term-future-fund-december-2021-grant-recommendations#Grants_evaluated_by_Evan_Hubinger) as well as an FTX Future Fund regrantor.
I also have a great deal of professional AGI safety research experience, including both very theoretical work (e.g. "Risks from Learned Optimization" (https://arxiv.org/abs/1906.01820) and other stuff I did at MIRI) and very empirical work (e.g. "Discovering Language Model Behaviors" (https://arxiv.org/abs/2212.09251) and other stuff I'm currently doing at Anthropic). As a result, I think I'm particularly well-placed to evaluate AI safety projects across the full spectrum of possible approaches.
I do a lot of mentorship (e.g. via the SERI MATS program I helped start (https://www.alignmentforum.org/posts/FpokmCnbP3CEZ5h4t/ml-alignment-theory-program-under-evan-hubinger)) and as a result I see a lot of early-stage researchers and projects that would often benefit greatly from additional funding but are too illegible for traditional funders.
| For | Date | Type | Amount |
|---|---|---|---|
| <584b961e-134c-47aa-895f-350d8524f14c> | over 1 year ago | profile donation | +25 |
| Lightcone Infrastructure | almost 2 years ago | project donation | 10000 |
| MATS Program | almost 2 years ago | project donation | 10000 |
| Apollo Research: Scale up interpretability & behavioral model evals research | almost 2 years ago | project donation | 10000 |
| Next Steps in Developmental Interpretability | about 2 years ago | project donation | 30000 |
| AI-Driven Market Alternatives for a post-AGI world | about 2 years ago | project donation | 5000 |
| Support for Deep Coverage of China and AI | about 2 years ago | project donation | 20000 |
| Evaluating the Effectiveness of Unlearning Techniques | over 2 years ago | project donation | 20000 |
| AI Policy work @ IAPS | over 2 years ago | project donation | 5000 |
| Research Staff for AI Safety Research Projects | over 2 years ago | project donation | 25000 |
| <a2e90f73-2e2b-4059-9e09-eb1000bc572e> | over 2 years ago | profile donation | +50 |
| MATS Program | over 2 years ago | project donation | 80000 |
| Manifund Bank | over 2 years ago | deposit | +230000 |
| <d4c24a4d-b393-4671-aae0-e6883fd0bc37> | over 2 years ago | profile donation | +10 |
| Long-Term Future Fund | over 2 years ago | project donation | 100000 |
| Athena - New Program for Women in AI Alignment Research | over 2 years ago | project donation | 20000 |
| MATS Program | almost 3 years ago | project donation | 17533 |
| Apollo Research: Scale up interpretability & behavioral model evals research | almost 3 years ago | project donation | 15000 |
| <02be5f43-1129-4025-b752-8127a793fd82> | almost 3 years ago | profile donation | +333 |
| Scaling Training Process Transparency | almost 3 years ago | project donation | 5000 |
| Manifund Bank | almost 3 years ago | deposit | +50000 |
| Exploring novel research directions in prosaic AI alignment | almost 3 years ago | project donation | 25000 |
| Medical Expenses for CHAI PhD Student | almost 3 years ago | project donation | 10000 |
| <8c5d3152-ffd8-4d0e-b447-95a31f51f9d3> | almost 3 years ago | profile donation | +100 |
| Avoiding Incentives for Performative Prediction in AI | about 3 years ago | project donation | 33000 |
| Apollo Research: Scale up interpretability & behavioral model evals research | about 3 years ago | project donation | 100000 |
| <d950592c-b002-4a71-8235-b92b66ab30ef> | about 3 years ago | profile donation | +100 |
| Scoping Developmental Interpretability | about 3 years ago | project donation | 100000 |
| Manifund Bank | about 3 years ago | deposit | +50000 |
| Activation vector steering with BCI | about 3 years ago | project donation | 15000 |
| Agency and (Dis)Empowerment | about 3 years ago | project donation | 60000 |
| Manifund Bank | about 3 years ago | deposit | +400000 |
No comments yet. Sign in to create one!