Manifund foxManifund
Home
Login
About
People
Categories
Newsletter
HomeAboutPeopleCategoriesLoginCreate
evhub avatarevhub avatar
Evan Hubinger

@evhub

regrantor

AGI safety Research Scientist at Anthropic. Previously Research Fellow at Machine Intelligence Research Institute.

https://www.alignmentforum.org/users/evhub

Donate

This is a donation to this user's regranting budget, which is not withdrawable.

Sign in to donate
$15,085total balance
$15,085charity balance
$0cash balance

$0 in pending offers

About Me

All of my grant-making will go towards reducing existential risks from artificial intelligence. Some reasons to think I'll do a good job at that:

  • I have a good deal of grant-making experience, having previously served as a Fund Manager for the EA Long-Term Future Fund (example grant write-ups: https://forum.effectivealtruism.org/posts/HYKDh2mLjapsgj9nB/long-term-future-fund-july-2021-grant-recommendations#Grant_reports_by_Evan_Hubinger and https://forum.effectivealtruism.org/posts/ddBLtdQjcjvZH5JvF/long-term-future-fund-december-2021-grant-recommendations#Grants_evaluated_by_Evan_Hubinger) as well as an FTX Future Fund regrantor.

  • I also have a great deal of professional AGI safety research experience, including both very theoretical work (e.g. "Risks from Learned Optimization" (https://arxiv.org/abs/1906.01820) and other stuff I did at MIRI) and very empirical work (e.g. "Discovering Language Model Behaviors" (https://arxiv.org/abs/2212.09251) and other stuff I'm currently doing at Anthropic). As a result, I think I'm particularly well-placed to evaluate AI safety projects across the full spectrum of possible approaches.

  • I do a lot of mentorship (e.g. via the SERI MATS program I helped start (https://www.alignmentforum.org/posts/FpokmCnbP3CEZ5h4t/ml-alignment-theory-program-under-evan-hubinger)) and as a result I see a lot of early-stage researchers and projects that would often benefit greatly from additional funding but are too illegible for traditional funders.

Outgoing donations

Lightcone Infrastructure
$10000
almost 2 years ago
MATS Program
$10000
almost 2 years ago
Apollo Research: Scale up interpretability & behavioral model evals research
$10000
almost 2 years ago
Next Steps in Developmental Interpretability
$30000
about 2 years ago
AI-Driven Market Alternatives for a post-AGI world
$5000
about 2 years ago
Support for Deep Coverage of China and AI
$20000
about 2 years ago
Evaluating the Effectiveness of Unlearning Techniques
$20000
over 2 years ago
AI Policy work @ IAPS
$5000
over 2 years ago
Research Staff for AI Safety Research Projects
$25000
over 2 years ago
MATS Program
$80000
over 2 years ago
Long-Term Future Fund
$100000
over 2 years ago
Athena - New Program for Women in AI Alignment Research
$20000
over 2 years ago
MATS Program
$17533
almost 3 years ago
Apollo Research: Scale up interpretability & behavioral model evals research
$15000
almost 3 years ago
Scaling Training Process Transparency
$5000
almost 3 years ago
Exploring novel research directions in prosaic AI alignment
$25000
almost 3 years ago
Medical Expenses for CHAI PhD Student
$10000
almost 3 years ago
Avoiding Incentives for Performative Prediction in AI
$33000
about 3 years ago
Apollo Research: Scale up interpretability & behavioral model evals research
$100000
about 3 years ago
Scoping Developmental Interpretability
$100000
about 3 years ago
Activation vector steering with BCI
$15000
about 3 years ago
Agency and (Dis)Empowerment
$60000
about 3 years ago
About EvanBy Evan7

No comments yet. Sign in to create one!

Transactions

ForDateTypeAmount
<584b961e-134c-47aa-895f-350d8524f14c>over 1 year agoprofile donation+25
Lightcone Infrastructurealmost 2 years agoproject donation10000
MATS Programalmost 2 years agoproject donation10000
Apollo Research: Scale up interpretability & behavioral model evals researchalmost 2 years agoproject donation10000
Next Steps in Developmental Interpretabilityabout 2 years agoproject donation30000
AI-Driven Market Alternatives for a post-AGI worldabout 2 years agoproject donation5000
Support for Deep Coverage of China and AIabout 2 years agoproject donation20000
Evaluating the Effectiveness of Unlearning Techniques over 2 years agoproject donation20000
AI Policy work @ IAPSover 2 years agoproject donation5000
Research Staff for AI Safety Research Projectsover 2 years agoproject donation25000
<a2e90f73-2e2b-4059-9e09-eb1000bc572e>over 2 years agoprofile donation+50
MATS Programover 2 years agoproject donation80000
Manifund Bankover 2 years agodeposit+230000
<d4c24a4d-b393-4671-aae0-e6883fd0bc37>over 2 years agoprofile donation+10
Long-Term Future Fundover 2 years agoproject donation100000
Athena - New Program for Women in AI Alignment Researchover 2 years agoproject donation20000
MATS Programalmost 3 years agoproject donation17533
Apollo Research: Scale up interpretability & behavioral model evals researchalmost 3 years agoproject donation15000
<02be5f43-1129-4025-b752-8127a793fd82>almost 3 years agoprofile donation+333
Scaling Training Process Transparencyalmost 3 years agoproject donation5000
Manifund Bankalmost 3 years agodeposit+50000
Exploring novel research directions in prosaic AI alignmentalmost 3 years agoproject donation25000
Medical Expenses for CHAI PhD Studentalmost 3 years agoproject donation10000
<8c5d3152-ffd8-4d0e-b447-95a31f51f9d3>almost 3 years agoprofile donation+100
Avoiding Incentives for Performative Prediction in AIabout 3 years agoproject donation33000
Apollo Research: Scale up interpretability & behavioral model evals researchabout 3 years agoproject donation100000
<d950592c-b002-4a71-8235-b92b66ab30ef>about 3 years agoprofile donation+100
Scoping Developmental Interpretabilityabout 3 years agoproject donation100000
Manifund Bankabout 3 years agodeposit+50000
Activation vector steering with BCIabout 3 years agoproject donation15000
Agency and (Dis)Empowermentabout 3 years agoproject donation60000
Manifund Bankabout 3 years agodeposit+400000