What progress have you made since your last update?
Added 4 new benchmark modules.
Evaluated all relevant AI models.
DystopiaBench helped me get an internship at the Center for AI Safety, building directly on work enabled by this grant.
What are your next steps?
Build a v2 of the benchmark, as recent frontier models such as Fable and Astra have saturated the current tests.
Design harder scenarios that better differentiate frontier models.
Continue testing new models as they are released.
Is there anything others could help you with?
Suggestions for harder or more realistic failure scenarios.