Elliot Tower
"ICML published ML researcher — evaluation validity, mechanistic interpretability applied to biology
Charan Gowda D
Testing when VLA world models (Robots) simulation doesn't matchs the reality before it becomes an incidents
udbhav bharti
First benchmark of AI sales agents under RERA and consumer-protection constraint: compliance hard-fails zero the score, human-validated LLM judges, 50% Hinglish
Xyra Sinclair
scry.io, epistemic infrastructure for existential security philanthropy
Nikhil Maturi
An open, cheap method that detects when an inoculation prompt inoculates against off-target traits, so labs and developers can catch undesired trait/persona cha
Nicholas Andrews
Malcolm Jackson
Support for a team of four & narrative rap about technological change
Shelby Abel
Tactile sensing for a robotic body with audio, visuals, and cooling systems - prototyping the sensor array for third android sense.
Justin Bianchini
Giorgos Tsimpoulis
Comparing decades of satellite imagery to predict and pevent coastal loss, starting in Greece
Justin Shenk
Increasing public awareness of AI risks and benefits through in-person, interactive experiences
Phil Palmer
Deploying customer screening software at DNA synthesis providers to reduce AI-enabled biothreats
陳鈺澔
An AI platform for crypto and stock analysis, news verification, scam detection and wallet safety, with an open-source Safety Kernel tested on TON.
Abhir Mehra
AI companies change the models behind the names you build on. Sometimes they say so. Sometimes they do not. Either way we write it down.
Christopher Head
Reproducible detector that reads whether a manipulation is still active in an LLMs stream. xfers across 6 model families.
Georgia Tech Research Corporation
We will test whether circuits in protein language models can detect function-preserving redesigns of known toxins that evade homology-based DNA-synthesis screen
Gabriel Sherman
A playbook to help AI safety policy advocates communicate with the U.S. government during the window of opportunity during an AI-related crisis.
Yunika Bajracharya
Five-week AI safety fellowship + 3-month project mentorship
David Yu
Adrian St. Vaughan
Validated on 1,200 cases (97.9–99.5%). The reasoning layer has a bug - we demonstrated the fix works. $9,800 / 90 days to ship the open-source toolkit.