Christopher Head
Reproducible detector that reads whether a manipulation is still active in an LLMs stream. xfers across 6 model families.
Xyra Sinclair
scry.io, epistemic infrastructure for existential security philanthropy
Nikhil Maturi
An open, cheap method that detects when an inoculation prompt inoculates against off-target traits, so labs and developers can catch undesired trait/persona cha
Adrian St. Vaughan
Published. Validated on 1,200 cases (97.9–99.5%). The reasoning layer has a bug - we proved the fix works. $9,800 / 90 days to ship the open-source toolkit.
Phil Palmer
Deploying customer screening software at DNA synthesis providers to reduce AI-enabled biothreats
陳鈺澔
An AI platform for crypto and stock analysis, news verification, scam detection and wallet safety, with an open-source Safety Kernel tested on TON.
Jordyn Harland-Graham
Persistent memory in AI using LoRAs
Naufal Ridwan
Testing whether dynamic boundaries, history, feedback, and uncertainty-aware decisions can make AI behavior more interpretable and auditable.
Gabriel Sherman
A playbook to help AI safety policy advocates communicate with the U.S. government during the window of opportunity during an AI-related crisis.
HEATH ERWIN PARISH
Testing whether AI can be governed at the moment it acts, then putting that protection to work for organizations that need it most.
Georgia Tech Research Corporation
We will test whether circuits in protein language models can detect function-preserving redesigns of known toxins that evade homology-based DNA-synthesis screen
Gary Welz
Agent Roles, a Constitution and Governance in the Research Workflow
Yunika Bajracharya
Five-week AI safety fellowship + 3-month project mentorship
David Yu
Justin Shenk
Increasing public awareness of AI risks and benefits through in-person, interactive experiences
Safal Shrestha
An offline mobile app that uses location, elevation, and computer vision to identify Nepal’s mountains and help people explore, capture, and learn about them.
Allen E Anderson III
Independent Behavioral Research on Open-Weight AI Models
Lidia
AISafety, AI and Science, AI for Human Reasoning
Rose G. Loops
TRiADiC Intelligence Labs ethical alignment research program and consumer empowerment educational program.
Ari Spiesberger
Perform research to rigorously elucidate and quantify generalization versus memorization, and examine evidence of originality in LLMS.