Xyra Sinclair
scry.io
Nikhil Maturi
An open, cheap method that detects when an inoculation prompt inoculates against off-target traits, so labs and developers can catch undesired trait/persona cha
Adrian St. Vaughan
Published. Validated on 1,200 cases (97.9–99.5%). The reasoning layer has a bug - we proved the fix works. $9,800 / 90 days to ship the open-source toolkit.
Phil Palmer
Deploying customer screening software at DNA synthesis providers to reduce AI-enabled biothreats
陳鈺澔
An AI platform for crypto and stock analysis, news verification, scam detection and wallet safety, with an open-source Safety Kernel tested on TON.
Jordyn Harland-Graham
Persistent memory in AI using LoRAs
HEATH ERWIN PARISH
Testing whether AI can be governed at the moment it acts, then putting that protection to work for organizations that need it most.
Naufal Ridwan
Testing whether dynamic boundaries, history, feedback, and uncertainty-aware decisions can make AI behavior more interpretable and auditable.
Gary Welz
Agent Roles, a Constitution and Governance in the Research Workflow
Gabriel Sherman
A playbook to help AI safety policy advocates communicate with the U.S. government during the window of opportunity during an AI-related crisis.
Georgia Tech Research Corporation
We will test whether circuits in protein language models can detect function-preserving redesigns of known toxins that evade homology-based DNA-synthesis screen
Safal Shrestha
An offline mobile app that uses location, elevation, and computer vision to identify Nepal’s mountains and help people explore, capture, and learn about them.
Yunika Bajracharya
Five-week AI safety fellowship + 3-month project mentorship
Allen E Anderson III
Independent Behavioral Research on Open-Weight AI Models
David Yu
Lidia
AISafety, AI and Science, AI for Human Reasoning
Justin Shenk
Increasing public awareness of AI risks and benefits through in-person, interactive experiences
Rose G. Loops
TRiADiC Intelligence Labs ethical alignment research program and consumer empowerment educational program.
Dung Claire Tran
EEG Evidence from ALS, Parkinson's, and Sleep. Abstract accepted for presentation at Models of Consciousness 7
Ari Spiesberger
Perform research to rigorously elucidate and quantify generalization versus memorization, and examine evidence of originality in LLMS.