Vladislav Vassilyev
A six-month pilot testing how AI can strengthen human reasoning without displacing judgment, responsibility, or ownership of decisions.
Elliot Arledge
Frontier-model sweeps, hardware-roofline scoring, and reward-hacking audits at kernelbench.com. Keeping an independent AI R&D automation benchmark alive.
Francisco Antonio Da Costa Barroso
A fixed-point reformulation of iterative reasoning in sparse MoE models, replacing a fixed step count with provable convergence. Builds on a validated v1
Rahul Billakanti
Ctrl+Z for AI agents. Rewind automatically snapshots the filesystem and the conversation together, so when your agent breaks something, one rollback fixes both.
Ceri John
Independent validators test a claim blind, seal their verdicts, and reveal together — a permanent, tamper-proof record of AI evals and scientific results.
Слава
A web service that makes your photos resistant to AI-powered non-consensual undressing
John Lunsford
An open schema, consent framework, and harm taxonomy bounding what embodied AI agents may do in the home, validated on a shipping humanoid robot.
Luke Hamond
A Dreamer 4-style multimodal world model with a grounded, self-generated reward, learning 24/7 on a $500 open arm — measuring wireheading, forgetting, drift
shivam dubey
Ahmed Rehan
Indus AI Node
Oliver Klingefjord
A publication about the institutions we need for powerful AI.
Nada Amin
Building LemmaScript, a verification toolchain for TypeScript
Mapping the attention heads that push LLMs toward refusal vs. compliance, and building an inference-time defense against both single- and multi-turn jailbreaks.
Anju Chhetri
Francisco Salcido
Open infrastructure helping anyone detect malicious links, fake QR codes and brand impersonation before becoming a victim.
Edward Izgorodin
An open test of whether AI-agent memory systems spread false claims, plus a provenance spec so an agent's memory can be audited, corrected, and forgotten.
Cyan Lynn Yun
Peter Boctor
A model already designed a 44-piece engine inside the verify loop. This funds measuring how reliably it does that, and scaling to harder machines.
Olivia McAllister
Ontoresonant Feedback Loops
Brendan Nestor
A live Reader that inspects what AI answers surfaced, missed, and shaped, turning each run into a record for a future inspection agent.