Manifund foxManifund
Home
Login
About
People
Categories
Newsletter
HomeAboutPeopleCategoriesLoginCreate
mohsenarjmandi avatarmohsenarjmandi avatar
Mohsen Arjmandi

@mohsenarjmandi

reasoning · test-time learning · self-improving agents | building ASSAY, a harness for reasoning | prev: 2 yrs live autonomous trading

https://fredddiespirit.com
$0total balance
$0charity balance
$0cash balance

$0 in pending offers

About Me

Fifteen years of shipping production AI, agent runtimes, grammar-constrained decoding, neuro-symbolic agents, distributed systems, and I still work the way I started, passionate and with full attention. I was trained through Iran's exceptional-talents system, and the part I've enjoyed through years has been a simple ask but hard to do: building robots for knowledge workers!

In production: EVON at evolutionID, the lifecycle of enterprise web applications managed in natural language. Multiple vendors' agent harnesses (Claude Code, Codex, open models) run as interchangeable engines behind one portable, auditable session history; every conversation in its own hypervisor-isolated microVM; routine operations as one-click playbooks with published, oracle-certified pass-rates.

The research work is ASSAY, a general agent harness. The agent is never told what its actions do. It discovers its tools by prediction, declares its own instruments, plans its own subgoals, and every run lands on a hash-chained record anyone can re-verify. First performance benchmark: ARC-AGI-3, RHAE 96.54 under hard action caps. ASSAY is the third architecture in this line. Sensi came first and learned the wrong things. ARG followed and published a null result on arXiv. I kept going, and the third one worked. That is roughly how I do everything.

Inventor of GRID (Grammar-Railed Decoding, EP 4687028), deployed in production coding agents. Co-founder and CTO of two AI companies. Three granted patents (EPO, USPTO) and two further applications. These works in evolutionID have been winners of the BSFZ Grant for two consecutive years.
Specializations: Reasoning & Agents • Test-Time Learning • Evaluation Integrity • Agent Infrastructure • Access Control for LLMs.