Three decision models on ~1,000 Berkeley probability exam questions
Abhijay stress-tested Jev, SemIf and Laya on adapted exam questions: Jev scored 83.7%, SemIf 61.6% and Laya 31.2%.
BUILDS
Agents, routers, scorers and games — built on models that decide. Every entry links to the models it uses, so you can pull the same one and start from there.
3 builds
Abhijay stress-tested Jev, SemIf and Laya on adapted exam questions: Jev scored 83.7%, SemIf 61.6% and Laya 31.2%.
Leonardo Stenico moved the Needle extension from hosted Jev to local Laya, so finding passages by meaning needs no API key.
Astrid's Postgres extension, written in Rust on Candle, runs the model in the database so a typed decision is a SQL call.
Many of the first builds here were collected by madewithlaya.com. Every entry links to its creator’s own post, repository or site.