A 0.6B replica of the System One idea trained on game environments. A Qwen3-0.6B backbone with an attention-based Choice head that scores a dynamic candidate set for Maze, Snake, ViZDoom and position prediction, shipped with its full training pipeline.
ORGANIZATION
TianyuCodings
@tianyu-codings
Author of NanoJev, a 0.6B replica trained on game environments, published with its full training pipeline.
1model