All models
COMPARE
2 models, side by side.
What each one decides, how well calibrated it is, how fast it answers and what it costs to pull. Up to 4 at a time; the better value in each row is marked.
| Property | mapika/decider | alibi-serikbay/jevk5 |
|---|---|---|
| Summary | The most-downloaded open System One reproduction. Merged Qwen3.5 fine-tunes, trained on about 95 public decision sets, then calibration-aware RL and a rank-64 LoRA, that softmax letter logits at an answer slot. | Qwen3.5-4B with a merged rank-16 LoRA distilled from 17,408 teacher questions and 30,000 public training rows, read out as a temperature-scaled softmax over answer-letter logits. Also in 9B, 2B and GGUF. |
| Decides | choice, score, noul, classify, route | choice, score, noul, classify, route |
| Architecture | decider | jevk5 |
| Fine-tuned from | qwen/qwen3.5-2b-base | qwen/qwen3.5-4b |
| License | apache-2.0 | apache-2.0 |
| Availability | Open weights | Open weights |
| Hosted by | — | — |
| Input price | — | — |
| Decision accuracy | 80.2% | 78.4% |
| Calibration error | 0.038 | 0.035 |
| Valid action rate | — | — |
| Median latency | 3.2 ms | 13.2 ms |
| p95 latency | — | — |
| Evaluation suite | Decider 67-task regression set | JevBench public hard tier |
| Latest version | 2026.09 | 0.3.0 |
| Variants | — | SHA256SUMS |
| Size of latest version | 3.5 GB | 7.9 GB |
| Files | 4 | 8 |
| Downloads | 0 | 0 |
| Stars | 0 | 0 |
| Tags | system-one, qwen, rl, calibrated, 2b | system-one, qwen, lora, distilled, 4b |
| Updated | Sep 25, 2026 | Sep 25, 2026 |