Loading comparison…
COMPARE
What each one decides, how well calibrated it is, how fast it answers and what it costs to pull. Up to 4 at a time; the better value in each row is marked.
| Property | bespoke-labs/bespoke-nimble-9b | mapika/decider |
|---|---|---|
| Summary | An open Jev-style LoRA on Qwen3.5-9B from Bespoke Labs, trained on 2,676 contrastively curated examples to score the allowed answer tokens directly for enums, booleans and rubric levels. Recipe, data and a public benchmark suite are released with it. | The most-downloaded open System One reproduction. Merged Qwen3.5 fine-tunes, trained on about 95 public decision sets, then calibration-aware RL and a rank-64 LoRA, that softmax letter logits at an answer slot. |
| Decides | choice, noul, score, classify, route | choice, score, noul, classify, route |
| Architecture | nimble | decider |
| Fine-tuned from | qwen/qwen3.5-9b | qwen/qwen3.5-2b-base |
| License | apache-2.0 | apache-2.0 |
| Availability | Open weights + hosted API | Open weights |
| Hosted by | Bespoke Labs | — |
| Input price | — | — |
| Decision accuracy | 90.1% | 80.2% |
| Calibration error | 0.054 | 0.038 |
| Valid action rate | — | — |
| Median latency | 106 ms | 3.2 ms |
| p95 latency | — | — |
| Evaluation suite | Bespoke held-out set (324 examples) | Decider 67-task regression set |
| Latest version | 2026.09 | 2026.09 |
| Variants | SHA256SUMS | — |
| Size of latest version | 184.3 MB | 3.5 GB |
| Files | 10 | 4 |
| Downloads | 0 | 0 |
| Stars | 0 | 0 |
| Tags | system-one, qwen, lora, curated-data, 9b | system-one, qwen, rl, calibrated, 2b |
| Updated | Sep 26, 2026 | Sep 26, 2026 |