All models

COMPARE

3 models, side by side.

What each one decides, how well calibrated it is, how fast it answers and what it costs to pull. Up to 4 at a time; the better value in each row is marked.

Propertymapika/deciderbespoke-labs/bespoke-nimble-9balibi-serikbay/jevk5
SummaryThe most-downloaded open System One reproduction. Merged Qwen3.5 fine-tunes, trained on about 95 public decision sets, then calibration-aware RL and a rank-64 LoRA, that softmax letter logits at an answer slot.An open Jev-style LoRA on Qwen3.5-9B from Bespoke Labs, trained on 2,676 contrastively curated examples to score the allowed answer tokens directly for enums, booleans and rubric levels. Recipe, data and a public benchmark suite are released with it.Qwen3.5-4B with a merged rank-16 LoRA distilled from 17,408 teacher questions and 30,000 public training rows, read out as a temperature-scaled softmax over answer-letter logits. Also in 9B, 2B and GGUF.
Decideschoice, score, noul, classify, routechoice, noul, score, classify, routechoice, score, noul, classify, route
Architecturedecidernimblejevk5
Fine-tuned fromqwen/qwen3.5-2b-baseqwen/qwen3.5-9bqwen/qwen3.5-4b
Licenseapache-2.0apache-2.0apache-2.0
AvailabilityOpen weightsOpen weights + hosted APIOpen weights
Hosted by—Bespoke Labs—
Input price———
Decision accuracy80.2%90.1%78.4%
Calibration error0.0380.0540.035
Valid action rate———
Median latency3.2 ms106 ms13.2 ms
p95 latency———
Evaluation suiteDecider 67-task regression setBespoke held-out set (324 examples)JevBench public hard tier
Latest version2026.092026.090.3.0
Variants—SHA256SUMSSHA256SUMS
Size of latest version3.5 GB184.3 MB7.9 GB
Files4108
Downloads000
Stars000
Tagssystem-one, qwen, rl, calibrated, 2bsystem-one, qwen, lora, curated-data, 9bsystem-one, qwen, lora, distilled, 4b
UpdatedSep 25, 2026Sep 25, 2026Sep 25, 2026