HEAD TO HEAD
AutoTrust AI Lab: jev-27b and Respan: span-01, compared on what they decide, where they run, what they cost and what their publishers report.
| Property | autotrust-ai/jev-27b | respan/span-01 |
|---|---|---|
| Summary | AutoTrust's student of TypeSafe Jev 1.13. A frozen Qwen3.8-27B plus a 108.9M-parameter decision block trained on Jev's own output distributions; one set of weights answers typed questions in one pass (System 1) or generates text with the untouched base (System 2). | Respan's behaviour-scoring model for evals, guardrails and monitoring. For each plain-language behaviour you define, it reads a conversation or agent trace and returns the probability the behaviour is present, absent or not observable, in one forward pass. |
| Decides | choice, score, noul | noul, classify |
| Architecture | blocks-of-experts | span |
| Fine-tuned from | qwen/qwen3.8-27b | — |
| License | apache-2.0 | proprietary |
| Availability | Open weights | Hosted API |
| Hosted by | — | Respan |
| Input price | — | $0.020/MTok |
| Decision accuracy | 88.7% | — |
| Calibration error | — | — |
| Valid action rate | — | — |
| Median latency | 137 ms | — |
| p95 latency | — | — |
| Evaluation suite | JevBench public set (231 items), family-macro score, maker's run | — |
| Latest version | 2026.09 | 2026.09 |
| Variants | adapter_vllm, adapter, reports | — |
| Size of latest version | 50.9 GB | — |
| Files | 42 | 0 |
| Downloads | 0 | 0 |
| Stars | 0 | 0 |
| Tags | system-one, qwen, distillation, vllm, dual-head, 27b | system-one, hosted, guardrails, evals |
| Updated | Sep 30, 2026 | Sep 30, 2026 |
Figures are from each model’s manifest; accuracy and latency are what the publishers report, on their own suites and hardware. Add a third model.
jev-27b is from AutoTrust AI Lab and span-01 from Respan. jev-27b has open weights you can download and run; span-01 is only available as a hosted API. Both answer noul questions. Only jev-27b answers choice and score. Only span-01 answers classify. jev-27b is licensed apache-2.0; span-01, proprietary.
Only jev-27b publishes an accuracy figure (88.7% on JevBench public set (231 items), family-macro score, maker's run); span-01 does not, so there is no comparison to make without your own test.
jev-27b: Free (open weights). span-01: $0.02 / $0 per 1M. Open weights cost nothing per call beyond your own hardware.
jev-27b yes — systemone pull autotrust-ai/jev-27b downloads its weights. The other is only served as a hosted API.