Invergent's decision model for text and images. A fine-tune of Gemma-4-26B-A4B (about 4B parameters active per token) that answers Choice, Noul and Score questions, with an optional thinking mode for harder questions.
meraGPT's hosted decision model (state-decider-1). Answers yes/no, choice and rubric-score questions over one state as calibrated distributions in a single pass, on the System One schema, so the typesafe-sdk works by changing its base URL.
Decides
choice, score, noul, classify, route
choice, score, noul, classify, route
Architecture
rune
state-decider
Fine-tuned from
google/gemma-4-26b-a4b-it
—
License
apache-2.0
proprietary
Availability
Open weights + hosted API
Hosted API
Hosted by
Invergent
meraGPT
Input price
—
$0.030/MTok
Decision accuracy
—
76.8%
Calibration error
—
—
Valid action rate
—
—
Median latency
—
—
p95 latency
—
—
Figures are from each model’s manifest; accuracy and latency are what the publishers report, on their own suites and hardware. Add a third model.
Questions
What is the difference between rune and state-decider-1?
rune is from Surogate (Invergent) and state-decider-1 from meraGPT. rune has open weights and a hosted API; state-decider-1 is only available as a hosted API. Both answer choice, score, noul, classify and route questions. rune reads up to 262K tokens of state, against 4K tokens for state-decider-1. rune is licensed apache-2.0; state-decider-1, proprietary.
Which is more accurate, rune or state-decider-1?
Only state-decider-1 publishes an accuracy figure (76.8% on typed-decisions benchmark (400 cases, 2,000 decisions; teacher-ensemble references), single run); rune does not, so there is no comparison to make without your own test.
Which is cheaper, rune or state-decider-1?
rune: Hosted, price not published, or free to self-host. state-decider-1: $0.03 / $0 per 1M. Open weights cost nothing per call beyond your own hardware.
Can I run rune or state-decider-1 locally?
rune yes — systemone pull surogate/rune downloads its weights. The other is only served as a hosted API.
Evaluation suite
—
typed-decisions benchmark (400 cases, 2,000 decisions; teacher-ensemble references), single run