Summary
Decider V1.1 27B scores 2.6 on AI BENCHY and ranks #395. It has 10.0 reliability, a 17.4% pass rate, $0.002 total cost, and 320ms average response time.
What makes Decider V1.1 27B unique: It is notably fast compared with similar models.
Model facts
Researched on 2026-10-07
- Parameters
- 27B
- Architecture
- Dense
- Availability
- Open source
- License
- Apache-2.0
The exact v1.1 model card identifies the Qwen3.8-27B dense backbone, Apache-2.0 weights, a separate decision readout, and the supplied inference implementation. 27B is the reported nominal model size; the model-card tensor summary rounds to 26B. OpenRouter serves this exact decision checkpoint with typed choices and probabilities, without generated text or exposed reasoning controls.
2.6
Consistency
5.2
10.0
$0.002
Total Output Tokens
69
Total Input Tokens
73,557
Input Price
$0.020 / 1M
Output Price
$0.000 / 1M
Cache Read Price
N/A
Cache Write Price
N/A
Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.
Wrong Tests: 8
Attempt pass rate: 17.4%
Benchmark coverage: 36/69 attempts. Choices supplied by adapters. Supports 12/23 tests. Other tests are unsupported, not failed. Score uses the full suite.
Flaky tests
0
Flaky tests had mixed outcomes across runs (at least one pass and one fail).
Charts
Choose the first model, then click a second model to open a side-by-side page.
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Quick Compare
Category Breakdown
| Category | Score | Consistency | Tests Correct |
|---|---|---|---|
| Agentic | 0.0 | 0.0 | |
| Anti-AI Tricks | 4.0 | 7.5 | |
| Coding | 4.9 | 6.7 | |
| Combined | 0.0 | 0.0 | |
| Data parsing and extraction | 6.5 | 10.0 | |
| Domain specific | 5.3 | 10.0 | |
| General Intelligence | 0.0 | 0.0 | |
| Instructions following | 1.5 | 5.0 | |
| Puzzle Solving | 1.7 | 3.3 | |
| Tool Calling | 0.0 | 0.0 | |
| Trivia | 0.0 | 0.0 |