Hollow markers & dashed spokes: axis not yet scored
G
94.5
A
99.5
U
—
N
—
T
95.7
L
—
E
—
T
13.6
Specification
- Parameters
- 27B dense
- Architecture
- qwen3_5
- Size on disk
- 17.74 GB
- Quantization
- Q4_K_M
- Format
- GGUF
- Reasoning (CoT)
- Yes — emits reasoning tokens
- Internal ID
- M25a
- Mean speed
- 22.2 tok/s across suites
- Stall census
- 1 stall in 23 observed tests (4.3%)
- Reasoning appetite
- 2,350 tokens mean · 16,383 max
- Model card
- huggingface.co
Suite results
| Suite | Score | Avg / 20 | Tests | tok/s |
|---|---|---|---|---|
| General capability (13-task real-workload suite) 1a | 245.7 / 260 | 18.9 | 13/13 | 22 |
| Agentic tool-calling & protocol adherence 1b | 159.2 / 160 | 19.9 | 8/8 | 22.5 |
Each test is LLM-judged 0–20 against a fixed rubric; suite max = tests × 20. Where fewer tests ran than the suite total, unrun tests count as zero toward the axis score. See methodology.