← Leaderboard

Qwen3.8-27B Q4_K_M GGUF

tested-offloaded Rank #14 of 33 · 4/8 GAUNTLET progress
Generalist — 94.5 / 100 G Agentic — 99.5 / 100 A Understanding — pending U Needle — pending N Thinking — 95.7 / 100 T Live — pending L Engineering — pending E Throughput — 13.6 / 100 T

Hollow markers & dashed spokes: axis not yet scored

G
94.5
A
99.5
U
N
T
95.7
L
E
T
13.6

Specification

Parameters
27B dense
Architecture
qwen3_5
Size on disk
17.74 GB
Quantization
Q4_K_M
Format
GGUF
Reasoning (CoT)
Yes — emits reasoning tokens
Internal ID
M25a
Mean speed
22.2 tok/s across suites
Stall census
1 stall in 23 observed tests (4.3%)
Reasoning appetite
2,350 tokens mean · 16,383 max
Model card
huggingface.co

Suite results

Suite Score Avg / 20 Tests tok/s
General capability (13-task real-workload suite) 1a 245.7 / 260 18.9 13/13 22
Agentic tool-calling & protocol adherence 1b 159.2 / 160 19.9 8/8 22.5

Each test is LLM-judged 0–20 against a fixed rubric; suite max = tests × 20. Where fewer tests ran than the suite total, unrun tests count as zero toward the axis score. See methodology.