← Leaderboard

Qwen3.8-27B MLX 8bit (lmstudio-community)

Tested Rank #39 of 39 · 2/8 GAUNTLET progress
Generalist — pending G Agentic — 87.5 / 100 A Understanding — pending U Needle — pending N Thinking — pending T Live — pending L Engineering — pending E Throughput — 6.6 / 100 T

Hollow markers & dashed spokes: axis not yet scored

G
A
87.5
U
N
T
n/d
L
E
T
6.6

Specification

Parameters
27B dense
Architecture
qwen3_5
Size on disk
28 GB
Quantization
8bit
Format
MLX
Class
Standard
Reasoning (CoT)
Yes — emits reasoning tokens
Internal ID
M25g
Mean speed
15.8 tok/s across suites
Model card
huggingface.co

Suite results

Suite Score Avg / 20 Tests tok/s Run
Agentic tool-calling & protocol adherence 1b 140 / 160 20 7/8 15.8 run 2026-08-20
Test Category Score tok/s
WA1 Tool Calling & Protocol Adherence 20/20 15.1
WA2 Tool Calling & Protocol Adherence 20/20 15.8
WA3 Tool Calling & Protocol Adherence 20/20 16.4
WA4 Tool Calling & Protocol Adherence 20/20 14.9
WA5 Tool Calling & Protocol Adherence 20/20 17.4
WB1 Agent Robustness & State Management 20/20 15.5
WB2 Agent Robustness & State Management 20/20 15.5

Category mean Tool Calling & Protocol Adherence 20Agent Robustness & State Management 20

Each test is LLM-judged 0–20 against a fixed rubric; suite max = tests × 20. Rows with a expand to the per-test breakdown. Where fewer tests ran than the suite total, unrun tests count as zero toward the axis score. See methodology.