Skip to content

Qwen3.6 27B MTP pi tune

Every measured variant of this model: how much quality each quant keeps, and the VRAM and speed it costs to run.

VRAM footprint vs Local Intelligence Index

LB-2026-07.2 | 25/22.5/22.5/22.5/7.5. Where this model and current-lane family runs land vs the frontier anchors.

Swipe horizontally to inspect the full chart →

Base modelDashed vertical lines mark common VRAM tiers.Amber points are synthetic demo preview data.

Variant profiles

Complete rows are ordered by Local Intelligence Index; partial rows show their measured axes but are not ranked. The VRAM/Fits columns (8Kcontext) tell you what your card needs. Ranks are within this family's variants.

Column guide

Variant — quant label plus, where sha-verified, the publisher of the exact weights file benchmarked; rows without a source line predate artifact identity records.

VRAM @8k — model weights, KV cache, and runtime headroom at 8k context.

Fits — the smallest common GPU VRAM tier above that estimate.

Prefill tok/s — prompt-processing speed.

Decode tok/s — generated-token speed after the prompt.

Overall tok/s — completion throughput across the full benchmark.

File size — the benchmarked model artifact on disk.

Runtime — the serving engine and version.

Run — the immutable benchmark receipt. Live rows link to the public submission record until the run is baked into the static site.

Pick the largest quant whose VRAM @8k fits your card.

Swipe horizontally for all variant metrics →

Rank (this family)VariantLocal Intelligence IndexLB-2026-07.2 | 25/22.5/22.5/22.5/7.5AgenticKnowledgeInstructionCodingMathVRAM @8kFitsPrefill tok/sDecode tok/sOverall tok/sFile sizeRuntimeRun
1
43.2±2.9
Agentic 8.3 / Knowledge 83.4 / Instruction 65.0 / Coding 25.5 / Math 26.6
8.3±7.4
83.4±3.9
65.0±5.4
25.5±7.1
26.6±7.2
19.5 GBn/a3,112.174.469.417.1 GBllama.cppb9852/fd1a05791receipt
no run yet31.4 GB32 GB29 GBbenchmark it
no run yet24.8 GB32 GB22.4 GBbenchmark it
Q5_K_Mby bytkim
no run yet22 GB24 GB19.6 GBbenchmark it
Q4_K_Mby bytkim
no run yet19.2 GB24 GB16.8 GBbenchmark it
Q3_K_Mby bytkim
no run yet15.9 GB16 GB13.5 GBbenchmark it
no run yet13.3 GB16 GB10.9 GBbenchmark it

vs base

Fine-tune comparison

Qwen3.6 27B MTP pi tune

Fine-tune of Qwen3.6 27B

composite n/acompare to base
fine-tune not yet benchmarked

Measured axis deltas appear after both rows have board data.