Skip to content

Qwen3.6 35B A3B Uncensored HauhauCS Aggressive

Every measured variant of this model: how much quality each quant keeps, and the VRAM and speed it costs to run.

VRAM footprint vs Local Intelligence Index

LB-2026-07.2 | 25/22.5/22.5/22.5/7.5. Where this model and current-lane family runs land vs the frontier anchors.

Swipe horizontally to inspect the full chart →

Base modelDashed vertical lines mark common VRAM tiers.Amber points are synthetic demo preview data.

Variant profiles

Complete rows are ordered by Local Intelligence Index; partial rows show their measured axes but are not ranked. The VRAM/Fits columns (8Kcontext) tell you what your card needs. Ranks are within this family's variants.

Column guide

Variant — quant label plus, where sha-verified, the publisher of the exact weights file benchmarked; rows without a source line predate artifact identity records.

VRAM @8k — model weights, KV cache, and runtime headroom at 8k context.

Fits — the smallest common GPU VRAM tier above that estimate.

Prefill tok/s — prompt-processing speed.

Decode tok/s — generated-token speed after the prompt.

Overall tok/s — completion throughput across the full benchmark.

File size — the benchmarked model artifact on disk.

Runtime — the serving engine and version.

Run — the immutable benchmark receipt. Live rows link to the public submission record until the run is baked into the static site.

Pick the largest quant whose VRAM @8k fits your card.

Swipe horizontally for all variant metrics →

Rank (this family)VariantLocal Intelligence IndexLB-2026-07.2 | 25/22.5/22.5/22.5/7.5AgenticKnowledgeInstructionCodingMathVRAM @8kFitsPrefill tok/sDecode tok/sOverall tok/sFile sizeRuntimeRun
1
base modelQwen3.6 35B A3BUD-Q4_K_Mbestby unsloth
41.0±2.7
Agentic 6.3 / Knowledge 82.0 / Instruction 57.5 / Coding 28.4 / Math 22.3
6.3±5.4
82.0±4.2
57.5±5.8
28.4±7.8
22.3±7.2
22.1 GB footprintn/a6,098228190.8llama.cppb9852/fd1a05791receipt
no run yet23.9 GB24 GB21.2 GBbenchmark it

vs base

Fine-tune comparison

Qwen3.6 35B A3B Uncensored HauhauCS Aggressive

Fine-tune of Qwen3.6 35B A3B

composite n/acompare to base
fine-tune not yet benchmarked

Measured axis deltas appear after both rows have board data.