DeepSWE Preview
Fine-tune of Qwen3 32B
Measured axis deltas appear after both rows have board data.
Every measured variant of this model: how much quality each quant keeps, and the VRAM and speed it costs to run.
LB-2026-07.2 | 25/22.5/22.5/22.5/7.5
This chart appears after the model's first measured run lands. Use the quant ladder below for file size and VRAM requirements.
Complete rows are ordered by Local Intelligence Index; partial rows show their measured axes but are not ranked. The VRAM/Fits columns (8Kcontext) tell you what your card needs. Ranks are within this family's variants.
Variant — quant label plus, where sha-verified, the publisher of the exact weights file benchmarked; rows without a source line predate artifact identity records.
VRAM @8k — model weights, KV cache, and runtime headroom at 8k context.
Fits — the smallest common GPU VRAM tier above that estimate.
Prefill tok/s — prompt-processing speed.
Decode tok/s — generated-token speed after the prompt.
Overall tok/s — completion throughput across the full benchmark.
File size — the benchmarked model artifact on disk.
Runtime — the serving engine and version.
Run — the immutable benchmark receipt. Live rows link to the public submission record until the run is baked into the static site.
Pick the largest quant whose VRAM @8k fits your card.
Swipe horizontally for all variant metrics →
| Rank (this family) | Variant | Local Intelligence Indexindex-v3.0 | 40/15/15/10/15/5 | VRAM @8k | Fits | Overall tok/s | File size | Runtime | Run |
|---|---|---|---|---|---|---|---|---|
| — | Q8_0by bartowski | no run yet | 37.4 GB | 48 GB | — | 34.8 GB | — | benchmark it |
| — | Q6_Kby bartowski | no run yet | 29.5 GB | 32 GB | — | 26.9 GB | — | benchmark it |
| — | Q5_K_Mby bartowski | no run yet | 25.8 GB | 32 GB | — | 23.2 GB | — | benchmark it |
| — | Q4_K_Mby bartowski | no run yet | 22.4 GB | 24 GB | — | 19.8 GB | — | benchmark it |
| — | Q3_K_Mby bartowski | no run yet | 18.6 GB | 24 GB | — | 16 GB | — | benchmark it |
| — | Q2_Kby bartowski | no run yet | 14.9 GB | 16 GB | — | 12.3 GB | — | benchmark it |
vs base
Fine-tune of Qwen3 32B
Measured axis deltas appear after both rows have board data.