Qwen2.5 7B Instruct abliterated v2
Fine-tune of Qwen2.5 7B Instruct
Measured axis deltas appear after both rows have board data.
Every measured variant of this model: how much quality each quant keeps, and the VRAM and speed it costs to run.
LB-2026-07.2 | 25/22.5/22.5/22.5/7.5
This chart appears after the model's first measured run lands. Use the quant ladder below for file size and VRAM requirements.
Complete rows are ordered by Local Intelligence Index; partial rows show their measured axes but are not ranked. The VRAM/Fits columns (8Kcontext) tell you what your card needs. Ranks are within this family's variants.
Variant — quant label plus, where sha-verified, the publisher of the exact weights file benchmarked; rows without a source line predate artifact identity records.
VRAM @8k — model weights, KV cache, and runtime headroom at 8k context.
Fits — the smallest common GPU VRAM tier above that estimate.
Prefill tok/s — prompt-processing speed.
Decode tok/s — generated-token speed after the prompt.
Overall tok/s — completion throughput across the full benchmark.
File size — the benchmarked model artifact on disk.
Runtime — the serving engine and version.
Run — the immutable benchmark receipt. Live rows link to the public submission record until the run is baked into the static site.
Pick the largest quant whose VRAM @8k fits your card.
Swipe horizontally for all variant metrics →
| Rank (this family) | Variant | Local Intelligence Indexindex-v3.0 | 40/15/15/10/15/5 | VRAM @8k | Fits | Overall tok/s | File size | Runtime | Run |
|---|---|---|---|---|---|---|---|---|
| — | Q8_0by mradermacher | no run yet | 9.5 GB | 12 GB | — | 8.1 GB | — | benchmark it |
| — | Q6_Kby mradermacher | no run yet | 7.7 GB | 8 GB | — | 6.3 GB | — | benchmark it |
| — | Q5_K_Mby mradermacher | no run yet | 6.8 GB | 8 GB | — | 5.4 GB | — | benchmark it |
| — | Q4_K_Mby mradermacher | no run yet | 6.1 GB | 8 GB | — | 4.7 GB | — | benchmark it |
| — | Q3_K_Mby mradermacher | no run yet | 5.2 GB | 8 GB | — | 3.8 GB | — | benchmark it |
| — | Q2_Kby mradermacher | no run yet | 4.4 GB | 8 GB | — | 3 GB | — | benchmark it |
vs base
Fine-tune of Qwen2.5 7B Instruct
Measured axis deltas appear after both rows have board data.