Skip to content

This receipt lets anyone verify this run against the frozen suite — see Methodology.

← back to Qwopus3.6 27B v2 MTP

suite-v2 | index-v4.2

Qwopus3.6 27B v2 MTP

qwopus3-6-27b-v2-mtp__qwopus3-6-27b-v2-mtp-q4km-s2v5

Local Intelligence Index
LB-2026-07.2 | 25/22.5/22.5/22.5/7.5
42.1
±3.0 95% CI
Profile: Agentic / Knowledge / Instruction / Coding / Math
Agentic 9.4 / Knowledge 80.4 / Instruction 57.8 / Coding 29.8 / Math 25.9
Weighted headline profile: Agentic 25%, Knowledge 22.5%, Instruction 22.5%, Coding 22.5%, Math 7.5%.
Total run time
23.6 h
Data quality note: this run has 0 error(s) and 49 no-answer item(s).
Data warnings
  • coding chance_corrected differs from item-derived mean; stale/inconsistent run JSON

Axis breakdown

Agenticworst axis
n=96 · errors=0 · no answer=0
9.4 ±8.0
Knowledge
n=400 · errors=0 · no answer=12
80.4 ±4.2
Instruction
n=294 · errors=0 · no answer=0
57.8 ±5.8
Coding
n=141 · errors=0 · no answer=0
29.8 ±7.8
Math
n=139 · errors=0 · no answer=37
25.9 ±7.2

Diagnostics · unweighted

Call formatting
n=330 · errors=0 · no answer=1
73.9 ±4.8
BFCL single-turn
— not measured
BFCL v3 multi-turn base — frozen snapshot
n=50 · errors=0 · no answer=0
24.0 ±12.0
BFCL multi-turn long-context
— not measured
RULER 32K
— not measured

Instruction-following decomposition

Strict = Termination × Conditional
Strict accuracy
57.8%
correct AND terminated / all
Termination rate
93.9%
terminated / all
Conditional accuracy
61.6%
correct / terminated

Outputs that hit the answer-token cap are counted incorrect; this prevents non-terminating generations from getting credit for matching required tokens inside a runaway response.

Manifest

model
qwen35
quant
Q4_K_M
runtime
llama.cpp b9852/fd1a05791 · KV k=f16,v=f16 · ctx 32,768
hardware
NVIDIA GeForce RTX 5090 (31.8 GB) · Windows-11-10.0.26200-SP0
os
Windows-11-10.0.26200-SP0
lane
bounded-final-v2
thinking_mode
n/a
caps
max_tokens_math: 0, max_tokens_mcq: 16384, thinking_budget: 8192
sampling
temp 0 | top_p n/a | top_k 1 | min_p n/a | seed 1234 | effort n/a | knowledge: max 16384 | instruction: max 16384 | coding: max 16384 | math: max 16384 | bfcl_multi_turn_base: max 16384 | tc_json_v1: max 16384
tokens
5,690,781 prompt / 5,429,223 completion / 11,120,004 total
tokens-to-answer
1,623 median / 16,384 p95
tok/s
63.8
total run time
23.6 h
est cost
n/a
n_items
1,457
n_errors
0
n_no_answer
49

Serving performance

NVIDIA GeForce RTX 5090 (31.8 GB) · Windows-11-10.0.26200-SP0

prefill
2,729.1 tok/s
decode
71 tok/s
TTFT proxy
396.1 ms
prompt processing before first token — non-streaming harness, lower bound
coverage
93.4%
prompt median / p95
396.1 ms / 1,874.1 ms
predicted median / p95
22,347.4 ms / 216,472.4 ms
benchprefilldecodeprompt mediann
amo1,250.6 tok/s66.8 tok/s197.1 ms39
bfcl_multi_turn_base3,324.4 tok/s72.5 tok/s1,976.4 ms50
bigcodebench_hard2,711.9 tok/s66.6 tok/s209.7 ms148
ifbench2,881.6 tok/s76.4 tok/s381.7 ms294
mmlu_pro2,582.8 tok/s69.5 tok/s362.4 ms400
olymmath_hard945.7 tok/s71.6 tok/s161.4 ms100
tc_json_v12,657.5 tok/s68.7 tok/s471.4 ms330

Source: llama.cpp server timings.

Provenance

suite_version: suite-v2

index_version: index-v4.2

source scorecard

version: 6

id: d7ba660a32a9abf9a3e6ef469f3e33e2c6b655de9331e3af7a4fed493fd0491c

Provenance drift: this receipt preserves its original scorecard metadata; the site projection is rendered under index-v4.2.

bfcl.jsonl26d990d589db8a8b2a70b23c592ea0aff9df8287d9509a2fadd98d1b72661e17
bfcl_multi_turn_base.jsonl41e5691e87d5c46f12ea19e2bcfe45cb019d83b279c6f1a640bf7ab203c3eec2
bfcl_multi_turn_long_context.jsonlc6f642197ab070573fc353bb7265e7efc813addc27eb59b66358830b68b6a1b2
coding.jsonl33635febb89ab6cb8f06e139bc33932ada89d90e32ce03820ad7f15712e19b8e
instruction.jsonl40dc0b3e14270d61e9deae13f30f70f04d1d65a304340a7b6fe29cf4a5c51257
knowledge.jsonl129b8d9726eab3676ca30d58fac23af4e07407eb537b9bfa10d4d24434b26ba4
math-part2.jsonl8126598901f0e2be27b2a4fed97fded7b2c43aa37ca3ecb580527ad11a15e53b
math.jsonl98e79f1da84680345224f48fc7d1ed8b220e76cfd0525da1c494633d1abd1904
ruler_32k.jsonl0bede1810663a7164e68f3008248d78ba247fb677440f06b3e1c63b8781b0540
tc_json_v1.jsonl571b3c4064b523174900883c786df4fdbb6c2a8924a148620a167415d67afd74