run-0149
performance
citable URL: https://halobench.com/records/run-0149/ — this address never moves; the anchor /records/#run-0149 keeps resolving
date 2026-08-12 · window 2026-08-12T21:28:38Z → 2026-08-12T21:31:58Z
suite llama-bench@10350 · config cfg-0036
cell depth 32,768 tokens (machine-written trailer)
Metrics
| metric | value |
|---|---|
| decode_tps | 51.020735 |
N=3 · stddev 0.022102 t/s — machine-written trailer, parsed at build time
Guard chain
waiver — No capability guard exists for Qwen3.6-35B on this backend/build/KV combination (run-0004's guard is a different suite and build; run-0005 guards the 122B, not this model). llama-bench measures throughput only and serves no endpoint for a guard to attach to. clm-0050/0051's methodology note records the closest available correctness check for these arms — 41/41 layers on GPU, CPU utilisation identical to stock, no stride bug — which is evidence against a gross correctness failure but is not a capability guard. Recorded as a gap.