Used over a million tokens in three sessions to test Qwen 3.6 35B MTP
Used over a million tokens in three separate sessions to test Qwen 3.6 35b (new Multi-token Prediction version)
A Reddit user tested Qwen3.6-35B-A3B MTP across three million-token-scale sessions, using 300k context and KV Q8_0, and reported about 1.5x the tok/sec of earlier tests.
Why it matters: HKR-H/K/R all pass: the million-token test is clickable, 300k context and KV Q8_0 add testable detail, and local speed maps to cost. Source is one Reddit post, so it stays below the high-importance band.