Qwen3.8-Flash-Next Is Posting Wins Over DeepSeek V4 Pro While llama.cpp Support Is Still an Unmerged PR
r/LocalLLaMA·medium signal
A 235-upvote r/LocalLLaMA gallery posted 2026-08-27 shows Qwen3.8-Flash-Next beating DeepSeek V4 Pro on comparison benchmarks, one day after release. The practitioner detail in the comments is the gap between benchmark and usable: one user compiled danielhanchen's PR and reports it 'rock solid' but with MTP not working and KV cache scaling oddities. Anyone planning to run this locally this week is compiling a PR, not pulling a release.