Issue 2026-09-29 · Industry · 开源 · 研究
User tests qwen3.8-flash-next on 4x R9700, beats expectations
A Reddit user shared results running qwen3.8-flash-next on four AMD R9700 GPUs with an MXFP4-FP8 quantized build for agentic coding tasks. Speed and quality exceeded expectations: single-stream generation above 150 tokens/s and prefill above 10,000 tokens/s.
Read original ↗