First few days of qwen3.8-flash-next on 4x R9700 - it's been really interesting so far
A user has been testing qwen3.8-flash-next on a system with 4x R9700, primarily for agentic coding use-cases. They report that the model, specifically the tcclaviger/Qwen3.8-Flash-Next-MXFP4-FP8 version, has significantly exceeded expectations in both speed and quality. Performance metrics include 3-5 concurrent streams each generating at ~100t/s, a single stream reaching 150+t/s, and prefill speeds of 10k+t/s.
This report offers specific, high-throughput performance metrics for qwen3.8-flash-next on consumer-grade hardware, unlike most benchmarks that focus on enterprise systems.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
PublishedOffset at this time: UTC+0Sep 29, 2026, 01:52 UTC
IngestedOffset at this time: UTC+0Sep 29, 2026, 11:00 UTC
- Published
- Sep 29, 2026, 01:52
- Ingested
- Sep 29, 2026, 11:00
- Source type
- Dev community
- Tier
- Community
- Source status
- Healthy
Tier is a per-source editorial setting, not a per-item score.
Discussion trend
The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.
Been testing with a variety of different agentic coding use-cases, mostly using a pi harness.
qwen3.8-flash-next has seriously exceeded my expectations (used https://huggingface.co/tcclaviger/Qwen3.8-Flash-Next-MXFP4-FP8)
Both speed and quality have surprised me, given that I can get 3-5 concurrent streams going with ~100t/s gen each, and single stream easily gets to 150+t/s. Prefill is 10k+t/s