Skip to content
RCreddit.com·

First few days of qwen3.8-flash-next on 4x R9700 - it's been really interesting so far

AI summary

A user has been testing qwen3.8-flash-next on a system with 4x R9700, primarily for agentic coding use-cases. They report that the model, specifically the tcclaviger/Qwen3.8-Flash-Next-MXFP4-FP8 version, has significantly exceeded expectations in both speed and quality. Performance metrics include 3-5 concurrent streams each generating at ~100t/s, a single stream reaching 150+t/s, and prefill speeds of 10k+t/s.

Why this one

This report offers specific, high-throughput performance metrics for qwen3.8-flash-next on consumer-grade hardware, unlike most benchmarks that focus on enterprise systems.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

PublishedOffset at this time: UTC+0Sep 29, 2026, 01:52 UTC

IngestedOffset at this time: UTC+0Sep 29, 2026, 11:00 UTC

Published
Sep 29, 2026, 01:52
Ingested
Sep 29, 2026, 11:00
Source type
Dev community
Tier
Community
Source status
Healthy

Tier is a per-source editorial setting, not a per-item score.

Discussion trend

No comparison yet
Latest 24h versus previous 24h snapshot means · 7-day curve

The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.

Been testing with a variety of different agentic coding use-cases, mostly using a pi harness.

qwen3.8-flash-next has seriously exceeded my expectations (used https://huggingface.co/tcclaviger/Qwen3.8-Flash-Next-MXFP4-FP8)

Both speed and quality have surprised me, given that I can get 3-5 concurrent streams going with ~100t/s gen each, and single stream easily gets to 150+t/s. Prefill is 10k+t/s

Source·reddit.com