Back
RCreddit.com
18
·1 days ago·RSS
Not on the current live radar

Qwen 3.8 Flash Next locally on simple mobile phone at 3.5 tok/s

View original
QwenModel release

Heat trend

↓ Cooling 17%
Latest 24h versus previous 24h · 7-day curve

The percentage is based on available heat signal, not comment count or independent people.

Why it matters

Qwen model activity is surfacing — worth tracking for capability changes, ecosystem impact, and availability.

AI summary

The Qwen 3.8 Flash Next model, which is 80GB, can now run locally on a mid-range Android phone with 12GB of RAM, achieving 3.5 tokens per second. This performance is attributed to optimizations and low quantization on the dense part of the model. This development demonstrates the feasibility of running such models on affordable mobile devices, specifically on phones costing $400–$500.