RCreddit.com
18
·1 days ago·RSS
Not on the current live radar
Qwen 3.8 Flash Next locally on simple mobile phone at 3.5 tok/s
QwenModel release
Heat trend
↓ Cooling 17%
The percentage is based on available heat signal, not comment count or independent people.
Qwen model activity is surfacing — worth tracking for capability changes, ecosystem impact, and availability.
The Qwen 3.8 Flash Next model, which is 80GB, can now run locally on a mid-range Android phone with 12GB of RAM, achieving 3.5 tokens per second. This performance is attributed to optimizations and low quantization on the dense part of the model. This development demonstrates the feasibility of running such models on affordable mobile devices, specifically on phones costing $400–$500.