Qwen 3.8 Flash Next locally on simple mobile phone at 3.5 tok/s
1.95T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1w2nz07/qwen_38_flash_next_locally_on_simple_mobile_phone/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryReport of running Qwen 3.8 Flash locally on a basic mobile phone, achieving 3.5 tokens per second inference.
Why it mattersConcrete on-device inference benchmark on minimal mobile hardware; useful baseline for anyone evaluating local LLM feasibility on phones.
Cited by
No citations on record.
