100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s
2.40T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1vqpc0f/100_worth_of_gpu_runs_qwen_38_27b_at_739_ts/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryA Reddit user on r/LocalLLaMA reports running a Qwen 27B-class model (titled 'qwen 3.8 27b') on approximately $100 of GPU hardware, achieving 7.39 tokens per second inference speed.
Why it mattersConcrete budget-hardware inference benchmark. Useful datapoint for anyone sizing a local LLM rig against a 27B model.
Cited by
No citations on record.
