Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro
2.25T2 sourcer/LocalLLaMA
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1wk9hze/qwen3827b_at_144_toks_on_an_m5_max_macbook_pro/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryA r/LocalLLaMA post reports a Qwen3.8-27B model running at 144 tokens per second on an M5 Max MacBook Pro. No benchmark methodology, hardware config, or context details are provided in the excerpt.
Why it mattersM5 Max local-LLM throughput numbers are still scarce; this gives a concrete data point for readers weighing Apple Silicon for on-device inference.
Cited by
No citations on record.
