Ran DS V4-Flash-0731 Locally on 3xMI50 32GB @ ~15 t/s TG
1.95T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1vd51ey/ran_ds_v4flash0731_locally_on_3xmi50_32gb_15_ts_tg/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryUser ran the DS V4-Flash-0731 model locally across three AMD MI50 32GB GPUs and reported approximately 15 tokens per second in text generation.
Why it mattersConcrete throughput figure for an uncommon AMD MI50 multi-GPU setup, useful as a reference point for anyone weighing older datacenter cards for local inference.
Cited by
No citations on record.
