A user has managed to run Kimi K3 on 80xRTX 5090, via 25GbE Ethernet.
2.55T2 sourcer/LocalLLaMA
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1v8hli2/a_user_has_managed_to_run_kimi_k3_on_80xrtx_5090/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryA user on r/LocalLLaMA reports running a Kimi model across 80 RTX 5090 GPUs interconnected via 25GbE Ethernet, demonstrating a large-scale consumer-GPU cluster for model inference.
Why it mattersDocuments a real 80-GPU RTX 5090 build with 25GbE fabric — a concrete data point for anyone evaluating Ethernet-based multi-node setups for MoE inference.
Cited by
No citations on record.
