Someone Ran Kimi K3 Across 80 RTX 5090s Linked by 25GbE Ethernet
r/LocalLLaMA·low signal
An r/LocalLLaMA post (109 upvotes, 52 comments) reports a working K3 deployment on 80x RTX 5090 connected over 25 gigabit Ethernet rather than NVLink or InfiniBand — roughly 2.5TB of aggregate VRAM against a ~594GB MXFP4 weight file, with the surplus absorbed by activation and KV overhead. The comment thread is where the value is: expert-parallel MoE over commodity Ethernet is exactly the regime vLLM's own blog warns about, where network bandwidth caps per-user output speed. Single-source and unaudited, but it is the first consumer-GPU K3 datapoint anyone has posted.