Did someone say MORE COMPUTE NEEDED? $NVDA Moonshot: "Since inference efficiency likewise benefits from larger high-bandwidth communication domains, we recommend deploying Kimi K3 on supernode configurations with 64 or more accelerators."
View on X ↗