From one 3090 to 20 DGX Sparks for local LLMs
A user on r/LocalLLaMA recounts scaling local LLM hardware from a single RTX 3090 to 16-GPU rigs and finally multiple ASUS GB10 (DGX Spark) nodes, running a 397B model in FP8 with 20% speedup and posting the first 8-node Spark setup on NVIDIA's forum.