
First-person account of scaling a personal GPU cluster from a single RTX 3090 to 32 GPUs over two years — the practical decisions, challenges, and lessons for anyone building serious home AI/ML infrastructure.
“Two years later I have 32 RTX 3090s across 4 servers, 768GB of VRAM total, connected with InfiniBand and running on solar power.”
“once you process billions of tokens a month, running your own hardware is cheaper than paying for an API, and your data stays with you.”
“One thing I'd recommend: clone your SSD right after you get your switch. Make the clone first, then experiment, so you can always roll back.”
“The question I get most: roughly 30k just for the GPUs and server components.”
Checking sign-in…
Loading comments…