Bought a RTX 5090 to avoid API fees; built a mini datacenter with two RTX 6000 Pros as next-gen releases loom.
Read the original at old.reddit.com→I bought an RTX 5090 last year just to run 27B models natively. I even fine-tuned it with my own data using LoRA, building RAGs and was pretty damn happy with the results at first. But, Q8 quantization 130k context...
Original headline: "Bought a 5090 to escape API fees. Ended up building a mini datacenter. Sound familiar?"
Coverage timeline
- Jul 29, 23:16 UTC r/LocalLLaMA lead source Bought a 5090 to escape API fees. Ended up building a mini datacenter. Sound familiar?