More params, less size? could larger models fit on smaller GPUs in a few years, such as a 30B+ model running on 16GB VRAM
Read the original at old.reddit.com→Is it likely that in a few years we'll have bigger models in sizes that may fit well within smaller GPUs? i.e., a 30B+ model running fast on 16GB VRAM, or even more than that. Among the clash of interests in the AI...
Original headline: "More params, less size?"
Coverage timeline
- Aug 19, 05:34 UTC r/LocalLLaMA lead source More params, less size?