[2506.13771] LittleBit: Ultra Low-Bit Quantization via Latent Factorization
Read the original at www.reddit.com→Interesting to see improvements and research into quantization aware training (QAT) that can make some really tiny models. submitted by /u/sn2006gy [link] [comments]
Coverage timeline
- Oct 8, 14:23 UTC r/LocalLLaMA lead source [2506.13771] LittleBit: Ultra Low-Bit Quantization via Latent Factorization