Gemma 4 E4B IQ2_XXS tensor level allocation recovers reasoning performance from 28.9 to 69.5 at the same 3.3 GB budget
Read the original at old.reddit.com→iq2_xxs tensor level allocation recovered reasoning from 28.9 -> 69.5 at the same 3.3gb budget. https://huggingface.co/ByteOtter/gemma-4-E4B-it-CADA-IQ2_XXS I posted my Gemma 4 12B q3 result a couple days ago, where...
Original headline: "Gemma 4 E4B IQ2_XXS: + 140.54% Reasoning Performance From Tensor Level Quantization Allocation"
Coverage timeline
- Aug 15, 13:29 UTC r/LocalLLaMA lead source Gemma 4 E4B IQ2_XXS: + 140.54% Reasoning Performance From Tensor Level Quantization Allocation