−2 logit bias on Bonsai 2 27B reduced accuracy from 44/50 to 43/50 on MATH-500; tokens slightly increased in run B
Read the original at www.reddit.com→A recent post here reported that a −2 logit bias on "wait", "maybe", and "perhaps" made Qwen3.5-4B more accurate and shorter on 50 MATH-500 questions. I tried it on Ternary Bonsai 2 27B (PTQ1_0, 5.53 GiB) on an RTX...
Original headline: "−2 logit bias on Bonsai 2 27B: 44/50 → 43/50 on MATH-500, +3% tokens"
Coverage timeline
- Oct 2, 18:13 UTC r/LocalLLaMA lead source −2 logit bias on Bonsai 2 27B: 44/50 → 43/50 on MATH-500, +3% tokens