1BIT Qwen 3.8 2.4T a95b: user reports 508 GB model usage, 50pp and 9.6 tgen with latency drop at 50k tokens; notes Unslloth Studio on Mac Studio Ultra performing.
Read the original at old.reddit.com→Processing img az99qopcg8jh1... So same as my prior post 1bit test... although this 1bit is a bit interesting you can read on unlsoth blog https://unsloth.ai/docs/models/qwen3.8 508 gigs being used I am using unsloth...
Original headline: "1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)"
Coverage timeline
- Aug 14, 01:54 UTC r/LocalLLaMA lead source 1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)