Inkling-Small by thinkingmachines: 276B total parameters, 12B active, 1M context window.
Read the original at old.reddit.com→276B total parameters, 12B active, 1M context window. Blog post: https://thinkingmachines.ai/news/inkling-small/ NVFP4: https://huggingface.co/thinkingmachines/Inkling-Small-NVFP4 GGUF's by Unsloth:...
Original headline: "Inkling-Small by thinkingmachines"
Coverage timeline
- Jul 30, 18:01 UTC r/LocalLLaMA lead source Inkling-Small by thinkingmachines
- Jul 30, 23:58 UTC r/LocalLLaMA Inkling-Small-276B-12B, effort "max" VS Qwen3.6-27B
- Jul 31, 21:46 UTC r/LocalLLaMA We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?
- Jul 31, 22:08 UTC r/LocalLLaMA Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?
- Aug 1, 02:29 UTC r/LocalLLaMA Me: Worn out from all the new model drops this week, but still hyped for all the great new releases.