Qwen3.8 Flash Next supports 1M context via YaRN, according to user discussion on model card page
Read the original at www.reddit.com→So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it...
Original headline: "Is anyone running Qwen3.8 Flash Next with a 1M context?"
Coverage timeline
- Oct 10, 17:49 UTC r/LocalLLaMA lead source Is anyone running Qwen3.8 Flash Next with a 1M context?