Meta and University of Washington researchers allow small models to edit their own context on every turn; results show performance gains with smaller models like Qwen 3.5 9b and Qwen 3.6 27b
Read the original at www.reddit.com→In this paper by Meta and University of Washington researchers, they threw out compaction. Instead, they allowed the model to edit its own context window on every "turn," using its shell skills to avoid re-generating...
Original headline: "Allowing small models to edit their own context on every turn = big win?"
Coverage timeline
- Oct 6, 10:33 UTC r/LocalLLaMA lead source Allowing small models to edit their own context on every turn = big win?