Open labs embrace continued post-training on existing models to achieve performance gains without retraining new base models
Read the original at old.reddit.com→A few years ago, every generation of model releases came entirely as new models, such as Qwen, Qwen 1.5, 2, 2.5, 3, and 3.5, llama, llama 2, llama 3, etc. But with the events of the recent few days, I think we can...
Original headline: "Open labs are finally embracing the power of continued post training"
Coverage timeline
- Aug 14, 15:26 UTC r/LocalLLaMA lead source Open labs are finally embracing the power of continued post training