Unsupervised post-training of foundation models: a survey
Read the original at arxiv.org→arXiv:2608.24982v1 Announce Type: new Abstract: Foundation-model post-training usually relies on human labels, preference data, stronger teachers, or executable verifiers. We study Unsupervised Post-Training (UPT):...
Original headline: "Unsupervised Post-Training of Foundation Models: A Survey"
Coverage timeline
- Aug 27, 04:00 UTC arXiv cs.CL lead source Unsupervised Post-Training of Foundation Models: A Survey