SFT-as-context mitigates forgetting in supervised fine-tuning
Read the original at arxiv.org→arXiv:2610.11132v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) equips large language models (LLMs) with specialized capabilities, but often comes at the cost of forgetting the general capabilities of...
Original headline: "SFT-as-Context Mitigates Forgetting in Supervised Fine-Tuning"
Coverage timeline
- Oct 9, 04:00 UTC arXiv cs.CL lead source SFT-as-Context Mitigates Forgetting in Supervised Fine-Tuning