LittleLearner: Language models trained on a U.S. elementary-school curriculum show limited generalization beyond the curriculum scope
Read the original at old.reddit.com→Modern LMs are trained on everything at once, so it is hard to tell whether a new skill was learned or merely elicited. We constrain the training distribution itself: an 88B-token corpus filtered to the U.S....
Original headline: "LittleLearner: Language Models Under Pedagogically-Controlled Knowledge Exposure"
Coverage timeline
- Aug 16, 09:12 UTC r/LocalLLaMA lead source LittleLearner: Language Models Under Pedagogically-Controlled Knowledge Exposure