From Weights to Words: Expressing and Editing Preference Model Inferences in Natural Language
Read the original at arxiv.org→arXiv:2607.16232v1 Announce Type: new Abstract: The growing use of statistical learning algorithms to infer human preferences from high-dimensional choice data runs up against a fundamental challenge: choice...
Coverage timeline
- Jul 21, 04:00 UTC arXiv cs.LG lead source From Weights to Words: Expressing and Editing Preference Model Inferences in Natural Language