AI revealed preferences: language models show stable task preferences in forced-choice experiments
Read the original at arxiv.org→arXiv:2608.26178v1 Announce Type: new Abstract: There is growing interest in whether language models have stable preferences, for technical, safety, and philosophical reasons. We test 20 language models and find a...
Original headline: "AI Revealed Preferences"
Coverage timeline
- Aug 28, 04:00 UTC arXiv cs.AI lead source AI Revealed Preferences