Beyond right and wrong: evaluating second-order social reasoning in large language models
Read the original at arxiv.org→arXiv:2609.05437v1 Announce Type: new Abstract: Previous AI alignment efforts have focused primarily on first-order social norms -- teaching models what is socially acceptable or unacceptable (e.g., `do not steal')....
Original headline: "Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models"
Coverage timeline
- Sep 9, 04:00 UTC arXiv cs.AI lead source Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models