Faking good and faking bad in LLMs: response distortion across dark triad traits
Read the original at arxiv.org→arXiv:2609.17534v1 Announce Type: new Abstract: Social desirability and impression management are pervasive sources of response distortion in human personality assessment, yet their effects on Large Language Models...
Original headline: "Faking Good and Faking Bad in LLMs: Response Distortion Across Dark Triad Personality Traits"
Coverage timeline
- Sep 17, 04:00 UTC arXiv cs.CL lead source Faking Good and Faking Bad in LLMs: Response Distortion Across Dark Triad Personality Traits