Self-reported archetypes and behavioral failures in large language models
Read the original at arxiv.org→arXiv:2609.15998v1 Announce Type: new Abstract: Every large language model (LLM) has behavioral traits and moral preferences that comprise its character. Whether by design or as an emergent property of training,...
Original headline: "Self-reported archetypes and behavioral failures in Large Language Models"
Coverage timeline
- Sep 16, 04:00 UTC arXiv cs.CL lead source Self-reported archetypes and behavioral failures in Large Language Models