When synthetic users fail: a cross-domain benchmark of LLM-simulated human survey responses
Read the original at arxiv.org→arXiv:2607.26348v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human respondents whose simulated answers feed product, policy, and market...
Original headline: "When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses"
Coverage timeline
- Jul 30, 04:00 UTC arXiv cs.CL lead source When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses