When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses
DGX agentarXiv:2607.26348v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human respondents whose simulated answers feed product, policy, and