When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses
最新研究揭示LLM模拟人类调查的局限性,跨领域基准测试发现合成用户存在系统性偏差,值得关注!
arXiv:2607.26348v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human responden…