Vibeleaderboard
← All Intel
Intel / article

When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

Source
arxiv.org
Author
Zihan Chen, Di Zhu, Lei Nico Zheng
Date
Why it matters

personas over-determine demographics and lose to simple statistical baselines at the individual level — and a larger, more capable model does not close the gap.

Terms in this piece · Glossary
  • LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
Recommended reads
Comments

Checking sign-in…

Loading comments…