When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses Paper • 2607.26348 • Published Jul 28