Why Scoring AI's Human Simulations Like a Math Test Gets It Wrong
A Renmin University team says grading AI social-simulation models against one "correct" human answer is fundamentally flawed, proposing a subjectivity coefficient and soft-label training method instead.

