Helpfulness Training Erodes Human Simulation | dailyai.report
23 stories from today
Model
91d ago
Helpfulness Training Erodes Human Simulation
A study of 26 million responses reveals that training LLMs to be helpful degrades their ability to mimic human behavior. This divergence worsens with each model generation. Even providing detailed demographic personas fails to improve individual predictions.
The Signal
Researchers conclude that the alignment process fundamentally alters how models simulate human decision-making, limiting their utility for behavioral science.