Valerio Capraro, Roberto Di Paolo, Veronica Pizziol. "A publicly available benchmark for assessing large language models’ ability to predict how humans balance self-interest and the interest of others." Scientific Reports, 2025. doi:10.1038/s41598-025-01715-7