Valerio Capraro, Roberto Di Paolo, Veronica Pizziol. "A publicly available benchmark for assessing large language models’ ability to predict how humans balance self-interest and the interest of others." Scientific Reports (2025). https://doi.org/10.1038/s41598-025-01715-7