Capraro, V., Paolo, R. D., Pizziol, V. (2025). A publicly available benchmark for assessing large language models’ ability to predict how humans balance self-interest and the interest of others. Scientific Reports. https://doi.org/10.1038/s41598-025-01715-7