Danyal, M., Roman, M., Shahid, A., Yahya, M. (2026). Contextual image caption creation using object positional embedding and generative models. PLoS ONE. https://doi.org/10.1371/journal.pone.0353466