Minh, L. T., Thinh, N. D., Loc, N. K. T., Quan, L. V., Tam, N. D., Son, L. H. (2026). ViSQA: A benchmark dataset and baseline models for Vietnamese spoken question answering. PLoS ONE. https://doi.org/10.1371/journal.pone.0340771