Jun Qin, Shenwei Chen, Zheng Ye, Jing Liu, Zhou Liu. "Video swin-CLSTM transformer: Enhancing human action recognition with optical flow and long-term dependencies." PLoS ONE (2025). https://doi.org/10.1371/journal.pone.0327717