B. Ankayarkanni, D. Usha Nandini, P. Sangeetha, R. Aroul Canessane, and Mary Livinsa Z. “MULTI-MODAL TRANSFORMER ARCHITECTURE WITH CROSS-ATTENTION FUSION FOR ROBUST AUDIO-VISUAL SENTIMENT ANALYSIS”. International Journal of Computer Information Systems and Industrial Management Applications 18, no. 11s (July 28, 2026): 1101–1112. Accessed August 18, 2026. https://cspub-ijcisim.org/index.php/ijcisim/article/view/3834.