B. ANKAYARKANNI; D. USHA NANDINI; P. SANGEETHA; R. AROUL CANESSANE; MARY LIVINSA Z. MULTI-MODAL TRANSFORMER ARCHITECTURE WITH CROSS-ATTENTION FUSION FOR ROBUST AUDIO-VISUAL SENTIMENT ANALYSIS. International Journal of Computer Information Systems and Industrial Management Applications, [S. l.], v. 18, n. 11s, p. 1101–1112, 2026. DOI: 10.70917/ijcisim-2026-3834. Disponível em: https://cspub-ijcisim.org/index.php/ijcisim/article/view/3834. Acesso em: 18 aug. 2026.