HARSHA SAI POTLURI. NEXT-GPT: Any-to-Any Multimodal LLM for Unified Text, Image, Audio, and Video Understanding. Acta Scientiae, [S. l.], v. 27, n. 2, p. 1–12, 2026. DOI: 10.22178/acta.27.2.1. Disponível em: https://www.periodicos.ulbra.org/index.php/acta/article/view/685. Acesso em: 20 sep. 2026.