V, Damodaran; M, Balasubramanian; M G, Jibukumar; I S, Smitha. A Hybrid Vision–Language Framework For Unified Image–Video Captioning And Semantic Visual Question Answering. International Journal of Artificial Intelligence and Machine Learning, [S. l.], v. 6, n. 7s, p. 790–807, 2026. Disponível em: https://svedbergopen.com/index.php/ijaiml/article/view/1125. Acesso em: 24 aug. 2026.