Rho, K., Lee, H., Iverson, V., & Chung, J. S. (2025). LAVCap: LLM-based Audio-Visual Captioning using Optimal Transport.
Chicago Style (17th ed.) CitationRho, Kyeongha, Hyeongkeun Lee, Valentio Iverson, and Joon Son Chung. LAVCap: LLM-based Audio-Visual Captioning Using Optimal Transport. 2025.
MLA (9th ed.) CitationRho, Kyeongha, et al. LAVCap: LLM-based Audio-Visual Captioning Using Optimal Transport. 2025.
Warning: These citations may not always be 100% accurate.