Wang, Y., Zhou, W., Feng, H., & Li, H. (2024). AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding.
Chicago Style (17th ed.) CitationWang, Yonghui, Wengang Zhou, Hao Feng, and Houqiang Li. AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding. 2024.
MLA (9th ed.) CitationWang, Yonghui, et al. AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding. 2024.
Warning: These citations may not always be 100% accurate.