Liu, H., Li, Y., Wang, Z., Zhang, S., Zhao, Z., Bo, Z., . . . He, K. (2026). ITO: Images and Texts as One via Synergizing Multiple Alignment and Training-Time Fusion.
Chicago Style (17th ed.) CitationLiu, Hanpeng, Yaqian Li, Zidan Wang, Shuoxi Zhang, Zonglin Zhao, Zihao Bo, Rinyoichi Takezoe, Kaiwen Long, and Kun He. ITO: Images and Texts as One via Synergizing Multiple Alignment and Training-Time Fusion. 2026.
MLA (9th ed.) CitationLiu, Hanpeng, et al. ITO: Images and Texts as One via Synergizing Multiple Alignment and Training-Time Fusion. 2026.
Warning: These citations may not always be 100% accurate.