APA (7th ed.) Citation

Li, M., Qu, F., Chen, Z., Su, N., Zhong, Z., Chen, Z., . . . Li, X. (2025). From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs.

Chicago Style (17th ed.) Citation

Li, Mingxiao, Fang Qu, Zhanpeng Chen, Na Su, Zhizhou Zhong, Ziyang Chen, Nan Du, and Xiaolong Li. From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs. 2025.

MLA (9th ed.) Citation

Li, Mingxiao, et al. From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs. 2025.

Warning: These citations may not always be 100% accurate.