Shen, Y., Fu, C., Dong, S., Wang, X., Zhang, Y., Chen, P., . . . Sun, X. (2025). Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy.
Chicago Style (17th ed.) CitationShen, Yunhang, et al. Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy. 2025.
MLA (9th ed.) CitationShen, Yunhang, et al. Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy. 2025.
Warning: These citations may not always be 100% accurate.