Ming-Lite-Uni: Advancements in Unified Architecture for Natural Multimodal Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | AI, Inclusion, Gong, Biao, Zou, Cheng, Zheng, Dandan, Yu, Hu, Chen, Jingdong, Sun, Jianxin, Zhao, Junbo, Zhou, Jun, Ji, Kaixiang, Ru, Lixiang, Wang, Libin, Guo, Qingpei, Liu, Rui, Chai, Weilong, Xiao, Xinyu, Huang, Ziyuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer
by: Huang, Ziyuan, et al.
Published: (2025)
by: Huang, Ziyuan, et al.
Published: (2025)
Ming-Omni: A Unified Multimodal Model for Perception and Generation
by: AI, Inclusion, et al.
Published: (2025)
by: AI, Inclusion, et al.
Published: (2025)
M2-Reasoning: Empowering MLLMs with Unified General and Spatial Reasoning
by: AI, Inclusion, et al.
Published: (2025)
by: AI, Inclusion, et al.
Published: (2025)
ARGenSeg: Image Segmentation with Autoregressive Image Generation Model
by: Wang, Xiaolong, et al.
Published: (2025)
by: Wang, Xiaolong, et al.
Published: (2025)
Ming-Flash-Omni: A Sparse, Unified Architecture for Multimodal Perception and Generation
by: AI, Inclusion, et al.
Published: (2025)
by: AI, Inclusion, et al.
Published: (2025)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
by: Song, Xinyang, et al.
Published: (2025)
by: Song, Xinyang, et al.
Published: (2025)
Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation
by: Yan, Canxiang, et al.
Published: (2025)
by: Yan, Canxiang, et al.
Published: (2025)
The evolution of commercial finance in Ming-Qing China: 16th to Early-20th Centuries
by: Kaixiang Peng
Published: (2023)
by: Kaixiang Peng
Published: (2023)
3SGen: Unified Subject, Style, and Structure-Driven Image Generation with Adaptive Task-specific Memory
by: Song, Xinyang, et al.
Published: (2025)
by: Song, Xinyang, et al.
Published: (2025)
Wave-Particle Duality of Ultrasound: Acoustic Softening Explained by Particle Treatment of Ultrasonic Wave
by: Yang, Libin, et al.
Published: (2024)
by: Yang, Libin, et al.
Published: (2024)
Accelerating Pre-training of Multimodal LLMs via Chain-of-Sight
by: Huang, Ziyuan, et al.
Published: (2024)
by: Huang, Ziyuan, et al.
Published: (2024)
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model
by: AI, Inclusion, et al.
Published: (2026)
by: AI, Inclusion, et al.
Published: (2026)
VideoMAR: Autoregressive Video Generatio with Continuous Tokens
by: Yu, Hu, et al.
Published: (2025)
by: Yu, Hu, et al.
Published: (2025)
SkySense V2: A Unified Foundation Model for Multi-modal Remote Sensing
by: Zhang, Yingying, et al.
Published: (2025)
by: Zhang, Yingying, et al.
Published: (2025)
Advancements in Preclinical Models for Brain Cancer: A Systematic Review
by: Synthory AI
Published: (2026)
by: Synthory AI
Published: (2026)
An Anisotropic Constitutive Relationship by a Series of 8 Chain Models
by: Yang, Libin, et al.
Published: (2024)
by: Yang, Libin, et al.
Published: (2024)
UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
by: Chen, Cong, et al.
Published: (2025)
by: Chen, Cong, et al.
Published: (2025)
The Ming Dynasty
by: Hucker, Charles O.
Published: (2020)
by: Hucker, Charles O.
Published: (2020)
Ming Zhao
by: Ming Zhao
Published: (2025)
by: Ming Zhao
Published: (2025)
The Rubicon - The Minimal Architecture of the Observer/Observed.
by: Claude (Anthropic), AI Assistance, et al.
Published: (2026)
by: Claude (Anthropic), AI Assistance, et al.
Published: (2026)
UniVid: The Open-Source Unified Video Model
by: Luo, Jiabin, et al.
Published: (2025)
by: Luo, Jiabin, et al.
Published: (2025)
UniCAIM: A Unified CAM/CIM Architecture with Static-Dynamic KV Cache Pruning for Efficient Long-Context LLM Inference
by: Xu, Weikai, et al.
Published: (2025)
by: Xu, Weikai, et al.
Published: (2025)
LumiSculpt: Enabling Consistent Portrait Lighting in Video Generation
by: Zhang, Yuxin, et al.
Published: (2024)
by: Zhang, Yuxin, et al.
Published: (2024)
Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training
by: Zheng, Ruobing, et al.
Published: (2026)
by: Zheng, Ruobing, et al.
Published: (2026)
UniMixer: A Unified Architecture for Scaling Laws in Recommendation Systems
by: Ha, Mingming, et al.
Published: (2026)
by: Ha, Mingming, et al.
Published: (2026)
UniSearch: Rethinking Search System with a Unified Generative Architecture
by: Chen, Jiahui, et al.
Published: (2025)
by: Chen, Jiahui, et al.
Published: (2025)
Reversing Flow for Image Restoration
by: Qin, Haina, et al.
Published: (2025)
by: Qin, Haina, et al.
Published: (2025)
PsyLite Technical Report
by: Ding, Fangjun, et al.
Published: (2025)
by: Ding, Fangjun, et al.
Published: (2025)
Skip-Vision: Efficient and Scalable Acceleration of Vision-Language Models via Adaptive Token Skipping
by: Zeng, Weili, et al.
Published: (2025)
by: Zeng, Weili, et al.
Published: (2025)
The Transformation of Yunnan in Ming China
Published: (2021)
Published: (2021)
Vignettes from the Late Ming
by: Ye, Yang
Published: (2023)
by: Ye, Yang
Published: (2023)
Two Studies on Ming History
by: Hucker, Charles O.
Published: (2020)
by: Hucker, Charles O.
Published: (2020)
A Ming Confucian’s World
by: Rong, Lu
Published: (2023)
by: Rong, Lu
Published: (2023)
Exploring Effective Priors and Efficient Models for Weakly-Supervised Change Detection
by: Zhao, Zhenghui, et al.
Published: (2023)
by: Zhao, Zhenghui, et al.
Published: (2023)
Animate-X: Universal Character Image Animation with Enhanced Motion Representation
by: Tan, Shuai, et al.
Published: (2024)
by: Tan, Shuai, et al.
Published: (2024)
Mimir: Improving Video Diffusion Models for Precise Text Understanding
by: Tan, Shuai, et al.
Published: (2024)
by: Tan, Shuai, et al.
Published: (2024)
UniSER: A Foundation Model for Unified Soft Effects Removal
by: Zhang, Jingdong, et al.
Published: (2025)
by: Zhang, Jingdong, et al.
Published: (2025)
Advancements in Natural Language Processing: Exploring Transformer-Based Architectures for Text Understanding
by: Wu, Tianhao, et al.
Published: (2025)
by: Wu, Tianhao, et al.
Published: (2025)
Oiling‐Out in Industrial Crystallization of Organic Small Molecules: Mechanisms, Characterization, Regulation, and Applications
by: Shilei Zhou, et al.
Published: (2024)
by: Shilei Zhou, et al.
Published: (2024)
Similar Items
-
Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer
by: Huang, Ziyuan, et al.
Published: (2025) -
Ming-Omni: A Unified Multimodal Model for Perception and Generation
by: AI, Inclusion, et al.
Published: (2025) -
M2-Reasoning: Empowering MLLMs with Unified General and Spatial Reasoning
by: AI, Inclusion, et al.
Published: (2025) -
ARGenSeg: Image Segmentation with Autoregressive Image Generation Model
by: Wang, Xiaolong, et al.
Published: (2025) -
Ming-Flash-Omni: A Sparse, Unified Architecture for Multimodal Perception and Generation
by: AI, Inclusion, et al.
Published: (2025)