UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yaqi, Lin, Wang, Zhang, Zijian, Yang, Miles, Chen, Jingyuan, Zhang, Wentao, Zhong, Zhao, Bo, Liefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding
von: Liu, Xiangyue, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyue, et al.
Veröffentlicht: (2026)
HYDRA: Unifying Multi-modal Generation and Understanding via Representation-Harmonized Tokenization
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024)
von: He, Junjie, et al.
Veröffentlicht: (2024)
UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition
von: Zhang, Shuai, et al.
Veröffentlicht: (2026)
von: Zhang, Shuai, et al.
Veröffentlicht: (2026)
UniCom: Towards a Unified and Cohesiveness-aware Framework for Community Search and Detection
von: Zhu, Yifan, et al.
Veröffentlicht: (2025)
von: Zhu, Yifan, et al.
Veröffentlicht: (2025)
UniAPO: Unified Multimodal Automated Prompt Optimization
von: Zhu, Qipeng, et al.
Veröffentlicht: (2025)
von: Zhu, Qipeng, et al.
Veröffentlicht: (2025)
UniNote: A Unified Embedding Model for Multimodal Representation and Ranking
von: Zhao, Jinghan, et al.
Veröffentlicht: (2026)
von: Zhao, Jinghan, et al.
Veröffentlicht: (2026)
Evolving Without Ending: Unifying Multimodal Incremental Learning for Continual Panoptic Perception
von: Yuan, Bo, et al.
Veröffentlicht: (2026)
von: Yuan, Bo, et al.
Veröffentlicht: (2026)
R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
UniRGB-IR: A Unified Framework for Visible-Infrared Semantic Tasks via Adapter Tuning
von: Yuan, Maoxun, et al.
Veröffentlicht: (2024)
von: Yuan, Maoxun, et al.
Veröffentlicht: (2024)
ParaUni: Enhance Generation in Unified Multimodal Model with Reinforcement-driven Hierarchical Parallel Information Interaction
von: Tan, Jiangtong, et al.
Veröffentlicht: (2025)
von: Tan, Jiangtong, et al.
Veröffentlicht: (2025)
Representation Forcing for Bottleneck-Free Unified Multimodal Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation
von: Li, Teng, et al.
Veröffentlicht: (2025)
von: Li, Teng, et al.
Veröffentlicht: (2025)
UniPPTBench: A Unified Benchmark for Presentation Generation Across Diverse Input Settings
von: Zhao, Bo, et al.
Veröffentlicht: (2026)
von: Zhao, Bo, et al.
Veröffentlicht: (2026)
UniLight: A Unified Representation for Lighting
von: Zhang, Zitian, et al.
Veröffentlicht: (2025)
von: Zhang, Zitian, et al.
Veröffentlicht: (2025)
UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
UniParser: Multi-Human Parsing with Unified Correlation Representation Learning
von: Chu, Jiaming, et al.
Veröffentlicht: (2023)
von: Chu, Jiaming, et al.
Veröffentlicht: (2023)
UniCTokens: Boosting Personalized Understanding and Generation via Unified Concept Tokens
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
Semantic Alignment for Multimodal Large Language Models
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
Demystifying Numerosity in Diffusion Models -- Limitations and Remedies
von: Zhao, Yaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Yaqi, et al.
Veröffentlicht: (2025)
UniG2U-Bench: Do Unified Models Advance Multimodal Understanding?
von: Wen, Zimo, et al.
Veröffentlicht: (2026)
von: Wen, Zimo, et al.
Veröffentlicht: (2026)
UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors
von: Chen, Houyuan, et al.
Veröffentlicht: (2026)
von: Chen, Houyuan, et al.
Veröffentlicht: (2026)
Learning at a Glance: Towards Interpretable Data-limited Continual Semantic Segmentation via Semantic-Invariance Modelling
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
UniCode$^2$: Cascaded Large-scale Codebooks for Unified Multimodal Understanding and Generation
von: Chen, Yanzhe, et al.
Veröffentlicht: (2025)
von: Chen, Yanzhe, et al.
Veröffentlicht: (2025)
UniMatch V2: Pushing the Limit of Semi-Supervised Semantic Segmentation
von: Yang, Lihe, et al.
Veröffentlicht: (2024)
von: Yang, Lihe, et al.
Veröffentlicht: (2024)
Implicit Guidance and Explicit Representation of Semantic Information in Points Cloud: A Survey
von: Tang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Tang, Jingyuan, et al.
Veröffentlicht: (2025)
UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding
von: Xu, Yueming, et al.
Veröffentlicht: (2025)
von: Xu, Yueming, et al.
Veröffentlicht: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
UniStitch: Unifying Semantic and Geometric Features for Image Stitching
von: Mei, Yuan, et al.
Veröffentlicht: (2026)
von: Mei, Yuan, et al.
Veröffentlicht: (2026)
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
von: He, Junjie, et al.
Veröffentlicht: (2025)
von: He, Junjie, et al.
Veröffentlicht: (2025)
iFSQ: Improving FSQ for Image Generation with 1 Line of Code
von: Lin, Bin, et al.
Veröffentlicht: (2026)
von: Lin, Bin, et al.
Veröffentlicht: (2026)
TUNA: Taming Unified Visual Representations for Native Unified Multimodal Models
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
Harmonizing Visual Representations for Unified Multimodal Understanding and Generation
von: Wu, Size, et al.
Veröffentlicht: (2025)
von: Wu, Size, et al.
Veröffentlicht: (2025)
Uni-LVC: A Unified Method for Intra- and Inter-Mode Learned Video Compression
von: Zhang, Yichi, et al.
Veröffentlicht: (2026)
von: Zhang, Yichi, et al.
Veröffentlicht: (2026)
UniVid: The Open-Source Unified Video Model
von: Luo, Jiabin, et al.
Veröffentlicht: (2025)
von: Luo, Jiabin, et al.
Veröffentlicht: (2025)
MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE
von: Li, Junzhe, et al.
Veröffentlicht: (2025)
von: Li, Junzhe, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding
von: Liu, Xiangyue, et al.
Veröffentlicht: (2026) -
HYDRA: Unifying Multi-modal Generation and Understanding via Representation-Harmonized Tokenization
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026) -
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024) -
UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition
von: Zhang, Shuai, et al.
Veröffentlicht: (2026) -
UniCom: Towards a Unified and Cohesiveness-aware Framework for Community Search and Detection
von: Zhu, Yifan, et al.
Veröffentlicht: (2025)