UniMIC: Token-Based Multimodal Interactive Coding for Human-AI Collaboration
Fuente:
arXiv
Saved in:
| Main Authors: | Mao, Qi, Yang, Tinghan, Li, Jiahao, Li, Bin, Jin, Libiao, Lu, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UniMIC: Towards Universal Multi-modality Perceptual Image Compression
by: Gao, Yixin, et al.
Published: (2024)
by: Gao, Yixin, et al.
Published: (2024)
Camera Artist: A Multi-Agent Framework for Cinematic Language Storytelling Video Generation
by: Hu, Haobo, et al.
Published: (2026)
by: Hu, Haobo, et al.
Published: (2026)
IC-Effect: Precise and Efficient Video Effects Editing via In-Context Learning
by: Li, Yuanhang, et al.
Published: (2025)
by: Li, Yuanhang, et al.
Published: (2025)
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
by: Yang, Panqi, et al.
Published: (2025)
by: Yang, Panqi, et al.
Published: (2025)
Generative Neural Video Compression via Video Diffusion Prior
by: Mao, Qi, et al.
Published: (2025)
by: Mao, Qi, et al.
Published: (2025)
Causal Responsibility Attribution for Human-AI Collaboration
by: Qi, Yahang, et al.
Published: (2024)
by: Qi, Yahang, et al.
Published: (2024)
UniWeTok: An Unified Binary Tokenizer with Codebook Size $\mathit{2^{128}}$ for Unified Multimodal Large Language Model
by: Zhuang, Shaobin, et al.
Published: (2026)
by: Zhuang, Shaobin, et al.
Published: (2026)
GIA-MIC: Multimodal Emotion Recognition with Gated Interactive Attention and Modality-Invariant Learning Constraints
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Interaction, Process, Infrastructure: A Unified Framework for Human-Agent Collaboration
by: Wang, Yun, et al.
Published: (2025)
by: Wang, Yun, et al.
Published: (2025)
Correctness Learning: Deductive Verification Guided Learning for Human-AI Collaboration
by: Jin, Zhao, et al.
Published: (2025)
by: Jin, Zhao, et al.
Published: (2025)
CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems
by: Wang, Yannan, et al.
Published: (2026)
by: Wang, Yannan, et al.
Published: (2026)
StarVid: Enhancing Semantic Alignment in Video Diffusion Models via Spatial and SynTactic Guided Attention Refocusing
by: Li, Yuanhang, et al.
Published: (2024)
by: Li, Yuanhang, et al.
Published: (2024)
UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding
by: Jiao, Yang, et al.
Published: (2025)
by: Jiao, Yang, et al.
Published: (2025)
UniFluids: Unified Neural Operator Learning with Conditional Flow-matching
by: Li, Haosen, et al.
Published: (2026)
by: Li, Haosen, et al.
Published: (2026)
Multimodal Medical Code Tokenizer
by: Su, Xiaorui, et al.
Published: (2025)
by: Su, Xiaorui, et al.
Published: (2025)
CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness
by: Liu, Jiayi, et al.
Published: (2024)
by: Liu, Jiayi, et al.
Published: (2024)
Deconstructing Human-AI Collaboration: Agency, Interaction, and Adaptation
by: Holter, Steffen, et al.
Published: (2024)
by: Holter, Steffen, et al.
Published: (2024)
UniTok: A Unified Tokenizer for Visual Generation and Understanding
by: Ma, Chuofan, et al.
Published: (2025)
by: Ma, Chuofan, et al.
Published: (2025)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
From Charts to Code: A Hierarchical Benchmark for Multimodal Models
by: Tang, Jiahao, et al.
Published: (2025)
by: Tang, Jiahao, et al.
Published: (2025)
The AI Collaborator: Bridging Human-AI Interaction in Educational and Professional Settings
by: Samadi, Mohammad Amin, et al.
Published: (2024)
by: Samadi, Mohammad Amin, et al.
Published: (2024)
KathDB: Explainable Multimodal Database Management System with Human-AI Collaboration
by: Xiao, Guorui, et al.
Published: (2025)
by: Xiao, Guorui, et al.
Published: (2025)
Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective
by: Liu, Hengyu, et al.
Published: (2026)
by: Liu, Hengyu, et al.
Published: (2026)
UniCode: Learning a Unified Codebook for Multimodal Large Language Models
by: Zheng, Sipeng, et al.
Published: (2024)
by: Zheng, Sipeng, et al.
Published: (2024)
Divide, then Ground: Adapting Frame Selection to Query Types for Long-Form Video Understanding
by: Li, Jialuo, et al.
Published: (2025)
by: Li, Jialuo, et al.
Published: (2025)
Human-Centered Human-AI Collaboration (HCHAC)
by: Gao, Qi, et al.
Published: (2025)
by: Gao, Qi, et al.
Published: (2025)
UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
by: Li, Yiheng, et al.
Published: (2024)
by: Li, Yiheng, et al.
Published: (2024)
Multimodal Contrastive Learning via Uni-Modal Coding and Cross-Modal Prediction for Multimodal Sentiment Analysis
by: Lin, Ronghao, et al.
Published: (2022)
by: Lin, Ronghao, et al.
Published: (2022)
AXIS: Efficient Human-Agent-Computer Interaction with API-First LLM-Based Agents
by: Lu, Junting, et al.
Published: (2024)
by: Lu, Junting, et al.
Published: (2024)
MiMIC: Mitigating Visual Modality Collapse in Universal Multimodal Retrieval While Avoiding Semantic Misalignment
by: Li, Juan, et al.
Published: (2026)
by: Li, Juan, et al.
Published: (2026)
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration
by: Liu, Jiaqi, et al.
Published: (2026)
by: Liu, Jiaqi, et al.
Published: (2026)
CutVerse: A Compositional GUI Agents Benchmark for Media Post-Production Editing
by: Hu, Haobo, et al.
Published: (2026)
by: Hu, Haobo, et al.
Published: (2026)
Deconstructing the Dual Black Box:A Plug-and-Play Cognitive Framework for Human-AI Collaborative Enhancement and Its Implications for AI Governance
by: Lu, Yiming
Published: (2025)
by: Lu, Yiming
Published: (2025)
Safe Multimodal Communication in Human-Robot Collaboration
by: Ferrari, Davide, et al.
Published: (2023)
by: Ferrari, Davide, et al.
Published: (2023)
Toward Trustworthy Evaluation of Sustainability Rating Methodologies: A Human-AI Collaborative Framework for Benchmark Dataset Construction
by: Cai, Xiaoran, et al.
Published: (2026)
by: Cai, Xiaoran, et al.
Published: (2026)
Agentic AI as Undercover Teammates: Argumentative Knowledge Construction in Hybrid Human-AI Collaborative Learning
by: Yan, Lixiang, et al.
Published: (2025)
by: Yan, Lixiang, et al.
Published: (2025)
Unbiased Collaborative Filtering with Fair Sampling
by: Liu, Jiahao, et al.
Published: (2025)
by: Liu, Jiahao, et al.
Published: (2025)
RECODE-H: A Benchmark for Research Code Development with Interactive Human Feedback
by: Miao, Chunyu, et al.
Published: (2025)
by: Miao, Chunyu, et al.
Published: (2025)
CATP: Cross-Attention Token Pruning for Accuracy Preserved Multimodal Model Inference
by: Liao, Ruqi, et al.
Published: (2024)
by: Liao, Ruqi, et al.
Published: (2024)
Mixed-Initiative Context: Structuring and Managing Context for Human-AI Collaboration
by: Li, Haichang, et al.
Published: (2026)
by: Li, Haichang, et al.
Published: (2026)
Similar Items
-
UniMIC: Towards Universal Multi-modality Perceptual Image Compression
by: Gao, Yixin, et al.
Published: (2024) -
Camera Artist: A Multi-Agent Framework for Cinematic Language Storytelling Video Generation
by: Hu, Haobo, et al.
Published: (2026) -
IC-Effect: Precise and Efficient Video Effects Editing via In-Context Learning
by: Li, Yuanhang, et al.
Published: (2025) -
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
by: Yang, Panqi, et al.
Published: (2025) -
Generative Neural Video Compression via Video Diffusion Prior
by: Mao, Qi, et al.
Published: (2025)