Saved in:
| Main Authors: | Chen, Wei, Du, Chaoqun, Gu, Feng, He, Wei, Li, Qizhen, Liu, Zide, Pan, Xuhao, Ren, Chang, Rao, Xudong, Wang, Chenfeng, Wei, Tao, Yu, Chengjun, Yu, Pengfei, Zheng, Yufei, Zhou, Chunpeng, Zhou, Pan, Zhu, Xuhan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.02895 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dual-Pathway Geometry-Aware MLLM for Spatial Intelligence
by: Zheng, Yufei, et al.
Published: (2026)
by: Zheng, Yufei, et al.
Published: (2026)
StreamingClaw Technical Report
by: Chen, Jiawei, et al.
Published: (2026)
by: Chen, Jiawei, et al.
Published: (2026)
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model
by: Wang, Chenfeng, et al.
Published: (2026)
by: Wang, Chenfeng, et al.
Published: (2026)
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation
by: Li, Pengzhi, et al.
Published: (2025)
by: Li, Pengzhi, et al.
Published: (2025)
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
by: Yu, Chengjun, et al.
Published: (2026)
by: Yu, Chengjun, et al.
Published: (2026)
MindWatcher: Toward Smarter Multimodal Tool-Integrated Reasoning
by: Chen, Jiawei, et al.
Published: (2025)
by: Chen, Jiawei, et al.
Published: (2025)
MindGPT: Advancing Human-AI Interaction with Non-Invasive fNIRS-Based Imagined Speech Decoding
by: Zhang, Suyi, et al.
Published: (2024)
by: Zhang, Suyi, et al.
Published: (2024)
LiSD: An Efficient Multi-Task Learning Framework for LiDAR Segmentation and Detection
by: Xu, Jiahua, et al.
Published: (2024)
by: Xu, Jiahua, et al.
Published: (2024)
A Modification to Two‐Stage Least Squares With Genetic Applications
by: Lei Fang, et al.
Published: (2025)
by: Lei Fang, et al.
Published: (2025)
Asymptotic behaviour of the weak inverse anisotropic mean curvature flow
by: Gao, Chaoqun, et al.
Published: (2025)
by: Gao, Chaoqun, et al.
Published: (2025)
Progressive Video Condensation with MLLM Agent for Long-form Video Understanding
by: Yin, Yufei, et al.
Published: (2026)
by: Yin, Yufei, et al.
Published: (2026)
Mind the Third Eye! Benchmarking Privacy Awareness in MLLM-powered Smartphone Agents
by: Lin, Zhixin, et al.
Published: (2025)
by: Lin, Zhixin, et al.
Published: (2025)
A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design Ideation
by: Shi, Weiyan, et al.
Published: (2026)
by: Shi, Weiyan, et al.
Published: (2026)
Visual Position Prompt for MLLM based Visual Grounding
by: Tang, Wei, et al.
Published: (2025)
by: Tang, Wei, et al.
Published: (2025)
Guiding Cross-Modal Representations with MLLM Priors via Preference Alignment
by: Zhao, Pengfei, et al.
Published: (2025)
by: Zhao, Pengfei, et al.
Published: (2025)
DHP: Efficient Scaling of MLLM Training with Dynamic Hybrid Parallelism
by: Niu, Yifan, et al.
Published: (2026)
by: Niu, Yifan, et al.
Published: (2026)
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
by: Fan, Yuankai, et al.
Published: (2025)
by: Fan, Yuankai, et al.
Published: (2025)
CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection
by: Li, Fanxiao, et al.
Published: (2025)
by: Li, Fanxiao, et al.
Published: (2025)
AutoJudger: An Agent-Driven Framework for Efficient Benchmarking of MLLMs
by: Ding, Xuanwen, et al.
Published: (2025)
by: Ding, Xuanwen, et al.
Published: (2025)
Large Language Model Federated Learning with Blockchain and Unlearning for Cross-Organizational Collaboration
by: Zuo, Xuhan, et al.
Published: (2024)
by: Zuo, Xuhan, et al.
Published: (2024)
Ostrakon-VL: Towards Domain-Expert MLLM for Food-Service and Retail Stores
by: Shen, Zhiyong, et al.
Published: (2026)
by: Shen, Zhiyong, et al.
Published: (2026)
MC-CoT: A Modular Collaborative CoT Framework for Zero-shot Medical-VQA with LLM and MLLM Integration
by: Wei, Lai, et al.
Published: (2024)
by: Wei, Lai, et al.
Published: (2024)
PnP-U3D: Plug-and-Play 3D Framework Bridging Autoregression and Diffusion for Unified Understanding and Generation
by: Chen, Yongwei, et al.
Published: (2026)
by: Chen, Yongwei, et al.
Published: (2026)
Changing climate reshapes age structure in China
by: Wei, Pan
Published: (2025)
by: Wei, Pan
Published: (2025)
Research on Collision Avoidance Methods for Logistics Unmanned Aerial Vehicle Based on Dynamic Controlled Interactive Collaborative Fusion
by: Yuetan Zhang, et al.
Published: (2026)
by: Yuetan Zhang, et al.
Published: (2026)
Label-guided Facial Retouching Reversion
by: Zhao, Guanhua, et al.
Published: (2024)
by: Zhao, Guanhua, et al.
Published: (2024)
Gradient Residual Connections
by: Pan, Yangchen, et al.
Published: (2026)
by: Pan, Yangchen, et al.
Published: (2026)
Two-Stage Voting for Robust and Efficient Suicide Risk Detection on Social Media
by: Song, Yukai, et al.
Published: (2025)
by: Song, Yukai, et al.
Published: (2025)
Curr-RLCER:Curriculum Reinforcement Learning For Coherence Explainable Recommendation
by: Pan, Xiangchen, et al.
Published: (2026)
by: Pan, Xiangchen, et al.
Published: (2026)
Joint Behavior-guided and Modality-coherence Conditional Graph Diffusion Denoising for Multi Modal Recommendation
by: Pan, Xiangchen, et al.
Published: (2026)
by: Pan, Xiangchen, et al.
Published: (2026)
MMP-Refer: Multimodal Path Retrieval-augmented LLMs For Explainable Recommendation
by: Pan, Xiangchen, et al.
Published: (2026)
by: Pan, Xiangchen, et al.
Published: (2026)
Trifluoromethylselenylative Difunctionalization of Arynes with [ Me 4 N ][ SeCF 3 ] †
by: Hao‐Nan Wang, et al.
Published: (2026)
by: Hao‐Nan Wang, et al.
Published: (2026)
MiniCPM-V: A GPT-4V Level MLLM on Your Phone
by: Yao, Yuan, et al.
Published: (2024)
by: Yao, Yuan, et al.
Published: (2024)
Comparison of Effective Dissipation Channels in Warm Higgs Inflation from Warm Background Evolution
by: Cheng, Wei, et al.
Published: (2026)
by: Cheng, Wei, et al.
Published: (2026)
MASRA: MLLM-Assisted Semantic-Relational Consistent Alignment for Video Temporal Grounding
by: Ran, Ran, et al.
Published: (2026)
by: Ran, Ran, et al.
Published: (2026)
SAM-SP: Self-Prompting Makes SAM Great Again
by: Zhou, Chunpeng, et al.
Published: (2024)
by: Zhou, Chunpeng, et al.
Published: (2024)
One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning
by: Zhou, Chunpeng, et al.
Published: (2025)
by: Zhou, Chunpeng, et al.
Published: (2025)
Less is More: A Closer Look at Semantic-based Few-Shot Learning
by: Zhou, Chunpeng, et al.
Published: (2024)
by: Zhou, Chunpeng, et al.
Published: (2024)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
Translation from Wearable PPG to 12-Lead ECG
by: Ji, Hui, et al.
Published: (2025)
by: Ji, Hui, et al.
Published: (2025)
Similar Items
-
Dual-Pathway Geometry-Aware MLLM for Spatial Intelligence
by: Zheng, Yufei, et al.
Published: (2026) -
StreamingClaw Technical Report
by: Chen, Jiawei, et al.
Published: (2026) -
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model
by: Wang, Chenfeng, et al.
Published: (2026) -
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation
by: Li, Pengzhi, et al.
Published: (2025) -
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
by: Yu, Chengjun, et al.
Published: (2026)