Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Wenbin, Ding, Liang, Zeng, Minyan, Zhou, Xiabin, Shen, Li, Luo, Yong, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG
von: Wang, Wenbin, et al.
Veröffentlicht: (2025)
von: Wang, Wenbin, et al.
Veröffentlicht: (2025)
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
WisdoM: Improving Multimodal Sentiment Analysis by Fusing Contextual World Knowledge
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
Control Large Language Models via Divide and Conquer
von: Li, Bingxuan, et al.
Veröffentlicht: (2024)
von: Li, Bingxuan, et al.
Veröffentlicht: (2024)
Divide-and-Conquer Inference for Large-Scale Visual Recognition with Multimodal Large Language Models
von: Ye, Zhipeng, et al.
Veröffentlicht: (2026)
von: Ye, Zhipeng, et al.
Veröffentlicht: (2026)
Try, Check and Retry: A Divide-and-Conquer Framework for Boosting Long-context Tool-Calling Performance of LLMs
von: Chen, Kunfeng, et al.
Veröffentlicht: (2026)
von: Chen, Kunfeng, et al.
Veröffentlicht: (2026)
Adapting and Evaluating Multimodal Large Language Models for Adolescent Idiopathic Scoliosis Self-Management: A Divide and Conquer Framework
von: Wu, Zhaolong, et al.
Veröffentlicht: (2025)
von: Wu, Zhaolong, et al.
Veröffentlicht: (2025)
Divide and Conquer: A Hybrid Strategy Defeats Multimodal Large Language Models
von: Mao, Yanxu, et al.
Veröffentlicht: (2024)
von: Mao, Yanxu, et al.
Veröffentlicht: (2024)
DCARL: A Divide-and-Conquer Framework for Autoregressive Long-Trajectory Video Generation
von: Ouyang, Junyi, et al.
Veröffentlicht: (2026)
von: Ouyang, Junyi, et al.
Veröffentlicht: (2026)
OOP: Object-Oriented Programming Evaluation Benchmark for Large Language Models
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models
von: Xiao, Junyuan, et al.
Veröffentlicht: (2026)
von: Xiao, Junyuan, et al.
Veröffentlicht: (2026)
Divide, Conquer, Combine Bayesian Decision Tree Sampling
von: Cochrane, Jodie A., et al.
Veröffentlicht: (2024)
von: Cochrane, Jodie A., et al.
Veröffentlicht: (2024)
Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability
von: Liang, Xiao, et al.
Veröffentlicht: (2026)
von: Liang, Xiao, et al.
Veröffentlicht: (2026)
Divide and Conquer Self-Supervised Learning for High-Content Imaging
von: Farndale, Lucas, et al.
Veröffentlicht: (2025)
von: Farndale, Lucas, et al.
Veröffentlicht: (2025)
An Examination on the Effectiveness of Divide-and-Conquer Prompting in Large Language Models
von: Zhang, Yizhou, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2024)
Deformable Medical Image Registration with Effective Anatomical Structure Representation and Divide-and-Conquer Network
von: Ma, Xinke, et al.
Veröffentlicht: (2025)
von: Ma, Xinke, et al.
Veröffentlicht: (2025)
MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation
von: Wang, Yanhui, et al.
Veröffentlicht: (2023)
von: Wang, Yanhui, et al.
Veröffentlicht: (2023)
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
von: Addanki, Vamsi, et al.
Veröffentlicht: (2024)
von: Addanki, Vamsi, et al.
Veröffentlicht: (2024)
Divide and Conquer: Accelerating Diffusion-Based Large Language Models via Adaptive Parallel Decoding
von: Luo, Xiangzhong, et al.
Veröffentlicht: (2026)
von: Luo, Xiangzhong, et al.
Veröffentlicht: (2026)
Divide and Conquer: Static-Dynamic Collaboration for Few-Shot Class-Incremental Learning
von: Bao, Kexin, et al.
Veröffentlicht: (2026)
von: Bao, Kexin, et al.
Veröffentlicht: (2026)
UDC: A Unified Neural Divide-and-Conquer Framework for Large-Scale Combinatorial Optimization Problems
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
Divide and Conquer: High-Resolution Industrial Anomaly Detection via Memory Efficient Tiled Ensemble
von: Rolih, Blaž, et al.
Veröffentlicht: (2024)
von: Rolih, Blaž, et al.
Veröffentlicht: (2024)
Divide, Harmonize, Then Conquer It: Shooting Multi-Commodity Flow Problems with Multimodal Language Models
von: Yuan, Xinyu, et al.
Veröffentlicht: (2026)
von: Yuan, Xinyu, et al.
Veröffentlicht: (2026)
Divide and Conquer: Rethinking the Training Paradigm of Neural Radiance Fields
von: Ma, Rongkai, et al.
Veröffentlicht: (2024)
von: Ma, Rongkai, et al.
Veröffentlicht: (2024)
Qubit Mapping: The Adaptive Divide-and-Conquer Approach
von: Huang, Yunqi, et al.
Veröffentlicht: (2024)
von: Huang, Yunqi, et al.
Veröffentlicht: (2024)
Divide (Text) and Conquer (Sentiment): Improved Sentiment Classification by Constituent Conflict Resolution
von: Kościałkowski, Jan, et al.
Veröffentlicht: (2025)
von: Kościałkowski, Jan, et al.
Veröffentlicht: (2025)
Divide-and-Conquer Approach to Holistic Cognition in High-Similarity Contexts with Limited Data
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
Accurate and Scalable Matrix Mechanisms via Divide and Conquer
von: He, Guanlin, et al.
Veröffentlicht: (2026)
von: He, Guanlin, et al.
Veröffentlicht: (2026)
Divide and Conquer: Multimodal Video Deepfake Detection via Cross-Modal Fusion and Localization
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
The Divide-and-Conquer Framework: A Suitable Setting for the DDM of the Future
von: Herrera-Revilla, Ismael, et al.
Veröffentlicht: (2019)
von: Herrera-Revilla, Ismael, et al.
Veröffentlicht: (2019)
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
von: Cui, Wendi, et al.
Veröffentlicht: (2024)
von: Cui, Wendi, et al.
Veröffentlicht: (2024)
Divide-and-Conquer: Dual-Hierarchical Optimization for Semantic 4D Gaussian Spatting
von: Yan, Zhiying, et al.
Veröffentlicht: (2025)
von: Yan, Zhiying, et al.
Veröffentlicht: (2025)
Edit Once, Update Everywhere: A Simple Framework for Cross-Lingual Knowledge Synchronization in LLMs
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
ICU: Conquering Language Barriers in Vision-and-Language Modeling by Dividing the Tasks into Image Captioning and Language Understanding
von: Wu, Guojun
Veröffentlicht: (2023)
von: Wu, Guojun
Veröffentlicht: (2023)
Dividing and Conquering the Van Vleck Catastrophe
von: Simon, Sophia, et al.
Veröffentlicht: (2025)
von: Simon, Sophia, et al.
Veröffentlicht: (2025)
Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
DCA: Dividing and Conquering Amnesia in Incremental Object Detection
von: Zhang, Aoting, et al.
Veröffentlicht: (2025)
von: Zhang, Aoting, et al.
Veröffentlicht: (2025)
Robust and Explainable Divide-and-Conquer Learning for Intrusion Detection
von: Zhou, Yan, et al.
Veröffentlicht: (2026)
von: Zhou, Yan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG
von: Wang, Wenbin, et al.
Veröffentlicht: (2025) -
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024) -
WisdoM: Improving Multimodal Sentiment Analysis by Fusing Contextual World Knowledge
von: Wang, Wenbin, et al.
Veröffentlicht: (2024) -
Control Large Language Models via Divide and Conquer
von: Li, Bingxuan, et al.
Veröffentlicht: (2024) -
Divide-and-Conquer Inference for Large-Scale Visual Recognition with Multimodal Large Language Models
von: Ye, Zhipeng, et al.
Veröffentlicht: (2026)