Mitigating Modal Imbalance in Multimodal Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Chen Henry, Kale, Neil, Raghunathan, Aditi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Trained Verification for Training- and Test-Time Self-Improvement
by: Wu, Chen Henry, et al.
Published: (2026)
by: Wu, Chen Henry, et al.
Published: (2026)
Understanding and Mitigating Premature Confidence for Better LLM Reasoning
by: Gai, Jingchu, et al.
Published: (2026)
by: Gai, Jingchu, et al.
Published: (2026)
Mode-Conditioning Unlocks Superior Test-Time Scaling
by: Wu, Chen Henry, et al.
Published: (2025)
by: Wu, Chen Henry, et al.
Published: (2025)
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
by: Nagarajan, Vaishnavh, et al.
Published: (2025)
by: Nagarajan, Vaishnavh, et al.
Published: (2025)
Reasoning as an Adaptive Defense for Safety
by: Kim, Taeyoun, et al.
Published: (2025)
by: Kim, Taeyoun, et al.
Published: (2025)
Hodoscope: Unsupervised Monitoring for AI Misbehaviors
by: Zhong, Ziqian, et al.
Published: (2026)
by: Zhong, Ziqian, et al.
Published: (2026)
Weight Ensembling Improves Reasoning in Language Models
by: Dang, Xingyu, et al.
Published: (2025)
by: Dang, Xingyu, et al.
Published: (2025)
Memorization Sinks: Isolating Memorization during LLM Training
by: Ghosal, Gaurav R., et al.
Published: (2025)
by: Ghosal, Gaurav R., et al.
Published: (2025)
Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination
by: Jiang, Jinrui, et al.
Published: (2026)
by: Jiang, Jinrui, et al.
Published: (2026)
Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization
by: Kong, Jiawei, et al.
Published: (2026)
by: Kong, Jiawei, et al.
Published: (2026)
Jailbreaking in the Haystack
by: Shah, Rishi Rajesh, et al.
Published: (2025)
by: Shah, Rishi Rajesh, et al.
Published: (2025)
MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models
by: Chung, Sangyun, et al.
Published: (2026)
by: Chung, Sangyun, et al.
Published: (2026)
Explaining and Mitigating the Modality Gap in Contrastive Multimodal Learning
by: Yaras, Can, et al.
Published: (2024)
by: Yaras, Can, et al.
Published: (2024)
Pando: Do Interpretability Methods Work When Models Won't Explain Themselves?
by: Zhong, Ziqian, et al.
Published: (2026)
by: Zhong, Ziqian, et al.
Published: (2026)
Test-Time Adaptation Induces Stronger Accuracy and Agreement-on-the-Line
by: Kim, Eungyeup, et al.
Published: (2023)
by: Kim, Eungyeup, et al.
Published: (2023)
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
by: Cai, Rui, et al.
Published: (2025)
by: Cai, Rui, et al.
Published: (2025)
Rethinking the Text-Vision Reasoning Imbalance in MLLMs through the Lens of Training Recipes
by: Yao, Guanyu, et al.
Published: (2025)
by: Yao, Guanyu, et al.
Published: (2025)
When One Modality Sabotages the Others: A Diagnostic Lens on Multimodal Reasoning
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
Mitigating Participation Imbalance Bias in Asynchronous Federated Learning
by: Chang, Xiangyu, et al.
Published: (2025)
by: Chang, Xiangyu, et al.
Published: (2025)
Hierarchical Contact-Rich Trajectory Optimization for Multi-Modal Manipulation using Tight Convex Relaxations
by: Shirai, Yuki, et al.
Published: (2025)
by: Shirai, Yuki, et al.
Published: (2025)
An Effective Automated Speaking Assessment Approach to Mitigating Data Scarcity and Imbalanced Distribution
by: Lo, Tien-Hong, et al.
Published: (2024)
by: Lo, Tien-Hong, et al.
Published: (2024)
The Role of Visual Modality in Multimodal Mathematical Reasoning: Challenges and Insights
by: Liu, Yufang, et al.
Published: (2025)
by: Liu, Yufang, et al.
Published: (2025)
S2D: Selective Spectral Decay for Quantization-Friendly Conditioning of Neural Activations
by: Chavan, Arnav, et al.
Published: (2026)
by: Chavan, Arnav, et al.
Published: (2026)
Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought
by: Li, Xuanchen, et al.
Published: (2026)
by: Li, Xuanchen, et al.
Published: (2026)
Mosaic of Modalities: A Comprehensive Benchmark for Multimodal Graph Learning
by: Zhu, Jing, et al.
Published: (2024)
by: Zhu, Jing, et al.
Published: (2024)
Look It Up: Analysing Internal Web Search Capabilities of Modern LLMs
by: Kale, Sahil
Published: (2025)
by: Kale, Sahil
Published: (2025)
Learning What to Attend First: Modality-Importance-Guided Reasoning for Reliable Multimodal Emotion Understanding
by: Rha, Hyeongseop, et al.
Published: (2025)
by: Rha, Hyeongseop, et al.
Published: (2025)
Generative AI and Power Imbalances in Global Education: Frameworks for Bias Mitigation
by: Nyaaba, Matthew, et al.
Published: (2024)
by: Nyaaba, Matthew, et al.
Published: (2024)
BLUR: A Benchmark for LLM Unlearning Robust to Forget-Retain Overlap
by: Hu, Shengyuan, et al.
Published: (2025)
by: Hu, Shengyuan, et al.
Published: (2025)
BalanceBenchmark: A Survey for Multimodal Imbalance Learning
by: Xu, Shaoxuan, et al.
Published: (2025)
by: Xu, Shaoxuan, et al.
Published: (2025)
Personalized Knowledge Tracing through Student Representation Reconstruction and Class Imbalance Mitigation
by: Chen, Zhiyu, et al.
Published: (2024)
by: Chen, Zhiyu, et al.
Published: (2024)
A Survey of Multimodal Mathematical Reasoning: From Perception, Alignment to Reasoning
by: Yang, Tianyu, et al.
Published: (2026)
by: Yang, Tianyu, et al.
Published: (2026)
Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion
by: Zhang, Yichi, et al.
Published: (2024)
by: Zhang, Yichi, et al.
Published: (2024)
Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning
by: Zhou, Hao, et al.
Published: (2026)
by: Zhou, Hao, et al.
Published: (2026)
Automated Assessment of Multimodal Answer Sheets in the STEM domain
by: Patil, Rajlaxmi, et al.
Published: (2024)
by: Patil, Rajlaxmi, et al.
Published: (2024)
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning
by: Xiang, Kun, et al.
Published: (2026)
by: Xiang, Kun, et al.
Published: (2026)
Base Models Look Human To AI Detectors
by: Xu, Yixuan Even, et al.
Published: (2026)
by: Xu, Yixuan Even, et al.
Published: (2026)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
by: Sun, Han, et al.
Published: (2026)
by: Sun, Han, et al.
Published: (2026)
Seeing with You: Perception-Reasoning Coevolution for Multimodal Reasoning
by: Miao, Ziqi, et al.
Published: (2026)
by: Miao, Ziqi, et al.
Published: (2026)
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
by: Zheng, Haojie, et al.
Published: (2024)
by: Zheng, Haojie, et al.
Published: (2024)
Similar Items
-
Self-Trained Verification for Training- and Test-Time Self-Improvement
by: Wu, Chen Henry, et al.
Published: (2026) -
Understanding and Mitigating Premature Confidence for Better LLM Reasoning
by: Gai, Jingchu, et al.
Published: (2026) -
Mode-Conditioning Unlocks Superior Test-Time Scaling
by: Wu, Chen Henry, et al.
Published: (2025) -
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
by: Nagarajan, Vaishnavh, et al.
Published: (2025) -
Reasoning as an Adaptive Defense for Safety
by: Kim, Taeyoun, et al.
Published: (2025)