Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Rujie, Zhao, Haozhe, Ci, Hai, Wang, Yizhou |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LongViTU: Instruction Tuning for Long-Form Video Understanding
by: Wu, Rujie, et al.
Published: (2025)
by: Wu, Rujie, et al.
Published: (2025)
FreeCloth: Free-form Generation Enhances Challenging Clothed Human Modeling
by: Ye, Hang, et al.
Published: (2024)
by: Ye, Hang, et al.
Published: (2024)
Maintaining Performance with Less Data
by: Sanderson, Dominic, et al.
Published: (2022)
by: Sanderson, Dominic, et al.
Published: (2022)
FREE: Faster and Better Data-Free Meta-Learning
by: Wei, Yongxian, et al.
Published: (2024)
by: Wei, Yongxian, et al.
Published: (2024)
GIST: Targeted Data Selection for Instruction Tuning via Coupled Optimization Geometry
by: Min, Guanghui, et al.
Published: (2026)
by: Min, Guanghui, et al.
Published: (2026)
Dynamic Context-oriented Decomposition for Task-aware Low-rank Adaptation with Less Forgetting and Faster Convergence
by: Yang, Yibo, et al.
Published: (2025)
by: Yang, Yibo, et al.
Published: (2025)
Progressive Data Dropout: An Embarrassingly Simple Approach to Faster Training
by: Sathiyanarayanan, Shriram M, et al.
Published: (2025)
by: Sathiyanarayanan, Shriram M, et al.
Published: (2025)
Pilot: Building the Federated Multimodal Instruction Tuning Framework
by: Xiong, Baochen, et al.
Published: (2025)
by: Xiong, Baochen, et al.
Published: (2025)
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning
by: Safaei, Bardia, et al.
Published: (2025)
by: Safaei, Bardia, et al.
Published: (2025)
Data Agent: Learning to Select Data via End-to-End Dynamic Optimization
by: Yang, Suorong, et al.
Published: (2026)
by: Yang, Suorong, et al.
Published: (2026)
Generative Modelling with High-Order Langevin Dynamics
by: Shi, Ziqiang, et al.
Published: (2024)
by: Shi, Ziqiang, et al.
Published: (2024)
Benchmarking Egocentric Multimodal Goal Inference for Assistive Wearable Agents
by: Veerabadran, Vijay, et al.
Published: (2025)
by: Veerabadran, Vijay, et al.
Published: (2025)
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
by: Davtyan, Aram, et al.
Published: (2026)
by: Davtyan, Aram, et al.
Published: (2026)
ProtoAda: Prototype-Guided Adaptive Adapter Expansion and Geometric Consolidation for Multimodal Continual Instruction Tuning
by: Shi, Yu-Cheng, et al.
Published: (2026)
by: Shi, Yu-Cheng, et al.
Published: (2026)
Prism: A Plug-in Reproducible Infrastructure for Scalable Multimodal Continual Instruction Tuning
by: Tang, Jun-Tao, et al.
Published: (2026)
by: Tang, Jun-Tao, et al.
Published: (2026)
TextSquare: Scaling up Text-Centric Visual Instruction Tuning
by: Tang, Jingqun, et al.
Published: (2024)
by: Tang, Jingqun, et al.
Published: (2024)
Data-Juicer Sandbox: A Feedback-Driven Suite for Multimodal Data-Model Co-development
by: Chen, Daoyuan, et al.
Published: (2024)
by: Chen, Daoyuan, et al.
Published: (2024)
Overcoming Catastrophic Forgetting in Visual Continual Learning with Reinforcement Fine-Tuning
by: Lou, Meng, et al.
Published: (2026)
by: Lou, Meng, et al.
Published: (2026)
Less is More: High-value Data Selection for Visual Instruction Tuning
by: Liu, Zikang, et al.
Published: (2024)
by: Liu, Zikang, et al.
Published: (2024)
Faster 3D Gaussian Splatting Convergence via Structure-Aware Densification
by: Lyu, Linjie, et al.
Published: (2026)
by: Lyu, Linjie, et al.
Published: (2026)
Reconstructive Visual Instruction Tuning
by: Wang, Haochen, et al.
Published: (2024)
by: Wang, Haochen, et al.
Published: (2024)
MMGPL: Multimodal Medical Data Analysis with Graph Prompt Learning
by: Peng, Liang, et al.
Published: (2023)
by: Peng, Liang, et al.
Published: (2023)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
Parrot: Multilingual Visual Instruction Tuning
by: Sun, Hai-Long, et al.
Published: (2024)
by: Sun, Hai-Long, et al.
Published: (2024)
Enhancing Compositional Generalization via Compositional Feature Alignment
by: Wang, Haoxiang, et al.
Published: (2024)
by: Wang, Haoxiang, et al.
Published: (2024)
A Faster Path to Continual Learning
by: Li, Wei, et al.
Published: (2026)
by: Li, Wei, et al.
Published: (2026)
Understanding Data Influence with Differential Approximation
by: Tan, Haoru, et al.
Published: (2025)
by: Tan, Haoru, et al.
Published: (2025)
Semantic Residual for Multimodal Unified Discrete Representation
by: Huang, Hai, et al.
Published: (2024)
by: Huang, Hai, et al.
Published: (2024)
Data-Free Quantization via Mixed-Precision Compensation without Fine-Tuning
by: Chen, Jun, et al.
Published: (2023)
by: Chen, Jun, et al.
Published: (2023)
Filter Like You Test: Data-Driven Data Filtering for CLIP Pretraining
by: Shechter, Mikey, et al.
Published: (2025)
by: Shechter, Mikey, et al.
Published: (2025)
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
by: Gupta, Sharut, et al.
Published: (2025)
by: Gupta, Sharut, et al.
Published: (2025)
Data-Driven Hierarchical Open Set Recognition
by: Hannum, Andrew, et al.
Published: (2024)
by: Hannum, Andrew, et al.
Published: (2024)
AIM-Fair: Advancing Algorithmic Fairness via Selectively Fine-Tuning Biased Models with Contextual Synthetic Data
by: Zhao, Zengqun, et al.
Published: (2025)
by: Zhao, Zengqun, et al.
Published: (2025)
CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning
by: Wang, Yiping, et al.
Published: (2024)
by: Wang, Yiping, et al.
Published: (2024)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
by: Wu, Shaojin, et al.
Published: (2025)
by: Wu, Shaojin, et al.
Published: (2025)
Data-Driven Stochastic Motion Evaluation and Optimization with Image by Spatially-Aligned Temporal Encoding
by: Oba, Takeru, et al.
Published: (2023)
by: Oba, Takeru, et al.
Published: (2023)
Cream of the Crop: Harvesting Rich, Scalable and Transferable Multi-Modal Data for Instruction Fine-Tuning
by: Lyu, Mengyao, et al.
Published: (2025)
by: Lyu, Mengyao, et al.
Published: (2025)
Multimodal Data Curation via Object Detection and Filter Ensembles
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
Driver Assistance System Based on Multimodal Data Hazard Detection
by: Zhouxiang, Long, et al.
Published: (2025)
by: Zhouxiang, Long, et al.
Published: (2025)
GeoMix: Towards Geometry-Aware Data Augmentation
by: Zhao, Wentao, et al.
Published: (2024)
by: Zhao, Wentao, et al.
Published: (2024)
Similar Items
-
LongViTU: Instruction Tuning for Long-Form Video Understanding
by: Wu, Rujie, et al.
Published: (2025) -
FreeCloth: Free-form Generation Enhances Challenging Clothed Human Modeling
by: Ye, Hang, et al.
Published: (2024) -
Maintaining Performance with Less Data
by: Sanderson, Dominic, et al.
Published: (2022) -
FREE: Faster and Better Data-Free Meta-Learning
by: Wei, Yongxian, et al.
Published: (2024) -
GIST: Targeted Data Selection for Instruction Tuning via Coupled Optimization Geometry
by: Min, Guanghui, et al.
Published: (2026)