Towards Robust Instruction Tuning on Multimodal Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Wei, Chen, Hui, Poria, Soujanya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toward Robust Multimodal Learning using Multimodal Foundational Models
von: Zhao, Xianbing, et al.
Veröffentlicht: (2024)
von: Zhao, Xianbing, et al.
Veröffentlicht: (2024)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
von: Bhardwaj, Rishabh, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Rishabh, et al.
Veröffentlicht: (2024)
Self-Adaptive Sampling for Efficient Video Question-Answering on Image--Text Models
von: Han, Wei, et al.
Veröffentlicht: (2023)
von: Han, Wei, et al.
Veröffentlicht: (2023)
Stacked from One: Multi-Scale Self-Injection for Context Window Extension
von: Han, Wei, et al.
Veröffentlicht: (2026)
von: Han, Wei, et al.
Veröffentlicht: (2026)
Large Language Models for Automated Open-domain Scientific Hypotheses Discovery
von: Yang, Zonglin, et al.
Veröffentlicht: (2023)
von: Yang, Zonglin, et al.
Veröffentlicht: (2023)
Can-Do! A Dataset and Neuro-Symbolic Grounded Framework for Embodied Planning with Large Multimodal Models
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
Two are better than one: Context window extension with multi-grained self-injection
von: Han, Wei, et al.
Veröffentlicht: (2024)
von: Han, Wei, et al.
Veröffentlicht: (2024)
Investigating Instruction Tuning Large Language Models on Graphs
von: Zhu, Kerui, et al.
Veröffentlicht: (2024)
von: Zhu, Kerui, et al.
Veröffentlicht: (2024)
The Jumping Reasoning Curve? Tracking the Evolution of Reasoning Performance in GPT-[n] and o-[n] Models on Multimodal Puzzles
von: Toh, Vernon Y. H., et al.
Veröffentlicht: (2025)
von: Toh, Vernon Y. H., et al.
Veröffentlicht: (2025)
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024)
von: Pang, Wei, et al.
Veröffentlicht: (2024)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
von: Pala, Tej Deep, et al.
Veröffentlicht: (2025)
von: Pala, Tej Deep, et al.
Veröffentlicht: (2025)
GraphGPT: Graph Instruction Tuning for Large Language Models
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
Toward Graph-Tokenizing Large Language Models with Reconstructive Graph Instruction Tuning
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
OctoPack: Instruction Tuning Code Large Language Models
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
von: Yang, Zonglin, et al.
Veröffentlicht: (2024)
von: Yang, Zonglin, et al.
Veröffentlicht: (2024)
DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors
von: Rakib, Tazeek Bin Abdur, et al.
Veröffentlicht: (2025)
von: Rakib, Tazeek Bin Abdur, et al.
Veröffentlicht: (2025)
WalledEval: A Comprehensive Safety Evaluation Toolkit for Large Language Models
von: Gupta, Prannaya, et al.
Veröffentlicht: (2024)
von: Gupta, Prannaya, et al.
Veröffentlicht: (2024)
MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
von: Liu, Fuxiao, et al.
Veröffentlicht: (2023)
von: Liu, Fuxiao, et al.
Veröffentlicht: (2023)
Optimizing Psychological Counseling with Instruction-Tuned Large Language Models
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
PREMISE: Matching-based Prediction for Accurate Review Recommendation
von: Han, Wei, et al.
Veröffentlicht: (2025)
von: Han, Wei, et al.
Veröffentlicht: (2025)
Unveiling the Impact of Coding Data Instruction Fine-Tuning on Large Language Models Reasoning
von: Zhang, Xinlu, et al.
Veröffentlicht: (2024)
von: Zhang, Xinlu, et al.
Veröffentlicht: (2024)
PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
von: Luo, Meng, et al.
Veröffentlicht: (2024)
von: Luo, Meng, et al.
Veröffentlicht: (2024)
Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
BioInstruct: Instruction Tuning of Large Language Models for Biomedical Natural Language Processing
von: Tran, Hieu, et al.
Veröffentlicht: (2023)
von: Tran, Hieu, et al.
Veröffentlicht: (2023)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
Is It Good Data for Multilingual Instruction Tuning or Just Bad Multilingual Evaluation for Large Language Models?
von: Chen, Pinzhen, et al.
Veröffentlicht: (2024)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2024)
Neuro-RIT: Neuron-Guided Instruction Tuning for Robust Retrieval-Augmented Language Model
von: Kim, Jaemin, et al.
Veröffentlicht: (2026)
von: Kim, Jaemin, et al.
Veröffentlicht: (2026)
Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning
von: Sun, Qi, et al.
Veröffentlicht: (2024)
von: Sun, Qi, et al.
Veröffentlicht: (2024)
Vikhr: The Family of Open-Source Instruction-Tuned Large Language Models for Russian
von: Nikolich, Aleksandr, et al.
Veröffentlicht: (2024)
von: Nikolich, Aleksandr, et al.
Veröffentlicht: (2024)
Graph-oriented Instruction Tuning of Large Language Models for Generic Graph Mining
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
Instruction Tuning for Large Language Models: A Survey
von: Zhang, Shengyu, et al.
Veröffentlicht: (2023)
von: Zhang, Shengyu, et al.
Veröffentlicht: (2023)
Lightweight Spatial Modeling for Combinatorial Information Extraction From Documents
von: Dong, Yanfei, et al.
Veröffentlicht: (2024)
von: Dong, Yanfei, et al.
Veröffentlicht: (2024)
Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models
von: Dima, George-Andrei, et al.
Veröffentlicht: (2025)
von: Dima, George-Andrei, et al.
Veröffentlicht: (2025)
Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
von: Xu, Jiashu, et al.
Veröffentlicht: (2023)
von: Xu, Jiashu, et al.
Veröffentlicht: (2023)
Keeping Yourself is Important in Downstream Tuning Multimodal Large Language Model
von: Huang, Wenke, et al.
Veröffentlicht: (2025)
von: Huang, Wenke, et al.
Veröffentlicht: (2025)
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
von: Rao, Jun, et al.
Veröffentlicht: (2024)
von: Rao, Jun, et al.
Veröffentlicht: (2024)
Federated Data-Efficient Instruction Tuning for Large Language Models
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
Leveraging Unstructured Text Data for Federated Instruction Tuning of Large Language Models
von: Ye, Rui, et al.
Veröffentlicht: (2024)
von: Ye, Rui, et al.
Veröffentlicht: (2024)
Learn from Downstream and Be Yourself in Multimodal Large Language Model Fine-Tuning
von: Huang, Wenke, et al.
Veröffentlicht: (2024)
von: Huang, Wenke, et al.
Veröffentlicht: (2024)
Mind the Gap: Conformative Decoding to Improve Output Diversity of Instruction-Tuned Large Language Models
von: Peeperkorn, Max, et al.
Veröffentlicht: (2025)
von: Peeperkorn, Max, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Toward Robust Multimodal Learning using Multimodal Foundational Models
von: Zhao, Xianbing, et al.
Veröffentlicht: (2024) -
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
von: Bhardwaj, Rishabh, et al.
Veröffentlicht: (2024) -
Self-Adaptive Sampling for Efficient Video Question-Answering on Image--Text Models
von: Han, Wei, et al.
Veröffentlicht: (2023) -
Stacked from One: Multi-Scale Self-Injection for Context Window Extension
von: Han, Wei, et al.
Veröffentlicht: (2026) -
Large Language Models for Automated Open-domain Scientific Hypotheses Discovery
von: Yang, Zonglin, et al.
Veröffentlicht: (2023)