MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yangning, Lan, Zihua, Qingsong, Lv, Li, Yinghui, Zheng, Hai-Tao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RAISE: Reinforced Adaptive Instruction Selection For Large Language Models
di: Lv, Qingsong, et al.
Pubblicazione: (2025)
di: Lv, Qingsong, et al.
Pubblicazione: (2025)
UltraWiki: Ultra-fine-grained Entity Set Expansion with Negative Seed Entities
di: Li, Yangning, et al.
Pubblicazione: (2024)
di: Li, Yangning, et al.
Pubblicazione: (2024)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
di: Li, Yangning, et al.
Pubblicazione: (2025)
di: Li, Yangning, et al.
Pubblicazione: (2025)
From Token to Line: Enhancing Code Generation with a Long-Term Perspective
di: Lu, Tingwei, et al.
Pubblicazione: (2025)
di: Lu, Tingwei, et al.
Pubblicazione: (2025)
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
di: Huang, Shulin, et al.
Pubblicazione: (2023)
di: Huang, Shulin, et al.
Pubblicazione: (2023)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
di: Li, Yangning, et al.
Pubblicazione: (2022)
di: Li, Yangning, et al.
Pubblicazione: (2022)
Exploring the Implicit Semantic Ability of Multimodal Large Language Models: A Pilot Study on Entity Set Expansion
di: Wang, Hebin, et al.
Pubblicazione: (2024)
di: Wang, Hebin, et al.
Pubblicazione: (2024)
When LLMs Meet Cunning Texts: A Fallacy Understanding Benchmark for Large Language Models
di: Li, Yinghui, et al.
Pubblicazione: (2024)
di: Li, Yinghui, et al.
Pubblicazione: (2024)
On the (In)Effectiveness of Large Language Models for Chinese Text Correction
di: Li, Yinghui, et al.
Pubblicazione: (2023)
di: Li, Yinghui, et al.
Pubblicazione: (2023)
Bidirectional End-to-End Learning of Retriever-Reader Paradigm for Entity Linking
di: Li, Yinghui, et al.
Pubblicazione: (2023)
di: Li, Yinghui, et al.
Pubblicazione: (2023)
Measuring Data Diversity for Instruction Tuning: A Systematic Analysis and A Reliable Metric
di: Yang, Yuming, et al.
Pubblicazione: (2025)
di: Yang, Yuming, et al.
Pubblicazione: (2025)
Rethinking the Roles of Large Language Models in Chinese Grammatical Error Correction
di: Li, Yinghui, et al.
Pubblicazione: (2024)
di: Li, Yinghui, et al.
Pubblicazione: (2024)
Correct Like Humans: Progressive Learning Framework for Chinese Text Error Correction
di: Li, Yinghui, et al.
Pubblicazione: (2023)
di: Li, Yinghui, et al.
Pubblicazione: (2023)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
di: Chen, Shaoshen, et al.
Pubblicazione: (2025)
di: Chen, Shaoshen, et al.
Pubblicazione: (2025)
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction
di: Ye, Jingheng, et al.
Pubblicazione: (2025)
di: Ye, Jingheng, et al.
Pubblicazione: (2025)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
di: Li, Yangning, et al.
Pubblicazione: (2025)
di: Li, Yangning, et al.
Pubblicazione: (2025)
Data Diversity Matters for Robust Instruction Tuning
di: Bukharin, Alexander, et al.
Pubblicazione: (2023)
di: Bukharin, Alexander, et al.
Pubblicazione: (2023)
Mitigating Catastrophic Forgetting in Multi-domain Chinese Spelling Correction by Multi-stage Knowledge Transfer Framework
di: Xing, Peng, et al.
Pubblicazione: (2024)
di: Xing, Peng, et al.
Pubblicazione: (2024)
MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models
di: Jin, Bohan, et al.
Pubblicazione: (2025)
di: Jin, Bohan, et al.
Pubblicazione: (2025)
ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions
di: Ye, Jingheng, et al.
Pubblicazione: (2024)
di: Ye, Jingheng, et al.
Pubblicazione: (2024)
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
di: Zhang, Ding, et al.
Pubblicazione: (2024)
di: Zhang, Ding, et al.
Pubblicazione: (2024)
Automatic Context Pattern Generation for Entity Set Expansion
di: Li, Yinghui, et al.
Pubblicazione: (2022)
di: Li, Yinghui, et al.
Pubblicazione: (2022)
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
di: Li, Yinghui, et al.
Pubblicazione: (2025)
di: Li, Yinghui, et al.
Pubblicazione: (2025)
Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
di: Li, Yangning, et al.
Pubblicazione: (2024)
di: Li, Yangning, et al.
Pubblicazione: (2024)
EvoConfig: Self-Evolving Multi-Agent Systems for Efficient Autonomous Environment Configuration
di: Guo, Xinshuai, et al.
Pubblicazione: (2026)
di: Guo, Xinshuai, et al.
Pubblicazione: (2026)
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
di: Liu, Daixian, et al.
Pubblicazione: (2026)
di: Liu, Daixian, et al.
Pubblicazione: (2026)
Beyond Similarity: A Gradient-based Graph Method for Instruction Tuning Data Selection
di: Zhao, Yang, et al.
Pubblicazione: (2025)
di: Zhao, Yang, et al.
Pubblicazione: (2025)
Importance-Aware Data Selection for Efficient LLM Instruction Tuning
di: Jiang, Tingyu, et al.
Pubblicazione: (2025)
di: Jiang, Tingyu, et al.
Pubblicazione: (2025)
Know the Unknown: An Uncertainty-Sensitive Method for LLM Instruction Tuning
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
MADS: Model-Aware Diverse Core Set Selection for Instruction Tuning
di: Bai, Yi, et al.
Pubblicazione: (2026)
di: Bai, Yi, et al.
Pubblicazione: (2026)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
di: Liu, Yilun, et al.
Pubblicazione: (2023)
di: Liu, Yilun, et al.
Pubblicazione: (2023)
GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models
di: Ruan, Zhiwen, et al.
Pubblicazione: (2026)
di: Ruan, Zhiwen, et al.
Pubblicazione: (2026)
M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets
di: Zhao, Chunguang, et al.
Pubblicazione: (2025)
di: Zhao, Chunguang, et al.
Pubblicazione: (2025)
Superfiltering: Weak-to-Strong Data Filtering for Fast Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
Dynamics of Instruction Fine-Tuning for Chinese Large Language Models
di: Song, Chiyu, et al.
Pubblicazione: (2023)
di: Song, Chiyu, et al.
Pubblicazione: (2023)
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
di: Yan, Jun, et al.
Pubblicazione: (2023)
di: Yan, Jun, et al.
Pubblicazione: (2023)
LLaVA-Video: Video Instruction Tuning With Synthetic Data
di: Zhang, Yuanhan, et al.
Pubblicazione: (2024)
di: Zhang, Yuanhan, et al.
Pubblicazione: (2024)
Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models
di: Li, Haoran, et al.
Pubblicazione: (2024)
di: Li, Haoran, et al.
Pubblicazione: (2024)
Data Selection for Multi-turn Dialogue Instruction Tuning
di: Li, Bo, et al.
Pubblicazione: (2026)
di: Li, Bo, et al.
Pubblicazione: (2026)
WeatherSyn: An Instruction Tuning MLLM For Weather Forecasting Report Generation
di: Zheng, Zinan, et al.
Pubblicazione: (2026)
di: Zheng, Zinan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
RAISE: Reinforced Adaptive Instruction Selection For Large Language Models
di: Lv, Qingsong, et al.
Pubblicazione: (2025) -
UltraWiki: Ultra-fine-grained Entity Set Expansion with Negative Seed Entities
di: Li, Yangning, et al.
Pubblicazione: (2024) -
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
di: Li, Yangning, et al.
Pubblicazione: (2025) -
From Token to Line: Enhancing Code Generation with a Long-Term Perspective
di: Lu, Tingwei, et al.
Pubblicazione: (2025) -
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
di: Huang, Shulin, et al.
Pubblicazione: (2023)