The Best Instruction-Tuning Data are Those That Fit
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Dylan, Dai, Qirun, Peng, Hao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
di: Dai, Qirun, et al.
Pubblicazione: (2025)
di: Dai, Qirun, et al.
Pubblicazione: (2025)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
di: Cao, Yihan, et al.
Pubblicazione: (2023)
di: Cao, Yihan, et al.
Pubblicazione: (2023)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
di: Zhang, Jipeng, et al.
Pubblicazione: (2024)
di: Zhang, Jipeng, et al.
Pubblicazione: (2024)
Instruction Diversity Drives Generalization To Unseen Tasks
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
SMART: Submodular Data Mixture Strategy for Instruction Tuning
di: Renduchintala, H S V N S Kowndinya, et al.
Pubblicazione: (2024)
di: Renduchintala, H S V N S Kowndinya, et al.
Pubblicazione: (2024)
LESS: Selecting Influential Data for Targeted Instruction Tuning
di: Xia, Mengzhou, et al.
Pubblicazione: (2024)
di: Xia, Mengzhou, et al.
Pubblicazione: (2024)
Inference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models
di: Chow, Yinlam, et al.
Pubblicazione: (2024)
di: Chow, Yinlam, et al.
Pubblicazione: (2024)
To be Continuous, or to be Discrete, Those are Bits of Questions
di: Wang, Yiran, et al.
Pubblicazione: (2024)
di: Wang, Yiran, et al.
Pubblicazione: (2024)
TableDreamer: Progressive and Weakness-guided Data Synthesis from Scratch for Table Instruction Tuning
di: Zheng, Mingyu, et al.
Pubblicazione: (2025)
di: Zheng, Mingyu, et al.
Pubblicazione: (2025)
Federated Data-Efficient Instruction Tuning for Large Language Models
di: Qin, Zhen, et al.
Pubblicazione: (2024)
di: Qin, Zhen, et al.
Pubblicazione: (2024)
Contrastive Instruction Tuning
di: Yan, Tianyi Lorena, et al.
Pubblicazione: (2024)
di: Yan, Tianyi Lorena, et al.
Pubblicazione: (2024)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
di: Yuan, Zhihang, et al.
Pubblicazione: (2026)
di: Yuan, Zhihang, et al.
Pubblicazione: (2026)
ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning
di: Wu, Yang, et al.
Pubblicazione: (2024)
di: Wu, Yang, et al.
Pubblicazione: (2024)
LongForm: Effective Instruction Tuning with Reverse Instructions
di: Köksal, Abdullatif, et al.
Pubblicazione: (2023)
di: Köksal, Abdullatif, et al.
Pubblicazione: (2023)
Instruction Tuning with Human Curriculum
di: Lee, Bruce W., et al.
Pubblicazione: (2023)
di: Lee, Bruce W., et al.
Pubblicazione: (2023)
Generative Representational Instruction Tuning
di: Muennighoff, Niklas, et al.
Pubblicazione: (2024)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2024)
Instruction Tuning for Large Language Models: A Survey
di: Zhang, Shengyu, et al.
Pubblicazione: (2023)
di: Zhang, Shengyu, et al.
Pubblicazione: (2023)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
di: Agarwal, Ishika, et al.
Pubblicazione: (2025)
di: Agarwal, Ishika, et al.
Pubblicazione: (2025)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
di: Tang, Zhengyang, et al.
Pubblicazione: (2024)
di: Tang, Zhengyang, et al.
Pubblicazione: (2024)
On the Effect of Instruction Tuning Loss on Generalization
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
di: Liu, Wei, et al.
Pubblicazione: (2023)
di: Liu, Wei, et al.
Pubblicazione: (2023)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
di: He, Bingxiang, et al.
Pubblicazione: (2024)
di: He, Bingxiang, et al.
Pubblicazione: (2024)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
Parameter Efficient Instruction Tuning: An Empirical Study
di: He, Pengfei
Pubblicazione: (2024)
di: He, Pengfei
Pubblicazione: (2024)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
di: Zhang, Yi-Kai, et al.
Pubblicazione: (2025)
di: Zhang, Yi-Kai, et al.
Pubblicazione: (2025)
Multilingual Instruction Tuning With Just a Pinch of Multilinguality
di: Shaham, Uri, et al.
Pubblicazione: (2024)
di: Shaham, Uri, et al.
Pubblicazione: (2024)
Instruction Fine-Tuning: Does Prompt Loss Matter?
di: Huerta-Enochian, Mathew, et al.
Pubblicazione: (2024)
di: Huerta-Enochian, Mathew, et al.
Pubblicazione: (2024)
Picky LLMs and Unreliable RMs: An Empirical Study on Safety Alignment after Instruction Tuning
di: Li, Guanlin, et al.
Pubblicazione: (2025)
di: Li, Guanlin, et al.
Pubblicazione: (2025)
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
di: Wu, Xuansheng, et al.
Pubblicazione: (2023)
di: Wu, Xuansheng, et al.
Pubblicazione: (2023)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
di: Köksal, Abdullatif, et al.
Pubblicazione: (2024)
di: Köksal, Abdullatif, et al.
Pubblicazione: (2024)
$\textbf{Only-IF}$:Revealing the Decisive Effect of Instruction Diversity on Generalization
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
di: Zou, Junyi
Pubblicazione: (2026)
di: Zou, Junyi
Pubblicazione: (2026)
DynamixSFT: Dynamic Mixture Optimization of Instruction Tuning Collections
di: Shin, Haebin, et al.
Pubblicazione: (2025)
di: Shin, Haebin, et al.
Pubblicazione: (2025)
CYCLE-INSTRUCT: Fully Seed-Free Instruction Tuning via Dual Self-Training and Cycle Consistency
di: Shen, Zhanming, et al.
Pubblicazione: (2025)
di: Shen, Zhanming, et al.
Pubblicazione: (2025)
Reconstructive Visual Instruction Tuning
di: Wang, Haochen, et al.
Pubblicazione: (2024)
di: Wang, Haochen, et al.
Pubblicazione: (2024)
Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
di: Dai, Qirun, et al.
Pubblicazione: (2025) -
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
di: Cao, Yihan, et al.
Pubblicazione: (2023) -
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
di: Zhang, Jipeng, et al.
Pubblicazione: (2024) -
Instruction Diversity Drives Generalization To Unseen Tasks
di: Zhang, Dylan, et al.
Pubblicazione: (2024) -
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)