SMART: Submodular Data Mixture Strategy for Instruction Tuning
Fuente:
arXiv
Guardado en:
| Autores principales: | Renduchintala, H S V N S Kowndinya, Bhatia, Sumit, Ramakrishnan, Ganesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?
por: Renduchintala, H S V N S Kowndinya, et al.
Publicado: (2026)
por: Renduchintala, H S V N S Kowndinya, et al.
Publicado: (2026)
On the Effect of Instruction Tuning Loss on Generalization
por: Chatterjee, Anwoy, et al.
Publicado: (2025)
por: Chatterjee, Anwoy, et al.
Publicado: (2025)
POSIX: A Prompt Sensitivity Index For Large Language Models
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
por: Gupta, Raavi, et al.
Publicado: (2025)
por: Gupta, Raavi, et al.
Publicado: (2025)
DynamixSFT: Dynamic Mixture Optimization of Instruction Tuning Collections
por: Shin, Haebin, et al.
Publicado: (2025)
por: Shin, Haebin, et al.
Publicado: (2025)
Tied-Lora: Enhancing parameter efficiency of LoRA with weight tying
por: Renduchintala, Adithya, et al.
Publicado: (2023)
por: Renduchintala, Adithya, et al.
Publicado: (2023)
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
por: Hui, Tingfeng, et al.
Publicado: (2024)
por: Hui, Tingfeng, et al.
Publicado: (2024)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
por: Cao, Yihan, et al.
Publicado: (2023)
por: Cao, Yihan, et al.
Publicado: (2023)
The Best Instruction-Tuning Data are Those That Fit
por: Zhang, Dylan, et al.
Publicado: (2025)
por: Zhang, Dylan, et al.
Publicado: (2025)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
por: Li, Ming, et al.
Publicado: (2024)
por: Li, Ming, et al.
Publicado: (2024)
LESS: Selecting Influential Data for Targeted Instruction Tuning
por: Xia, Mengzhou, et al.
Publicado: (2024)
por: Xia, Mengzhou, et al.
Publicado: (2024)
Inducing Robustness in a 2 Dimensional Direct Preference Optimization Paradigm
por: Shashidhar, Sarvesh, et al.
Publicado: (2025)
por: Shashidhar, Sarvesh, et al.
Publicado: (2025)
Federated Data-Efficient Instruction Tuning for Large Language Models
por: Qin, Zhen, et al.
Publicado: (2024)
por: Qin, Zhen, et al.
Publicado: (2024)
Contrastive Instruction Tuning
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
por: Yuan, Zhihang, et al.
Publicado: (2026)
por: Yuan, Zhihang, et al.
Publicado: (2026)
LongForm: Effective Instruction Tuning with Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2023)
por: Köksal, Abdullatif, et al.
Publicado: (2023)
Generative Representational Instruction Tuning
por: Muennighoff, Niklas, et al.
Publicado: (2024)
por: Muennighoff, Niklas, et al.
Publicado: (2024)
Instruction Tuning with Human Curriculum
por: Lee, Bruce W., et al.
Publicado: (2023)
por: Lee, Bruce W., et al.
Publicado: (2023)
Diversity as a Reward: Fine-Tuning LLMs on a Mixture of Domain-Undetermined Data
por: Ling, Zhenqing, et al.
Publicado: (2025)
por: Ling, Zhenqing, et al.
Publicado: (2025)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
por: Zhang, Jipeng, et al.
Publicado: (2024)
por: Zhang, Jipeng, et al.
Publicado: (2024)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
por: Agarwal, Ishika, et al.
Publicado: (2025)
por: Agarwal, Ishika, et al.
Publicado: (2025)
XFT: Unlocking the Power of Code Instruction Tuning by Simply Merging Upcycled Mixture-of-Experts
por: Ding, Yifeng, et al.
Publicado: (2024)
por: Ding, Yifeng, et al.
Publicado: (2024)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
por: Liu, Wei, et al.
Publicado: (2023)
por: Liu, Wei, et al.
Publicado: (2023)
Parameter-Efficient Routed Fine-Tuning: Mixture-of-Experts Demands Mixture of Adaptation Modules
por: Liu, Yilun, et al.
Publicado: (2025)
por: Liu, Yilun, et al.
Publicado: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
por: He, Bingxiang, et al.
Publicado: (2024)
por: He, Bingxiang, et al.
Publicado: (2024)
Parameter Efficient Instruction Tuning: An Empirical Study
por: He, Pengfei
Publicado: (2024)
por: He, Pengfei
Publicado: (2024)
ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning
por: Wu, Yang, et al.
Publicado: (2024)
por: Wu, Yang, et al.
Publicado: (2024)
TableDreamer: Progressive and Weakness-guided Data Synthesis from Scratch for Table Instruction Tuning
por: Zheng, Mingyu, et al.
Publicado: (2025)
por: Zheng, Mingyu, et al.
Publicado: (2025)
SMART: Self-Aware Agent for Tool Overuse Mitigation
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
MoR: Mixture of Ranks for Low-Rank Adaptation Tuning
por: Tang, Chuanyu, et al.
Publicado: (2024)
por: Tang, Chuanyu, et al.
Publicado: (2024)
MEPT: Mixture of Expert Prompt Tuning as a Manifold Mapper
por: Zeng, Runjia, et al.
Publicado: (2025)
por: Zeng, Runjia, et al.
Publicado: (2025)
Multilingual Instruction Tuning With Just a Pinch of Multilinguality
por: Shaham, Uri, et al.
Publicado: (2024)
por: Shaham, Uri, et al.
Publicado: (2024)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
por: Tang, Zhengyang, et al.
Publicado: (2024)
por: Tang, Zhengyang, et al.
Publicado: (2024)
Instruction Fine-Tuning: Does Prompt Loss Matter?
por: Huerta-Enochian, Mathew, et al.
Publicado: (2024)
por: Huerta-Enochian, Mathew, et al.
Publicado: (2024)
Instruction Tuning for Large Language Models: A Survey
por: Zhang, Shengyu, et al.
Publicado: (2023)
por: Zhang, Shengyu, et al.
Publicado: (2023)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2024)
por: Köksal, Abdullatif, et al.
Publicado: (2024)
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
por: Wu, Xuansheng, et al.
Publicado: (2023)
por: Wu, Xuansheng, et al.
Publicado: (2023)
Improving Model Evaluation using SMART Filtering of Benchmark Datasets
por: Gupta, Vipul, et al.
Publicado: (2024)
por: Gupta, Vipul, et al.
Publicado: (2024)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
por: Zou, Junyi
Publicado: (2026)
por: Zou, Junyi
Publicado: (2026)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
por: Xiao, Yuxin, et al.
Publicado: (2024)
por: Xiao, Yuxin, et al.
Publicado: (2024)
Ejemplares similares
-
Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?
por: Renduchintala, H S V N S Kowndinya, et al.
Publicado: (2026) -
On the Effect of Instruction Tuning Loss on Generalization
por: Chatterjee, Anwoy, et al.
Publicado: (2025) -
POSIX: A Prompt Sensitivity Index For Large Language Models
por: Chatterjee, Anwoy, et al.
Publicado: (2024) -
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
por: Gupta, Raavi, et al.
Publicado: (2025) -
DynamixSFT: Dynamic Mixture Optimization of Instruction Tuning Collections
por: Shin, Haebin, et al.
Publicado: (2025)