LESS: Selecting Influential Data for Targeted Instruction Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xia, Mengzhou, Malladi, Sadhika, Gururangan, Suchin, Arora, Sanjeev, Chen, Danqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization
von: Razin, Noam, et al.
Veröffentlicht: (2024)
von: Razin, Noam, et al.
Veröffentlicht: (2024)
Trainable Transformer in Transformer
von: Panigrahi, Abhishek, et al.
Veröffentlicht: (2023)
von: Panigrahi, Abhishek, et al.
Veröffentlicht: (2023)
Diversity-driven Data Selection for Language Model Tuning through Sparse Autoencoder
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
Fine-Tuning Language Models with Just Forward Passes
von: Malladi, Sadhika, et al.
Veröffentlicht: (2023)
von: Malladi, Sadhika, et al.
Veröffentlicht: (2023)
Provable unlearning in topic modeling and downstream tasks
von: Wei, Stanley, et al.
Veröffentlicht: (2024)
von: Wei, Stanley, et al.
Veröffentlicht: (2024)
Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
Preference Learning Algorithms Do Not Learn Preference Rankings
von: Chen, Angelica, et al.
Veröffentlicht: (2024)
von: Chen, Angelica, et al.
Veröffentlicht: (2024)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
von: Li, Ming, et al.
Veröffentlicht: (2024)
von: Li, Ming, et al.
Veröffentlicht: (2024)
SILO Language Models: Isolating Legal Risk In a Nonparametric Datastore
von: Min, Sewon, et al.
Veröffentlicht: (2023)
von: Min, Sewon, et al.
Veröffentlicht: (2023)
Influential Language Data Selection via Gradient Trajectory Pursuit
von: Deng, Zhiwei, et al.
Veröffentlicht: (2024)
von: Deng, Zhiwei, et al.
Veröffentlicht: (2024)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
What is in Your Safe Data? Identifying Benign Data that Breaks Safety
von: He, Luxi, et al.
Veröffentlicht: (2024)
von: He, Luxi, et al.
Veröffentlicht: (2024)
SimPO: Simple Preference Optimization with a Reference-Free Reward
von: Meng, Yu, et al.
Veröffentlicht: (2024)
von: Meng, Yu, et al.
Veröffentlicht: (2024)
LitSearch: A Retrieval Benchmark for Scientific Literature Search
von: Ajith, Anirudh, et al.
Veröffentlicht: (2024)
von: Ajith, Anirudh, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
von: Zhang, Jipeng, et al.
Veröffentlicht: (2024)
von: Zhang, Jipeng, et al.
Veröffentlicht: (2024)
The Coverage Principle: How Pre-Training Enables Post-Training
von: Chen, Fan, et al.
Veröffentlicht: (2025)
von: Chen, Fan, et al.
Veröffentlicht: (2025)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
von: Liu, Wei, et al.
Veröffentlicht: (2023)
von: Liu, Wei, et al.
Veröffentlicht: (2023)
Rethinking Thinking Tokens: LLMs as Improvement Operators
von: Madaan, Lovish, et al.
Veröffentlicht: (2025)
von: Madaan, Lovish, et al.
Veröffentlicht: (2025)
ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning
von: Wu, Yang, et al.
Veröffentlicht: (2024)
von: Wu, Yang, et al.
Veröffentlicht: (2024)
In Good GRACEs: Principled Teacher Selection for Knowledge Distillation
von: Panigrahi, Abhishek, et al.
Veröffentlicht: (2025)
von: Panigrahi, Abhishek, et al.
Veröffentlicht: (2025)
Self-Generated Critiques Boost Reward Modeling for Language Models
von: Yu, Yue, et al.
Veröffentlicht: (2024)
von: Yu, Yue, et al.
Veröffentlicht: (2024)
Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning
von: Li, Xiaochuan, et al.
Veröffentlicht: (2024)
von: Li, Xiaochuan, et al.
Veröffentlicht: (2024)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
von: Wang, Zige, et al.
Veröffentlicht: (2025)
von: Wang, Zige, et al.
Veröffentlicht: (2025)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
The Best Instruction-Tuning Data are Those That Fit
von: Zhang, Dylan, et al.
Veröffentlicht: (2025)
von: Zhang, Dylan, et al.
Veröffentlicht: (2025)
Contrastive Instruction Tuning
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
Why is Your Language Model a Poor Implicit Reward Model?
von: Razin, Noam, et al.
Veröffentlicht: (2025)
von: Razin, Noam, et al.
Veröffentlicht: (2025)
Contextual Drag: How Errors in the Context Affect LLM Reasoning
von: Cheng, Yun, et al.
Veröffentlicht: (2026)
von: Cheng, Yun, et al.
Veröffentlicht: (2026)
SMART: Submodular Data Mixture Strategy for Instruction Tuning
von: Renduchintala, H S V N S Kowndinya, et al.
Veröffentlicht: (2024)
von: Renduchintala, H S V N S Kowndinya, et al.
Veröffentlicht: (2024)
Lory: Fully Differentiable Mixture-of-Experts for Autoregressive Language Model Pre-training
von: Zhong, Zexuan, et al.
Veröffentlicht: (2024)
von: Zhong, Zexuan, et al.
Veröffentlicht: (2024)
Representing Rule-based Chatbots with Transformers
von: Friedman, Dan, et al.
Veröffentlicht: (2024)
von: Friedman, Dan, et al.
Veröffentlicht: (2024)
Federated Data-Efficient Instruction Tuning for Large Language Models
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
On the SDEs and Scaling Rules for Adaptive Gradient Algorithms
von: Malladi, Sadhika, et al.
Veröffentlicht: (2022)
von: Malladi, Sadhika, et al.
Veröffentlicht: (2022)
LongForm: Effective Instruction Tuning with Reverse Instructions
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2023)
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2023)
Can Models Learn Skill Composition from Examples?
von: Zhao, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2024)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
Instruction Tuning with Human Curriculum
von: Lee, Bruce W., et al.
Veröffentlicht: (2023)
von: Lee, Bruce W., et al.
Veröffentlicht: (2023)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
von: Agarwal, Ishika, et al.
Veröffentlicht: (2025)
von: Agarwal, Ishika, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization
von: Razin, Noam, et al.
Veröffentlicht: (2024) -
Trainable Transformer in Transformer
von: Panigrahi, Abhishek, et al.
Veröffentlicht: (2023) -
Diversity-driven Data Selection for Language Model Tuning through Sparse Autoencoder
von: Yang, Xianjun, et al.
Veröffentlicht: (2025) -
Fine-Tuning Language Models with Just Forward Passes
von: Malladi, Sadhika, et al.
Veröffentlicht: (2023) -
Provable unlearning in topic modeling and downstream tasks
von: Wei, Stanley, et al.
Veröffentlicht: (2024)