Diversity Measurement and Subset Selection for Instruction Tuning Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Peiqi, Shen, Yikang, Guo, Zhen, Stallone, Matthew, Kim, Yoon, Golland, Polina, Panda, Rameswar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gated Linear Attention Transformers with Hardware-Efficient Training
von: Yang, Songlin, et al.
Veröffentlicht: (2023)
von: Yang, Songlin, et al.
Veröffentlicht: (2023)
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation
von: Guo, Zhen, et al.
Veröffentlicht: (2024)
von: Guo, Zhen, et al.
Veröffentlicht: (2024)
Calibrating Expressions of Certainty
von: Wang, Peiqi, et al.
Veröffentlicht: (2024)
von: Wang, Peiqi, et al.
Veröffentlicht: (2024)
FlashFormer: Whole-Model Kernels for Efficient Low-Batch Inference
von: Nrusimha, Aniruddha, et al.
Veröffentlicht: (2025)
von: Nrusimha, Aniruddha, et al.
Veröffentlicht: (2025)
Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler
von: Shen, Yikang, et al.
Veröffentlicht: (2024)
von: Shen, Yikang, et al.
Veröffentlicht: (2024)
PaTH Attention: Position Encoding via Accumulating Householder Transformations
von: Yang, Songlin, et al.
Veröffentlicht: (2025)
von: Yang, Songlin, et al.
Veröffentlicht: (2025)
Scaling Stick-Breaking Attention: An Efficient Implementation and In-depth Study
von: Tan, Shawn, et al.
Veröffentlicht: (2024)
von: Tan, Shawn, et al.
Veröffentlicht: (2024)
Mitigating the Impact of Outlier Channels for Language Model Quantization with Activation Regularization
von: Nrusimha, Aniruddha, et al.
Veröffentlicht: (2024)
von: Nrusimha, Aniruddha, et al.
Veröffentlicht: (2024)
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
von: Pan, Bowen, et al.
Veröffentlicht: (2024)
von: Pan, Bowen, et al.
Veröffentlicht: (2024)
Scattered Mixture-of-Experts Implementation
von: Tan, Shawn, et al.
Veröffentlicht: (2024)
von: Tan, Shawn, et al.
Veröffentlicht: (2024)
Parallelizing Linear Transformers with the Delta Rule over Sequence Length
von: Yang, Songlin, et al.
Veröffentlicht: (2024)
von: Yang, Songlin, et al.
Veröffentlicht: (2024)
SHED: Shapley-Based Automated Dataset Refinement for Instruction Fine-Tuning
von: He, Yexiao, et al.
Veröffentlicht: (2024)
von: He, Yexiao, et al.
Veröffentlicht: (2024)
Data Diversity Matters for Robust Instruction Tuning
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
von: Brandon, William, et al.
Veröffentlicht: (2024)
von: Brandon, William, et al.
Veröffentlicht: (2024)
LCSB: Layer-Cyclic Selective Backpropagation for Memory-Efficient On-Device LLM Fine-Tuning
von: Park, Juneyoung, et al.
Veröffentlicht: (2026)
von: Park, Juneyoung, et al.
Veröffentlicht: (2026)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
Learning to Generate Instruction Tuning Datasets for Zero-Shot Task Adaptation
von: Nayak, Nihal V., et al.
Veröffentlicht: (2024)
von: Nayak, Nihal V., et al.
Veröffentlicht: (2024)
Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts
von: Kang, Junmo, et al.
Veröffentlicht: (2024)
von: Kang, Junmo, et al.
Veröffentlicht: (2024)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
von: Li, Ming, et al.
Veröffentlicht: (2024)
von: Li, Ming, et al.
Veröffentlicht: (2024)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
LESS: Selecting Influential Data for Targeted Instruction Tuning
von: Xia, Mengzhou, et al.
Veröffentlicht: (2024)
von: Xia, Mengzhou, et al.
Veröffentlicht: (2024)
Dynamic Subset Tuning: Expanding the Operational Range of Parameter-Efficient Training for Large Language Models
von: Stahlberg, Felix, et al.
Veröffentlicht: (2024)
von: Stahlberg, Felix, et al.
Veröffentlicht: (2024)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
Improving Multilingual Instruction Finetuning via Linguistically Natural and Diverse Datasets
von: Indurthi, Sathish Reddy, et al.
Veröffentlicht: (2024)
von: Indurthi, Sathish Reddy, et al.
Veröffentlicht: (2024)
Synthetic Data RL: Task Definition Is All You Need
von: Guo, Yiduo, et al.
Veröffentlicht: (2025)
von: Guo, Yiduo, et al.
Veröffentlicht: (2025)
Federated Data-Efficient Instruction Tuning for Large Language Models
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
CoMMIT: Coordinated Multimodal Instruction Tuning
von: Li, Xintong, et al.
Veröffentlicht: (2024)
von: Li, Xintong, et al.
Veröffentlicht: (2024)
ESD: Expected Squared Difference as a Tuning-Free Trainable Calibration Measure
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2023)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2023)
Contrastive Instruction Tuning
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
Multi-Scale Heterogeneous Text-Attributed Graph Datasets From Diverse Domains
von: Liu, Yunhui, et al.
Veröffentlicht: (2024)
von: Liu, Yunhui, et al.
Veröffentlicht: (2024)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
von: Zhang, Jipeng, et al.
Veröffentlicht: (2024)
von: Zhang, Jipeng, et al.
Veröffentlicht: (2024)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
HiDe-LLaVA: Hierarchical Decoupling for Continual Instruction Tuning of Multimodal Large Language Model
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments
von: Xu, Zhangchen, et al.
Veröffentlicht: (2025)
von: Xu, Zhangchen, et al.
Veröffentlicht: (2025)
DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition
von: Kim, Siun, et al.
Veröffentlicht: (2026)
von: Kim, Siun, et al.
Veröffentlicht: (2026)
CIDAR: Culturally Relevant Instruction Dataset For Arabic
von: Alyafeai, Zaid, et al.
Veröffentlicht: (2024)
von: Alyafeai, Zaid, et al.
Veröffentlicht: (2024)
Diversity-driven Data Selection for Language Model Tuning through Sparse Autoencoder
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Gated Linear Attention Transformers with Hardware-Efficient Training
von: Yang, Songlin, et al.
Veröffentlicht: (2023) -
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation
von: Guo, Zhen, et al.
Veröffentlicht: (2024) -
Calibrating Expressions of Certainty
von: Wang, Peiqi, et al.
Veröffentlicht: (2024) -
FlashFormer: Whole-Model Kernels for Efficient Low-Batch Inference
von: Nrusimha, Aniruddha, et al.
Veröffentlicht: (2025) -
Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler
von: Shen, Yikang, et al.
Veröffentlicht: (2024)