GradientSpace: Unsupervised Data Clustering for Improved Instruction Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Sridharan, Shrihari, Ravikumar, Deepak, Raghunathan, Anand, Roy, Kaushik |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Experts are all you need: A Composable Framework for Large Language Model Inference
by: Sridharan, Shrihari, et al.
Published: (2025)
by: Sridharan, Shrihari, et al.
Published: (2025)
Ev-Edge: Efficient Execution of Event-based Vision Algorithms on Commodity Edge Platforms
by: Sridharan, Shrihari, et al.
Published: (2024)
by: Sridharan, Shrihari, et al.
Published: (2024)
KV-CAR: KV Cache Compression using Autoencoders and KV Reuse in Large Language Models
by: Roy, Sourjya, et al.
Published: (2025)
by: Roy, Sourjya, et al.
Published: (2025)
Finding the Muses: Identifying Coresets through Loss Trajectories
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
SAP: Corrective Machine Unlearning with Scaled Activation Projection for Label Noise Robustness
by: Kodge, Sangamesh, et al.
Published: (2024)
by: Kodge, Sangamesh, et al.
Published: (2024)
Unveiling Privacy, Memorization, and Input Curvature Links
by: Ravikumar, Deepak, et al.
Published: (2024)
by: Ravikumar, Deepak, et al.
Published: (2024)
TRIM: Token-wise Attention-Derived Saliency for Data-Efficient Instruction Tuning
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
by: Zhang, Jipeng, et al.
Published: (2024)
by: Zhang, Jipeng, et al.
Published: (2024)
Analysis of Long Range Dependency Understanding in State Space Models
by: Ravikumar, Srividya, et al.
Published: (2026)
by: Ravikumar, Srividya, et al.
Published: (2026)
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
by: Roy, Arjun, et al.
Published: (2026)
by: Roy, Arjun, et al.
Published: (2026)
MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning
by: Awasthi, Ankita, et al.
Published: (2026)
by: Awasthi, Ankita, et al.
Published: (2026)
Coresets from Trajectories: Selecting Data via Correlation of Loss Differences
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
XAMBA: Enabling Efficient State Space Models on Resource-Constrained Neural Processing Units
by: Das, Arghadip, et al.
Published: (2025)
by: Das, Arghadip, et al.
Published: (2025)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
by: Yuan, Zhihang, et al.
Published: (2026)
by: Yuan, Zhihang, et al.
Published: (2026)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
by: Wang, Zige, et al.
Published: (2025)
by: Wang, Zige, et al.
Published: (2025)
T-SHIRT: Token-Selective Hierarchical Data Selection for Instruction Tuning
by: Fu, Yanjun, et al.
Published: (2025)
by: Fu, Yanjun, et al.
Published: (2025)
Curvature Clues: Decoding Deep Learning Privacy with Input Loss Curvature
by: Ravikumar, Deepak, et al.
Published: (2024)
by: Ravikumar, Deepak, et al.
Published: (2024)
Deep Unlearning: Fast and Efficient Gradient-free Approach to Class Forgetting
by: Kodge, Sangamesh, et al.
Published: (2023)
by: Kodge, Sangamesh, et al.
Published: (2023)
S2D: Selective Spectral Decay for Quantization-Friendly Conditioning of Neural Activations
by: Chavan, Arnav, et al.
Published: (2026)
by: Chavan, Arnav, et al.
Published: (2026)
Unsupervised Optimisation of GNNs for Node Clustering
by: Leeney, William, et al.
Published: (2024)
by: Leeney, William, et al.
Published: (2024)
Weight Ensembling Improves Reasoning in Language Models
by: Dang, Xingyu, et al.
Published: (2025)
by: Dang, Xingyu, et al.
Published: (2025)
Federated Continual Instruction Tuning
by: Guo, Haiyang, et al.
Published: (2025)
by: Guo, Haiyang, et al.
Published: (2025)
Power side-channel leakage localization through adversarial training of deep neural networks
by: Gammell, Jimmy, et al.
Published: (2024)
by: Gammell, Jimmy, et al.
Published: (2024)
Eigen Attention: Attention in Low-Rank Space for KV Cache Compression
by: Saxena, Utkarsh, et al.
Published: (2024)
by: Saxena, Utkarsh, et al.
Published: (2024)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
by: Cao, Yihan, et al.
Published: (2023)
by: Cao, Yihan, et al.
Published: (2023)
The Best Instruction-Tuning Data are Those That Fit
by: Zhang, Dylan, et al.
Published: (2025)
by: Zhang, Dylan, et al.
Published: (2025)
Reward Learning from Best-of-$N$ Preference Data: Targets, Tradeoffs, and Design Principles
by: Pukdee, Rattana, et al.
Published: (2026)
by: Pukdee, Rattana, et al.
Published: (2026)
LEAD: Iterative Data Selection for Efficient LLM Instruction Tuning
by: Lin, Xiaotian, et al.
Published: (2025)
by: Lin, Xiaotian, et al.
Published: (2025)
Diffusion Instruction Tuning
by: Jin, Chen, et al.
Published: (2025)
by: Jin, Chen, et al.
Published: (2025)
Advancing Compressed Video Action Recognition through Progressive Knowledge Distillation
by: Soufleri, Efstathia, et al.
Published: (2024)
by: Soufleri, Efstathia, et al.
Published: (2024)
Adapt-$\infty$: Scalable Continual Multimodal Instruction Tuning via Dynamic Data Selection
by: Maharana, Adyasha, et al.
Published: (2024)
by: Maharana, Adyasha, et al.
Published: (2024)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
SMART: Submodular Data Mixture Strategy for Instruction Tuning
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2024)
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2024)
LESS: Selecting Influential Data for Targeted Instruction Tuning
by: Xia, Mengzhou, et al.
Published: (2024)
by: Xia, Mengzhou, et al.
Published: (2024)
Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data
by: Rallapalli, Swati, et al.
Published: (2025)
by: Rallapalli, Swati, et al.
Published: (2025)
LogicTree: Structured Proof Exploration for Coherent and Rigorous Logical Reasoning with Large Language Models
by: He, Kang, et al.
Published: (2025)
by: He, Kang, et al.
Published: (2025)
2D-ThermAl: Physics-Informed Framework for Thermal Analysis of Circuits using Generative AI
by: Chandra, Soumyadeep, et al.
Published: (2025)
by: Chandra, Soumyadeep, et al.
Published: (2025)
Federated Data-Efficient Instruction Tuning for Large Language Models
by: Qin, Zhen, et al.
Published: (2024)
by: Qin, Zhen, et al.
Published: (2024)
Refining Filter Global Feature Weighting for Fully-Unsupervised Clustering
by: Galis, Fabian, et al.
Published: (2025)
by: Galis, Fabian, et al.
Published: (2025)
Federated Unsupervised Domain Generalization using Global and Local Alignment of Gradients
by: Pourpanah, Farhad, et al.
Published: (2024)
by: Pourpanah, Farhad, et al.
Published: (2024)
Similar Items
-
Experts are all you need: A Composable Framework for Large Language Model Inference
by: Sridharan, Shrihari, et al.
Published: (2025) -
Ev-Edge: Efficient Execution of Event-based Vision Algorithms on Commodity Edge Platforms
by: Sridharan, Shrihari, et al.
Published: (2024) -
KV-CAR: KV Cache Compression using Autoencoders and KV Reuse in Large Language Models
by: Roy, Sourjya, et al.
Published: (2025) -
Finding the Muses: Identifying Coresets through Loss Trajectories
by: Nagaraj, Manish, et al.
Published: (2025) -
SAP: Corrective Machine Unlearning with Scaled Activation Projection for Label Noise Robustness
by: Kodge, Sangamesh, et al.
Published: (2024)