Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Fangxin, Baghershahi, Peyman, He, Langzhou, Zou, Henry Peng, Medya, Sourav, Yu, Philip S. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
by: Chaturvedi, Rochana, et al.
Published: (2025)
by: Chaturvedi, Rochana, et al.
Published: (2025)
Unsupervised Prompting for Graph Neural Networks
by: Baghershahi, Peyman, et al.
Published: (2025)
by: Baghershahi, Peyman, et al.
Published: (2025)
GRAPHLCP: Structure-Aware Localized Conformal Prediction on Graphs
by: Baghershahi, Peyman, et al.
Published: (2026)
by: Baghershahi, Peyman, et al.
Published: (2026)
Colorful Talks with Graphs: Human-Interpretable Graph Encodings for Large Language Models
by: Zangari, Angelo, et al.
Published: (2026)
by: Zangari, Angelo, et al.
Published: (2026)
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation
by: Zhang, Weizhi, et al.
Published: (2025)
by: Zhang, Weizhi, et al.
Published: (2025)
Selection of LLM Fine-Tuning Data based on Orthogonal Rules
by: Li, Xiaomin, et al.
Published: (2024)
by: Li, Xiaomin, et al.
Published: (2024)
From Nodes to Narratives: Explaining Graph Neural Networks with LLMs and Graph Context
by: Baghershahi, Peyman, et al.
Published: (2025)
by: Baghershahi, Peyman, et al.
Published: (2025)
Utility-Diversity Aware Online Batch Selection for LLM Supervised Fine-tuning
by: Zou, Heming, et al.
Published: (2025)
by: Zou, Heming, et al.
Published: (2025)
Design Requirements for Human-Centered Graph Neural Network Explanations
by: Habibi, Pantea, et al.
Published: (2024)
by: Habibi, Pantea, et al.
Published: (2024)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
MSSR: Memory-Aware Adaptive Replay for Continual LLM Fine-Tuning
by: Lu, Yiyang, et al.
Published: (2026)
by: Lu, Yiyang, et al.
Published: (2026)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
by: Sun, Yifan, et al.
Published: (2025)
by: Sun, Yifan, et al.
Published: (2025)
Mitigating Training Imbalance in LLM Fine-Tuning via Selective Parameter Merging
by: Ju, Yiming, et al.
Published: (2024)
by: Ju, Yiming, et al.
Published: (2024)
SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
by: Shen, Han, et al.
Published: (2024)
by: Shen, Han, et al.
Published: (2024)
Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-Tuning
by: Pang, Jinlong, et al.
Published: (2025)
by: Pang, Jinlong, et al.
Published: (2025)
ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning
by: Wu, Yang, et al.
Published: (2024)
by: Wu, Yang, et al.
Published: (2024)
BANGS: Game-Theoretic Node Selection for Graph Self-Training
by: Wang, Fangxin, et al.
Published: (2024)
by: Wang, Fangxin, et al.
Published: (2024)
PoliTune: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in Large Language Models
by: Agiza, Ahmed, et al.
Published: (2024)
by: Agiza, Ahmed, et al.
Published: (2024)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
by: Wang, Zige, et al.
Published: (2025)
by: Wang, Zige, et al.
Published: (2025)
Get more for less: Principled Data Selection for Warming Up Fine-Tuning in LLMs
by: Kang, Feiyang, et al.
Published: (2024)
by: Kang, Feiyang, et al.
Published: (2024)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
by: Ananta, Moses, et al.
Published: (2025)
by: Ananta, Moses, et al.
Published: (2025)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
by: Xia, Yuchen, et al.
Published: (2024)
by: Xia, Yuchen, et al.
Published: (2024)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
by: Liu, Wei, et al.
Published: (2023)
by: Liu, Wei, et al.
Published: (2023)
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
by: Li, Sijia, et al.
Published: (2026)
by: Li, Sijia, et al.
Published: (2026)
DataShield: Safety-degrading Data Filtering for LLM Benign Instruction Fine-Tuning
by: Zhang, Junbo, et al.
Published: (2026)
by: Zhang, Junbo, et al.
Published: (2026)
AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air
by: Yang, Shiyi, et al.
Published: (2025)
by: Yang, Shiyi, et al.
Published: (2025)
TACOS: Open Tagging and Comparative Scoring for Instruction Fine-Tuning Data Selection
by: He, Xixiang, et al.
Published: (2025)
by: He, Xixiang, et al.
Published: (2025)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
by: Wang, Xinyu, et al.
Published: (2026)
by: Wang, Xinyu, et al.
Published: (2026)
Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs
by: Ban, Hao, et al.
Published: (2025)
by: Ban, Hao, et al.
Published: (2025)
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning
by: Zou, Jiaru, et al.
Published: (2025)
by: Zou, Jiaru, et al.
Published: (2025)
Taming Sensitive Weights : Noise Perturbation Fine-tuning for Robust LLM Quantization
by: Wang, Dongwei, et al.
Published: (2024)
by: Wang, Dongwei, et al.
Published: (2024)
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
by: Wang, Xu, et al.
Published: (2025)
by: Wang, Xu, et al.
Published: (2025)
CURLoRA: Stable LLM Continual Fine-Tuning and Catastrophic Forgetting Mitigation
by: Fawi, Muhammad
Published: (2024)
by: Fawi, Muhammad
Published: (2024)
Fine-Tune an SLM or Prompt an LLM? The Case of Generating Low-Code Workflows
by: Ayala, Orlando Marquez, et al.
Published: (2025)
by: Ayala, Orlando Marquez, et al.
Published: (2025)
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning
by: Qian, Livia, et al.
Published: (2026)
by: Qian, Livia, et al.
Published: (2026)
Federated In-Context LLM Agent Learning
by: Wu, Panlong, et al.
Published: (2024)
by: Wu, Panlong, et al.
Published: (2024)
DoGE: Domain Reweighting with Generalization Estimation
by: Fan, Simin, et al.
Published: (2023)
by: Fan, Simin, et al.
Published: (2023)
LIFT the Veil for the Truth: Principal Weights Emerge after Rank Reduction for Reasoning-Focused Supervised Fine-Tuning
by: Liu, Zihang, et al.
Published: (2025)
by: Liu, Zihang, et al.
Published: (2025)
LESS: Selecting Influential Data for Targeted Instruction Tuning
by: Xia, Mengzhou, et al.
Published: (2024)
by: Xia, Mengzhou, et al.
Published: (2024)
Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data
by: Rallapalli, Swati, et al.
Published: (2025)
by: Rallapalli, Swati, et al.
Published: (2025)
Similar Items
-
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
by: Chaturvedi, Rochana, et al.
Published: (2025) -
Unsupervised Prompting for Graph Neural Networks
by: Baghershahi, Peyman, et al.
Published: (2025) -
GRAPHLCP: Structure-Aware Localized Conformal Prediction on Graphs
by: Baghershahi, Peyman, et al.
Published: (2026) -
Colorful Talks with Graphs: Human-Interpretable Graph Encodings for Large Language Models
by: Zangari, Angelo, et al.
Published: (2026) -
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation
by: Zhang, Weizhi, et al.
Published: (2025)