For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Deng, Wenlong, Zeng, Qi, Zhang, Jiaming, Chen, Minghui, Ding, Zixin, Thrampoulidis, Christos, Gong, Boying, Li, Xiaoxiao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On Group Relative Policy Optimization Collapse in Agent Search: The Lazy Likelihood-Displacement
por: Deng, Wenlong, et al.
Publicado: (2025)
por: Deng, Wenlong, et al.
Publicado: (2025)
Token Hidden Reward: Steering Exploration-Exploitation in Group Relative Deep Reinforcement Learning
por: Deng, Wenlong, et al.
Publicado: (2025)
por: Deng, Wenlong, et al.
Publicado: (2025)
DARE the Extreme: Revisiting Delta-Parameter Pruning For Fine-Tuned Models
por: Deng, Wenlong, et al.
Publicado: (2024)
por: Deng, Wenlong, et al.
Publicado: (2024)
LLM-Assisted Content Conditional Debiasing for Fair Text Embedding
por: Deng, Wenlong, et al.
Publicado: (2024)
por: Deng, Wenlong, et al.
Publicado: (2024)
Unlocking the Potential of Prompt-Tuning in Bridging Generalized and Personalized Federated Learning
por: Deng, Wenlong, et al.
Publicado: (2023)
por: Deng, Wenlong, et al.
Publicado: (2023)
On the Effect of Negative Gradient in Group Relative Deep Reinforcement Optimization
por: Deng, Wenlong, et al.
Publicado: (2025)
por: Deng, Wenlong, et al.
Publicado: (2025)
Implicit Optimization Bias of Next-Token Prediction in Linear Models
por: Thrampoulidis, Christos
Publicado: (2024)
por: Thrampoulidis, Christos
Publicado: (2024)
Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models
por: Deng, Wenlong, et al.
Publicado: (2026)
por: Deng, Wenlong, et al.
Publicado: (2026)
Geometry of Semantics in Next-Token Prediction: How Optimization Implicitly Organizes Linguistic Representations
por: Zhao, Yize, et al.
Publicado: (2025)
por: Zhao, Yize, et al.
Publicado: (2025)
Enhancing Clinical Multiple-Choice Questions Benchmarks with Knowledge Graph Guided Distractor Generation
por: Yang, Running, et al.
Publicado: (2025)
por: Yang, Running, et al.
Publicado: (2025)
Advantage Shaping as Surrogate Reward Maximization: Unifying Pass@K Policy Gradients
por: Thrampoulidis, Christos, et al.
Publicado: (2025)
por: Thrampoulidis, Christos, et al.
Publicado: (2025)
Facts in Stats: Impacts of Pretraining Diversity on Language Model Generalization
por: Behnia, Tina, et al.
Publicado: (2025)
por: Behnia, Tina, et al.
Publicado: (2025)
Low-rank finetuning for LLMs: A fairness perspective
por: Das, Saswat, et al.
Publicado: (2024)
por: Das, Saswat, et al.
Publicado: (2024)
ReMix: Reinforcement routing for mixtures of LoRAs in LLM finetuning
por: Qiu, Ruizhong, et al.
Publicado: (2026)
por: Qiu, Ruizhong, et al.
Publicado: (2026)
Implicit Geometry of Next-token Prediction: From Language Sparsity Patterns to Model Representations
por: Zhao, Yize, et al.
Publicado: (2024)
por: Zhao, Yize, et al.
Publicado: (2024)
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs
por: Betley, Jan, et al.
Publicado: (2025)
por: Betley, Jan, et al.
Publicado: (2025)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
por: Deora, Puneesh, et al.
Publicado: (2025)
por: Deora, Puneesh, et al.
Publicado: (2025)
Short-Context Dominance: How Much Local Context Natural Language Actually Needs?
por: Vakilian, Vala, et al.
Publicado: (2025)
por: Vakilian, Vala, et al.
Publicado: (2025)
CSE-SFP: Enabling Unsupervised Sentence Representation Learning via a Single Forward Pass
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
Do LLMs and VLMs Share Neurons for Inference? Evidence and Mechanisms of Cross-Modal Transfer
por: Cui, Chenhang, et al.
Publicado: (2026)
por: Cui, Chenhang, et al.
Publicado: (2026)
BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning
por: Wang, Shengao, et al.
Publicado: (2025)
por: Wang, Shengao, et al.
Publicado: (2025)
Understanding Contextual Recall in Transformers: How Finetuning Enables In-Context Reasoning over Pretraining Knowledge
por: Vasudeva, Bhavya, et al.
Publicado: (2026)
por: Vasudeva, Bhavya, et al.
Publicado: (2026)
Improving Fine-grained Visual Understanding in VLMs through Text-Only Training
por: Choi, Dasol, et al.
Publicado: (2024)
por: Choi, Dasol, et al.
Publicado: (2024)
FANNO: Augmenting High-Quality Instruction Data with Open-Sourced LLMs Only
por: Zhu, He, et al.
Publicado: (2024)
por: Zhu, He, et al.
Publicado: (2024)
LLaMA based Punctuation Restoration With Forward Pass Only Decoding
por: Pang, Yutong, et al.
Publicado: (2024)
por: Pang, Yutong, et al.
Publicado: (2024)
Transformers as Support Vector Machines
por: Tarzanagh, Davoud Ataee, et al.
Publicado: (2023)
por: Tarzanagh, Davoud Ataee, et al.
Publicado: (2023)
Prompt Valuation Based on Shapley Values
por: Liu, Hanxi, et al.
Publicado: (2023)
por: Liu, Hanxi, et al.
Publicado: (2023)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
por: Wang, Zige, et al.
Publicado: (2025)
por: Wang, Zige, et al.
Publicado: (2025)
SpeCache: Speculative Key-Value Caching for Efficient Generation of LLMs
por: Jie, Shibo, et al.
Publicado: (2025)
por: Jie, Shibo, et al.
Publicado: (2025)
Recursive Think-Answer Process for LLMs and VLMs
por: Lee, Byung-Kwan, et al.
Publicado: (2026)
por: Lee, Byung-Kwan, et al.
Publicado: (2026)
Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs
por: Carbune, Victor, et al.
Publicado: (2024)
por: Carbune, Victor, et al.
Publicado: (2024)
PreFT: Prefill-only finetuning for efficient inference
por: Lanpouthakoun, Andrew, et al.
Publicado: (2026)
por: Lanpouthakoun, Andrew, et al.
Publicado: (2026)
REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning
por: Lee, Seungmin, et al.
Publicado: (2026)
por: Lee, Seungmin, et al.
Publicado: (2026)
Gated Tree Cross-Attention for Checkpoint-Compatible Syntax Injection in Decoder-Only LLMs
por: Gao, Xinyu, et al.
Publicado: (2026)
por: Gao, Xinyu, et al.
Publicado: (2026)
Implicit Bias of Spectral Descent and Muon on Multiclass Separable Data
por: Fan, Chen, et al.
Publicado: (2025)
por: Fan, Chen, et al.
Publicado: (2025)
Be Cautious When Merging Unfamiliar LLMs: A Phishing Model Capable of Stealing Privacy
por: Guo, Zhenyuan, et al.
Publicado: (2025)
por: Guo, Zhenyuan, et al.
Publicado: (2025)
Debiased Noise Editing on Foundation Models for Fair Medical Image Classification
por: Jin, Ruinan, et al.
Publicado: (2024)
por: Jin, Ruinan, et al.
Publicado: (2024)
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation
por: Deng, Wei, et al.
Publicado: (2025)
por: Deng, Wei, et al.
Publicado: (2025)
Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only
por: Yao, Jihan, et al.
Publicado: (2024)
por: Yao, Jihan, et al.
Publicado: (2024)
Ejemplares similares
-
On Group Relative Policy Optimization Collapse in Agent Search: The Lazy Likelihood-Displacement
por: Deng, Wenlong, et al.
Publicado: (2025) -
Token Hidden Reward: Steering Exploration-Exploitation in Group Relative Deep Reinforcement Learning
por: Deng, Wenlong, et al.
Publicado: (2025) -
DARE the Extreme: Revisiting Delta-Parameter Pruning For Fine-Tuned Models
por: Deng, Wenlong, et al.
Publicado: (2024) -
LLM-Assisted Content Conditional Debiasing for Fair Text Embedding
por: Deng, Wenlong, et al.
Publicado: (2024) -
Unlocking the Potential of Prompt-Tuning in Bridging Generalized and Personalized Federated Learning
por: Deng, Wenlong, et al.
Publicado: (2023)