Rethinking Data Shapley for Data Selection Tasks: Misleads and Merits
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Jiachen T., Yang, Tianji, Zou, James, Kwon, Yongchan, Jia, Ruoxi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Efficient Data Shapley for Weighted Nearest Neighbor Algorithms
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
Data Shapley in One Training Run
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
di: Kwon, Yongchan, et al.
Pubblicazione: (2023)
di: Kwon, Yongchan, et al.
Pubblicazione: (2023)
Capturing the Temporal Dependence of Training Data Influence
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
di: Wang, Jiachen T., et al.
Pubblicazione: (2024)
What LLMs Think When You Don't Tell Them What to Think About?
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice
di: Wang, Jiachen T., et al.
Pubblicazione: (2025)
di: Wang, Jiachen T., et al.
Pubblicazione: (2025)
Group Shapley Value and Counterfactual Simulations in a Structural Model
di: Kwon, Yongchan, et al.
Pubblicazione: (2024)
di: Kwon, Yongchan, et al.
Pubblicazione: (2024)
Towards Data Valuation via Asymmetric Data Shapley
di: Zheng, Xi, et al.
Pubblicazione: (2024)
di: Zheng, Xi, et al.
Pubblicazione: (2024)
Uncertainty Quantification of Data Shapley via Statistical Inference
di: Wu, Mengmeng, et al.
Pubblicazione: (2024)
di: Wu, Mengmeng, et al.
Pubblicazione: (2024)
Certified Data Removal Under High-dimensional Settings
di: Zou, Haolin, et al.
Pubblicazione: (2025)
di: Zou, Haolin, et al.
Pubblicazione: (2025)
2D-OOB: Attributing Data Contribution Through Joint Valuation Framework
di: Sun, Yifan, et al.
Pubblicazione: (2024)
di: Sun, Yifan, et al.
Pubblicazione: (2024)
TimeInf: Time Series Data Contribution via Influence Functions
di: Zhang, Yizi, et al.
Pubblicazione: (2024)
di: Zhang, Yizi, et al.
Pubblicazione: (2024)
Data Valuation and Selection in a Federated Model Marketplace
di: Li, Wenqian, et al.
Pubblicazione: (2025)
di: Li, Wenqian, et al.
Pubblicazione: (2025)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
di: Just, Hoang Anh, et al.
Pubblicazione: (2025)
di: Just, Hoang Anh, et al.
Pubblicazione: (2025)
ReasonIF: Large Reasoning Models Fail to Follow Instructions During Reasoning
di: Kwon, Yongchan, et al.
Pubblicazione: (2025)
di: Kwon, Yongchan, et al.
Pubblicazione: (2025)
Is Data Shapley Not Better than Random in Data Selection? Ask NASH
di: Tian, Xiao, et al.
Pubblicazione: (2026)
di: Tian, Xiao, et al.
Pubblicazione: (2026)
Distributionally Robust Instrumental Variables Estimation
di: Qu, Zhaonan, et al.
Pubblicazione: (2024)
di: Qu, Zhaonan, et al.
Pubblicazione: (2024)
TokenShapley: Token Level Context Attribution with Shapley Value
di: Xiao, Yingtai, et al.
Pubblicazione: (2025)
di: Xiao, Yingtai, et al.
Pubblicazione: (2025)
DMRL: Data- and Model-aware Reward Learning for Data Extraction
di: Wang, Zhiqiang, et al.
Pubblicazione: (2025)
di: Wang, Zhiqiang, et al.
Pubblicazione: (2025)
Proper Dataset Valuation by Pointwise Mutual Information
di: Zheng, Shuran, et al.
Pubblicazione: (2024)
di: Zheng, Shuran, et al.
Pubblicazione: (2024)
Newfluence: Boosting Model interpretability and Understanding in High Dimensions
di: Zou, Haolin, et al.
Pubblicazione: (2025)
di: Zou, Haolin, et al.
Pubblicazione: (2025)
In-Run Data Shapley for Adam Optimizer
di: Ding, Meng, et al.
Pubblicazione: (2026)
di: Ding, Meng, et al.
Pubblicazione: (2026)
A Sustainable AI Economy Needs Data Deals That Work for Generators
di: Jia, Ruoxi, et al.
Pubblicazione: (2026)
di: Jia, Ruoxi, et al.
Pubblicazione: (2026)
Fast-DataShapley: Neural Modeling for Training Data Valuation
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
Action Shapley: A Training Data Selection Metric for World Model in Reinforcement Learning
di: Ghosh, Rajat, et al.
Pubblicazione: (2026)
di: Ghosh, Rajat, et al.
Pubblicazione: (2026)
Losing is for Cherishing: Data Valuation Based on Machine Unlearning and Shapley Value
di: Ma, Le, et al.
Pubblicazione: (2025)
di: Ma, Le, et al.
Pubblicazione: (2025)
Data-Efficient and Robust Task Selection for Meta-Learning
di: Zhan, Donglin, et al.
Pubblicazione: (2024)
di: Zhan, Donglin, et al.
Pubblicazione: (2024)
Voice "Cloning" is Style Transfer
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2026)
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2026)
Quagmires in SFT-RL Post-Training: When High SFT Scores Mislead and What to Use Instead
di: Kang, Feiyang, et al.
Pubblicazione: (2025)
di: Kang, Feiyang, et al.
Pubblicazione: (2025)
CHG Shapley: Efficient Data Valuation and Selection towards Trustworthy Machine Learning
di: Cai, Huaiguang
Pubblicazione: (2024)
di: Cai, Huaiguang
Pubblicazione: (2024)
Federated Learning for Data Market: Shapley-UCB for Seller Selection and Incentives
di: Chen, Kongyang, et al.
Pubblicazione: (2024)
di: Chen, Kongyang, et al.
Pubblicazione: (2024)
A Comprehensive Study of Shapley Value in Data Analytics
di: Lin, Hong, et al.
Pubblicazione: (2024)
di: Lin, Hong, et al.
Pubblicazione: (2024)
Thresholding Data Shapley for Data Cleansing Using Multi-Armed Bandits
di: Namba, Hiroyuki, et al.
Pubblicazione: (2024)
di: Namba, Hiroyuki, et al.
Pubblicazione: (2024)
Data-Centric Human Preference with Rationales for Direct Preference Alignment
di: Just, Hoang Anh, et al.
Pubblicazione: (2024)
di: Just, Hoang Anh, et al.
Pubblicazione: (2024)
Tab-Shapley: Identifying Top-k Tabular Data Quality Insights
di: Padala, Manisha, et al.
Pubblicazione: (2025)
di: Padala, Manisha, et al.
Pubblicazione: (2025)
Understanding Impact of Human Feedback via Influence Functions
di: Min, Taywon, et al.
Pubblicazione: (2025)
di: Min, Taywon, et al.
Pubblicazione: (2025)
Rethinking the Intermediate Features in Adversarial Attacks: Misleading Robotic Models via Adversarial Distillation
di: Zhao, Ke, et al.
Pubblicazione: (2024)
di: Zhao, Ke, et al.
Pubblicazione: (2024)
Rethinking Data Value: Asymmetric Data Shapley for Structure-Aware Valuation in Data Markets and Machine Learning Pipelines
di: Zheng, Xi, et al.
Pubblicazione: (2025)
di: Zheng, Xi, et al.
Pubblicazione: (2025)
NeuroLoRA: Context-Aware Neuromodulation for Parameter-Efficient Multi-Task Adaptation
di: Yang, Yuxin, et al.
Pubblicazione: (2026)
di: Yang, Yuxin, et al.
Pubblicazione: (2026)
The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes
di: Ko, Myeongseob, et al.
Pubblicazione: (2024)
di: Ko, Myeongseob, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Efficient Data Shapley for Weighted Nearest Neighbor Algorithms
di: Wang, Jiachen T., et al.
Pubblicazione: (2024) -
Data Shapley in One Training Run
di: Wang, Jiachen T., et al.
Pubblicazione: (2024) -
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
di: Kwon, Yongchan, et al.
Pubblicazione: (2023) -
Capturing the Temporal Dependence of Training Data Influence
di: Wang, Jiachen T., et al.
Pubblicazione: (2024) -
What LLMs Think When You Don't Tell Them What to Think About?
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)