Exploring Training Data Attribution under Limited Access Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shiyuan, Deng, Junwei, Bae, Juhan, Ma, Jiaqi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Ensembles Improve Training Data Attribution
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
Training Data Attribution via Approximate Unrolled Differentiation
by: Bae, Juhan, et al.
Published: (2024)
by: Bae, Juhan, et al.
Published: (2024)
Adversarial Attacks on Data Attribution
by: Wang, Xinhe, et al.
Published: (2024)
by: Wang, Xinhe, et al.
Published: (2024)
Taming Hyperparameter Sensitivity in Data Attribution: Practical Selection Without Costly Retraining
by: Wang, Weiyi, et al.
Published: (2025)
by: Wang, Weiyi, et al.
Published: (2025)
A Versatile Influence Function for Data Attribution with Non-Decomposable Loss
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
$\texttt{dattri}$: A Library for Efficient Data Attribution
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
Better Training Data Attribution via Better Inverse Hessian-Vector Products
by: Wang, Andrew, et al.
Published: (2025)
by: Wang, Andrew, et al.
Published: (2025)
How Faithful Is Trajectory-Based Data Attribution? Error Sources, Remedies, and Practical Guidelines
by: Deng, Junwei, et al.
Published: (2026)
by: Deng, Junwei, et al.
Published: (2026)
Influence Functions for Scalable Data Attribution in Diffusion Models
by: Mlodozeniec, Bruno, et al.
Published: (2024)
by: Mlodozeniec, Bruno, et al.
Published: (2024)
Detecting and Filtering Unsafe Training Data via Data Attribution with Denoised Representation
by: Pan, Yijun, et al.
Published: (2025)
by: Pan, Yijun, et al.
Published: (2025)
Daunce: Data Attribution through Uncertainty Estimation
by: Pan, Xingyuan, et al.
Published: (2025)
by: Pan, Xingyuan, et al.
Published: (2025)
OATS: Online Data Augmentation for Time Series Foundation Models
by: Deng, Junwei, et al.
Published: (2026)
by: Deng, Junwei, et al.
Published: (2026)
Spectral-factorized Positive-definite Curvature Learning for NN Training
by: Lin, Wu, et al.
Published: (2025)
by: Lin, Wu, et al.
Published: (2025)
Measuring Fine-Grained Relatedness in Multitask Learning via Data Attribution
by: Tu, Yiwen, et al.
Published: (2025)
by: Tu, Yiwen, et al.
Published: (2025)
IF-GUIDE: Influence Function-Guided Detoxification of LLMs
by: Coalson, Zachary, et al.
Published: (2025)
by: Coalson, Zachary, et al.
Published: (2025)
Enhancing Training Data Attribution with Representational Optimization
by: Sun, Weiwei, et al.
Published: (2025)
by: Sun, Weiwei, et al.
Published: (2025)
GraSS: Scalable Data Attribution with Gradient Sparsification and Sparse Projection
by: Hu, Pingbang, et al.
Published: (2025)
by: Hu, Pingbang, et al.
Published: (2025)
Predicting User Intents and Musical Attributes from Music Discovery Conversations
by: Kwon, Daeyong, et al.
Published: (2024)
by: Kwon, Daeyong, et al.
Published: (2024)
Using Large Language Models for Hyperparameter Optimization
by: Zhang, Michael R., et al.
Published: (2023)
by: Zhang, Michael R., et al.
Published: (2023)
S2Aligner: Pair-Efficient and Transferable Pre-Training for Sparse Text-Attributed Graphs
by: Wang, Yuhan, et al.
Published: (2026)
by: Wang, Yuhan, et al.
Published: (2026)
Efficient Sketches for Training Data Attribution and Studying the Loss Landscape
by: Schioppa, Andrea
Published: (2024)
by: Schioppa, Andrea
Published: (2024)
A Snapshot of Influence: A Local Data Attribution Framework for Online Reinforcement Learning
by: Hu, Yuzheng, et al.
Published: (2025)
by: Hu, Yuzheng, et al.
Published: (2025)
Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems
by: Zhang, Shichang, et al.
Published: (2025)
by: Zhang, Shichang, et al.
Published: (2025)
Learning from the Right Rollouts: Data Attribution for PPO-based LLM Post-Training
by: Shu, Dong, et al.
Published: (2026)
by: Shu, Dong, et al.
Published: (2026)
Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training
by: Hu, Pingbang, et al.
Published: (2026)
by: Hu, Pingbang, et al.
Published: (2026)
Attribute Graphs Underlying Molecular Generative Models: Path to Learning with Limited Data
by: Hoffman, Samuel C., et al.
Published: (2022)
by: Hoffman, Samuel C., et al.
Published: (2022)
Online Joint Assortment-Inventory Optimization under MNL Choices
by: Liang, Yong, et al.
Published: (2023)
by: Liang, Yong, et al.
Published: (2023)
Training Data Attribution (TDA): Examining Its Adoption & Use Cases
by: Cheng, Deric, et al.
Published: (2025)
by: Cheng, Deric, et al.
Published: (2025)
Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making
by: Chen, Ruoyu, et al.
Published: (2026)
by: Chen, Ruoyu, et al.
Published: (2026)
On the Limitation and Experience Replay for GNNs in Continual Learning
by: Su, Junwei, et al.
Published: (2023)
by: Su, Junwei, et al.
Published: (2023)
DFF: Decision-Focused Fine-tuning for Smarter Predict-then-Optimize with Limited Data
by: Yang, Jiaqi, et al.
Published: (2025)
by: Yang, Jiaqi, et al.
Published: (2025)
GNNMerge: Merging of GNN Models Without Accessing Training Data
by: Garg, Vipul, et al.
Published: (2025)
by: Garg, Vipul, et al.
Published: (2025)
In-Context Data Distillation with TabPFN
by: Ma, Junwei, et al.
Published: (2024)
by: Ma, Junwei, et al.
Published: (2024)
ExPLAIND: Unifying Model, Data, and Training Attribution to Study Model Behavior
by: Eichin, Florian, et al.
Published: (2025)
by: Eichin, Florian, et al.
Published: (2025)
LoRIF: Low-Rank Influence Functions for Scalable Training Data Attribution
by: Li, Shuangqi, et al.
Published: (2026)
by: Li, Shuangqi, et al.
Published: (2026)
Learning to Weight Parameters for Training Data Attribution
by: Li, Shuangqi, et al.
Published: (2025)
by: Li, Shuangqi, et al.
Published: (2025)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Scalable Data Attribution via Forward-Only Test-Time Inference
by: Ma, Sibo, et al.
Published: (2025)
by: Ma, Sibo, et al.
Published: (2025)
Offline Clustering of Linear Bandits: The Power of Clusters under Limited Data
by: Liu, Jingyuan, et al.
Published: (2025)
by: Liu, Jingyuan, et al.
Published: (2025)
Energy-Weighted Flow Matching for Offline Reinforcement Learning
by: Zhang, Shiyuan, et al.
Published: (2025)
by: Zhang, Shiyuan, et al.
Published: (2025)
Similar Items
-
Efficient Ensembles Improve Training Data Attribution
by: Deng, Junwei, et al.
Published: (2024) -
Training Data Attribution via Approximate Unrolled Differentiation
by: Bae, Juhan, et al.
Published: (2024) -
Adversarial Attacks on Data Attribution
by: Wang, Xinhe, et al.
Published: (2024) -
Taming Hyperparameter Sensitivity in Data Attribution: Practical Selection Without Costly Retraining
by: Wang, Weiyi, et al.
Published: (2025) -
A Versatile Influence Function for Data Attribution with Non-Decomposable Loss
by: Deng, Junwei, et al.
Published: (2024)