Training Data Influence Analysis and Estimation: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hammoudeh, Zayd, Lowd, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Provable Robustness Against a Union of $\ell_0$ Adversarial Attacks
von: Hammoudeh, Zayd, et al.
Veröffentlicht: (2023)
von: Hammoudeh, Zayd, et al.
Veröffentlicht: (2023)
The Ultimate Cookbook for Invisible Poison: Crafting Subtle Clean-Label Text Backdoors with Style Attributes
von: You, Wencong, et al.
Veröffentlicht: (2025)
von: You, Wencong, et al.
Veröffentlicht: (2025)
Deep learning in medical image registration: introduction and survey
von: Hammoudeh, Ahmad, et al.
Veröffentlicht: (2023)
von: Hammoudeh, Ahmad, et al.
Veröffentlicht: (2023)
Predicting the Order of Upcoming Tokens Improves Language Modeling
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025)
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025)
Softpick: No Attention Sink, No Massive Activations with Rectified Softmax
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025)
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025)
Capturing the Temporal Dependence of Training Data Influence
von: Wang, Jiachen T., et al.
Veröffentlicht: (2024)
von: Wang, Jiachen T., et al.
Veröffentlicht: (2024)
DMin: Scalable Training Data Influence Estimation for Diffusion Models
von: Lin, Huawei, et al.
Veröffentlicht: (2024)
von: Lin, Huawei, et al.
Veröffentlicht: (2024)
On Training Data Influence of GPT Models
von: Chai, Yekun, et al.
Veröffentlicht: (2024)
von: Chai, Yekun, et al.
Veröffentlicht: (2024)
The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes
von: Ko, Myeongseob, et al.
Veröffentlicht: (2024)
von: Ko, Myeongseob, et al.
Veröffentlicht: (2024)
MLKV: Multi-Layer Key-Value Heads for Memory Efficient Transformer Decoding
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods
von: Kharma, Mohammed, et al.
Veröffentlicht: (2026)
von: Kharma, Mohammed, et al.
Veröffentlicht: (2026)
Accumulative SGD Influence Estimation for Data Attribution
von: Shi, Yunxiao, et al.
Veröffentlicht: (2025)
von: Shi, Yunxiao, et al.
Veröffentlicht: (2025)
Self-Training: A Survey
von: Amini, Massih-Reza, et al.
Veröffentlicht: (2022)
von: Amini, Massih-Reza, et al.
Veröffentlicht: (2022)
Data Attribution for Diffusion Models: Timestep-induced Bias in Influence Estimation
von: Xie, Tong, et al.
Veröffentlicht: (2024)
von: Xie, Tong, et al.
Veröffentlicht: (2024)
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies
von: Ma, Yi, et al.
Veröffentlicht: (2025)
von: Ma, Yi, et al.
Veröffentlicht: (2025)
The Approximate Fisher Influence Function: Faster Estimation of Data Influence in Statistical Models
von: Lev, Omri, et al.
Veröffentlicht: (2024)
von: Lev, Omri, et al.
Veröffentlicht: (2024)
f-INE: A Hypothesis Testing Framework for Estimating Influence under Training Randomness
von: Panda, Subhodip, et al.
Veröffentlicht: (2025)
von: Panda, Subhodip, et al.
Veröffentlicht: (2025)
LoRIF: Low-Rank Influence Functions for Scalable Training Data Attribution
von: Li, Shuangqi, et al.
Veröffentlicht: (2026)
von: Li, Shuangqi, et al.
Veröffentlicht: (2026)
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
von: Kwon, Yongchan, et al.
Veröffentlicht: (2023)
von: Kwon, Yongchan, et al.
Veröffentlicht: (2023)
Layer-Aware Influence for Online Data Valuation Estimation
von: Yang, Ziao, et al.
Veröffentlicht: (2025)
von: Yang, Ziao, et al.
Veröffentlicht: (2025)
HyperINF: Unleashing the HyperPower of the Schulz's Method for Data Influence Estimation
von: Zhou, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2024)
Data-Dependent Stability Analysis of Adversarial Training
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
A Comparative Analysis of Influence Signals for Data Debugging
von: Myrtakis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Myrtakis, Nikolaos, et al.
Veröffentlicht: (2025)
Error Estimate and Convergence Analysis for Data Valuation
von: Liang, Zhangyong, et al.
Veröffentlicht: (2025)
von: Liang, Zhangyong, et al.
Veröffentlicht: (2025)
Data Descriptions from Large Language Models with Influence Estimation
von: Kim, Chaeri, et al.
Veröffentlicht: (2025)
von: Kim, Chaeri, et al.
Veröffentlicht: (2025)
Diffusion Attribution Score: Evaluating Training Data Influence in Diffusion Models
von: Lin, Jinxu, et al.
Veröffentlicht: (2024)
von: Lin, Jinxu, et al.
Veröffentlicht: (2024)
Distributional Training Data Attribution: What do Influence Functions Sample?
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2025)
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2025)
Adversarial Training: A Survey
von: Zhao, Mengnan, et al.
Veröffentlicht: (2024)
von: Zhao, Mengnan, et al.
Veröffentlicht: (2024)
A Survey on Pre-Trained Diffusion Model Distillations
von: Fan, Xuhui, et al.
Veröffentlicht: (2025)
von: Fan, Xuhui, et al.
Veröffentlicht: (2025)
Survey of Data-driven Newsvendor: Unified Analysis and Spectrum of Achievable Regrets
von: Chen, Zhuoxin, et al.
Veröffentlicht: (2024)
von: Chen, Zhuoxin, et al.
Veröffentlicht: (2024)
Concept Influence: Leveraging Interpretability to Improve Performance and Efficiency in Training Data Attribution
von: Kowal, Matthew, et al.
Veröffentlicht: (2026)
von: Kowal, Matthew, et al.
Veröffentlicht: (2026)
PLUMAGE: Probabilistic Low rank Unbiased Min Variance Gradient Estimator for Efficient Large Model Training
von: Haroush, Matan, et al.
Veröffentlicht: (2025)
von: Haroush, Matan, et al.
Veröffentlicht: (2025)
Empowering Time Series Analysis with Synthetic Data: A Survey and Outlook in the Era of Foundation Models
von: Liu, Xu, et al.
Veröffentlicht: (2025)
von: Liu, Xu, et al.
Veröffentlicht: (2025)
The Influence of Faulty Labels in Data Sets on Human Pose Estimation
von: Schwarz, Arnold, et al.
Veröffentlicht: (2024)
von: Schwarz, Arnold, et al.
Veröffentlicht: (2024)
Machine Learning and Data Analysis Using Posets: A Survey
von: Mwafise, Arnauld Mesinga
Veröffentlicht: (2024)
von: Mwafise, Arnauld Mesinga
Veröffentlicht: (2024)
An Analysis of Causal Effect Estimation using Outcome Invariant Data Augmentation
von: Akbar, Uzair, et al.
Veröffentlicht: (2025)
von: Akbar, Uzair, et al.
Veröffentlicht: (2025)
Querying Kernel Methods Suffices for Reconstructing their Training Data
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
Scalable Multiagent Reinforcement Learning with Collective Influence Estimation
von: Luo, Zhenglong, et al.
Veröffentlicht: (2026)
von: Luo, Zhenglong, et al.
Veröffentlicht: (2026)
A Survey on Archetypal Analysis
von: Alcacer, Aleix, et al.
Veröffentlicht: (2025)
von: Alcacer, Aleix, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Provable Robustness Against a Union of $\ell_0$ Adversarial Attacks
von: Hammoudeh, Zayd, et al.
Veröffentlicht: (2023) -
The Ultimate Cookbook for Invisible Poison: Crafting Subtle Clean-Label Text Backdoors with Style Attributes
von: You, Wencong, et al.
Veröffentlicht: (2025) -
Deep learning in medical image registration: introduction and survey
von: Hammoudeh, Ahmad, et al.
Veröffentlicht: (2023) -
Predicting the Order of Upcoming Tokens Improves Language Modeling
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025) -
Softpick: No Attention Sink, No Massive Activations with Rectified Softmax
von: Zuhri, Zayd M. K., et al.
Veröffentlicht: (2025)