Better Training Data Attribution via Better Inverse Hessian-Vector Products
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Andrew, Nguyen, Elisa, Yang, Runshi, Bae, Juhan, McIlraith, Sheila A., Grosse, Roger |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gauss-Newton Unlearning for the LLM Era
von: McKinney, Lev, et al.
Veröffentlicht: (2026)
von: McKinney, Lev, et al.
Veröffentlicht: (2026)
Training Data Attribution via Approximate Unrolled Differentiation
von: Bae, Juhan, et al.
Veröffentlicht: (2024)
von: Bae, Juhan, et al.
Veröffentlicht: (2024)
Pluralistic Alignment Over Time
von: Klassen, Toryn Q., et al.
Veröffentlicht: (2024)
von: Klassen, Toryn Q., et al.
Veröffentlicht: (2024)
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2023)
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2023)
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
von: Alamdari, Parand A., et al.
Veröffentlicht: (2026)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2026)
Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data
von: Li, Andrew C., et al.
Veröffentlicht: (2025)
von: Li, Andrew C., et al.
Veröffentlicht: (2025)
Exploring Training Data Attribution under Limited Access Constraints
von: Zhang, Shiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Shiyuan, et al.
Veröffentlicht: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
von: Alamdari, Parand A., et al.
Veröffentlicht: (2023)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2023)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
Pushdown Reward Machines for Reinforcement Learning
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
Spectral-factorized Positive-definite Curvature Learning for NN Training
von: Lin, Wu, et al.
Veröffentlicht: (2025)
von: Lin, Wu, et al.
Veröffentlicht: (2025)
Better World Models Can Lead to Better Post-Training Performance
von: Gupta, Prakhar, et al.
Veröffentlicht: (2025)
von: Gupta, Prakhar, et al.
Veröffentlicht: (2025)
Better Hessians Matter: Studying the Impact of Curvature Approximations in Influence Functions
von: Hong, Steve, et al.
Veröffentlicht: (2025)
von: Hong, Steve, et al.
Veröffentlicht: (2025)
Influence Functions for Scalable Data Attribution in Diffusion Models
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2024)
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2024)
Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2025)
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2025)
Towards Better Understanding Attribution Methods
von: Rao, Sukrut, et al.
Veröffentlicht: (2022)
von: Rao, Sukrut, et al.
Veröffentlicht: (2022)
Reward Machines for Deep RL in Noisy and Uncertain Environments
von: Li, Andrew C., et al.
Veröffentlicht: (2024)
von: Li, Andrew C., et al.
Veröffentlicht: (2024)
Bayesian Influence Functions for Hessian-Free Data Attribution
von: Kreer, Philipp Alexander, et al.
Veröffentlicht: (2025)
von: Kreer, Philipp Alexander, et al.
Veröffentlicht: (2025)
Distributional Training Data Attribution: What do Influence Functions Sample?
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2025)
von: Mlodozeniec, Bruno, et al.
Veröffentlicht: (2025)
OptiMer: Optimal Distribution Vector Merging Is Better than Data Mixing for Continual Pre-Training
von: Song, Haiyue, et al.
Veröffentlicht: (2026)
von: Song, Haiyue, et al.
Veröffentlicht: (2026)
Better Representations via Adversarial Training in Pre-Training: A Theoretical Perspective
von: Xing, Yue, et al.
Veröffentlicht: (2024)
von: Xing, Yue, et al.
Veröffentlicht: (2024)
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
von: Horvitz, Eric, et al.
Veröffentlicht: (2024)
von: Horvitz, Eric, et al.
Veröffentlicht: (2024)
Better Understanding Differences in Attribution Methods via Systematic Evaluations
von: Rao, Sukrut, et al.
Veröffentlicht: (2023)
von: Rao, Sukrut, et al.
Veröffentlicht: (2023)
Two Heads are Actually Better than One: Towards Better Adversarial Robustness via Transduction and Rejection
von: Palumbo, Nils, et al.
Veröffentlicht: (2023)
von: Palumbo, Nils, et al.
Veröffentlicht: (2023)
Train Faster, Perform Better: Modular Adaptive Training in Over-Parameterized Models
von: Shi, Yubin, et al.
Veröffentlicht: (2024)
von: Shi, Yubin, et al.
Veröffentlicht: (2024)
Grokked Models are Better Unlearners
von: Liang, Yuanbang, et al.
Veröffentlicht: (2025)
von: Liang, Yuanbang, et al.
Veröffentlicht: (2025)
Provably Better Explanations with Optimized Aggregation of Feature Attributions
von: Decker, Thomas, et al.
Veröffentlicht: (2024)
von: Decker, Thomas, et al.
Veröffentlicht: (2024)
Two Heads Are Better Than One: Boosting Graph Sparse Training via Semantic and Topological Awareness
von: Zhang, Guibin, et al.
Veröffentlicht: (2024)
von: Zhang, Guibin, et al.
Veröffentlicht: (2024)
Data Imputation by Pursuing Better Classification: A Supervised Kernel-Based Method
von: Yang, Ruikai, et al.
Veröffentlicht: (2024)
von: Yang, Ruikai, et al.
Veröffentlicht: (2024)
Learning with Confidence: Training Better Classifiers from Soft Labels
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2024)
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2024)
Benchmark Data Repositories for Better Benchmarking
von: Longjohn, Rachel, et al.
Veröffentlicht: (2024)
von: Longjohn, Rachel, et al.
Veröffentlicht: (2024)
Round-trip Reinforcement Learning: Self-Consistent Training for Better Chemical LLMs
von: Kong, Lecheng, et al.
Veröffentlicht: (2025)
von: Kong, Lecheng, et al.
Veröffentlicht: (2025)
Towards Better Performance in Incomplete LDL: Addressing Data Imbalance
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2024)
Series of Hessian-Vector Products for Tractable Saddle-Free Newton Optimisation of Neural Networks
von: Oldewage, Elre T., et al.
Veröffentlicht: (2023)
von: Oldewage, Elre T., et al.
Veröffentlicht: (2023)
Satisficing and Optimal Generalised Planning via Goal Regression (Extended Version)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2025)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2025)
Generative Data Refinement: Just Ask for Better Data
von: Jiang, Minqi, et al.
Veröffentlicht: (2025)
von: Jiang, Minqi, et al.
Veröffentlicht: (2025)
lpNTK: Better Generalisation with Less Data via Sample Interaction During Learning
von: Guo, Shangmin, et al.
Veröffentlicht: (2024)
von: Guo, Shangmin, et al.
Veröffentlicht: (2024)
Making Better Use of Unlabelled Data in Bayesian Active Learning
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
Lower Difficulty and Better Robustness: A Bregman Divergence Perspective for Adversarial Training
von: Wu, Zihui, et al.
Veröffentlicht: (2022)
von: Wu, Zihui, et al.
Veröffentlicht: (2022)
First Attentions Last: Better Exploiting First Attentions for Efficient Transformer Training
von: Kim, Gyudong, et al.
Veröffentlicht: (2025)
von: Kim, Gyudong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Gauss-Newton Unlearning for the LLM Era
von: McKinney, Lev, et al.
Veröffentlicht: (2026) -
Training Data Attribution via Approximate Unrolled Differentiation
von: Bae, Juhan, et al.
Veröffentlicht: (2024) -
Pluralistic Alignment Over Time
von: Klassen, Toryn Q., et al.
Veröffentlicht: (2024) -
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2023) -
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
von: Alamdari, Parand A., et al.
Veröffentlicht: (2026)