Saved in:
| Main Authors: | Davarmanesh, Parmida, Wilson, Ashia, Radhakrishnan, Adityanarayanan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.00333 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Catapults in SGD: spikes in the training loss and their impact on generalization through feature learning
by: Zhu, Libin, et al.
Published: (2023)
by: Zhu, Libin, et al.
Published: (2023)
xRFM: Accurate, scalable, and interpretable feature learning models for tabular data
by: Beaglehole, Daniel, et al.
Published: (2025)
by: Beaglehole, Daniel, et al.
Published: (2025)
Toward universal steering and monitoring of AI models
by: Beaglehole, Daniel, et al.
Published: (2025)
by: Beaglehole, Daniel, et al.
Published: (2025)
Linear Recursive Feature Machines provably recover low-rank matrices
by: Radhakrishnan, Adityanarayanan, et al.
Published: (2024)
by: Radhakrishnan, Adityanarayanan, et al.
Published: (2024)
Context-Scaling versus Task-Scaling in In-Context Learning
by: Abedsoltan, Amirhesam, et al.
Published: (2024)
by: Abedsoltan, Amirhesam, et al.
Published: (2024)
The Weight Gram Matrix Captures Sequential Feature Linearization in Deep Networks
by: Cha, Taehun, et al.
Published: (2026)
by: Cha, Taehun, et al.
Published: (2026)
Quadratic models for understanding catapult dynamics of neural networks
by: Zhu, Libin, et al.
Published: (2022)
by: Zhu, Libin, et al.
Published: (2022)
The Approximate Fisher Influence Function: Faster Estimation of Data Influence in Statistical Models
by: Lev, Omri, et al.
Published: (2024)
by: Lev, Omri, et al.
Published: (2024)
Emergence in non-neural models: grokking modular arithmetic via average gradient outer product
by: Mallinar, Neil, et al.
Published: (2024)
by: Mallinar, Neil, et al.
Published: (2024)
Position: AI Evaluations Should be Grounded on a Theory of Capability
by: Jo, Nathanael, et al.
Published: (2025)
by: Jo, Nathanael, et al.
Published: (2025)
Mean-field underdamped Langevin dynamics and its spacetime discretization
by: Fu, Qiang, et al.
Published: (2023)
by: Fu, Qiang, et al.
Published: (2023)
Semivalue-based data valuation is arbitrary and gameable
by: Diehl, Hannah, et al.
Published: (2025)
by: Diehl, Hannah, et al.
Published: (2025)
From Cross-Validation to SURE: Asymptotic Risk of Tuned Regularized Estimators
by: Adusumilli, Karun, et al.
Published: (2026)
by: Adusumilli, Karun, et al.
Published: (2026)
High-accuracy sampling from constrained spaces with the Metropolis-adjusted Preconditioned Langevin Algorithm
by: Srinivasan, Vishwak, et al.
Published: (2024)
by: Srinivasan, Vishwak, et al.
Published: (2024)
Fast sampling from constrained spaces using the Metropolis-adjusted Mirror Langevin algorithm
by: Srinivasan, Vishwak, et al.
Published: (2023)
by: Srinivasan, Vishwak, et al.
Published: (2023)
UCD: Unlearning in LLMs via Contrastive Decoding
by: Suriyakumar, Vinith M., et al.
Published: (2025)
by: Suriyakumar, Vinith M., et al.
Published: (2025)
Adaptive Kernel Selection for Stein Variational Gradient Descent
by: Melcher, Moritz, et al.
Published: (2025)
by: Melcher, Moritz, et al.
Published: (2025)
Layered Unlearning for Adversarial Relearning
by: Qian, Timothy, et al.
Published: (2025)
by: Qian, Timothy, et al.
Published: (2025)
Adaptive Backtracking Line Search
by: Cavalcanti, Joao V., et al.
Published: (2024)
by: Cavalcanti, Joao V., et al.
Published: (2024)
An analysis of the combination of feature selection and machine learning methods for an accurate and timely detection of lung cancer
by: Shahriyar, Omid, et al.
Published: (2025)
by: Shahriyar, Omid, et al.
Published: (2025)
Clinician input steers frontier AI models toward both accurate and harmful decisions
by: Lopez, Ivan, et al.
Published: (2026)
by: Lopez, Ivan, et al.
Published: (2026)
Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search
by: Flores, Gerardo A., et al.
Published: (2026)
by: Flores, Gerardo A., et al.
Published: (2026)
DiffusionWorldViewer: Exposing and Broadening the Worldview Reflected by Generative Text-to-Image Models
by: De Simone, Zoe, et al.
Published: (2023)
by: De Simone, Zoe, et al.
Published: (2023)
The Fast Mixing Mechanism for Differential Privacy
by: Lev, Omri, et al.
Published: (2026)
by: Lev, Omri, et al.
Published: (2026)
The Gaussian Mixing Mechanism: Renyi Differential Privacy via Gaussian Sketches
by: Lev, Omri, et al.
Published: (2025)
by: Lev, Omri, et al.
Published: (2025)
Unstable Unlearning: The Hidden Risk of Concept Resurgence in Diffusion Models
by: Suriyakumar, Vinith M., et al.
Published: (2024)
by: Suriyakumar, Vinith M., et al.
Published: (2024)
Contextual Linear Activation Steering of Language Models
by: Hsu, Brandon, et al.
Published: (2026)
by: Hsu, Brandon, et al.
Published: (2026)
Aligning Evaluation with Clinical Priorities: Calibration, Label Shift, and Error Costs
by: Flores, Gerardo A., et al.
Published: (2025)
by: Flores, Gerardo A., et al.
Published: (2025)
Supervised learning pays attention
by: Craig, Erin, et al.
Published: (2025)
by: Craig, Erin, et al.
Published: (2025)
A Consequentialist Critique of Binary Classification Evaluation: Theory, Practice, and Tools
by: Flores, Gerardo, et al.
Published: (2025)
by: Flores, Gerardo, et al.
Published: (2025)
An extension of linear self-attention for in-context learning
by: Hagiwara, Katsuyuki
Published: (2025)
by: Hagiwara, Katsuyuki
Published: (2025)
In-Context Clustering with Large Language Models
by: Wang, Ying, et al.
Published: (2025)
by: Wang, Ying, et al.
Published: (2025)
Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM
by: Suriyakumar, Vinith M., et al.
Published: (2026)
by: Suriyakumar, Vinith M., et al.
Published: (2026)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Easydiagnos: a framework for accurate feature selection for automatic diagnosis in smart healthcare
by: Maji, Prasenjit, et al.
Published: (2024)
by: Maji, Prasenjit, et al.
Published: (2024)
Redundant feature screening method for human activity recognition based on attention purification mechanism
by: Li, Xiaoyang, et al.
Published: (2025)
by: Li, Xiaoyang, et al.
Published: (2025)
Language steering in latent space to mitigate unintended code-switching
by: Goncharov, Andrey, et al.
Published: (2025)
by: Goncharov, Andrey, et al.
Published: (2025)
Beyond Barren Plateaus: A Scalable Quantum Convolutional Architecture for High-Fidelity Image Classification
by: Delhibabu, Radhakrishnan
Published: (2026)
by: Delhibabu, Radhakrishnan
Published: (2026)
Material synthesis through simulations guided by machine learning: a position paper
by: Syed, Usman, et al.
Published: (2024)
by: Syed, Usman, et al.
Published: (2024)
ESM-NBR: fast and accurate nucleic acid-binding residue prediction via protein language model feature representation and multi-task learning
by: Zeng, Wenwu, et al.
Published: (2023)
by: Zeng, Wenwu, et al.
Published: (2023)
Similar Items
-
Catapults in SGD: spikes in the training loss and their impact on generalization through feature learning
by: Zhu, Libin, et al.
Published: (2023) -
xRFM: Accurate, scalable, and interpretable feature learning models for tabular data
by: Beaglehole, Daniel, et al.
Published: (2025) -
Toward universal steering and monitoring of AI models
by: Beaglehole, Daniel, et al.
Published: (2025) -
Linear Recursive Feature Machines provably recover low-rank matrices
by: Radhakrishnan, Adityanarayanan, et al.
Published: (2024) -
Context-Scaling versus Task-Scaling in In-Context Learning
by: Abedsoltan, Amirhesam, et al.
Published: (2024)