Stochastic Resetting Mitigates Latent Gradient Bias of SGD from Label Noise
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bae, Youngkyoung, Song, Yeongwoo, Jeong, Hawoong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inferring the Langevin Equation with Uncertainty via Bayesian Neural Networks
von: Bae, Youngkyoung, et al.
Veröffentlicht: (2024)
von: Bae, Youngkyoung, et al.
Veröffentlicht: (2024)
Towards Cross Domain Generalization of Hamiltonian Representation via Meta Learning
von: Song, Yeongwoo, et al.
Veröffentlicht: (2022)
von: Song, Yeongwoo, et al.
Veröffentlicht: (2022)
Gaussian Universality in Neural Network Dynamics with Generalized Structured Input Distributions
von: Bae, Jaeyong, et al.
Veröffentlicht: (2024)
von: Bae, Jaeyong, et al.
Veröffentlicht: (2024)
Uncovering Emergent Physics Representations Learned In-Context by Large Language Models
von: Song, Yeongwoo, et al.
Veröffentlicht: (2025)
von: Song, Yeongwoo, et al.
Veröffentlicht: (2025)
Spontaneous symmetry breaking and Goldstone modes for deep information propagation
von: Iqbal, Nabil, et al.
Veröffentlicht: (2026)
von: Iqbal, Nabil, et al.
Veröffentlicht: (2026)
Generative Stochastic Optimal Transport: Guided Harmonic Path-Integral Diffusion
von: Chertkov, Michael
Veröffentlicht: (2025)
von: Chertkov, Michael
Veröffentlicht: (2025)
On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective
von: Li, Yuhao, et al.
Veröffentlicht: (2026)
von: Li, Yuhao, et al.
Veröffentlicht: (2026)
Temporal Memory for Resource-Constrained Agents: Continual Learning via Stochastic Compress-Add-Smooth
von: Chertkov, Michael
Veröffentlicht: (2026)
von: Chertkov, Michael
Veröffentlicht: (2026)
Trees to Flows and Back: Unifying Decision Trees and Diffusion Models
von: Ramachandran, Sai Niranjan, et al.
Veröffentlicht: (2026)
von: Ramachandran, Sai Niranjan, et al.
Veröffentlicht: (2026)
Symmetry and Generalisation in Neural Approximations of Renormalisation Transformations
von: Ashworth, Cassidy, et al.
Veröffentlicht: (2025)
von: Ashworth, Cassidy, et al.
Veröffentlicht: (2025)
Grokking as a Falsifiable Finite-Size Transition
von: Bi, Yuda, et al.
Veröffentlicht: (2026)
von: Bi, Yuda, et al.
Veröffentlicht: (2026)
Thermodynamic Irreversibility of Training Algorithms
von: Ziyin, Liu, et al.
Veröffentlicht: (2026)
von: Ziyin, Liu, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints
von: Choi, Hoyun, et al.
Veröffentlicht: (2026)
von: Choi, Hoyun, et al.
Veröffentlicht: (2026)
Cross-fluctuation phase transitions reveal sampling dynamics in diffusion models
von: Ramachandran, Sai Niranjan, et al.
Veröffentlicht: (2025)
von: Ramachandran, Sai Niranjan, et al.
Veröffentlicht: (2025)
Ridge Regression from Poisson Resetting: A Renewal Perspective on Spectral Regularization
von: Jolakoski, Petar
Veröffentlicht: (2026)
von: Jolakoski, Petar
Veröffentlicht: (2026)
Maximum diffusion reinforcement learning
von: Berrueta, Thomas A., et al.
Veröffentlicht: (2023)
von: Berrueta, Thomas A., et al.
Veröffentlicht: (2023)
A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics
von: Yao, Louie Hong, et al.
Veröffentlicht: (2026)
von: Yao, Louie Hong, et al.
Veröffentlicht: (2026)
Exact Identity Linking Entropy Production and Mutual Information
von: Cho, Doohyeong, et al.
Veröffentlicht: (2025)
von: Cho, Doohyeong, et al.
Veröffentlicht: (2025)
Tensor tree learns hidden relational structures in data to construct generative models
von: Harada, Kenji, et al.
Veröffentlicht: (2024)
von: Harada, Kenji, et al.
Veröffentlicht: (2024)
Phase Transitions in the Output Distribution of Large Language Models
von: Arnold, Julian, et al.
Veröffentlicht: (2024)
von: Arnold, Julian, et al.
Veröffentlicht: (2024)
Mixing Artificial and Natural Intelligence: From Statistical Mechanics to AI and Back to Turbulence
von: Chertkov, Michael
Veröffentlicht: (2024)
von: Chertkov, Michael
Veröffentlicht: (2024)
Entropy, concentration, and learning: a statistical mechanics primer
von: Balsubramani, Akshay
Veröffentlicht: (2024)
von: Balsubramani, Akshay
Veröffentlicht: (2024)
Sampling Decisions
von: Chertkov, Michael, et al.
Veröffentlicht: (2025)
von: Chertkov, Michael, et al.
Veröffentlicht: (2025)
Adaptive Path Integral Diffusion: AdaPID
von: Chertkov, Michael, et al.
Veröffentlicht: (2025)
von: Chertkov, Michael, et al.
Veröffentlicht: (2025)
Scalable Discrete Diffusion Samplers: Combinatorial Optimization and Statistical Physics
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
IsingFormer: Augmenting Parallel Tempering With Learned Proposals
von: Bunaiyan, Saleh, et al.
Veröffentlicht: (2025)
von: Bunaiyan, Saleh, et al.
Veröffentlicht: (2025)
Relaxation-assisted reverse annealing on nonnegative/binary matrix factorization
von: Haba, Renichiro, et al.
Veröffentlicht: (2025)
von: Haba, Renichiro, et al.
Veröffentlicht: (2025)
On the Separability of Information in Diffusion Models
von: Premkumar, Akhil
Veröffentlicht: (2025)
von: Premkumar, Akhil
Veröffentlicht: (2025)
Exploring how deep learning decodes anomalous diffusion via Grad-CAM
von: Bae, Jaeyong, et al.
Veröffentlicht: (2024)
von: Bae, Jaeyong, et al.
Veröffentlicht: (2024)
Stochastic Resetting and Large Deviations
von: Evans, Martin R., et al.
Veröffentlicht: (2024)
von: Evans, Martin R., et al.
Veröffentlicht: (2024)
Autonomous Ratcheting by Stochastic Resetting
von: Ghosh, Pulak K., et al.
Veröffentlicht: (2023)
von: Ghosh, Pulak K., et al.
Veröffentlicht: (2023)
Role of volatility mixing in wealth condensation transition
von: Hur, Jaeseok, et al.
Veröffentlicht: (2026)
von: Hur, Jaeseok, et al.
Veröffentlicht: (2026)
Anomaly, class division, and decoupling in income dynamics
von: Hur, Jaeseok, et al.
Veröffentlicht: (2025)
von: Hur, Jaeseok, et al.
Veröffentlicht: (2025)
Stochastic Resetting Accelerates Policy Convergence in Reinforcement Learning
von: Zhou, Jello, et al.
Veröffentlicht: (2026)
von: Zhou, Jello, et al.
Veröffentlicht: (2026)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
von: Chen, Feng, et al.
Veröffentlicht: (2023)
von: Chen, Feng, et al.
Veröffentlicht: (2023)
Label Noise Robustness of Conformal Prediction
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2022)
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2022)
Emergence of Nonequilibrium Latent Cycles in Unsupervised Generative Modeling
von: Baiesi, Marco, et al.
Veröffentlicht: (2025)
von: Baiesi, Marco, et al.
Veröffentlicht: (2025)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
von: Lattimore, Tor
Veröffentlicht: (2026)
von: Lattimore, Tor
Veröffentlicht: (2026)
Detailed balance in large language model-driven agents
von: Song, Zhuo-Yang, et al.
Veröffentlicht: (2025)
von: Song, Zhuo-Yang, et al.
Veröffentlicht: (2025)
Stein-Rule Shrinkage for Stochastic Gradient Estimation in High Dimensions
von: Arashi, M., et al.
Veröffentlicht: (2026)
von: Arashi, M., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Inferring the Langevin Equation with Uncertainty via Bayesian Neural Networks
von: Bae, Youngkyoung, et al.
Veröffentlicht: (2024) -
Towards Cross Domain Generalization of Hamiltonian Representation via Meta Learning
von: Song, Yeongwoo, et al.
Veröffentlicht: (2022) -
Gaussian Universality in Neural Network Dynamics with Generalized Structured Input Distributions
von: Bae, Jaeyong, et al.
Veröffentlicht: (2024) -
Uncovering Emergent Physics Representations Learned In-Context by Large Language Models
von: Song, Yeongwoo, et al.
Veröffentlicht: (2025) -
Spontaneous symmetry breaking and Goldstone modes for deep information propagation
von: Iqbal, Nabil, et al.
Veröffentlicht: (2026)