Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
Fuente:
arXiv
Salvato in:
| Autori principali: | Galashov, Alexandre, Titsias, Michalis K., György, András, Lyle, Clare, Pascanu, Razvan, Teh, Yee Whye, Sahani, Maneesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Kalman Filter for Online Classification of Non-Stationary Data
di: Titsias, Michalis K., et al.
Pubblicazione: (2023)
di: Titsias, Michalis K., et al.
Pubblicazione: (2023)
Revisiting Dynamic Evaluation: Online Adaptation for Large Language Models
di: Rannen-Triki, Amal, et al.
Pubblicazione: (2024)
di: Rannen-Triki, Amal, et al.
Pubblicazione: (2024)
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
di: Li, Qinyu, et al.
Pubblicazione: (2025)
di: Li, Qinyu, et al.
Pubblicazione: (2025)
What Can Grokking Teach Us About Learning Under Nonstationarity?
di: Lyle, Clare, et al.
Pubblicazione: (2025)
di: Lyle, Clare, et al.
Pubblicazione: (2025)
The Illusion of Stochasticity in LLMs
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
New Bounds for Sparse Variational Gaussian Processes
di: Titsias, Michalis K.
Pubblicazione: (2025)
di: Titsias, Michalis K.
Pubblicazione: (2025)
Incorporating Unlabelled Data into Bayesian Neural Networks
di: Sharma, Mrinank, et al.
Pubblicazione: (2023)
di: Sharma, Mrinank, et al.
Pubblicazione: (2023)
Disentangling the Causes of Plasticity Loss in Neural Networks
di: Lyle, Clare, et al.
Pubblicazione: (2024)
di: Lyle, Clare, et al.
Pubblicazione: (2024)
Demystifying Diffusion Objectives: Reweighted Losses are Better Variational Bounds
di: Shi, Jiaxin, et al.
Pubblicazione: (2025)
di: Shi, Jiaxin, et al.
Pubblicazione: (2025)
A solution for the mean parametrization of the von Mises-Fisher distribution
di: Nonnenmacher, Marcel, et al.
Pubblicazione: (2024)
di: Nonnenmacher, Marcel, et al.
Pubblicazione: (2024)
Deep Grokking: Would Deep Neural Networks Generalize Better?
di: Fan, Simin, et al.
Pubblicazione: (2024)
di: Fan, Simin, et al.
Pubblicazione: (2024)
Personalized Federated Learning with Exact Stochastic Gradient Descent
di: Nikoloutsopoulos, Sotirios, et al.
Pubblicazione: (2022)
di: Nikoloutsopoulos, Sotirios, et al.
Pubblicazione: (2022)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
di: Zheng, Zhi, et al.
Pubblicazione: (2025)
di: Zheng, Zhi, et al.
Pubblicazione: (2025)
Sparse Gaussian Processes: Structured Approximations and Power-EP Revisited
di: Bui, Thang D., et al.
Pubblicazione: (2025)
di: Bui, Thang D., et al.
Pubblicazione: (2025)
Sampling on Discrete Spaces with Temporal Point Processes
di: Stewart, Cameron A., et al.
Pubblicazione: (2026)
di: Stewart, Cameron A., et al.
Pubblicazione: (2026)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
di: Veness, Joel, et al.
Pubblicazione: (2025)
di: Veness, Joel, et al.
Pubblicazione: (2025)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
di: Sims, Anya, et al.
Pubblicazione: (2024)
di: Sims, Anya, et al.
Pubblicazione: (2024)
LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories
di: Kang, Liwei, et al.
Pubblicazione: (2026)
di: Kang, Liwei, et al.
Pubblicazione: (2026)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
di: Nguyen-Hien, T. Duy, et al.
Pubblicazione: (2025)
di: Nguyen-Hien, T. Duy, et al.
Pubblicazione: (2025)
From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing
di: Liu, Wei, et al.
Pubblicazione: (2026)
di: Liu, Wei, et al.
Pubblicazione: (2026)
Sparse Orthogonal Variational Inference for Gaussian Processes
di: Shi, Jiaxin, et al.
Pubblicazione: (2019)
di: Shi, Jiaxin, et al.
Pubblicazione: (2019)
Variance Reduction for the Independent Metropolis Sampler
di: Liu, Siran, et al.
Pubblicazione: (2024)
di: Liu, Siran, et al.
Pubblicazione: (2024)
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
di: Lai, Yuhang, et al.
Pubblicazione: (2026)
di: Lai, Yuhang, et al.
Pubblicazione: (2026)
L3Ms -- Lagrange Large Language Models
di: Dhillon, Guneet S., et al.
Pubblicazione: (2024)
di: Dhillon, Guneet S., et al.
Pubblicazione: (2024)
SymDiff: Equivariant Diffusion via Stochastic Symmetrisation
di: Zhang, Leo, et al.
Pubblicazione: (2024)
di: Zhang, Leo, et al.
Pubblicazione: (2024)
Latent Space Representations of Neural Algorithmic Reasoners
di: Mirjanić, Vladimir V., et al.
Pubblicazione: (2023)
di: Mirjanić, Vladimir V., et al.
Pubblicazione: (2023)
Lattice: Learning to Efficiently Compress the Memory
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
Fine-Tuned In-Context Learners for Efficient Adaptation
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
Normalization and effective learning rates in reinforcement learning
di: Lyle, Clare, et al.
Pubblicazione: (2024)
di: Lyle, Clare, et al.
Pubblicazione: (2024)
Meta-Learning Objectives for Preference Optimization
di: Alfano, Carlo, et al.
Pubblicazione: (2024)
di: Alfano, Carlo, et al.
Pubblicazione: (2024)
EvIL: Evolution Strategies for Generalisable Imitation Learning
di: Sapora, Silvia, et al.
Pubblicazione: (2024)
di: Sapora, Silvia, et al.
Pubblicazione: (2024)
Manifold Aware Denoising Score Matching (MAD)
di: Levy-Jurgenson, Alona, et al.
Pubblicazione: (2026)
di: Levy-Jurgenson, Alona, et al.
Pubblicazione: (2026)
Revisiting Adam for Streaming Reinforcement Learning
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
Gaussian Invariant Markov Chain Monte Carlo
di: Titsias, Michalis K., et al.
Pubblicazione: (2025)
di: Titsias, Michalis K., et al.
Pubblicazione: (2025)
Learning-Order Autoregressive Models with Application to Molecular Graph Generation
di: Wang, Zhe, et al.
Pubblicazione: (2025)
di: Wang, Zhe, et al.
Pubblicazione: (2025)
Maximum Likelihood Learning of Latent Dynamics Without Reconstruction
di: Hromadka, Samo, et al.
Pubblicazione: (2025)
di: Hromadka, Samo, et al.
Pubblicazione: (2025)
Rao-Blackwellised Reparameterisation Gradients
di: Lam, Kevin H., et al.
Pubblicazione: (2025)
di: Lam, Kevin H., et al.
Pubblicazione: (2025)
Successor-Predecessor Intrinsic Exploration
di: Yu, Changmin, et al.
Pubblicazione: (2023)
di: Yu, Changmin, et al.
Pubblicazione: (2023)
Selective Safety Steering via Value-Filtered Decoding
di: Einbinder, Bat-Sheva, et al.
Pubblicazione: (2026)
di: Einbinder, Bat-Sheva, et al.
Pubblicazione: (2026)
Simplified and Generalized Masked Diffusion for Discrete Data
di: Shi, Jiaxin, et al.
Pubblicazione: (2024)
di: Shi, Jiaxin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Kalman Filter for Online Classification of Non-Stationary Data
di: Titsias, Michalis K., et al.
Pubblicazione: (2023) -
Revisiting Dynamic Evaluation: Online Adaptation for Large Language Models
di: Rannen-Triki, Amal, et al.
Pubblicazione: (2024) -
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
di: Li, Qinyu, et al.
Pubblicazione: (2025) -
What Can Grokking Teach Us About Learning Under Nonstationarity?
di: Lyle, Clare, et al.
Pubblicazione: (2025) -
The Illusion of Stochasticity in LLMs
di: Gu, Xiangming, et al.
Pubblicazione: (2026)