Avoiding Catastrophe in Online Learning by Asking for Help
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Plaut, Benjamin, Zhu, Hanlin, Russell, Stuart |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Safe Learning Under Irreversible Dynamics via Asking for Help
von: Plaut, Benjamin, et al.
Veröffentlicht: (2025)
von: Plaut, Benjamin, et al.
Veröffentlicht: (2025)
Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards
von: Liaw, Sarah, et al.
Veröffentlicht: (2025)
von: Liaw, Sarah, et al.
Veröffentlicht: (2025)
Getting By Goal Misgeneralization With a Little Help From a Mentor
von: Trinh, Tu, et al.
Veröffentlicht: (2024)
von: Trinh, Tu, et al.
Veröffentlicht: (2024)
Safety Training Persists Through Helpfulness Optimization in LLM Agents
von: Plaut, Benjamin
Veröffentlicht: (2026)
von: Plaut, Benjamin
Veröffentlicht: (2026)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
von: Plaut, Benjamin, et al.
Veröffentlicht: (2024)
von: Plaut, Benjamin, et al.
Veröffentlicht: (2024)
Learning Identifiable Structures Helps Avoid Bias in DNN-based Supervised Causal Learning
von: Zhang, Jiaru, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaru, et al.
Veröffentlicht: (2025)
Consequentialist Objectives and Catastrophe
von: Marklund, Henrik, et al.
Veröffentlicht: (2026)
von: Marklund, Henrik, et al.
Veröffentlicht: (2026)
GSM-Agent: Understanding Agentic Reasoning Using Controllable Environments
von: Zhu, Hanlin, et al.
Veröffentlicht: (2025)
von: Zhu, Hanlin, et al.
Veröffentlicht: (2025)
How to Leverage Predictive Uncertainty Estimates for Reducing Catastrophic Forgetting in Online Continual Learning
von: Serra, Giuseppe, et al.
Veröffentlicht: (2024)
von: Serra, Giuseppe, et al.
Veröffentlicht: (2024)
Sequencing to Mitigate Catastrophic Forgetting in Continual Learning
von: Moussa, Hesham G., et al.
Veröffentlicht: (2025)
von: Moussa, Hesham G., et al.
Veröffentlicht: (2025)
The Effective Horizon Explains Deep RL Performance in Stochastic Environments
von: Laidlaw, Cassidy, et al.
Veröffentlicht: (2023)
von: Laidlaw, Cassidy, et al.
Veröffentlicht: (2023)
Addressing Loss of Plasticity and Catastrophic Forgetting in Continual Learning
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
On the Implicit Adversariality of Catastrophic Forgetting in Deep Continual Learning
von: Peng, Ze, et al.
Veröffentlicht: (2025)
von: Peng, Ze, et al.
Veröffentlicht: (2025)
BAMDP Shaping: a Unified Framework for Intrinsic Motivation and Reward Shaping
von: Lidayan, Aly, et al.
Veröffentlicht: (2024)
von: Lidayan, Aly, et al.
Veröffentlicht: (2024)
Cross-Domain Imitation Learning via Optimal Transport
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2021)
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2021)
Natural Mitigation of Catastrophic Interference: Continual Learning in Power-Law Learning Environments
von: Gandhi, Atith, et al.
Veröffentlicht: (2024)
von: Gandhi, Atith, et al.
Veröffentlicht: (2024)
CORE: Mitigating Catastrophic Forgetting in Continual Learning through Cognitive Replay
von: Zhang, Jianshu, et al.
Veröffentlicht: (2024)
von: Zhang, Jianshu, et al.
Veröffentlicht: (2024)
Catastrophic Forgetting in Kolmogorov-Arnold Networks
von: Rahman, Mohammad Marufur, et al.
Veröffentlicht: (2025)
von: Rahman, Mohammad Marufur, et al.
Veröffentlicht: (2025)
Learning the Preferences of a Learning Agent
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2026)
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2026)
Evidence of Learned Look-Ahead in a Chess-Playing Neural Network
von: Jenner, Erik, et al.
Veröffentlicht: (2024)
von: Jenner, Erik, et al.
Veröffentlicht: (2024)
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
von: Hoang, Huy, et al.
Veröffentlicht: (2023)
von: Hoang, Huy, et al.
Veröffentlicht: (2023)
Learning to Help in Multi-Class Settings
von: Wu, Yu, et al.
Veröffentlicht: (2025)
von: Wu, Yu, et al.
Veröffentlicht: (2025)
Can LLMs Alleviate Catastrophic Forgetting in Graph Continual Learning? A Systematic Study
von: Cheng, Ziyang, et al.
Veröffentlicht: (2025)
von: Cheng, Ziyang, et al.
Veröffentlicht: (2025)
Active teacher selection for reward learning
von: Freedman, Rachel, et al.
Veröffentlicht: (2023)
von: Freedman, Rachel, et al.
Veröffentlicht: (2023)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
When Your AIs Deceive You: Challenges of Partial Observability in Reinforcement Learning from Human Feedback
von: Lang, Leon, et al.
Veröffentlicht: (2024)
von: Lang, Leon, et al.
Veröffentlicht: (2024)
Distribution-Level Memory Recall for Continual Learning: Preserving Knowledge and Avoiding Confusion
von: Cheng, Shaoxu, et al.
Veröffentlicht: (2024)
von: Cheng, Shaoxu, et al.
Veröffentlicht: (2024)
Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding
von: Pons, Gerard, et al.
Veröffentlicht: (2026)
von: Pons, Gerard, et al.
Veröffentlicht: (2026)
Quantifying Catastrophic Forgetting in IoT Intrusion Detection Systems
von: Banerjee, Sourasekhar, et al.
Veröffentlicht: (2026)
von: Banerjee, Sourasekhar, et al.
Veröffentlicht: (2026)
A Conformal Predictive Measure for Assessing Catastrophic Forgetting
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2025)
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2025)
On Representation Complexity of Model-based and Model-free Reinforcement Learning
von: Zhu, Hanlin, et al.
Veröffentlicht: (2023)
von: Zhu, Hanlin, et al.
Veröffentlicht: (2023)
Synthetic Error Injection Fails to Elicit Self-Correction In Language Models
von: Wu, David X., et al.
Veröffentlicht: (2025)
von: Wu, David X., et al.
Veröffentlicht: (2025)
Crafting Interpretable Embeddings by Asking LLMs Questions
von: Benara, Vinamra, et al.
Veröffentlicht: (2024)
von: Benara, Vinamra, et al.
Veröffentlicht: (2024)
Intelligent Learning Rate Distribution to reduce Catastrophic Forgetting in Transformers
von: Kenneweg, Philip, et al.
Veröffentlicht: (2024)
von: Kenneweg, Philip, et al.
Veröffentlicht: (2024)
Explaining Robustness to Catastrophic Forgetting Through Incremental Concept Formation
von: Barari, Nicki, et al.
Veröffentlicht: (2025)
von: Barari, Nicki, et al.
Veröffentlicht: (2025)
Catastrophic Forgetting Mitigation Through Plateau Phase Activity Profiling
von: Mashiach, Idan, et al.
Veröffentlicht: (2025)
von: Mashiach, Idan, et al.
Veröffentlicht: (2025)
A Generalized Acquisition Function for Preference-based Reward Learning
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
Curriculum-enhanced GroupDRO: Challenging the Norm of Avoiding Curriculum Learning in Subpopulation Shift Setups
von: Barbalau, Antonio
Veröffentlicht: (2024)
von: Barbalau, Antonio
Veröffentlicht: (2024)
AI Alignment with Changing and Influenceable Reward Functions
von: Carroll, Micah, et al.
Veröffentlicht: (2024)
von: Carroll, Micah, et al.
Veröffentlicht: (2024)
Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning
von: Cheng, Ziheng, et al.
Veröffentlicht: (2026)
von: Cheng, Ziheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Safe Learning Under Irreversible Dynamics via Asking for Help
von: Plaut, Benjamin, et al.
Veröffentlicht: (2025) -
Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards
von: Liaw, Sarah, et al.
Veröffentlicht: (2025) -
Getting By Goal Misgeneralization With a Little Help From a Mentor
von: Trinh, Tu, et al.
Veröffentlicht: (2024) -
Safety Training Persists Through Helpfulness Optimization in LLM Agents
von: Plaut, Benjamin
Veröffentlicht: (2026) -
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
von: Plaut, Benjamin, et al.
Veröffentlicht: (2024)