Reset & Distill: A Recipe for Overcoming Negative Transfer in Continual Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Ahn, Hongjoon, Hyeon, Jinu, Oh, Youngmin, Hwang, Bosun, Moon, Taesup |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Action-Sufficient Goal Representations
by: Hyeon, Jinu, et al.
Published: (2026)
by: Hyeon, Jinu, et al.
Published: (2026)
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
by: Choi, Heewoong, et al.
Published: (2024)
by: Choi, Heewoong, et al.
Published: (2024)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025)
by: Ahn, Hongjoon, et al.
Published: (2025)
Forget Forgetting: Continual Learning in a World of Abundant Memory
by: Cho, Dongkyu, et al.
Published: (2025)
by: Cho, Dongkyu, et al.
Published: (2025)
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
by: Moon, Seungjae, et al.
Published: (2026)
by: Moon, Seungjae, et al.
Published: (2026)
The Power of Resets in Online Reinforcement Learning
by: Mhammedi, Zakaria, et al.
Published: (2024)
by: Mhammedi, Zakaria, et al.
Published: (2024)
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
by: Hwang, Junseo, et al.
Published: (2025)
by: Hwang, Junseo, et al.
Published: (2025)
Self-Normalized Resets for Plasticity in Continual Learning
by: Farias, Vivek F., et al.
Published: (2024)
by: Farias, Vivek F., et al.
Published: (2024)
Regularizing with Pseudo-Negatives for Continual Self-Supervised Learning
by: Cha, Sungmin, et al.
Published: (2023)
by: Cha, Sungmin, et al.
Published: (2023)
PRISP: Privacy-Safe Few-Shot Personalization via Lightweight Adaptation
by: Park, Junho, et al.
Published: (2026)
by: Park, Junho, et al.
Published: (2026)
Overcoming Data Inequality across Domains with Semi-Supervised Domain Generalization
by: Park, Jinha, et al.
Published: (2024)
by: Park, Jinha, et al.
Published: (2024)
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning
by: Singireddy, Suraj, et al.
Published: (2023)
by: Singireddy, Suraj, et al.
Published: (2023)
Stable On-Policy Distillation through Adaptive Target Reformulation
by: Jang, Ijun, et al.
Published: (2026)
by: Jang, Ijun, et al.
Published: (2026)
Addressing Negative Transfer in Diffusion Models
by: Go, Hyojun, et al.
Published: (2023)
by: Go, Hyojun, et al.
Published: (2023)
A Recipe for Stable Offline Multi-agent Reinforcement Learning
by: Lee, Dongsu, et al.
Published: (2026)
by: Lee, Dongsu, et al.
Published: (2026)
CADENT: Gated Hybrid Distillation for Sample-Efficient Transfer in Reinforcement Learning
by: Alinejad, Mahyar, et al.
Published: (2026)
by: Alinejad, Mahyar, et al.
Published: (2026)
Multi-granularity Knowledge Transfer for Continual Reinforcement Learning
by: Pan, Chaofan, et al.
Published: (2024)
by: Pan, Chaofan, et al.
Published: (2024)
M3: Mamba-assisted Multi-Circuit Optimization via MBRL with Effective Scheduling
by: Oh, Youngmin, et al.
Published: (2024)
by: Oh, Youngmin, et al.
Published: (2024)
Efficient Epistemic Uncertainty Estimation for Large Language Models via Knowledge Distillation
by: Park, Seonghyeon, et al.
Published: (2026)
by: Park, Seonghyeon, et al.
Published: (2026)
Overcoming Slow Decision Frequencies in Continuous Control: Model-Based Sequence Reinforcement Learning for Model-Free Control
by: Patel, Devdhar, et al.
Published: (2024)
by: Patel, Devdhar, et al.
Published: (2024)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
by: Li, Chunxiao, et al.
Published: (2024)
by: Li, Chunxiao, et al.
Published: (2024)
A Practical Two-Stage Recipe for Mathematical LLMs: Maximizing Accuracy with SFT and Efficiency with Reinforcement Learning
by: Yoshihara, Hiroshi, et al.
Published: (2025)
by: Yoshihara, Hiroshi, et al.
Published: (2025)
Learning to Act Robustly with View-Invariant Latent Actions
by: Jeong, Youngjoon, et al.
Published: (2026)
by: Jeong, Youngjoon, et al.
Published: (2026)
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
by: Yin, Bojian, et al.
Published: (2024)
by: Yin, Bojian, et al.
Published: (2024)
Overcoming Growth-Induced Forgetting in Task-Agnostic Continual Learning
by: Zhao, Yuqing, et al.
Published: (2024)
by: Zhao, Yuqing, et al.
Published: (2024)
Diffusion-Assisted Distillation for Self-Supervised Graph Representation Learning with MLPs
by: Ahn, Seong Jin, et al.
Published: (2025)
by: Ahn, Seong Jin, et al.
Published: (2025)
DoMIX: An Efficient Framework for Exploiting Domain Knowledge in Fine-Tuning
by: Kim, Dohoon, et al.
Published: (2025)
by: Kim, Dohoon, et al.
Published: (2025)
Dataset Distillation for Offline Reinforcement Learning
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
Reinforcement Learning via Self-Distillation
by: Hübotter, Jonas, et al.
Published: (2026)
by: Hübotter, Jonas, et al.
Published: (2026)
Overcoming Overfitting in Reinforcement Learning via Gaussian Process Diffusion Policy
by: Horprasert, Amornyos, et al.
Published: (2025)
by: Horprasert, Amornyos, et al.
Published: (2025)
Directed Structural Adaptation to Overcome Statistical Conflicts and Enable Continual Learning
by: Erden, Zeki Doruk, et al.
Published: (2024)
by: Erden, Zeki Doruk, et al.
Published: (2024)
Low-redundancy Distillation for Continual Learning
by: Liu, RuiQi, et al.
Published: (2023)
by: Liu, RuiQi, et al.
Published: (2023)
A Cross Modal Knowledge Distillation & Data Augmentation Recipe for Improving Transcriptomics Representations through Morphological Features
by: Bendidi, Ihab, et al.
Published: (2025)
by: Bendidi, Ihab, et al.
Published: (2025)
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
Contrastive Residual Energy Test-time Adaptation
by: Han, Yewon, et al.
Published: (2025)
by: Han, Yewon, et al.
Published: (2025)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
by: Li, Lanpei, et al.
Published: (2024)
by: Li, Lanpei, et al.
Published: (2024)
A Ridge Too Far: Correcting Over-Shrinkage via Negative Regularization
by: Kim, Dongseok, et al.
Published: (2025)
by: Kim, Dongseok, et al.
Published: (2025)
X-PEFT: eXtremely Parameter-Efficient Fine-Tuning for Extreme Multi-Profile Scenarios
by: Kwak, Namju, et al.
Published: (2024)
by: Kwak, Namju, et al.
Published: (2024)
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
by: Oh, Youngmin, et al.
Published: (2026)
by: Oh, Youngmin, et al.
Published: (2026)
Tabular Transfer Learning via Prompting LLMs
by: Nam, Jaehyun, et al.
Published: (2024)
by: Nam, Jaehyun, et al.
Published: (2024)
Similar Items
-
Action-Sufficient Goal Representations
by: Hyeon, Jinu, et al.
Published: (2026) -
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
by: Choi, Heewoong, et al.
Published: (2024) -
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025) -
Forget Forgetting: Continual Learning in a World of Abundant Memory
by: Cho, Dongkyu, et al.
Published: (2025) -
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
by: Moon, Seungjae, et al.
Published: (2026)