Reset & Distill: A Recipe for Overcoming Negative Transfer in Continual Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Ahn, Hongjoon, Hyeon, Jinu, Oh, Youngmin, Hwang, Bosun, Moon, Taesup |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Action-Sufficient Goal Representations
di: Hyeon, Jinu, et al.
Pubblicazione: (2026)
di: Hyeon, Jinu, et al.
Pubblicazione: (2026)
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
di: Choi, Heewoong, et al.
Pubblicazione: (2024)
di: Choi, Heewoong, et al.
Pubblicazione: (2024)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
di: Ahn, Hongjoon, et al.
Pubblicazione: (2025)
di: Ahn, Hongjoon, et al.
Pubblicazione: (2025)
Forget Forgetting: Continual Learning in a World of Abundant Memory
di: Cho, Dongkyu, et al.
Pubblicazione: (2025)
di: Cho, Dongkyu, et al.
Pubblicazione: (2025)
The Power of Resets in Online Reinforcement Learning
di: Mhammedi, Zakaria, et al.
Pubblicazione: (2024)
di: Mhammedi, Zakaria, et al.
Pubblicazione: (2024)
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
di: Moon, Seungjae, et al.
Pubblicazione: (2026)
di: Moon, Seungjae, et al.
Pubblicazione: (2026)
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
di: Hwang, Junseo, et al.
Pubblicazione: (2025)
di: Hwang, Junseo, et al.
Pubblicazione: (2025)
Self-Normalized Resets for Plasticity in Continual Learning
di: Farias, Vivek F., et al.
Pubblicazione: (2024)
di: Farias, Vivek F., et al.
Pubblicazione: (2024)
Regularizing with Pseudo-Negatives for Continual Self-Supervised Learning
di: Cha, Sungmin, et al.
Pubblicazione: (2023)
di: Cha, Sungmin, et al.
Pubblicazione: (2023)
PRISP: Privacy-Safe Few-Shot Personalization via Lightweight Adaptation
di: Park, Junho, et al.
Pubblicazione: (2026)
di: Park, Junho, et al.
Pubblicazione: (2026)
Overcoming Data Inequality across Domains with Semi-Supervised Domain Generalization
di: Park, Jinha, et al.
Pubblicazione: (2024)
di: Park, Jinha, et al.
Pubblicazione: (2024)
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning
di: Singireddy, Suraj, et al.
Pubblicazione: (2023)
di: Singireddy, Suraj, et al.
Pubblicazione: (2023)
Stable On-Policy Distillation through Adaptive Target Reformulation
di: Jang, Ijun, et al.
Pubblicazione: (2026)
di: Jang, Ijun, et al.
Pubblicazione: (2026)
Addressing Negative Transfer in Diffusion Models
di: Go, Hyojun, et al.
Pubblicazione: (2023)
di: Go, Hyojun, et al.
Pubblicazione: (2023)
A Recipe for Stable Offline Multi-agent Reinforcement Learning
di: Lee, Dongsu, et al.
Pubblicazione: (2026)
di: Lee, Dongsu, et al.
Pubblicazione: (2026)
CADENT: Gated Hybrid Distillation for Sample-Efficient Transfer in Reinforcement Learning
di: Alinejad, Mahyar, et al.
Pubblicazione: (2026)
di: Alinejad, Mahyar, et al.
Pubblicazione: (2026)
Multi-granularity Knowledge Transfer for Continual Reinforcement Learning
di: Pan, Chaofan, et al.
Pubblicazione: (2024)
di: Pan, Chaofan, et al.
Pubblicazione: (2024)
M3: Mamba-assisted Multi-Circuit Optimization via MBRL with Effective Scheduling
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
Efficient Epistemic Uncertainty Estimation for Large Language Models via Knowledge Distillation
di: Park, Seonghyeon, et al.
Pubblicazione: (2026)
di: Park, Seonghyeon, et al.
Pubblicazione: (2026)
Overcoming Slow Decision Frequencies in Continuous Control: Model-Based Sequence Reinforcement Learning for Model-Free Control
di: Patel, Devdhar, et al.
Pubblicazione: (2024)
di: Patel, Devdhar, et al.
Pubblicazione: (2024)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
di: Li, Chunxiao, et al.
Pubblicazione: (2024)
di: Li, Chunxiao, et al.
Pubblicazione: (2024)
A Practical Two-Stage Recipe for Mathematical LLMs: Maximizing Accuracy with SFT and Efficiency with Reinforcement Learning
di: Yoshihara, Hiroshi, et al.
Pubblicazione: (2025)
di: Yoshihara, Hiroshi, et al.
Pubblicazione: (2025)
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
di: Yin, Bojian, et al.
Pubblicazione: (2024)
di: Yin, Bojian, et al.
Pubblicazione: (2024)
Learning to Act Robustly with View-Invariant Latent Actions
di: Jeong, Youngjoon, et al.
Pubblicazione: (2026)
di: Jeong, Youngjoon, et al.
Pubblicazione: (2026)
Overcoming Growth-Induced Forgetting in Task-Agnostic Continual Learning
di: Zhao, Yuqing, et al.
Pubblicazione: (2024)
di: Zhao, Yuqing, et al.
Pubblicazione: (2024)
Diffusion-Assisted Distillation for Self-Supervised Graph Representation Learning with MLPs
di: Ahn, Seong Jin, et al.
Pubblicazione: (2025)
di: Ahn, Seong Jin, et al.
Pubblicazione: (2025)
DoMIX: An Efficient Framework for Exploiting Domain Knowledge in Fine-Tuning
di: Kim, Dohoon, et al.
Pubblicazione: (2025)
di: Kim, Dohoon, et al.
Pubblicazione: (2025)
Dataset Distillation for Offline Reinforcement Learning
di: Light, Jonathan, et al.
Pubblicazione: (2024)
di: Light, Jonathan, et al.
Pubblicazione: (2024)
Reinforcement Learning via Self-Distillation
di: Hübotter, Jonas, et al.
Pubblicazione: (2026)
di: Hübotter, Jonas, et al.
Pubblicazione: (2026)
Overcoming Overfitting in Reinforcement Learning via Gaussian Process Diffusion Policy
di: Horprasert, Amornyos, et al.
Pubblicazione: (2025)
di: Horprasert, Amornyos, et al.
Pubblicazione: (2025)
Directed Structural Adaptation to Overcome Statistical Conflicts and Enable Continual Learning
di: Erden, Zeki Doruk, et al.
Pubblicazione: (2024)
di: Erden, Zeki Doruk, et al.
Pubblicazione: (2024)
Low-redundancy Distillation for Continual Learning
di: Liu, RuiQi, et al.
Pubblicazione: (2023)
di: Liu, RuiQi, et al.
Pubblicazione: (2023)
A Cross Modal Knowledge Distillation & Data Augmentation Recipe for Improving Transcriptomics Representations through Morphological Features
di: Bendidi, Ihab, et al.
Pubblicazione: (2025)
di: Bendidi, Ihab, et al.
Pubblicazione: (2025)
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
di: Li, Yaxuan, et al.
Pubblicazione: (2026)
di: Li, Yaxuan, et al.
Pubblicazione: (2026)
Contrastive Residual Energy Test-time Adaptation
di: Han, Yewon, et al.
Pubblicazione: (2025)
di: Han, Yewon, et al.
Pubblicazione: (2025)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
di: Li, Lanpei, et al.
Pubblicazione: (2024)
di: Li, Lanpei, et al.
Pubblicazione: (2024)
A Ridge Too Far: Correcting Over-Shrinkage via Negative Regularization
di: Kim, Dongseok, et al.
Pubblicazione: (2025)
di: Kim, Dongseok, et al.
Pubblicazione: (2025)
X-PEFT: eXtremely Parameter-Efficient Fine-Tuning for Extreme Multi-Profile Scenarios
di: Kwak, Namju, et al.
Pubblicazione: (2024)
di: Kwak, Namju, et al.
Pubblicazione: (2024)
Tabular Transfer Learning via Prompting LLMs
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
di: Oh, Youngmin, et al.
Pubblicazione: (2026)
di: Oh, Youngmin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Action-Sufficient Goal Representations
di: Hyeon, Jinu, et al.
Pubblicazione: (2026) -
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
di: Choi, Heewoong, et al.
Pubblicazione: (2024) -
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
di: Ahn, Hongjoon, et al.
Pubblicazione: (2025) -
Forget Forgetting: Continual Learning in a World of Abundant Memory
di: Cho, Dongkyu, et al.
Pubblicazione: (2025) -
The Power of Resets in Online Reinforcement Learning
di: Mhammedi, Zakaria, et al.
Pubblicazione: (2024)