Data Augmentation for Continual RL via Adversarial Gradient Episodic Memory
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Sihao, Zhao, Xingyu, Huang, Xiaowei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continuous Geometry-Aware Graph Diffusion via Hyperbolic Neural PDE
by: Liu, Jiaxu, et al.
Published: (2024)
by: Liu, Jiaxu, et al.
Published: (2024)
Momentum Boosted Episodic Memory for Improving Learning in Long-Tailed RL Environments
by: Fernandes, Dolton, et al.
Published: (2025)
by: Fernandes, Dolton, et al.
Published: (2025)
Koopman Learning with Episodic Memory
by: Redman, William T., et al.
Published: (2023)
by: Redman, William T., et al.
Published: (2023)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026)
by: Liu, Xin, et al.
Published: (2026)
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration
by: Zhao, Yang, et al.
Published: (2026)
by: Zhao, Yang, et al.
Published: (2026)
Augmenting Online RL with Offline Data is All You Need: A Unified Hybrid RL Algorithm Design and Analysis
by: Huang, Ruiquan, et al.
Published: (2025)
by: Huang, Ruiquan, et al.
Published: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Fast Non-Episodic Finite-Horizon RL with K-Step Lookahead Thresholding
by: Xu, Jiamin, et al.
Published: (2026)
by: Xu, Jiamin, et al.
Published: (2026)
Offline RL via Feature-Occupancy Gradient Ascent
by: Neu, Gergely, et al.
Published: (2024)
by: Neu, Gergely, et al.
Published: (2024)
MGSER-SAM: Memory-Guided Soft Experience Replay with Sharpness-Aware Optimization for Enhanced Continual Learning
by: Li, Xingyu, et al.
Published: (2024)
by: Li, Xingyu, et al.
Published: (2024)
Adversarial Training for Graph Neural Networks via Graph Subspace Energy Optimization
by: Liu, Ganlin, et al.
Published: (2024)
by: Liu, Ganlin, et al.
Published: (2024)
Scaling In-Context Online Learning Capability of LLMs via Cross-Episode Meta-RL
by: Lin, Xiaofeng, et al.
Published: (2026)
by: Lin, Xiaofeng, et al.
Published: (2026)
Adapting to Stochastic and Adversarial Losses in Episodic MDPs with Aggregate Bandit Feedback
by: Ito, Shinji, et al.
Published: (2025)
by: Ito, Shinji, et al.
Published: (2025)
Enhancing Robust Fairness via Confusional Spectral Regularization
by: Jin, Gaojie, et al.
Published: (2025)
by: Jin, Gaojie, et al.
Published: (2025)
Provably Efficient RL under Episode-Wise Safety in Constrained MDPs with Linear Function Approximation
by: Kitamura, Toshinori, et al.
Published: (2025)
by: Kitamura, Toshinori, et al.
Published: (2025)
Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution
by: Qiu, Haiyun, et al.
Published: (2026)
by: Qiu, Haiyun, et al.
Published: (2026)
Rethinking Cost-sensitive Classification in Deep Learning via Adversarial Data Augmentation
by: Chen, Qiyuan, et al.
Published: (2022)
by: Chen, Qiyuan, et al.
Published: (2022)
Promoting Generalization for Exact Solvers via Adversarial Instance Augmentation
by: Liu, Haoyang, et al.
Published: (2023)
by: Liu, Haoyang, et al.
Published: (2023)
Lightweight Safety Guardrails via Synthetic Data and RL-guided Adversarial Training
by: Ilin, Aleksei, et al.
Published: (2025)
by: Ilin, Aleksei, et al.
Published: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
by: Belaire, Roman, et al.
Published: (2024)
by: Belaire, Roman, et al.
Published: (2024)
SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
by: Xia, Peng, et al.
Published: (2026)
by: Xia, Peng, et al.
Published: (2026)
Linking In-context Learning in Transformers to Human Episodic Memory
by: Ji-An, Li, et al.
Published: (2024)
by: Ji-An, Li, et al.
Published: (2024)
Larimar: Large Language Models with Episodic Memory Control
by: Das, Payel, et al.
Published: (2024)
by: Das, Payel, et al.
Published: (2024)
Planning from Imagination: Episodic Simulation and Episodic Memory for Vision-and-Language Navigation
by: Pan, Yiyuan, et al.
Published: (2024)
by: Pan, Yiyuan, et al.
Published: (2024)
Diffusion-based Episodes Augmentation for Offline Multi-Agent Reinforcement Learning
by: Oh, Jihwan, et al.
Published: (2024)
by: Oh, Jihwan, et al.
Published: (2024)
GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control
by: Xu, Haofeng, et al.
Published: (2026)
by: Xu, Haofeng, et al.
Published: (2026)
Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL
by: Wu, Ian, et al.
Published: (2026)
by: Wu, Ian, et al.
Published: (2026)
Maintaining Adversarial Robustness in Continuous Learning
by: Ru, Xiaolei, et al.
Published: (2024)
by: Ru, Xiaolei, et al.
Published: (2024)
ContextRL: Enhancing MLLM's Knowledge Discovery Efficiency with Context-Augmented RL
by: Lu, Xingyu, et al.
Published: (2026)
by: Lu, Xingyu, et al.
Published: (2026)
S2O: Enhancing Adversarial Training with Second-Order Statistics of Weights
by: Jin, Gaojie, et al.
Published: (2026)
by: Jin, Gaojie, et al.
Published: (2026)
MAD-TD: Model-Augmented Data stabilizes High Update Ratio RL
by: Voelcker, Claas A, et al.
Published: (2024)
by: Voelcker, Claas A, et al.
Published: (2024)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
by: Pink, Mathis, et al.
Published: (2024)
by: Pink, Mathis, et al.
Published: (2024)
ER-MIA: Black-Box Adversarial Memory Injection Attacks on Long-Term Memory-Augmented Large Language Models
by: Piehl, Mitchell, et al.
Published: (2026)
by: Piehl, Mitchell, et al.
Published: (2026)
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
by: Karine, Karine, et al.
Published: (2024)
by: Karine, Karine, et al.
Published: (2024)
Rethinking the Intermediate Features in Adversarial Attacks: Misleading Robotic Models via Adversarial Distillation
by: Zhao, Ke, et al.
Published: (2024)
by: Zhao, Ke, et al.
Published: (2024)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
by: Li, Sijia, et al.
Published: (2025)
by: Li, Sijia, et al.
Published: (2025)
Episodic Memory in Agentic Frameworks: Suggesting Next Tasks
by: Fiorini, Sandro Rama, et al.
Published: (2025)
by: Fiorini, Sandro Rama, et al.
Published: (2025)
VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory
by: Lei, Yuheng, et al.
Published: (2026)
by: Lei, Yuheng, et al.
Published: (2026)
FedLGA: Towards System-Heterogeneity of Federated Learning via Local Gradient Approximation
by: Li, Xingyu, et al.
Published: (2021)
by: Li, Xingyu, et al.
Published: (2021)
Similar Items
-
Continuous Geometry-Aware Graph Diffusion via Hyperbolic Neural PDE
by: Liu, Jiaxu, et al.
Published: (2024) -
Momentum Boosted Episodic Memory for Improving Learning in Long-Tailed RL Environments
by: Fernandes, Dolton, et al.
Published: (2025) -
Koopman Learning with Episodic Memory
by: Redman, William T., et al.
Published: (2023) -
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026) -
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)