Momentum Boosted Episodic Memory for Improving Learning in Long-Tailed RL Environments
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Fernandes, Dolton, Kaushik, Pramod, Shukla, Harsh, Surampudi, Bapi Raju |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
HyperGALE: ASD Classification via Hypergraph Gated Attention with Learnable Hyperedges
par: Arora, Mehul, et autres
Publié: (2024)
par: Arora, Mehul, et autres
Publié: (2024)
Transparency in Sleep Staging: Deep Learning Method for EEG Sleep Stage Classification with Model Interpretability
par: Sharma, Shivam, et autres
Publié: (2023)
par: Sharma, Shivam, et autres
Publié: (2023)
Taming the Tail: NoI Topology Synthesis for Mixed DL Workloads on Chiplet-Based Accelerators
par: Shukla, Arnav, et autres
Publié: (2025)
par: Shukla, Arnav, et autres
Publié: (2025)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
par: Bhatia, Abhinav, et autres
Publié: (2023)
par: Bhatia, Abhinav, et autres
Publié: (2023)
Class Confidence Aware Reweighting for Long Tailed Learning
par: Jagati, Brainard Philemon, et autres
Publié: (2026)
par: Jagati, Brainard Philemon, et autres
Publié: (2026)
Painless Federated Learning: An Interplay of Line-Search and Extrapolation
par: Geetika, et autres
Publié: (2024)
par: Geetika, et autres
Publié: (2024)
AI Generalisation Gap In Comorbid Sleep Disorder Staging
par: Bose, Saswata, et autres
Publié: (2026)
par: Bose, Saswata, et autres
Publié: (2026)
Larimar: Large Language Models with Episodic Memory Control
par: Das, Payel, et autres
Publié: (2024)
par: Das, Payel, et autres
Publié: (2024)
Dynamic Vocabulary Pruning: Stable LLM-RL by Taming the Tail
par: Li, Yingru, et autres
Publié: (2025)
par: Li, Yingru, et autres
Publié: (2025)
Scaling In-Context Online Learning Capability of LLMs via Cross-Episode Meta-RL
par: Lin, Xiaofeng, et autres
Publié: (2026)
par: Lin, Xiaofeng, et autres
Publié: (2026)
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
par: Karine, Karine, et autres
Publié: (2024)
par: Karine, Karine, et autres
Publié: (2024)
Taming the Long-Tail: Efficient Reasoning RL Training with Adaptive Drafter
par: Hu, Qinghao, et autres
Publié: (2025)
par: Hu, Qinghao, et autres
Publié: (2025)
TrueGradeAI: Retrieval-Augmented and Bias-Resistant AI for Transparent and Explainable Digital Assessments
par: Thakur, Rakesh, et autres
Publié: (2025)
par: Thakur, Rakesh, et autres
Publié: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
par: Cherepanov, Egor, et autres
Publié: (2025)
par: Cherepanov, Egor, et autres
Publié: (2025)
VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory
par: Lei, Yuheng, et autres
Publié: (2026)
par: Lei, Yuheng, et autres
Publié: (2026)
Federated Instrumental Variable Analysis via Federated Generalized Method of Moments
par: Geetika, et autres
Publié: (2025)
par: Geetika, et autres
Publié: (2025)
Heterogeneous Learning Rate Scheduling for Neural Architecture Search on Long-Tailed Datasets
par: Tang, Chenxia
Publié: (2024)
par: Tang, Chenxia
Publié: (2024)
Decoupled Federated Learning on Long-Tailed and Non-IID data with Feature Statistics
par: Chen, Zhuoxin, et autres
Publié: (2024)
par: Chen, Zhuoxin, et autres
Publié: (2024)
Episodic Memory in Agentic Frameworks: Suggesting Next Tasks
par: Fiorini, Sandro Rama, et autres
Publié: (2025)
par: Fiorini, Sandro Rama, et autres
Publié: (2025)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
par: Li, Sijia, et autres
Publié: (2025)
par: Li, Sijia, et autres
Publié: (2025)
SMMF: Square-Matricized Momentum Factorization for Memory-Efficient Optimization
par: Park, Kwangryeol, et autres
Publié: (2024)
par: Park, Kwangryeol, et autres
Publié: (2024)
SOAP-RL: Sequential Option Advantage Propagation for Reinforcement Learning in POMDP Environments
par: Ishida, Shu, et autres
Publié: (2024)
par: Ishida, Shu, et autres
Publié: (2024)
SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning
par: Wang, Jichao, et autres
Publié: (2026)
par: Wang, Jichao, et autres
Publié: (2026)
Towards Improving Long-Tail Entity Predictions in Temporal Knowledge Graphs through Global Similarity and Weighted Sampling
par: Mirtaheri, Mehrnoosh, et autres
Publié: (2025)
par: Mirtaheri, Mehrnoosh, et autres
Publié: (2025)
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
par: Rangwani, Harsh, et autres
Publié: (2024)
par: Rangwani, Harsh, et autres
Publié: (2024)
Muon Outperforms Adam in Tail-End Associative Memory Learning
par: Wang, Shuche, et autres
Publié: (2025)
par: Wang, Shuche, et autres
Publié: (2025)
Improving Long-Tailed Object Detection with Balanced Group Softmax and Metric Learning
par: Gaba, Satyam
Publié: (2025)
par: Gaba, Satyam
Publié: (2025)
Space Alignment Matters: The Missing Piece for Inducing Neural Collapse in Long-Tailed Learning
par: Wang, Jinping, et autres
Publié: (2025)
par: Wang, Jinping, et autres
Publié: (2025)
Episodic Memories Generation and Evaluation Benchmark for Large Language Models
par: Huet, Alexis, et autres
Publié: (2025)
par: Huet, Alexis, et autres
Publié: (2025)
Echo: A Large Language Model with Temporal Episodic Memory
par: Liu, WenTao, et autres
Publié: (2025)
par: Liu, WenTao, et autres
Publié: (2025)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
par: Pink, Mathis, et autres
Publié: (2024)
par: Pink, Mathis, et autres
Publié: (2024)
Episodic Reinforcement Learning with Expanded State-reward Space
par: Liang, Dayang, et autres
Publié: (2024)
par: Liang, Dayang, et autres
Publié: (2024)
Episodic-free Task Selection for Few-shot Learning
par: Zhang, Tao
Publié: (2024)
par: Zhang, Tao
Publié: (2024)
Curiosity & Entropy Driven Unsupervised RL in Multiple Environments
par: Dewan, Shaurya, et autres
Publié: (2024)
par: Dewan, Shaurya, et autres
Publié: (2024)
BubbleSpec: Turning Long-Tail Bubbles into Speculative Rollout Drafts for Synchronous Reinforcement Learning
par: Xu, Yuhang, et autres
Publié: (2026)
par: Xu, Yuhang, et autres
Publié: (2026)
Ensemble of Pre-Trained Models for Long-Tailed Trajectory Prediction
par: Thuremella, Divya, et autres
Publié: (2025)
par: Thuremella, Divya, et autres
Publié: (2025)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
par: Muslimani, Calarina, et autres
Publié: (2025)
par: Muslimani, Calarina, et autres
Publié: (2025)
Logarithmic Memory Networks (LMNs): Efficient Long-Range Sequence Modeling for Resource-Constrained Environments
par: Taha, Mohamed A.
Publié: (2025)
par: Taha, Mohamed A.
Publié: (2025)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
par: Cherepanov, Egor, et autres
Publié: (2024)
par: Cherepanov, Egor, et autres
Publié: (2024)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
par: Schmied, Thomas, et autres
Publié: (2024)
par: Schmied, Thomas, et autres
Publié: (2024)
Documents similaires
-
HyperGALE: ASD Classification via Hypergraph Gated Attention with Learnable Hyperedges
par: Arora, Mehul, et autres
Publié: (2024) -
Transparency in Sleep Staging: Deep Learning Method for EEG Sleep Stage Classification with Model Interpretability
par: Sharma, Shivam, et autres
Publié: (2023) -
Taming the Tail: NoI Topology Synthesis for Mixed DL Workloads on Chiplet-Based Accelerators
par: Shukla, Arnav, et autres
Publié: (2025) -
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
par: Bhatia, Abhinav, et autres
Publié: (2023) -
Class Confidence Aware Reweighting for Long Tailed Learning
par: Jagati, Brainard Philemon, et autres
Publié: (2026)