Squeezing More from the Stream : Learning Representation Online for Streaming Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nilaksh, Clavaud, Antoine, Reymond, Mathieu, Rivest, François, Chandar, Sarath |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CrystalGym: A New Benchmark for Materials Discovery Using Reinforcement Learning
von: Govindarajan, Prashant, et al.
Veröffentlicht: (2025)
von: Govindarajan, Prashant, et al.
Veröffentlicht: (2025)
GRPO-$λ$: Credit Assignment improves LLM Reasoning
von: Parthasarathi, Prasanna, et al.
Veröffentlicht: (2025)
von: Parthasarathi, Prasanna, et al.
Veröffentlicht: (2025)
Revisiting Adam for Streaming Reinforcement Learning
von: Gogianu, Florin, et al.
Veröffentlicht: (2026)
von: Gogianu, Florin, et al.
Veröffentlicht: (2026)
Intentional Updates for Streaming Reinforcement Learning
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2026)
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2026)
Streaming Deep Reinforcement Learning Finally Works
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
Towards Practical Tool Usage for Continually Learning LLMs
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Should We Attend More or Less? Modulating Attention for Fairness
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2023)
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
Neural Coherence : Find higher performance to out-of-distribution tasks from few samples
von: Guiroy, Simon, et al.
Veröffentlicht: (2025)
von: Guiroy, Simon, et al.
Veröffentlicht: (2025)
Towards Batch-to-Streaming Deep Reinforcement Learning for Continuous Control
von: De Monte, Riccardo, et al.
Veröffentlicht: (2026)
von: De Monte, Riccardo, et al.
Veröffentlicht: (2026)
Intelligent Switching for Reset-Free RL
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
Context-Aware Assistant Selection for Improved Inference Acceleration with Large Language Models
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Lookbehind-SAM: k steps back, 1 step forward
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
Large Language Model Integration with Reinforcement Learning to Augment Decision-Making in Autonomous Cyber Operations
von: Tholl, Konur, et al.
Veröffentlicht: (2025)
von: Tholl, Konur, et al.
Veröffentlicht: (2025)
Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models
von: Nilaksh, et al.
Veröffentlicht: (2026)
von: Nilaksh, et al.
Veröffentlicht: (2026)
Stackelberg Coupling of Online Representation Learning and Reinforcement Learning
von: Martinez, Fernando, et al.
Veröffentlicht: (2025)
von: Martinez, Fernando, et al.
Veröffentlicht: (2025)
Structured Reinforcement Learning for Media Streaming at the Wireless Edge
von: Bura, Archana, et al.
Veröffentlicht: (2024)
von: Bura, Archana, et al.
Veröffentlicht: (2024)
StreamFP: Learnable Fingerprint-guided Data Selection for Efficient Stream Learning
von: Shi, Tongjun, et al.
Veröffentlicht: (2024)
von: Shi, Tongjun, et al.
Veröffentlicht: (2024)
Extension OL-MDISF: Online Learning from Mix-Typed, Drifted, and Incomplete Streaming Features
von: Zhuo, Shengda, et al.
Veröffentlicht: (2025)
von: Zhuo, Shengda, et al.
Veröffentlicht: (2025)
Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming
von: He, Zhiqiang, et al.
Veröffentlicht: (2025)
von: He, Zhiqiang, et al.
Veröffentlicht: (2025)
CoPeP: Benchmarking Continual Pretraining for Protein Language Models
von: Patil, Darshan, et al.
Veröffentlicht: (2026)
von: Patil, Darshan, et al.
Veröffentlicht: (2026)
Learning from Streaming Data when Users Choose
von: Su, Jinyan, et al.
Veröffentlicht: (2024)
von: Su, Jinyan, et al.
Veröffentlicht: (2024)
Effect of Document Packing on the Latent Multi-Hop Reasoning Capabilities of Large Language Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
Near-Optimal Online Deployment and Routing for Streaming LLMs
von: Li, Shaoang, et al.
Veröffentlicht: (2025)
von: Li, Shaoang, et al.
Veröffentlicht: (2025)
Steering Large Language Model Activations in Sparse Spaces
von: Bayat, Reza, et al.
Veröffentlicht: (2025)
von: Bayat, Reza, et al.
Veröffentlicht: (2025)
Shielded Controller Units for RL with Operational Constraints Applied to Remote Microgrids
von: Nekoei, Hadi, et al.
Veröffentlicht: (2025)
von: Nekoei, Hadi, et al.
Veröffentlicht: (2025)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
von: Wang, Changhong, et al.
Veröffentlicht: (2024)
von: Wang, Changhong, et al.
Veröffentlicht: (2024)
Autonomous Drift Learning in Data Streams: A Unified Perspective
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
In-context Learning of Evolving Data Streams with Tabular Foundational Models
von: Lourenço, Afonso, et al.
Veröffentlicht: (2025)
von: Lourenço, Afonso, et al.
Veröffentlicht: (2025)
Multi-Label Transfer Learning in Non-Stationary Data Streams
von: Du, Honghui, et al.
Veröffentlicht: (2025)
von: Du, Honghui, et al.
Veröffentlicht: (2025)
Online Sparse Feature Selection in Data Streams via Differential Evolution
von: Xu, Ruiyang
Veröffentlicht: (2025)
von: Xu, Ruiyang
Veröffentlicht: (2025)
PROL : Rehearsal Free Continual Learning in Streaming Data via Prompt Online Learning
von: Ma'sum, M. Anwar, et al.
Veröffentlicht: (2025)
von: Ma'sum, M. Anwar, et al.
Veröffentlicht: (2025)
StreamEnsemble: Predictive Queries over Spatiotemporal Streaming Data
von: Chaves, Anderson, et al.
Veröffentlicht: (2024)
von: Chaves, Anderson, et al.
Veröffentlicht: (2024)
Continuous Fair SMOTE -- Fairness-Aware Stream Learning from Imbalanced Data
von: Lammers, Kathrin, et al.
Veröffentlicht: (2025)
von: Lammers, Kathrin, et al.
Veröffentlicht: (2025)
Why Don't Prompt-Based Fairness Metrics Correlate?
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2024)
D-SPEAR: Dual-Stream Prioritized Experience Adaptive Replay for Stable Reinforcement Learning in Robotic Manipulation
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
Do Large Language Models Know How Much They Know?
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Adaptive Hoeffding Tree with Transfer Learning for Streaming Synchrophasor Data Sets
von: Mrabet, Zakaria El, et al.
Veröffentlicht: (2025)
von: Mrabet, Zakaria El, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CrystalGym: A New Benchmark for Materials Discovery Using Reinforcement Learning
von: Govindarajan, Prashant, et al.
Veröffentlicht: (2025) -
GRPO-$λ$: Credit Assignment improves LLM Reasoning
von: Parthasarathi, Prasanna, et al.
Veröffentlicht: (2025) -
Revisiting Adam for Streaming Reinforcement Learning
von: Gogianu, Florin, et al.
Veröffentlicht: (2026) -
Intentional Updates for Streaming Reinforcement Learning
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2026) -
Streaming Deep Reinforcement Learning Finally Works
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)