Model-Free Active Exploration in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Russo, Alessio, Proutiere, Alexandre |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2026)
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2026)
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024)
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024)
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
von: Dubail, Bastien, et al.
Veröffentlicht: (2025)
von: Dubail, Bastien, et al.
Veröffentlicht: (2025)
Optimal Centered Active Excitation in Linear System Identification
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
Adaptive Reinforcement Learning for Unobservable Random Delays
von: Wikman, John, et al.
Veröffentlicht: (2025)
von: Wikman, John, et al.
Veröffentlicht: (2025)
Curvature-Guided LoRA: Steering in the pretrained NTK subspace
von: Zheng, Frédéric, et al.
Veröffentlicht: (2026)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2026)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2026)
von: Foffano, Daniele, et al.
Veröffentlicht: (2026)
In-Context Learning for Pure Exploration
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
Conformal Predictions under Markovian Data
von: Zheng, Frédéric, et al.
Veröffentlicht: (2024)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2024)
Pure Exploration with Feedback Graphs
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
Adaptive Exploration for Multi-Reward Multi-Policy Evaluation
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
On Universally Optimal Algorithms for A/B Testing
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
In-Context Learning for Pure Exploration in Continuous Spaces
von: Russo, Alessio, et al.
Veröffentlicht: (2026)
von: Russo, Alessio, et al.
Veröffentlicht: (2026)
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
Receding-Horizon Control via Drifting Models
von: Foffano, Daniele, et al.
Veröffentlicht: (2026)
von: Foffano, Daniele, et al.
Veröffentlicht: (2026)
Near-optimal Rank Adaptive Inference of High Dimensional Matrices
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
Adaptive Active Learning for Regression via Reinforcement Learning
von: Nguyen, Simon D., et al.
Veröffentlicht: (2026)
von: Nguyen, Simon D., et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning and Sequence Modeling for Downlink Link Adaptation
von: Peri, Samuele, et al.
Veröffentlicht: (2024)
von: Peri, Samuele, et al.
Veröffentlicht: (2024)
Low-Rank Bandits via Tight Two-to-Infinity Singular Subspace Recovery
von: Jedra, Yassir, et al.
Veröffentlicht: (2024)
von: Jedra, Yassir, et al.
Veröffentlicht: (2024)
Minimizing Human Intervention in Online Classification
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
Active Exploration in Bayesian Model-based Reinforcement Learning for Robot Manipulation
von: Plou, Carlos, et al.
Veröffentlicht: (2024)
von: Plou, Carlos, et al.
Veröffentlicht: (2024)
Active Exploration via Autoregressive Generation of Missing Data
von: Cai, Tiffany Tianhui, et al.
Veröffentlicht: (2024)
von: Cai, Tiffany Tianhui, et al.
Veröffentlicht: (2024)
Optimal Clustering from Noisy Binary Feedback
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
Explainable Reinforcement Learning via Temporal Policy Decomposition
von: Ruggeri, Franco, et al.
Veröffentlicht: (2025)
von: Ruggeri, Franco, et al.
Veröffentlicht: (2025)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Policy Testing in Markov Decision Processes
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
Efficient Learning of POMDPs with Known Observation Model in Average-Reward Setting
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
Near-Optimal Clustering in Mixture of Markov Chains
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
On Efficient Bayesian Exploration in Model-Based Reinforcement Learning
von: Caron, Alberto, et al.
Veröffentlicht: (2025)
von: Caron, Alberto, et al.
Veröffentlicht: (2025)
Learning-Driven Exploration for Reinforcement Learning
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
Community-based Multi-Agent Reinforcement Learning with Transfer and Active Exploration
von: Shi, Zhaoyang
Veröffentlicht: (2025)
von: Shi, Zhaoyang
Veröffentlicht: (2025)
Achieving $\widetilde{\mathcal{O}}(\sqrt{T})$ Regret in Average-Reward POMDPs with Known Observation Models
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
Predictive Representations for Skill Transfer in Reinforcement Learning
von: Vereecken, Ruben, et al.
Veröffentlicht: (2026)
von: Vereecken, Ruben, et al.
Veröffentlicht: (2026)
Fair Best Arm Identification with Fixed Confidence
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
Search Inspired Exploration in Reinforcement Learning
von: Sotirchos, Georgios, et al.
Veröffentlicht: (2026)
von: Sotirchos, Georgios, et al.
Veröffentlicht: (2026)
Offline Model-Based Reinforcement Learning with Anti-Exploration
von: Srinivasan, Padmanaba, et al.
Veröffentlicht: (2024)
von: Srinivasan, Padmanaba, et al.
Veröffentlicht: (2024)
Random Latent Exploration for Deep Reinforcement Learning
von: Mahankali, Srinath, et al.
Veröffentlicht: (2024)
von: Mahankali, Srinath, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025) -
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023) -
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2026) -
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024) -
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
von: Dubail, Bastien, et al.
Veröffentlicht: (2025)