Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement
Fuente:
arXiv
Saved in:
| Main Authors: | Kallel, Mahdi, Tölle, Johannes, Hendawy, Ahmed, D'Eramo, Carlo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023)
by: Hendawy, Ahmed, et al.
Published: (2023)
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
by: Hendawy, Ahmed, et al.
Published: (2025)
by: Hendawy, Ahmed, et al.
Published: (2025)
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024)
by: Kallel, Mahdi, et al.
Published: (2024)
Deterministic Exploration via Stationary Bellman Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2024)
by: Griesbach, Sebastian, et al.
Published: (2024)
Learning to Explore in Diverse Reward Settings via Temporal-Difference-Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2025)
by: Griesbach, Sebastian, et al.
Published: (2025)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
On the Benefit of Optimal Transport for Curriculum Reinforcement Learning
by: Klink, Pascal, et al.
Published: (2023)
by: Klink, Pascal, et al.
Published: (2023)
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
by: Farr, Noah, et al.
Published: (2026)
by: Farr, Noah, et al.
Published: (2026)
Dynamic Obstacle Avoidance with Bounded Rationality Adversarial Reinforcement Learning
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
by: Reddi, Aryaman, et al.
Published: (2025)
by: Reddi, Aryaman, et al.
Published: (2025)
Sharing Knowledge in Multi-Task Deep Reinforcement Learning
by: D'Eramo, Carlo, et al.
Published: (2024)
by: D'Eramo, Carlo, et al.
Published: (2024)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Gradient Iterated Temporal-Difference Learning
by: Vincent, Théo, et al.
Published: (2026)
by: Vincent, Théo, et al.
Published: (2026)
Domain Randomization via Entropy Maximization
by: Tiboni, Gabriele, et al.
Published: (2023)
by: Tiboni, Gabriele, et al.
Published: (2023)
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
Machine Learning with Physics Knowledge for Prediction: A Survey
by: Watson, Joe, et al.
Published: (2024)
by: Watson, Joe, et al.
Published: (2024)
Light to Heavy, Brief to Eternal: An Axion for Every Occasion (in the Early Universe)
by: D'Eramo, Francesco
Published: (2026)
by: D'Eramo, Francesco
Published: (2026)
Probing Non-Minimal Dark Sectors via the 21 cm Line at Cosmic Dawn
by: Cima, Federico, et al.
Published: (2025)
by: Cima, Federico, et al.
Published: (2025)
Back to the phase space: thermal axion dark radiation via couplings to standard model fermions
by: D'Eramo, Francesco, et al.
Published: (2024)
by: D'Eramo, Francesco, et al.
Published: (2024)
Symbolic Regression via Latent Iterative Refinement
by: Chu, Xieting, et al.
Published: (2026)
by: Chu, Xieting, et al.
Published: (2026)
Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
Continual Learning Should Move Beyond Incremental Classification
by: Mitchell, Rupert, et al.
Published: (2025)
by: Mitchell, Rupert, et al.
Published: (2025)
Axion Portal to Scalar Dark Matter: Unveiling Stabilizing Symmetry Footprints
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Reinforcement Learning via Implicit Imitation Guidance
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
From Imitation to Refinement -- Residual RL for Precise Assembly
by: Ankile, Lars, et al.
Published: (2024)
by: Ankile, Lars, et al.
Published: (2024)
Going Beyond Expert Performance via Deep Implicit Imitation Reinforcement Learning
by: Chrysomallis, Iason, et al.
Published: (2025)
by: Chrysomallis, Iason, et al.
Published: (2025)
Imitation Bootstrapped Reinforcement Learning
by: Hu, Hengyuan, et al.
Published: (2023)
by: Hu, Hengyuan, et al.
Published: (2023)
An Introduction to Deep Reinforcement and Imitation Learning
by: Santana, Pedro
Published: (2025)
by: Santana, Pedro
Published: (2025)
Neural Global Optimization via Iterative Refinement from Noisy Samples
by: Muzaffar, Qusay, et al.
Published: (2026)
by: Muzaffar, Qusay, et al.
Published: (2026)
Exploring Synthesizable Chemical Space with Iterative Pathway Refinements
by: Lee, Seul, et al.
Published: (2025)
by: Lee, Seul, et al.
Published: (2025)
Self-Refining Diffusion Samplers: Enabling Parallelization via Parareal Iterations
by: Selvam, Nikil Roashan, et al.
Published: (2024)
by: Selvam, Nikil Roashan, et al.
Published: (2024)
Cosmic-Ray Signatures of Annihilating and Semi-Annihilating Dark Matter via One-Step Cascades
by: D'Eramo, Francesco, et al.
Published: (2026)
by: D'Eramo, Francesco, et al.
Published: (2026)
Imitating Language via Scalable Inverse Reinforcement Learning
by: Wulfmeier, Markus, et al.
Published: (2024)
by: Wulfmeier, Markus, et al.
Published: (2024)
Improving Next Tokens via Second-to-Last Predictions with Generate and Refine
by: Schneider, Johannes
Published: (2024)
by: Schneider, Johannes
Published: (2024)
RILe: Reinforced Imitation Learning
by: Albaba, Mert, et al.
Published: (2024)
by: Albaba, Mert, et al.
Published: (2024)
Directly Forecasting Belief for Reinforcement Learning with Delays
by: Wu, Qingyuan, et al.
Published: (2025)
by: Wu, Qingyuan, et al.
Published: (2025)
Federated In-Context Learning: Iterative Refinement for Improved Answer Quality
by: Wang, Ruhan, et al.
Published: (2025)
by: Wang, Ruhan, et al.
Published: (2025)
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024)
by: Jelley, Adam, et al.
Published: (2024)
Similar Items
-
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023) -
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
by: Hendawy, Ahmed, et al.
Published: (2025) -
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024) -
Deterministic Exploration via Stationary Bellman Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2024) -
Learning to Explore in Diverse Reward Settings via Temporal-Difference-Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2025)