Upside-Down Reinforcement Learning for More Interpretable Optimal Control
Fuente:
arXiv
Saved in:
| Main Authors: | Cardenas-Cartagena, Juan, Falzari, Massimiliano, Zullich, Marco, Sabatelli, Matthia |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025)
by: Falzari, Massimiliano, et al.
Published: (2025)
Sparsity-Driven Plasticity in Multi-Task Reinforcement Learning
by: Todorov, Aleksandar, et al.
Published: (2025)
by: Todorov, Aleksandar, et al.
Published: (2025)
On the Generalisation of Koopman Representations for Chaotic System Control
by: Hjikakou, Kyriakos, et al.
Published: (2025)
by: Hjikakou, Kyriakos, et al.
Published: (2025)
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025)
by: Veselý, Viktor, et al.
Published: (2025)
Smaller Batches, Bigger Gains? Investigating the Impact of Batch Sizes on Reinforcement Learning Based Real-World Production Scheduling
by: Müller, Arthur, et al.
Published: (2024)
by: Müller, Arthur, et al.
Published: (2024)
$ε$-Optimally Solving Zero-Sum POSGs
by: Escudie, Erwan, et al.
Published: (2024)
by: Escudie, Erwan, et al.
Published: (2024)
Large-image Object Detection for Fine-grained Recognition of Punches Patterns in Medieval Panel Painting
by: Bruegger, Josh, et al.
Published: (2025)
by: Bruegger, Josh, et al.
Published: (2025)
VDSC: Enhancing Exploration Timing with Value Discrepancy and State Counts
by: Captari, Marius, et al.
Published: (2024)
by: Captari, Marius, et al.
Published: (2024)
Video-Driven Graph Network-Based Simulators
by: Szewczyk, Franciszek, et al.
Published: (2024)
by: Szewczyk, Franciszek, et al.
Published: (2024)
Upside Down Reinforcement Learning with Policy Generators
by: Di Ventura, Jacopo, et al.
Published: (2025)
by: Di Ventura, Jacopo, et al.
Published: (2025)
Measuring Orthogonality as the Blind-Spot of Uncertainty Disentanglement
by: de Jong, Ivo Pascal, et al.
Published: (2024)
by: de Jong, Ivo Pascal, et al.
Published: (2024)
Scaling Multimodal Search and Recommendation with Small Language Models via Upside-Down Reinforcement Learning
by: Lin, Yu-Chen, et al.
Published: (2025)
by: Lin, Yu-Chen, et al.
Published: (2025)
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation
by: Tashu, Tsegaye Misikir, et al.
Published: (2024)
by: Tashu, Tsegaye Misikir, et al.
Published: (2024)
Forecasting Smog Clouds With Deep Learning
by: Oldenburg, Valentijn, et al.
Published: (2024)
by: Oldenburg, Valentijn, et al.
Published: (2024)
On the Convergence and Stability of Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning, and Online Decision Transformers
by: Štrupl, Miroslav, et al.
Published: (2025)
by: Štrupl, Miroslav, et al.
Published: (2025)
Can Bayesian Neural Networks Explicitly Model Input Uncertainty?
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
Deep Gaussian Process Proximal Policy Optimization
by: van der Lende, Matthijs, et al.
Published: (2025)
by: van der Lende, Matthijs, et al.
Published: (2025)
An $ε$-Optimal Sequential Approach for Solving zs-POSGs
by: Escudie, Erwan C., et al.
Published: (2026)
by: Escudie, Erwan C., et al.
Published: (2026)
Uncertainty in Semantic Language Modeling with PIXELS
by: Radu, Stefania, et al.
Published: (2025)
by: Radu, Stefania, et al.
Published: (2025)
Turn That Frown Upside Down: FaceID Customization via Cross-Training Data
by: Wang, Shuhe, et al.
Published: (2025)
by: Wang, Shuhe, et al.
Published: (2025)
Unified Uncertainties: Combining Input, Data and Model Uncertainty into a Single Formulation
by: Valdenegro-Toro, Matias, et al.
Published: (2024)
by: Valdenegro-Toro, Matias, et al.
Published: (2024)
No Single Metric Tells the Whole Story: A Multi-Dimensional Evaluation Framework for Uncertainty Attributions
by: Schiller, Emily, et al.
Published: (2026)
by: Schiller, Emily, et al.
Published: (2026)
Operator World Models for Reinforcement Learning
by: Novelli, Pietro, et al.
Published: (2024)
by: Novelli, Pietro, et al.
Published: (2024)
On the Properties of Feature Attribution for Supervised Contrastive Learning
by: Arrighi, Leonardo, et al.
Published: (2026)
by: Arrighi, Leonardo, et al.
Published: (2026)
Memory Allocation in Resource-Constrained Reinforcement Learning
by: Tamborski, Massimiliano, et al.
Published: (2025)
by: Tamborski, Massimiliano, et al.
Published: (2025)
ε-Optimally Solving Two-Player Zero-Sum POSGs
by: Escudie, Erwan Christian, et al.
Published: (2025)
by: Escudie, Erwan Christian, et al.
Published: (2025)
A Reinforcement Learning Approach for Optimal Control in Microgrids
by: Salaorni, Davide, et al.
Published: (2025)
by: Salaorni, Davide, et al.
Published: (2025)
Interpretable Learning Dynamics in Unsupervised Reinforcement Learning
by: Pandey, Shashwat
Published: (2025)
by: Pandey, Shashwat
Published: (2025)
Mechanistic Interpretability of Reinforcement Learning Agents
by: Trim, Tristan, et al.
Published: (2024)
by: Trim, Tristan, et al.
Published: (2024)
Interpreting Reinforcement Learning Agents with Susceptibilities
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
Actively Learning Reinforcement Learning: A Stochastic Optimal Control Approach
by: Ramadan, Mohammad S., et al.
Published: (2023)
by: Ramadan, Mohammad S., et al.
Published: (2023)
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
by: Brouwer, Eric, et al.
Published: (2024)
by: Brouwer, Eric, et al.
Published: (2024)
Unifying Model Predictive Path Integral Control, Reinforcement Learning, and Diffusion Models for Optimal Control and Planning
by: Li, Yankai, et al.
Published: (2025)
by: Li, Yankai, et al.
Published: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
Optimizing Operation Recipes with Reinforcement Learning for Safe and Interpretable Control of Chemical Processes
by: Brandner, Dean, et al.
Published: (2025)
by: Brandner, Dean, et al.
Published: (2025)
Attributions All the Way Down? The Metagame of Interpretability
by: Baniecki, Hubert, et al.
Published: (2026)
by: Baniecki, Hubert, et al.
Published: (2026)
Beyond the Neural Fog: Interpretable Learning for AC Optimal Power Flow
by: Pineda, Salvador, et al.
Published: (2024)
by: Pineda, Salvador, et al.
Published: (2024)
More Benefits of Being Distributional: Second-Order Bounds for Reinforcement Learning
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
Optimizing Interpretable Decision Tree Policies for Reinforcement Learning
by: Vos, Daniël, et al.
Published: (2024)
by: Vos, Daniël, et al.
Published: (2024)
Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies
by: Gross, Dennis, et al.
Published: (2024)
by: Gross, Dennis, et al.
Published: (2024)
Similar Items
-
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025) -
Sparsity-Driven Plasticity in Multi-Task Reinforcement Learning
by: Todorov, Aleksandar, et al.
Published: (2025) -
On the Generalisation of Koopman Representations for Chaotic System Control
by: Hjikakou, Kyriakos, et al.
Published: (2025) -
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025) -
Smaller Batches, Bigger Gains? Investigating the Impact of Batch Sizes on Reinforcement Learning Based Real-World Production Scheduling
by: Müller, Arthur, et al.
Published: (2024)