On the Importance of Multistability for Horizon Generalization in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bakija, Asad, De Geeter, Florent, Brandoit, Julien, Sacré, Pierre, Drion, Guillaume |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spike-based computation using classical recurrent neural networks
von: De Geeter, Florent, et al.
Veröffentlicht: (2023)
von: De Geeter, Florent, et al.
Veröffentlicht: (2023)
Parallelizable memory recurrent units
von: De Geeter, Florent, et al.
Veröffentlicht: (2026)
von: De Geeter, Florent, et al.
Veröffentlicht: (2026)
Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications
von: Brandoit, Julien, et al.
Veröffentlicht: (2026)
von: Brandoit, Julien, et al.
Veröffentlicht: (2026)
Fast reconstruction of degenerate populations of conductance-based neuron models from spike times
von: Brandoit, Julien, et al.
Veröffentlicht: (2025)
von: Brandoit, Julien, et al.
Veröffentlicht: (2025)
Energy-Efficient Implementation of Spiking Recurrent Cells on FPGA
von: Harmeling, Pascal, et al.
Veröffentlicht: (2026)
von: Harmeling, Pascal, et al.
Veröffentlicht: (2026)
Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations
von: Fyon, Arthur, et al.
Veröffentlicht: (2026)
von: Fyon, Arthur, et al.
Veröffentlicht: (2026)
Context-dependent manifold learning: A neuromodulated constrained autoencoder approach
von: Adriaens, Jérôme, et al.
Veröffentlicht: (2026)
von: Adriaens, Jérôme, et al.
Veröffentlicht: (2026)
Inward rectifier potassium channels interact with calcium channels to promote robust and physiological bistability
von: De Worm, Anaëlle, et al.
Veröffentlicht: (2024)
von: De Worm, Anaëlle, et al.
Veröffentlicht: (2024)
Neuromodulation supports robust rhythmic pattern transitions in degenerate central pattern generators with fixed connectivity
von: Fyon, Arthur, et al.
Veröffentlicht: (2026)
von: Fyon, Arthur, et al.
Veröffentlicht: (2026)
Dimensionality reduction of neuronal degeneracy reveals two interfering physiological mechanisms
von: Fyon, Arthur, et al.
Veröffentlicht: (2024)
von: Fyon, Arthur, et al.
Veröffentlicht: (2024)
Horizon Generalization in Reinforcement Learning
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism
von: Ni, Tianwei, et al.
Veröffentlicht: (2025)
von: Ni, Tianwei, et al.
Veröffentlicht: (2025)
Online Learning and Information Exponents: On The Importance of Batch size, and Time/Complexity Tradeoffs
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Infinite Horizon Mean Field Problems in Continuous Spaces
von: Angiuli, Andrea, et al.
Veröffentlicht: (2023)
von: Angiuli, Andrea, et al.
Veröffentlicht: (2023)
Stochastic Decision Horizons for Constrained Reinforcement Learning
von: Milosevic, Nikola, et al.
Veröffentlicht: (2026)
von: Milosevic, Nikola, et al.
Veröffentlicht: (2026)
Inverse Reinforcement Learning with Multiple Planning Horizons
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
FedSPDnet: Geometry-Aware Federated Deep Learning with SPDnet
von: Pautrel, Thibault, et al.
Veröffentlicht: (2026)
von: Pautrel, Thibault, et al.
Veröffentlicht: (2026)
Effective Reward Specification in Deep Reinforcement Learning
von: Roy, Julien
Veröffentlicht: (2024)
von: Roy, Julien
Veröffentlicht: (2024)
On the Effective Horizon of Inverse Reinforcement Learning
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
Horizon Reduction as Information Loss in Offline Reinforcement Learning
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning
von: Zhao, Hanye, et al.
Veröffentlicht: (2024)
von: Zhao, Hanye, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Random Time Horizons
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with Universal Horizon Models
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
Statistical Inference of the Value Function for Reinforcement Learning in Infinite Horizon Settings
von: Shi, C., et al.
Veröffentlicht: (2020)
von: Shi, C., et al.
Veröffentlicht: (2020)
Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks
von: Yu, Yang
Veröffentlicht: (2025)
von: Yu, Yang
Veröffentlicht: (2025)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
von: Chen, Kevin, et al.
Veröffentlicht: (2025)
von: Chen, Kevin, et al.
Veröffentlicht: (2025)
Test Where Decisions Matter: Importance-driven Testing for Deep Reinforcement Learning
von: Pranger, Stefan, et al.
Veröffentlicht: (2024)
von: Pranger, Stefan, et al.
Veröffentlicht: (2024)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Beyond Distribution Sharpening: The Importance of Task Rewards
von: Mittal, Sarthak, et al.
Veröffentlicht: (2026)
von: Mittal, Sarthak, et al.
Veröffentlicht: (2026)
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation
von: Park, Jaehyun, et al.
Veröffentlicht: (2024)
von: Park, Jaehyun, et al.
Veröffentlicht: (2024)
Policy Zooming: Adaptive Discretization-based Infinite-Horizon Average-Reward Reinforcement Learning
von: Kar, Avik, et al.
Veröffentlicht: (2024)
von: Kar, Avik, et al.
Veröffentlicht: (2024)
Hybrid Training for Enhanced Multi-task Generalization in Multi-agent Reinforcement Learning
von: Zhang, Mingliang, et al.
Veröffentlicht: (2024)
von: Zhang, Mingliang, et al.
Veröffentlicht: (2024)
Model-Free versus Model-Based Reinforcement Learning for Fixed-Wing UAV Attitude Control Under Varying Wind Conditions
von: Olivares, David, et al.
Veröffentlicht: (2024)
von: Olivares, David, et al.
Veröffentlicht: (2024)
Know your Trajectory -- Trustworthy Reinforcement Learning deployment through Importance-Based Trajectory Analysis
von: F, Clifford, et al.
Veröffentlicht: (2025)
von: F, Clifford, et al.
Veröffentlicht: (2025)
Maximum Entropy Semi-Supervised Inverse Reinforcement Learning
von: Audiffren, Julien, et al.
Veröffentlicht: (2026)
von: Audiffren, Julien, et al.
Veröffentlicht: (2026)
Multistability of Self-Attention Dynamics in Transformers
von: Altafini, Claudio
Veröffentlicht: (2025)
von: Altafini, Claudio
Veröffentlicht: (2025)
A Zero-Shot Reinforcement Learning Strategy for Autonomous Guidewire Navigation
von: Scarponi, Valentina, et al.
Veröffentlicht: (2024)
von: Scarponi, Valentina, et al.
Veröffentlicht: (2024)
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation
von: Dai, Juntao, et al.
Veröffentlicht: (2024)
von: Dai, Juntao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Spike-based computation using classical recurrent neural networks
von: De Geeter, Florent, et al.
Veröffentlicht: (2023) -
Parallelizable memory recurrent units
von: De Geeter, Florent, et al.
Veröffentlicht: (2026) -
Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications
von: Brandoit, Julien, et al.
Veröffentlicht: (2026) -
Fast reconstruction of degenerate populations of conductance-based neuron models from spike times
von: Brandoit, Julien, et al.
Veröffentlicht: (2025) -
Energy-Efficient Implementation of Spiking Recurrent Cells on FPGA
von: Harmeling, Pascal, et al.
Veröffentlicht: (2026)