Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Leeftink, David, Hinne, Max, van Gerven, Marcel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Optimal Control of Probabilistic Dynamics Models via Mean Hamiltonian Minimization
por: Leeftink, David, et al.
Publicado: (2025)
por: Leeftink, David, et al.
Publicado: (2025)
Automated Discovery of Laser Dicing Processes with Bayesian Optimization for Semiconductor Manufacturing
por: Leeftink, David, et al.
Publicado: (2025)
por: Leeftink, David, et al.
Publicado: (2025)
Robust Inference of Dynamic Covariance Using Wishart Processes and Sequential Monte Carlo
por: Huijsdens, Hester, et al.
Publicado: (2024)
por: Huijsdens, Hester, et al.
Publicado: (2024)
Gradient-Free Training of Recurrent Neural Networks using Random Perturbations
por: Fernandez, Jesus Garcia, et al.
Publicado: (2024)
por: Fernandez, Jesus Garcia, et al.
Publicado: (2024)
Energy-Efficient Spiking Recurrent Neural Network for Gesture Recognition on Embedded GPUs
por: Varposhti, Marzieh Hassanshahi, et al.
Publicado: (2024)
por: Varposhti, Marzieh Hassanshahi, et al.
Publicado: (2024)
Unraveling the Hidden Dynamical Structure in Recurrent Neural Policies
por: Li, Jin, et al.
Publicado: (2026)
por: Li, Jin, et al.
Publicado: (2026)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization
por: Küçükoğlu, Burcu, et al.
Publicado: (2022)
por: Küçükoğlu, Burcu, et al.
Publicado: (2022)
A Unified Perspective on Optimization in Machine Learning and Neuroscience: From Gradient Descent to Neural Adaptation
por: Fernández, Jesús García, et al.
Publicado: (2025)
por: Fernández, Jesús García, et al.
Publicado: (2025)
Node Perturbation Can Effectively Train Multi-Layer Neural Networks
por: Dalm, Sander, et al.
Publicado: (2023)
por: Dalm, Sander, et al.
Publicado: (2023)
Generative Modeling of Neural Dynamics via Latent Stochastic Differential Equations
por: ElGazzar, Ahmed, et al.
Publicado: (2024)
por: ElGazzar, Ahmed, et al.
Publicado: (2024)
Spiking Neural Networks for Continuous Control via End-to-End Model-Based Learning
por: Huebotter, Justus, et al.
Publicado: (2025)
por: Huebotter, Justus, et al.
Publicado: (2025)
Discovering Continuous-Time Memory-Based Symbolic Policies using Genetic Programming
por: de Vries, Sigur, et al.
Publicado: (2024)
por: de Vries, Sigur, et al.
Publicado: (2024)
Probabilistic Forecasting via Autoregressive Flow Matching
por: ElGazzar, Ahmed, et al.
Publicado: (2025)
por: ElGazzar, Ahmed, et al.
Publicado: (2025)
Ornstein-Uhlenbeck Adaptation as a Mechanism for Learning in Brains and Machines
por: Fernandez, Jesus Garcia, et al.
Publicado: (2024)
por: Fernandez, Jesus Garcia, et al.
Publicado: (2024)
Efficient Deep Learning with Decorrelated Backpropagation
por: Dalm, Sander, et al.
Publicado: (2024)
por: Dalm, Sander, et al.
Publicado: (2024)
Fractional Order Distributed Optimization
por: Lixandru, Andrei, et al.
Publicado: (2024)
por: Lixandru, Andrei, et al.
Publicado: (2024)
Probabilistic Prediction of Neural Dynamics via Autoregressive Flow Matching
por: Rogalla, Nicole, et al.
Publicado: (2026)
por: Rogalla, Nicole, et al.
Publicado: (2026)
Generative Modeling of Clinical Time Series via Latent Stochastic Differential Equations
por: Aslanimoghanloo, Muhammad, et al.
Publicado: (2025)
por: Aslanimoghanloo, Muhammad, et al.
Publicado: (2025)
Noise-based reward-modulated learning
por: Fernández, Jesús García, et al.
Publicado: (2025)
por: Fernández, Jesús García, et al.
Publicado: (2025)
Recurrent Off-Policy Deep Reinforcement Learning Doesn't Have to be Slow
por: Clark, Tyler, et al.
Publicado: (2025)
por: Clark, Tyler, et al.
Publicado: (2025)
Investigating Action Encodings in Recurrent Neural Networks in Reinforcement Learning
por: Schlegel, Matthew, et al.
Publicado: (2026)
por: Schlegel, Matthew, et al.
Publicado: (2026)
Subspace Node Pruning
por: Offergeld, Joshua, et al.
Publicado: (2024)
por: Offergeld, Joshua, et al.
Publicado: (2024)
Real-Time Decorrelation-Based Anomaly Detection for Multivariate Time Series
por: Sadough, Amirhossein, et al.
Publicado: (2025)
por: Sadough, Amirhossein, et al.
Publicado: (2025)
Graph Structure Learning with Interpretable Bayesian Neural Networks
por: Wasserman, Max, et al.
Publicado: (2024)
por: Wasserman, Max, et al.
Publicado: (2024)
Structured Imitation Learning of Interactive Policies through Inverse Games
por: Sun, Max M., et al.
Publicado: (2025)
por: Sun, Max M., et al.
Publicado: (2025)
Rethinking Recurrent Neural Networks for Time Series Forecasting: A Reinforced Recurrent Encoder with Prediction-Oriented Proximal Policy Optimization
por: Lai, Xin, et al.
Publicado: (2026)
por: Lai, Xin, et al.
Publicado: (2026)
Survival Dynamics of Neural and Programmatic Policies in Evolutionary Reinforcement Learning
por: Roupassov-Ruiz, Anton, et al.
Publicado: (2026)
por: Roupassov-Ruiz, Anton, et al.
Publicado: (2026)
Recurrent Reinforcement Learning with Memoroids
por: Morad, Steven, et al.
Publicado: (2024)
por: Morad, Steven, et al.
Publicado: (2024)
Recurrent State Encoders for Efficient Neural Combinatorial Optimization
por: Dernedde, Tim, et al.
Publicado: (2025)
por: Dernedde, Tim, et al.
Publicado: (2025)
Action-Graph Policies: Learning Action Co-dependencies in Multi-Agent Reinforcement Learning
por: Gupta, Nikunj, et al.
Publicado: (2026)
por: Gupta, Nikunj, et al.
Publicado: (2026)
BARNN: A Bayesian Autoregressive and Recurrent Neural Network
por: Coscia, Dario, et al.
Publicado: (2025)
por: Coscia, Dario, et al.
Publicado: (2025)
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
por: Weltevrede, Max, et al.
Publicado: (2025)
por: Weltevrede, Max, et al.
Publicado: (2025)
Learning Sequence Attractors in Recurrent Networks with Hidden Neurons
por: Lu, Yao, et al.
Publicado: (2024)
por: Lu, Yao, et al.
Publicado: (2024)
Learning with Hidden Factorial Structure
por: Arnal, Charles, et al.
Publicado: (2024)
por: Arnal, Charles, et al.
Publicado: (2024)
Playing Hex and Counter Wargames using Reinforcement Learning and Recurrent Neural Networks
por: Palma, Guilherme, et al.
Publicado: (2025)
por: Palma, Guilherme, et al.
Publicado: (2025)
Recurrent Neural Networks with Linear Structures for Electricity Price Forecasting
por: Amor, Souhir Ben, et al.
Publicado: (2025)
por: Amor, Souhir Ben, et al.
Publicado: (2025)
Neural Co-Optimization of Structural Topology, Manufacturable Layers, and Path Orientations for Fiber-Reinforced Composites
por: Liu, Tao, et al.
Publicado: (2025)
por: Liu, Tao, et al.
Publicado: (2025)
Deep State Space Recurrent Neural Networks for Time Series Forecasting
por: Inzirillo, Hugo
Publicado: (2024)
por: Inzirillo, Hugo
Publicado: (2024)
Bayesian Meta-Reinforcement Learning with Laplace Variational Recurrent Networks
por: de Vries, Joery A., et al.
Publicado: (2025)
por: de Vries, Joery A., et al.
Publicado: (2025)
Ejemplares similares
-
Optimal Control of Probabilistic Dynamics Models via Mean Hamiltonian Minimization
por: Leeftink, David, et al.
Publicado: (2025) -
Automated Discovery of Laser Dicing Processes with Bayesian Optimization for Semiconductor Manufacturing
por: Leeftink, David, et al.
Publicado: (2025) -
Robust Inference of Dynamic Covariance Using Wishart Processes and Sequential Monte Carlo
por: Huijsdens, Hester, et al.
Publicado: (2024) -
Gradient-Free Training of Recurrent Neural Networks using Random Perturbations
por: Fernandez, Jesus Garcia, et al.
Publicado: (2024) -
Energy-Efficient Spiking Recurrent Neural Network for Gesture Recognition on Embedded GPUs
por: Varposhti, Marzieh Hassanshahi, et al.
Publicado: (2024)