Bilinear Mamba-Koopman Neural MPC for Varying Dynamics
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Pagi, Matan, Sorek, Zohar |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
par: Yousaf, Iqra
Publié: (2024)
par: Yousaf, Iqra
Publié: (2024)
Deep Reinforcement Learning for Day-to-day Dynamic Tolling in Tradable Credit Schemes
par: Wu, Xiaoyi, et autres
Publié: (2025)
par: Wu, Xiaoyi, et autres
Publié: (2025)
Accuracy, Memory Efficiency and Generalization: A Comparative Study on Liquid Neural Networks and Recurrent Neural Networks
par: Zong, Shilong, et autres
Publié: (2025)
par: Zong, Shilong, et autres
Publié: (2025)
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
par: Tomashevskiy, Timofey
Publié: (2026)
par: Tomashevskiy, Timofey
Publié: (2026)
Predicting and improving test-time scaling laws via reward tail-guided search
par: Li, Muheng, et autres
Publié: (2026)
par: Li, Muheng, et autres
Publié: (2026)
Zero-Shot Context Generalization in Reinforcement Learning from Few Training Contexts
par: Chapman, James, et autres
Publié: (2025)
par: Chapman, James, et autres
Publié: (2025)
FlowRL: Flow-Augmented Few-Shot Reinforcement Learning for Semi-Structured Sensor Data
par: Pivezhandi, Mohammad, et autres
Publié: (2024)
par: Pivezhandi, Mohammad, et autres
Publié: (2024)
Fractional Policy Gradients: Reinforcement Learning with Long-Term Memory
par: Pawar, Urvi, et autres
Publié: (2025)
par: Pawar, Urvi, et autres
Publié: (2025)
Autopilot-Preserving Residual Q-Learning with HJB-Inspired Finite-Action Risk Filtering for Fixed-Wing UAV Command Supervision
par: Iscan, Mehmet, et autres
Publié: (2026)
par: Iscan, Mehmet, et autres
Publié: (2026)
Score-informed Neural Operator for Enhancing Ordering-based Causal Discovery
par: Kang, Jiyeon, et autres
Publié: (2025)
par: Kang, Jiyeon, et autres
Publié: (2025)
Differentiable Symbolic Planning: A Neural Architecture for Constraint Reasoning with Learned Feasibility
par: Oruganti, Venkatakrishna Reddy
Publié: (2026)
par: Oruganti, Venkatakrishna Reddy
Publié: (2026)
The Final-Stage Bottleneck: A Systematic Dissection of the R-Learner for Network Causal Inference
par: Sairam, S, et autres
Publié: (2025)
par: Sairam, S, et autres
Publié: (2025)
Are We Winning the Wrong Game? Revisiting Evaluation Practices for Long-Term Time Series Forecasting
par: Phungtua-eng, Thanapol, et autres
Publié: (2026)
par: Phungtua-eng, Thanapol, et autres
Publié: (2026)
Towards Systematic Generalization for Power Grid Optimization Problems
par: Memon, Zeeshan, et autres
Publié: (2026)
par: Memon, Zeeshan, et autres
Publié: (2026)
Machine Learning Based Path Planning for Improved Rover Navigation (Pre-Print Version)
par: Abcouwer, Neil, et autres
Publié: (2020)
par: Abcouwer, Neil, et autres
Publié: (2020)
Safe Reinforcement Learning with Preference-based Constraint Inference
par: Li, Chenglin, et autres
Publié: (2026)
par: Li, Chenglin, et autres
Publié: (2026)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
par: Hiremath, Prakul Sunil
Publié: (2026)
par: Hiremath, Prakul Sunil
Publié: (2026)
Regret-Aware Policy Optimization: Environment-Level Memory for Replay Suppression under Delayed Harm
par: Hiremath, Prakul Sunil
Publié: (2026)
par: Hiremath, Prakul Sunil
Publié: (2026)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
par: Belcamino, Valerio, et autres
Publié: (2026)
par: Belcamino, Valerio, et autres
Publié: (2026)
An Aircraft Upset Recovery System with Reinforcement Learning
par: Demir, Mahir, et autres
Publié: (2026)
par: Demir, Mahir, et autres
Publié: (2026)
Not All Transitions Matter: Evidence from PPO
par: Basnet, Ajhesh
Publié: (2026)
par: Basnet, Ajhesh
Publié: (2026)
AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites
par: Zhang, Qinshi, et autres
Publié: (2026)
par: Zhang, Qinshi, et autres
Publié: (2026)
What Do World Models Learn in RL? Probing Latent Representations in Learned Environment Simulators
par: Zhang, Xinyu
Publié: (2026)
par: Zhang, Xinyu
Publié: (2026)
APC-GNN++: An Adaptive Patient-Centric GNN with Context-Aware Attention and Mini-Graph Explainability for Diabetes Classification
par: Berkani, Khaled
Publié: (2025)
par: Berkani, Khaled
Publié: (2025)
From Theory to Practice with RAVEN-UCB: Addressing Non-Stationarity in Multi-Armed Bandits through Variance Adaptation
par: Fang, Junyi, et autres
Publié: (2025)
par: Fang, Junyi, et autres
Publié: (2025)
Embedded Safety-Aligned Intelligence via Differentiable Internal Alignment Embeddings
par: Rathva, Harsh, et autres
Publié: (2025)
par: Rathva, Harsh, et autres
Publié: (2025)
Working Paper: Active Causal Structure Learning with Latent Variables: Towards Learning to Detour in Autonomous Robots
par: Riscos, Pablo de los, et autres
Publié: (2024)
par: Riscos, Pablo de los, et autres
Publié: (2024)
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
par: Zawalski, Michał, et autres
Publié: (2022)
par: Zawalski, Michał, et autres
Publié: (2022)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
par: Nuzhin, Egor E., et autres
Publié: (2024)
par: Nuzhin, Egor E., et autres
Publié: (2024)
Joint Combinatorial Node Selection and Resource Allocations in the Lightning Network using Attention-based Reinforcement Learning
par: Salahshour, Mahdi, et autres
Publié: (2024)
par: Salahshour, Mahdi, et autres
Publié: (2024)
Evolving machine learning workflows through interactive AutoML
par: Barbudo, Rafael, et autres
Publié: (2024)
par: Barbudo, Rafael, et autres
Publié: (2024)
Incentives for Responsiveness, Instrumental Control and Impact
par: Carey, Ryan, et autres
Publié: (2020)
par: Carey, Ryan, et autres
Publié: (2020)
Adaptable Hindsight Experience Replay for Search-Based Learning
par: Vazaios, Alexandros, et autres
Publié: (2025)
par: Vazaios, Alexandros, et autres
Publié: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
par: Sauter, Andreas W. M., et autres
Publié: (2024)
par: Sauter, Andreas W. M., et autres
Publié: (2024)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
par: Xu, Zhe
Publié: (2026)
par: Xu, Zhe
Publié: (2026)
Deep Policy Iteration with Integer Programming for Inventory Management
par: Harsha, Pavithra, et autres
Publié: (2021)
par: Harsha, Pavithra, et autres
Publié: (2021)
COMET: Codebook-based Online-adaptive Multi-scale Embedding for Time-series Anomaly Detection
par: Park, Jinwoo, et autres
Publié: (2026)
par: Park, Jinwoo, et autres
Publié: (2026)
Simple yet Effective Node Property Prediction on Edge Streams under Distribution Shifts
par: Lee, Jongha, et autres
Publié: (2025)
par: Lee, Jongha, et autres
Publié: (2025)
The CRITICAL Records Integrated Standardization Pipeline (CRISP): End-to-End Processing of Large-scale Multi-institutional OMOP CDM Data
par: Luo, Xiaolong, et autres
Publié: (2025)
par: Luo, Xiaolong, et autres
Publié: (2025)
Improved Exploration in GFlownets via Enhanced Epistemic Neural Networks
par: Muhammad, Sajan, et autres
Publié: (2025)
par: Muhammad, Sajan, et autres
Publié: (2025)
Documents similaires
-
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
par: Yousaf, Iqra
Publié: (2024) -
Deep Reinforcement Learning for Day-to-day Dynamic Tolling in Tradable Credit Schemes
par: Wu, Xiaoyi, et autres
Publié: (2025) -
Accuracy, Memory Efficiency and Generalization: A Comparative Study on Liquid Neural Networks and Recurrent Neural Networks
par: Zong, Shilong, et autres
Publié: (2025) -
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
par: Tomashevskiy, Timofey
Publié: (2026) -
Predicting and improving test-time scaling laws via reward tail-guided search
par: Li, Muheng, et autres
Publié: (2026)