Zero-Shot Context Generalization in Reinforcement Learning from Few Training Contexts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chapman, James, Karhadkar, Kedar, Montufar, Guido |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FlowRL: Flow-Augmented Few-Shot Reinforcement Learning for Semi-Structured Sensor Data
von: Pivezhandi, Mohammad, et al.
Veröffentlicht: (2024)
von: Pivezhandi, Mohammad, et al.
Veröffentlicht: (2024)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
von: Yousaf, Iqra
Veröffentlicht: (2024)
von: Yousaf, Iqra
Veröffentlicht: (2024)
Fractional Policy Gradients: Reinforcement Learning with Long-Term Memory
von: Pawar, Urvi, et al.
Veröffentlicht: (2025)
von: Pawar, Urvi, et al.
Veröffentlicht: (2025)
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
von: Tomashevskiy, Timofey
Veröffentlicht: (2026)
von: Tomashevskiy, Timofey
Veröffentlicht: (2026)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
APC-GNN++: An Adaptive Patient-Centric GNN with Context-Aware Attention and Mini-Graph Explainability for Diabetes Classification
von: Berkani, Khaled
Veröffentlicht: (2025)
von: Berkani, Khaled
Veröffentlicht: (2025)
An Aircraft Upset Recovery System with Reinforcement Learning
von: Demir, Mahir, et al.
Veröffentlicht: (2026)
von: Demir, Mahir, et al.
Veröffentlicht: (2026)
Predicting and improving test-time scaling laws via reward tail-guided search
von: Li, Muheng, et al.
Veröffentlicht: (2026)
von: Li, Muheng, et al.
Veröffentlicht: (2026)
Safe Reinforcement Learning with Preference-based Constraint Inference
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Day-to-day Dynamic Tolling in Tradable Credit Schemes
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
von: Xu, Zhe
Veröffentlicht: (2026)
von: Xu, Zhe
Veröffentlicht: (2026)
Joint Combinatorial Node Selection and Resource Allocations in the Lightning Network using Attention-based Reinforcement Learning
von: Salahshour, Mahdi, et al.
Veröffentlicht: (2024)
von: Salahshour, Mahdi, et al.
Veröffentlicht: (2024)
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
von: Silue, Bram, et al.
Veröffentlicht: (2025)
von: Silue, Bram, et al.
Veröffentlicht: (2025)
Combining Trained Models in Reinforcement Learning
von: Patil, Ujjwal, et al.
Veröffentlicht: (2026)
von: Patil, Ujjwal, et al.
Veröffentlicht: (2026)
Towards Systematic Generalization for Power Grid Optimization Problems
von: Memon, Zeeshan, et al.
Veröffentlicht: (2026)
von: Memon, Zeeshan, et al.
Veröffentlicht: (2026)
The Final-Stage Bottleneck: A Systematic Dissection of the R-Learner for Network Causal Inference
von: Sairam, S, et al.
Veröffentlicht: (2025)
von: Sairam, S, et al.
Veröffentlicht: (2025)
SMOSE: Sparse Mixture of Shallow Experts for Interpretable Reinforcement Learning in Continuous Control Tasks
von: Vincze, Mátyás, et al.
Veröffentlicht: (2024)
von: Vincze, Mátyás, et al.
Veröffentlicht: (2024)
What Do World Models Learn in RL? Probing Latent Representations in Learned Environment Simulators
von: Zhang, Xinyu
Veröffentlicht: (2026)
von: Zhang, Xinyu
Veröffentlicht: (2026)
Working Paper: Active Causal Structure Learning with Latent Variables: Towards Learning to Detour in Autonomous Robots
von: Riscos, Pablo de los, et al.
Veröffentlicht: (2024)
von: Riscos, Pablo de los, et al.
Veröffentlicht: (2024)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
von: Belcamino, Valerio, et al.
Veröffentlicht: (2026)
von: Belcamino, Valerio, et al.
Veröffentlicht: (2026)
Adaptable Hindsight Experience Replay for Search-Based Learning
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
Machine Learning Based Path Planning for Improved Rover Navigation (Pre-Print Version)
von: Abcouwer, Neil, et al.
Veröffentlicht: (2020)
von: Abcouwer, Neil, et al.
Veröffentlicht: (2020)
Differentiable Symbolic Planning: A Neural Architecture for Constraint Reasoning with Learned Feasibility
von: Oruganti, Venkatakrishna Reddy
Veröffentlicht: (2026)
von: Oruganti, Venkatakrishna Reddy
Veröffentlicht: (2026)
Are We Winning the Wrong Game? Revisiting Evaluation Practices for Long-Term Time Series Forecasting
von: Phungtua-eng, Thanapol, et al.
Veröffentlicht: (2026)
von: Phungtua-eng, Thanapol, et al.
Veröffentlicht: (2026)
An Automatic Ground Collision Avoidance System with Reinforcement Learning
von: Sevgili, Seyyid Osman, et al.
Veröffentlicht: (2026)
von: Sevgili, Seyyid Osman, et al.
Veröffentlicht: (2026)
Social Interpretable Reinforcement Learning
von: Custode, Leonardo Lucio, et al.
Veröffentlicht: (2024)
von: Custode, Leonardo Lucio, et al.
Veröffentlicht: (2024)
Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning
von: Wang, Zizhao, et al.
Veröffentlicht: (2024)
von: Wang, Zizhao, et al.
Veröffentlicht: (2024)
Not All Transitions Matter: Evidence from PPO
von: Basnet, Ajhesh
Veröffentlicht: (2026)
von: Basnet, Ajhesh
Veröffentlicht: (2026)
Tackling Decision Processes with Non-Cumulative Objectives using Reinforcement Learning
von: Nägele, Maximilian, et al.
Veröffentlicht: (2024)
von: Nägele, Maximilian, et al.
Veröffentlicht: (2024)
From Theory to Practice with RAVEN-UCB: Addressing Non-Stationarity in Multi-Armed Bandits through Variance Adaptation
von: Fang, Junyi, et al.
Veröffentlicht: (2025)
von: Fang, Junyi, et al.
Veröffentlicht: (2025)
Embedded Safety-Aligned Intelligence via Differentiable Internal Alignment Embeddings
von: Rathva, Harsh, et al.
Veröffentlicht: (2025)
von: Rathva, Harsh, et al.
Veröffentlicht: (2025)
Score-informed Neural Operator for Enhancing Ordering-based Causal Discovery
von: Kang, Jiyeon, et al.
Veröffentlicht: (2025)
von: Kang, Jiyeon, et al.
Veröffentlicht: (2025)
Bilinear Mamba-Koopman Neural MPC for Varying Dynamics
von: Pagi, Matan, et al.
Veröffentlicht: (2026)
von: Pagi, Matan, et al.
Veröffentlicht: (2026)
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
von: Zawalski, Michał, et al.
Veröffentlicht: (2022)
von: Zawalski, Michał, et al.
Veröffentlicht: (2022)
Evolving machine learning workflows through interactive AutoML
von: Barbudo, Rafael, et al.
Veröffentlicht: (2024)
von: Barbudo, Rafael, et al.
Veröffentlicht: (2024)
Incentives for Responsiveness, Instrumental Control and Impact
von: Carey, Ryan, et al.
Veröffentlicht: (2020)
von: Carey, Ryan, et al.
Veröffentlicht: (2020)
Regret-Aware Policy Optimization: Environment-Level Memory for Replay Suppression under Delayed Harm
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites
von: Zhang, Qinshi, et al.
Veröffentlicht: (2026)
von: Zhang, Qinshi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FlowRL: Flow-Augmented Few-Shot Reinforcement Learning for Semi-Structured Sensor Data
von: Pivezhandi, Mohammad, et al.
Veröffentlicht: (2024) -
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
von: Yousaf, Iqra
Veröffentlicht: (2024) -
Fractional Policy Gradients: Reinforcement Learning with Long-Term Memory
von: Pawar, Urvi, et al.
Veröffentlicht: (2025) -
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
von: Tomashevskiy, Timofey
Veröffentlicht: (2026) -
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)