Toward Learning POMDPs Beyond Full-Rank Actions and State Observability
Fuente:
arXiv
Guardado en:
| Autores principales: | Shaw, Seiji, Manderson, Travis, Kessens, Chad, Roy, Nicholas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
por: Zhang, Yunuo, et al.
Publicado: (2025)
por: Zhang, Yunuo, et al.
Publicado: (2025)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Guided Policy Optimization under Partial Observability
por: Li, Yueheng, et al.
Publicado: (2025)
por: Li, Yueheng, et al.
Publicado: (2025)
Symmetry-aware Reinforcement Learning for Robotic Assembly under Partial Observability with a Soft Wrist
por: Nguyen, Hai, et al.
Publicado: (2024)
por: Nguyen, Hai, et al.
Publicado: (2024)
Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learning
por: Harmel, Moritz, et al.
Publicado: (2023)
por: Harmel, Moritz, et al.
Publicado: (2023)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Unsupervised Learning of Effective Actions in Robotics
por: Zaric, Marko, et al.
Publicado: (2024)
por: Zaric, Marko, et al.
Publicado: (2024)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
por: Lanier, Michael, et al.
Publicado: (2024)
por: Lanier, Michael, et al.
Publicado: (2024)
Autoregressive Action Sequence Learning for Robotic Manipulation
por: Zhang, Xinyu, et al.
Publicado: (2024)
por: Zhang, Xinyu, et al.
Publicado: (2024)
Equivariant Action Sampling for Reinforcement Learning and Planning
por: Zhao, Linfeng, et al.
Publicado: (2024)
por: Zhao, Linfeng, et al.
Publicado: (2024)
A Finite-State Controller Based Offline Solver for Deterministic POMDPs
por: Schutz, Alex, et al.
Publicado: (2025)
por: Schutz, Alex, et al.
Publicado: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
por: Zeng, Qixin, et al.
Publicado: (2026)
por: Zeng, Qixin, et al.
Publicado: (2026)
Learning to Act Robustly with View-Invariant Latent Actions
por: Jeong, Youngjoon, et al.
Publicado: (2026)
por: Jeong, Youngjoon, et al.
Publicado: (2026)
Vintix: Action Model via In-Context Reinforcement Learning
por: Polubarov, Andrey, et al.
Publicado: (2025)
por: Polubarov, Andrey, et al.
Publicado: (2025)
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
por: Zhang, Yunuo, et al.
Publicado: (2025)
por: Zhang, Yunuo, et al.
Publicado: (2025)
Jump-Start Reinforcement Learning with Vision-Language-Action Regularization
por: Moroncelli, Angelo, et al.
Publicado: (2026)
por: Moroncelli, Angelo, et al.
Publicado: (2026)
Dual Action Policy for Robust Sim-to-Real Reinforcement Learning
por: Terence, Ng Wen Zheng, et al.
Publicado: (2024)
por: Terence, Ng Wen Zheng, et al.
Publicado: (2024)
Toward Accurate Long-Horizon Robotic Manipulation: Language-to-Action with Foundation Models via Scene Graphs
por: Dinesh, Sushil Samuel, et al.
Publicado: (2025)
por: Dinesh, Sushil Samuel, et al.
Publicado: (2025)
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
Space Robotics Bench: Robot Learning Beyond Earth
por: Orsula, Andrej, et al.
Publicado: (2025)
por: Orsula, Andrej, et al.
Publicado: (2025)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
por: Bu, Qingwen, et al.
Publicado: (2025)
por: Bu, Qingwen, et al.
Publicado: (2025)
RDAR: Reward-Driven Agent Relevance Estimation for Autonomous Driving
por: Bosio, Carlo, et al.
Publicado: (2025)
por: Bosio, Carlo, et al.
Publicado: (2025)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
por: Liu, Huihan, et al.
Publicado: (2026)
por: Liu, Huihan, et al.
Publicado: (2026)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
por: Seo, Younggyo, et al.
Publicado: (2024)
por: Seo, Younggyo, et al.
Publicado: (2024)
Scheduling Drone and Mobile Charger via Hybrid-Action Deep Reinforcement Learning
por: Dou, Jizhe, et al.
Publicado: (2024)
por: Dou, Jizhe, et al.
Publicado: (2024)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
por: Guo, Jian-Ting, et al.
Publicado: (2025)
por: Guo, Jian-Ting, et al.
Publicado: (2025)
CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations
por: Liang, Anthony, et al.
Publicado: (2025)
por: Liang, Anthony, et al.
Publicado: (2025)
Efficient Multi-Task Reinforcement Learning via Task-Specific Action Correction
por: Feng, Jinyuan, et al.
Publicado: (2024)
por: Feng, Jinyuan, et al.
Publicado: (2024)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
por: Chen, Xiaoyu, et al.
Publicado: (2025)
por: Chen, Xiaoyu, et al.
Publicado: (2025)
Adapting World Models with Latent-State Dynamics Residuals
por: Lanier, JB, et al.
Publicado: (2025)
por: Lanier, JB, et al.
Publicado: (2025)
Learning with Language-Guided State Abstractions
por: Peng, Andi, et al.
Publicado: (2024)
por: Peng, Andi, et al.
Publicado: (2024)
Beyond Hard Constraints: Budget-Conditioned Reachability For Safe Offline Reinforcement Learning
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
OAT: Ordered Action Tokenization
por: Liu, Chaoqi, et al.
Publicado: (2026)
por: Liu, Chaoqi, et al.
Publicado: (2026)
Towards Robust Zero-Shot Reinforcement Learning
por: Zheng, Kexin, et al.
Publicado: (2025)
por: Zheng, Kexin, et al.
Publicado: (2025)
Behavior Generation with Latent Actions
por: Lee, Seungjae, et al.
Publicado: (2024)
por: Lee, Seungjae, et al.
Publicado: (2024)
Redundancy-aware Action Spaces for Robot Learning
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
Improving Dribbling, Passing, and Marking Actions in Soccer Simulation 2D Games Using Machine Learning
por: Zare, Nader, et al.
Publicado: (2024)
por: Zare, Nader, et al.
Publicado: (2024)
Towards General Purpose Robots at Scale: Lifelong Learning and Learning to Use Memory
por: Yue, William
Publicado: (2024)
por: Yue, William
Publicado: (2024)
Continuous Reasoning for Vision-Language-Action
por: Wu, Yueh-Hua, et al.
Publicado: (2026)
por: Wu, Yueh-Hua, et al.
Publicado: (2026)
Ejemplares similares
-
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
por: Zhang, Yunuo, et al.
Publicado: (2025) -
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
por: Azeem, Muqsit, et al.
Publicado: (2024) -
Guided Policy Optimization under Partial Observability
por: Li, Yueheng, et al.
Publicado: (2025) -
Symmetry-aware Reinforcement Learning for Robotic Assembly under Partial Observability with a Soft Wrist
por: Nguyen, Hai, et al.
Publicado: (2024) -
Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learning
por: Harmel, Moritz, et al.
Publicado: (2023)