ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yunuo, Luo, Baiting, Mukhopadhyay, Ayan, Karsai, Gabor, Dubey, Abhishek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
by: Zhang, Yunuo, et al.
Published: (2025)
by: Zhang, Yunuo, et al.
Published: (2025)
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
by: Luo, Baiting, et al.
Published: (2025)
by: Luo, Baiting, et al.
Published: (2025)
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
by: Luo, Baiting, et al.
Published: (2024)
by: Luo, Baiting, et al.
Published: (2024)
In-Context Planning with Latent Temporal Abstractions
by: Luo, Baiting, et al.
Published: (2026)
by: Luo, Baiting, et al.
Published: (2026)
Shrinking POMCP: A Framework for Real-Time UAV Search and Rescue
by: Zhang, Yunuo, et al.
Published: (2024)
by: Zhang, Yunuo, et al.
Published: (2024)
Decision Making in Non-Stationary Environments with Policy-Augmented Search
by: Pettet, Ava, et al.
Published: (2024)
by: Pettet, Ava, et al.
Published: (2024)
NS-Gym: Open-Source Simulation Environments and Benchmarks for Non-Stationary Markov Decision Processes
by: Keplinger, Nathaniel S., et al.
Published: (2025)
by: Keplinger, Nathaniel S., et al.
Published: (2025)
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Toward Learning POMDPs Beyond Full-Rank Actions and State Observability
by: Shaw, Seiji, et al.
Published: (2026)
by: Shaw, Seiji, et al.
Published: (2026)
No Compromise in Solution Quality: Speeding Up Belief-dependent Continuous POMDPs via Adaptive Multilevel Simplification
by: Zhitnikov, Andrey, et al.
Published: (2023)
by: Zhitnikov, Andrey, et al.
Published: (2023)
Consistency Trajectory Planning: High-Quality and Efficient Trajectory Optimization for Offline Model-Based Reinforcement Learning
by: Wang, Guanquan, et al.
Published: (2025)
by: Wang, Guanquan, et al.
Published: (2025)
Reconciling Spatial and Temporal Abstractions for Goal Representation
by: Zadem, Mehdi, et al.
Published: (2024)
by: Zadem, Mehdi, et al.
Published: (2024)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
by: Hong, Matthew M., et al.
Published: (2026)
by: Hong, Matthew M., et al.
Published: (2026)
Deep Learning Warm Starts for Trajectory Optimization on the International Space Station
by: Banerjee, Somrita, et al.
Published: (2025)
by: Banerjee, Somrita, et al.
Published: (2025)
Spectral Alignment in Forward-Backward Representations via Temporal Abstraction
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
by: Tabakov, Stefan, et al.
Published: (2025)
by: Tabakov, Stefan, et al.
Published: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
by: Cai, Shizhe, et al.
Published: (2025)
by: Cai, Shizhe, et al.
Published: (2025)
How to Explore with Belief: State Entropy Maximization in POMDPs
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
Optimizing Sensor Redundancy in Sequential Decision-Making Problems
by: Nüßlein, Jonas, et al.
Published: (2024)
by: Nüßlein, Jonas, et al.
Published: (2024)
Can We Really Learn One Representation to Optimize All Rewards?
by: Zheng, Chongyi, et al.
Published: (2026)
by: Zheng, Chongyi, et al.
Published: (2026)
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
by: Farhadi, MH, et al.
Published: (2026)
by: Farhadi, MH, et al.
Published: (2026)
Belief Aided Navigation using Bayesian Reinforcement Learning for Avoiding Humans in Blind Spots
by: Kim, Jinyeob, et al.
Published: (2024)
by: Kim, Jinyeob, et al.
Published: (2024)
Premier-TACO is a Few-Shot Policy Learner: Pretraining Multitask Representation via Temporal Action-Driven Contrastive Loss
by: Zheng, Ruijie, et al.
Published: (2024)
by: Zheng, Ruijie, et al.
Published: (2024)
Forecasting and Mitigating Disruptions in Public Bus Transit Services
by: Han, Chaeeun, et al.
Published: (2024)
by: Han, Chaeeun, et al.
Published: (2024)
Enabling MCTS Explainability for Sequential Planning Through Computation Tree Logic
by: An, Ziyan, et al.
Published: (2024)
by: An, Ziyan, et al.
Published: (2024)
Temporally Consistent Object-Centric Learning by Contrasting Slots
by: Manasyan, Anna, et al.
Published: (2024)
by: Manasyan, Anna, et al.
Published: (2024)
Memory-Consistent Neural Networks for Imitation Learning
by: Sridhar, Kaustubh, et al.
Published: (2023)
by: Sridhar, Kaustubh, et al.
Published: (2023)
Robot Policy Learning with Temporal Optimal Transport Reward
by: Fu, Yuwei, et al.
Published: (2024)
by: Fu, Yuwei, et al.
Published: (2024)
Tru-POMDP: Task Planning Under Uncertainty via Tree of Hypotheses and Open-Ended POMDPs
by: Tang, Wenjing, et al.
Published: (2025)
by: Tang, Wenjing, et al.
Published: (2025)
SliceIt! -- A Dual Simulator Framework for Learning Robot Food Slicing
by: Beltran-Hernandez, Cristian C., et al.
Published: (2024)
by: Beltran-Hernandez, Cristian C., et al.
Published: (2024)
A Finite-State Controller Based Offline Solver for Deterministic POMDPs
by: Schutz, Alex, et al.
Published: (2025)
by: Schutz, Alex, et al.
Published: (2025)
EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly
by: Lai, Shih-Yu, et al.
Published: (2026)
by: Lai, Shih-Yu, et al.
Published: (2026)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks
by: Meng, Yue, et al.
Published: (2025)
by: Meng, Yue, et al.
Published: (2025)
PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations
by: Zhang, Yang, et al.
Published: (2026)
by: Zhang, Yang, et al.
Published: (2026)
Temporal Transfer Learning for Traffic Optimization with Coarse-grained Advisory Autonomy
by: Cho, Jung-Hoon, et al.
Published: (2023)
by: Cho, Jung-Hoon, et al.
Published: (2023)
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023)
by: Bobu, Andreea, et al.
Published: (2023)
Foundation Policies with Hilbert Representations
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
Similar Items
-
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
by: Zhang, Yunuo, et al.
Published: (2025) -
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
by: Luo, Baiting, et al.
Published: (2025) -
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
by: Luo, Baiting, et al.
Published: (2024) -
In-Context Planning with Latent Temporal Abstractions
by: Luo, Baiting, et al.
Published: (2026) -
Shrinking POMCP: A Framework for Real-Time UAV Search and Rescue
by: Zhang, Yunuo, et al.
Published: (2024)