Searching in Space and Time: Unified Memory-Action Loops for Open-World Object Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Taijing, Kumar, Sateesh, Xu, Junhong, Pavlakos, Georgios, Biswas, Joydeep, Martín-Martín, Roberto |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
by: Kumar, Sateesh, et al.
Published: (2025)
by: Kumar, Sateesh, et al.
Published: (2025)
Creating and Repairing Robot Programs in Open-World Domains
by: Schlesinger, Claire, et al.
Published: (2024)
by: Schlesinger, Claire, et al.
Published: (2024)
VOFA: Visual Object Goal Pushing with Force-Adaptive Control for Humanoids
by: Hu, Zichao, et al.
Published: (2026)
by: Hu, Zichao, et al.
Published: (2026)
SLAC: Simulation-Pretrained Latent Action Space for Whole-Body Real-World RL
by: Hu, Jiaheng, et al.
Published: (2025)
by: Hu, Jiaheng, et al.
Published: (2025)
CLOVER: Context-aware Long-term Object Viewpoint- and Environment- Invariant Representation Learning
by: Lee, Dongmyeong, et al.
Published: (2024)
by: Lee, Dongmyeong, et al.
Published: (2024)
KinScene: Model-Based Mobile Manipulation of Articulated Scenes
by: Hsu, Cheng-Chun, et al.
Published: (2024)
by: Hsu, Cheng-Chun, et al.
Published: (2024)
POMDP-based Object Search with Growing State Space and Hybrid Action Domain
by: Chen, Yongbo, et al.
Published: (2026)
by: Chen, Yongbo, et al.
Published: (2026)
OpenGuide: Assistive Object Retrieval in Indoor Spaces for Individuals with Visual Impairments
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos
by: Shah, Rutav, et al.
Published: (2025)
by: Shah, Rutav, et al.
Published: (2025)
SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation
by: Hsu, Cheng-Chun, et al.
Published: (2024)
by: Hsu, Cheng-Chun, et al.
Published: (2024)
Long-Term Memory for VLA-based Agents in Open-World Task Execution
by: Huang, Xu, et al.
Published: (2026)
by: Huang, Xu, et al.
Published: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
by: jia, Feiyang, et al.
Published: (2026)
by: jia, Feiyang, et al.
Published: (2026)
ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation
by: Anwar, Abrar, et al.
Published: (2024)
by: Anwar, Abrar, et al.
Published: (2024)
Building Explicit World Model for Zero-Shot Open-World Object Manipulation
by: Li, Xiaotong, et al.
Published: (2026)
by: Li, Xiaotong, et al.
Published: (2026)
Terrain Costmap Generation via Scaled Preference Conditioning
by: Mao, Luisa, et al.
Published: (2025)
by: Mao, Luisa, et al.
Published: (2025)
Open Scene Graphs for Open-World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2025)
by: Loo, Joel, et al.
Published: (2025)
Open Scene Graphs for Open World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2024)
by: Loo, Joel, et al.
Published: (2024)
X-MOBILITY: End-To-End Generalizable Navigation via World Modeling
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
OVAMOS: A Framework for Open-Vocabulary Multi-Object Search in Unknown Environments
by: Wang, Qianwei, et al.
Published: (2025)
by: Wang, Qianwei, et al.
Published: (2025)
Lift, Splat, Map: Lifting Foundation Masks for Label-Free Semantic Scene Completion
by: Zhang, Arthur, et al.
Published: (2024)
by: Zhang, Arthur, et al.
Published: (2024)
Multi-Agent Inverse Reinforcement Learning in Real World Unstructured Pedestrian Crowds
by: Chandra, Rohan, et al.
Published: (2024)
by: Chandra, Rohan, et al.
Published: (2024)
Enhancing Object Search in Indoor Spaces via Personalized Object-factored Ontologies
by: Chikhalikar, Akash, et al.
Published: (2025)
by: Chikhalikar, Akash, et al.
Published: (2025)
$τ_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
by: Zhou, Pengfei, et al.
Published: (2026)
by: Zhou, Pengfei, et al.
Published: (2026)
RynnVLA-002: A Unified Vision-Language-Action and World Model
by: Cen, Jun, et al.
Published: (2025)
by: Cen, Jun, et al.
Published: (2025)
OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
by: Liu, Yushan, et al.
Published: (2026)
by: Liu, Yushan, et al.
Published: (2026)
Rethinking Social Robot Navigation: Leveraging the Best of Two Worlds
by: Raj, Amir Hossain, et al.
Published: (2023)
by: Raj, Amir Hossain, et al.
Published: (2023)
LILAC: Language-Conditioned Object-Centric Optical Flow for Open-Loop Trajectory Generation
by: Kambara, Motonari, et al.
Published: (2026)
by: Kambara, Motonari, et al.
Published: (2026)
Semantic Masking and Visual Feature Matching for Robust Localization
by: Mao, Luisa, et al.
Published: (2024)
by: Mao, Luisa, et al.
Published: (2024)
Rapid Adaptation of Particle Dynamics for Generalized Deformable Object Mobile Manipulation
by: Wu, Bohan, et al.
Published: (2026)
by: Wu, Bohan, et al.
Published: (2026)
WorldPlanner: Monte Carlo Tree Search and MPC with Action-Conditioned Visual World Models
by: Khorrambakht, R., et al.
Published: (2025)
by: Khorrambakht, R., et al.
Published: (2025)
Towards Open-World Grasping with Large Vision-Language Models
by: Tziafas, Georgios, et al.
Published: (2024)
by: Tziafas, Georgios, et al.
Published: (2024)
World Guidance: World Modeling in Condition Space for Action Generation
by: Su, Yue, et al.
Published: (2026)
by: Su, Yue, et al.
Published: (2026)
OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation
by: Pei, Jiahua, et al.
Published: (2026)
by: Pei, Jiahua, et al.
Published: (2026)
PACER: Preference-conditioned All-terrain Costmap Generation
by: Mao, Luisa, et al.
Published: (2024)
by: Mao, Luisa, et al.
Published: (2024)
UniSaT: Unified-Objective Belief Model and Planner to Search for and Track Multiple Objects
by: Santos, Leonardo, et al.
Published: (2024)
by: Santos, Leonardo, et al.
Published: (2024)
VENTURA: Adapting Image Diffusion Models for Unified Task Conditioned Navigation
by: Zhang, Arthur, et al.
Published: (2025)
by: Zhang, Arthur, et al.
Published: (2025)
Language-Grounded Dynamic Scene Graphs for Interactive Object Search with Mobile Manipulation
by: Honerkamp, Daniel, et al.
Published: (2024)
by: Honerkamp, Daniel, et al.
Published: (2024)
Immersive Human-in-the-Loop Control: Real-Time 3D Surface Meshing and Physics Simulation
by: Akturk, Sait, et al.
Published: (2024)
by: Akturk, Sait, et al.
Published: (2024)
Open-Loop Planning, Closed-Loop Verification: Speculative Verification for VLA
by: Wang, Zihua, et al.
Published: (2026)
by: Wang, Zihua, et al.
Published: (2026)
C-NAV: Towards Self-Evolving Continual Object Navigation in Open World
by: Yu, Ming-Ming, et al.
Published: (2025)
by: Yu, Ming-Ming, et al.
Published: (2025)
Similar Items
-
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
by: Kumar, Sateesh, et al.
Published: (2025) -
Creating and Repairing Robot Programs in Open-World Domains
by: Schlesinger, Claire, et al.
Published: (2024) -
VOFA: Visual Object Goal Pushing with Force-Adaptive Control for Humanoids
by: Hu, Zichao, et al.
Published: (2026) -
SLAC: Simulation-Pretrained Latent Action Space for Whole-Body Real-World RL
by: Hu, Jiaheng, et al.
Published: (2025) -
CLOVER: Context-aware Long-term Object Viewpoint- and Environment- Invariant Representation Learning
by: Lee, Dongmyeong, et al.
Published: (2024)