P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Levy, Mara, Haldar, Siddhant, Pinto, Lerrel, Shirivastava, Abhinav |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control
by: Cui, Zichen Jeff, et al.
Published: (2024)
by: Cui, Zichen Jeff, et al.
Published: (2024)
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
by: Haldar, Siddhant, et al.
Published: (2025)
by: Haldar, Siddhant, et al.
Published: (2025)
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
by: Zhao, Zifan, et al.
Published: (2025)
by: Zhao, Zifan, et al.
Published: (2025)
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
by: Liu, Peiqi, et al.
Published: (2024)
by: Liu, Peiqi, et al.
Published: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
by: Jahangard, Simindokht, et al.
Published: (2025)
by: Jahangard, Simindokht, et al.
Published: (2025)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
by: Liu, Yijun, et al.
Published: (2025)
by: Liu, Yijun, et al.
Published: (2025)
V-HOP: Visuo-Haptic 6D Object Pose Tracking
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
by: Yuan, Wentao, et al.
Published: (2024)
by: Yuan, Wentao, et al.
Published: (2024)
SpatialPoint: Spatial-aware Point Prediction for Embodied Localization
by: Zhu, Qiming, et al.
Published: (2026)
by: Zhu, Qiming, et al.
Published: (2026)
Learning 3D Robotics Perception using Inductive Priors
by: Irshad, Muhammad Zubair
Published: (2024)
by: Irshad, Muhammad Zubair
Published: (2024)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
by: Lancaster, Patrick, et al.
Published: (2023)
by: Lancaster, Patrick, et al.
Published: (2023)
Learning Precise, Contact-Rich Manipulation through Uncalibrated Tactile Skins
by: Pattabiraman, Venkatesh, et al.
Published: (2024)
by: Pattabiraman, Venkatesh, et al.
Published: (2024)
V-VIPE: Variational View Invariant Pose Embedding
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
EgoZero: Robot Learning from Smart Glasses
by: Liu, Vincent, et al.
Published: (2025)
by: Liu, Vincent, et al.
Published: (2025)
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
by: Chen, Xinyi, et al.
Published: (2025)
by: Chen, Xinyi, et al.
Published: (2025)
Multi-Scale Neighborhood Occupancy Masked Autoencoder for Self-Supervised Learning in LiDAR Point Clouds
by: Abdelsamad, Mohamed, et al.
Published: (2025)
by: Abdelsamad, Mohamed, et al.
Published: (2025)
MPVO: Motion-Prior based Visual Odometry for PointGoal Navigation
by: Paul, Sayan, et al.
Published: (2024)
by: Paul, Sayan, et al.
Published: (2024)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
by: Huang, Wenlong, et al.
Published: (2026)
by: Huang, Wenlong, et al.
Published: (2026)
Point2Graph: An End-to-end Point Cloud-based 3D Open-Vocabulary Scene Graph for Robot Navigation
by: Xu, Yifan, et al.
Published: (2024)
by: Xu, Yifan, et al.
Published: (2024)
SEM: Enhancing Spatial Understanding for Robust Robot Manipulation
by: Lin, Xuewu, et al.
Published: (2025)
by: Lin, Xuewu, et al.
Published: (2025)
Spatially Visual Perception for End-to-End Robotic Learning
by: Davies, Travis, et al.
Published: (2024)
by: Davies, Travis, et al.
Published: (2024)
Leveraging Object Priors for Point Tracking
by: Boote, Bikram, et al.
Published: (2024)
by: Boote, Bikram, et al.
Published: (2024)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
by: Wu, Yiming, et al.
Published: (2025)
by: Wu, Yiming, et al.
Published: (2025)
Fixed External Cameras as Common Prior Maps for Active 3D Scene Graph Generation
by: Modi, Giorgia, et al.
Published: (2026)
by: Modi, Giorgia, et al.
Published: (2026)
Toward General Object-level Mapping from Sparse Views with 3D Diffusion Priors
by: Liao, Ziwei, et al.
Published: (2024)
by: Liao, Ziwei, et al.
Published: (2024)
Dynamic Robot-Assisted Surgery with Hierarchical Class-Incremental Semantic Segmentation
by: Hindel, Julia, et al.
Published: (2025)
by: Hindel, Julia, et al.
Published: (2025)
PhotoAgent: A Robotic Photographer with Spatial and Aesthetic Understanding
by: Che, Lirong, et al.
Published: (2026)
by: Che, Lirong, et al.
Published: (2026)
ARDuP: Active Region Video Diffusion for Universal Policies
by: Huang, Shuaiyi, et al.
Published: (2024)
by: Huang, Shuaiyi, et al.
Published: (2024)
Rectified Point Flow: Generic Point Cloud Pose Estimation
by: Sun, Tao, et al.
Published: (2025)
by: Sun, Tao, et al.
Published: (2025)
Sparse3DTrack: Monocular 3D Object Tracking Using Sparse Supervision
by: Gosala, Nikhil, et al.
Published: (2026)
by: Gosala, Nikhil, et al.
Published: (2026)
Online-Adaptive Anomaly Detection for Defect Identification in Aircraft Assembly
by: Shete, Siddhant, et al.
Published: (2024)
by: Shete, Siddhant, et al.
Published: (2024)
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation
by: Xing, Youguang, et al.
Published: (2025)
by: Xing, Youguang, et al.
Published: (2025)
Efficient Robotic Policy Learning via Latent Space Backward Planning
by: Liu, Dongxiu, et al.
Published: (2025)
by: Liu, Dongxiu, et al.
Published: (2025)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
by: Giannakakis, Nikos, et al.
Published: (2025)
by: Giannakakis, Nikos, et al.
Published: (2025)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
by: Zhou, Enshen, et al.
Published: (2025)
by: Zhou, Enshen, et al.
Published: (2025)
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
by: Li, Zaijing, et al.
Published: (2026)
by: Li, Zaijing, et al.
Published: (2026)
BAKU: An Efficient Transformer for Multi-Task Policy Learning
by: Haldar, Siddhant, et al.
Published: (2024)
by: Haldar, Siddhant, et al.
Published: (2024)
From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
by: Zhang, Zhengshen, et al.
Published: (2025)
by: Zhang, Zhengshen, et al.
Published: (2025)
Differentiable Inverse Graphics for Zero-shot Scene Reconstruction and Robot Grasping
by: Arriaga, Octavio, et al.
Published: (2026)
by: Arriaga, Octavio, et al.
Published: (2026)
NeRF-Aug: Data Augmentation for Robotics with Neural Radiance Fields
by: Zhu, Eric, et al.
Published: (2024)
by: Zhu, Eric, et al.
Published: (2024)
Similar Items
-
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control
by: Cui, Zichen Jeff, et al.
Published: (2024) -
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
by: Haldar, Siddhant, et al.
Published: (2025) -
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
by: Zhao, Zifan, et al.
Published: (2025) -
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
by: Liu, Peiqi, et al.
Published: (2024) -
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
by: Jahangard, Simindokht, et al.
Published: (2025)