Any-point Trajectory Modeling for Policy Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wen, Chuan, Lin, Xingyu, So, John, Chen, Kai, Dou, Qi, Gao, Yang, Abbeel, Pieter |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rodrigues Network for Learning Robot Actions
by: Zhang, Jialiang, et al.
Published: (2025)
by: Zhang, Jialiang, et al.
Published: (2025)
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
by: Wang, Renhao, et al.
Published: (2025)
by: Wang, Renhao, et al.
Published: (2025)
Closing the Visual Sim-to-Real Gap with Object-Composable NeRFs
by: Mishra, Nikhil, et al.
Published: (2024)
by: Mishra, Nikhil, et al.
Published: (2024)
Object-centric 3D Motion Field for Robot Learning from Human Videos
by: Yin, Zhao-Heng, et al.
Published: (2025)
by: Yin, Zhao-Heng, et al.
Published: (2025)
Can Transformers Capture Spatial Relations between Objects?
by: Wen, Chuan, et al.
Published: (2024)
by: Wen, Chuan, et al.
Published: (2024)
Any6D: Model-free 6D Pose Estimation of Novel Objects
by: Lee, Taeyeop, et al.
Published: (2025)
by: Lee, Taeyeop, et al.
Published: (2025)
Twisting Lids Off with Two Hands
by: Lin, Toru, et al.
Published: (2024)
by: Lin, Toru, et al.
Published: (2024)
SimEndoGS: Efficient Data-driven Scene Simulation using Robotic Surgery Videos via Physics-embedded 3D Gaussians
by: Yang, Zhenya, et al.
Published: (2024)
by: Yang, Zhenya, et al.
Published: (2024)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
by: Huang, Huang, et al.
Published: (2025)
by: Huang, Huang, et al.
Published: (2025)
Enhanced Scale-aware Depth Estimation for Monocular Endoscopic Scenes with Geometric Modeling
by: Wei, Ruofeng, et al.
Published: (2024)
by: Wei, Ruofeng, et al.
Published: (2024)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
by: Wang, Yuran, et al.
Published: (2025)
by: Wang, Yuran, et al.
Published: (2025)
General Flow as Foundation Affordance for Scalable Robot Learning
by: Yuan, Chengbo, et al.
Published: (2024)
by: Yuan, Chengbo, et al.
Published: (2024)
Lightning Grasp: High Performance Procedural Grasp Synthesis with Contact Fields
by: Yin, Zhao-Heng, et al.
Published: (2025)
by: Yin, Zhao-Heng, et al.
Published: (2025)
World Model for Robot Learning: A Comprehensive Survey
by: Hou, Bohan, et al.
Published: (2026)
by: Hou, Bohan, et al.
Published: (2026)
Hand-Object Interaction Pretraining from Videos
by: Singh, Himanshu Gaurav, et al.
Published: (2024)
by: Singh, Himanshu Gaurav, et al.
Published: (2024)
Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling
by: Qiu, Xiaowen, et al.
Published: (2025)
by: Qiu, Xiaowen, et al.
Published: (2025)
AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents
by: Cui, Jieming, et al.
Published: (2024)
by: Cui, Jieming, et al.
Published: (2024)
Visual Representation Learning with Stochastic Frame Prediction
by: Jang, Huiwon, et al.
Published: (2024)
by: Jang, Huiwon, et al.
Published: (2024)
D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping
by: Lou, Haozhe, et al.
Published: (2026)
by: Lou, Haozhe, et al.
Published: (2026)
MetaTra: Meta-Learning for Generalized Trajectory Prediction in Unseen Domain
by: Li, Xiaohe, et al.
Published: (2024)
by: Li, Xiaohe, et al.
Published: (2024)
Social-Transmotion: Promptable Human Trajectory Prediction
by: Saadatnejad, Saeed, et al.
Published: (2023)
by: Saadatnejad, Saeed, et al.
Published: (2023)
A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
Towards Predicting Any Human Trajectory In Context
by: Fujii, Ryo, et al.
Published: (2025)
by: Fujii, Ryo, et al.
Published: (2025)
Referring-Aware Visuomotor Policy Learning for Closed-Loop Manipulation
by: Ma, Jiahua, et al.
Published: (2026)
by: Ma, Jiahua, et al.
Published: (2026)
TripleMixer: A 3D Point Cloud Denoising Model for Adverse Weather
by: Zhao, Xiongwei, et al.
Published: (2024)
by: Zhao, Xiongwei, et al.
Published: (2024)
Multi-modal Motion Prediction using Temporal Ensembling with Learning-based Aggregation
by: Hong, Kai-Yin, et al.
Published: (2024)
by: Hong, Kai-Yin, et al.
Published: (2024)
Robot Learning from Any Images
by: Zhao, Siheng, et al.
Published: (2025)
by: Zhao, Siheng, et al.
Published: (2025)
The RoboDrive Challenge: Drive Anytime Anywhere in Any Condition
by: Kong, Lingdong, et al.
Published: (2024)
by: Kong, Lingdong, et al.
Published: (2024)
Unified Human Localization and Trajectory Prediction with Monocular Vision
by: Luan, Po-Chien, et al.
Published: (2025)
by: Luan, Po-Chien, et al.
Published: (2025)
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2026)
by: Deng, Yijie, et al.
Published: (2026)
DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
by: Egbe, ThankGod, et al.
Published: (2025)
by: Egbe, ThankGod, et al.
Published: (2025)
ET-SEED: Efficient Trajectory-Level SE(3) Equivariant Diffusion Policy
by: Tie, Chenrui, et al.
Published: (2024)
by: Tie, Chenrui, et al.
Published: (2024)
Large Video Planner Enables Generalizable Robot Control
by: Chen, Boyuan, et al.
Published: (2025)
by: Chen, Boyuan, et al.
Published: (2025)
MAS-SAM: Segment Any Marine Animal with Aggregated Features
by: Yan, Tianyu, et al.
Published: (2024)
by: Yan, Tianyu, et al.
Published: (2024)
Visual Imitation Enables Contextual Humanoid Control
by: Allshire, Arthur, et al.
Published: (2025)
by: Allshire, Arthur, et al.
Published: (2025)
AnyView: Synthesizing Any Novel View in Dynamic Scenes
by: Van Hoorick, Basile, et al.
Published: (2026)
by: Van Hoorick, Basile, et al.
Published: (2026)
How to Peel with a Knife: Aligning Fine-Grained Manipulation with Human Preference
by: Lin, Toru, et al.
Published: (2026)
by: Lin, Toru, et al.
Published: (2026)
TAPTR: Tracking Any Point with Transformers as Detection
by: Li, Hongyang, et al.
Published: (2024)
by: Li, Hongyang, et al.
Published: (2024)
Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
by: Chen, Hongyi, et al.
Published: (2026)
by: Chen, Hongyi, et al.
Published: (2026)
AnyDexGrasp: General Dexterous Grasping for Different Hands with Human-level Learning Efficiency
by: Fang, Hao-Shu, et al.
Published: (2025)
by: Fang, Hao-Shu, et al.
Published: (2025)
Similar Items
-
Rodrigues Network for Learning Robot Actions
by: Zhang, Jialiang, et al.
Published: (2025) -
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
by: Wang, Renhao, et al.
Published: (2025) -
Closing the Visual Sim-to-Real Gap with Object-Composable NeRFs
by: Mishra, Nikhil, et al.
Published: (2024) -
Object-centric 3D Motion Field for Robot Learning from Human Videos
by: Yin, Zhao-Heng, et al.
Published: (2025) -
Can Transformers Capture Spatial Relations between Objects?
by: Wen, Chuan, et al.
Published: (2024)