ObjectForesight: Predicting Future 3D Object Trajectories from Human Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Soraki, Rustin, Bharadhwaj, Homanga, Farhadi, Ali, Mottaghi, Roozbeh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
by: Bharadhwaj, Homanga, et al.
Published: (2024)
by: Bharadhwaj, Homanga, et al.
Published: (2024)
Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
by: Chen, Hongyi, et al.
Published: (2026)
by: Chen, Hongyi, et al.
Published: (2026)
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
by: Bao, Chen, et al.
Published: (2024)
by: Bao, Chen, et al.
Published: (2024)
Controllable Human-Object Interaction Synthesis
by: Li, Jiaman, et al.
Published: (2023)
by: Li, Jiaman, et al.
Published: (2023)
CrossFusion: A Multi-Scale Cross-Attention Convolutional Fusion Model for Cancer Survival Prediction
by: Soraki, Rustin, et al.
Published: (2025)
by: Soraki, Rustin, et al.
Published: (2025)
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
by: Chen, Mingfei, et al.
Published: (2025)
by: Chen, Mingfei, et al.
Published: (2025)
Web2Grasp: Learning Functional Grasps from Web Images of Hand-Object Interactions
by: Chen, Hongyi, et al.
Published: (2025)
by: Chen, Hongyi, et al.
Published: (2025)
From an Image to a Scene: Learning to Imagine the World from a Million 360 Videos
by: Wallingford, Matthew, et al.
Published: (2024)
by: Wallingford, Matthew, et al.
Published: (2024)
Walk through Paintings: Egocentric World Models from Internet Priors
by: Bagchi, Anurag, et al.
Published: (2026)
by: Bagchi, Anurag, et al.
Published: (2026)
WildDet3D: Scaling Promptable 3D Detection in the Wild
by: Huang, Weikai, et al.
Published: (2026)
by: Huang, Weikai, et al.
Published: (2026)
LiLo-VLA: Compositional Long-Horizon Manipulation via Linked Object-Centric Policies
by: Yang, Yue, et al.
Published: (2026)
by: Yang, Yue, et al.
Published: (2026)
TrajectoryMover: Generative Movement of Object Trajectories in Videos
by: Chhatre, Kiran, et al.
Published: (2026)
by: Chhatre, Kiran, et al.
Published: (2026)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Foresight in Motion: Reinforcing Trajectory Prediction with Reward Heuristics
by: Pei, Muleilan, et al.
Published: (2025)
by: Pei, Muleilan, et al.
Published: (2025)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
by: Mao, Aihua, et al.
Published: (2026)
by: Mao, Aihua, et al.
Published: (2026)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
DINO-Foresight: Looking into the Future with DINO
by: Karypidis, Efstathios, et al.
Published: (2024)
by: Karypidis, Efstathios, et al.
Published: (2024)
RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos
by: Xia, Hongchi, et al.
Published: (2024)
by: Xia, Hongchi, et al.
Published: (2024)
Future Success Prediction in Open-Vocabulary Object Manipulation Tasks Based on End-Effector Trajectories
by: Kambara, Motonari, et al.
Published: (2024)
by: Kambara, Motonari, et al.
Published: (2024)
Trajectory Prediction in Dynamic Object Tracking: A Critical Study
by: Dong, Zhongping, et al.
Published: (2025)
by: Dong, Zhongping, et al.
Published: (2025)
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
by: Phung, Quynh, et al.
Published: (2026)
by: Phung, Quynh, et al.
Published: (2026)
InTraGen: Trajectory-controlled Video Generation for Object Interactions
by: Liu, Zuhao, et al.
Published: (2024)
by: Liu, Zuhao, et al.
Published: (2024)
A Simple Video Segmenter by Tracking Objects Along Axial Trajectories
by: He, Ju, et al.
Published: (2023)
by: He, Ju, et al.
Published: (2023)
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
by: Wang, Hanqing, et al.
Published: (2026)
by: Wang, Hanqing, et al.
Published: (2026)
ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos
by: Chen, Yuantao, et al.
Published: (2026)
by: Chen, Yuantao, et al.
Published: (2026)
Reconstructing Hand-Held Objects in 3D from Images and Videos
by: Wu, Jane, et al.
Published: (2024)
by: Wu, Jane, et al.
Published: (2024)
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
by: Huang, Kuan-Chih, et al.
Published: (2023)
by: Huang, Kuan-Chih, et al.
Published: (2023)
TrajSSL: Trajectory-Enhanced Semi-Supervised 3D Object Detection
by: Jacobson, Philip, et al.
Published: (2024)
by: Jacobson, Philip, et al.
Published: (2024)
HOIMotion: Forecasting Human Motion During Human-Object Interactions Using Egocentric 3D Object Bounding Boxes
by: Hu, Zhiming, et al.
Published: (2024)
by: Hu, Zhiming, et al.
Published: (2024)
Learning Temporal Cues by Predicting Objects Move for Multi-camera 3D Object Detection
by: Moon, Seokha, et al.
Published: (2024)
by: Moon, Seokha, et al.
Published: (2024)
Video Spatial Reasoning with Object-Centric 3D Rollout
by: Tang, Haoran, et al.
Published: (2025)
by: Tang, Haoran, et al.
Published: (2025)
StreamMOTP: Streaming and Unified Framework for Joint 3D Multi-Object Tracking and Trajectory Prediction
by: Zhuang, Jiaheng, et al.
Published: (2024)
by: Zhuang, Jiaheng, et al.
Published: (2024)
Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
by: Bharadhwaj, Homanga, et al.
Published: (2024)
by: Bharadhwaj, Homanga, et al.
Published: (2024)
CORE4D: A 4D Human-Object-Human Interaction Dataset for Collaborative Object REarrangement
by: Liu, Yun, et al.
Published: (2024)
by: Liu, Yun, et al.
Published: (2024)
SOTFormer: A Minimal Transformer for Unified Object Tracking and Trajectory Prediction
by: Dong, Zhongping, et al.
Published: (2025)
by: Dong, Zhongping, et al.
Published: (2025)
Learning Group Interactions and Semantic Intentions for Multi-Object Trajectory Prediction
by: Qi, Mengshi, et al.
Published: (2024)
by: Qi, Mengshi, et al.
Published: (2024)
Survey on Modeling of Human-made Articulated Objects
by: Liu, Jiayi, et al.
Published: (2024)
by: Liu, Jiayi, et al.
Published: (2024)
Putting the Object Back into Video Object Segmentation
by: Cheng, Ho Kei, et al.
Published: (2023)
by: Cheng, Ho Kei, et al.
Published: (2023)
MoDA: Modeling Deformable 3D Objects from Casual Videos
by: Song, Chaoyue, et al.
Published: (2023)
by: Song, Chaoyue, et al.
Published: (2023)
TRAM: Global Trajectory and Motion of 3D Humans from in-the-wild Videos
by: Wang, Yufu, et al.
Published: (2024)
by: Wang, Yufu, et al.
Published: (2024)
Similar Items
-
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
by: Bharadhwaj, Homanga, et al.
Published: (2024) -
Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
by: Chen, Hongyi, et al.
Published: (2026) -
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
by: Bao, Chen, et al.
Published: (2024) -
Controllable Human-Object Interaction Synthesis
by: Li, Jiaman, et al.
Published: (2023) -
CrossFusion: A Multi-Scale Cross-Attention Convolutional Fusion Model for Cancer Survival Prediction
by: Soraki, Rustin, et al.
Published: (2025)