SSL-Interactions: Pretext Tasks for Interactive Trajectory Prediction
Fuente:
arXiv
Guardado en:
| Autores principales: | Bhattacharyya, Prarthana, Huang, Chengjie, Czarnecki, Krzysztof |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Risk-aware Trajectory Prediction by Incorporating Spatio-temporal Traffic Interaction Analysis
por: Thuremella, Divya, et al.
Publicado: (2024)
por: Thuremella, Divya, et al.
Publicado: (2024)
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
por: Chen, Mingfei, et al.
Publicado: (2025)
por: Chen, Mingfei, et al.
Publicado: (2025)
Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition
por: Wen, Yuhang, et al.
Publicado: (2023)
por: Wen, Yuhang, et al.
Publicado: (2023)
Learning the Pedestrian-Vehicle Interaction for Pedestrian Trajectory Prediction
por: Zhang, Chi, et al.
Publicado: (2022)
por: Zhang, Chi, et al.
Publicado: (2022)
MFSeg: Efficient Multi-frame 3D Semantic Segmentation
por: Huang, Chengjie, et al.
Publicado: (2025)
por: Huang, Chengjie, et al.
Publicado: (2025)
Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy
por: Mandil, Willow, et al.
Publicado: (2023)
por: Mandil, Willow, et al.
Publicado: (2023)
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
por: Fang, Huang, et al.
Publicado: (2025)
por: Fang, Huang, et al.
Publicado: (2025)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
por: Ji, Yihong, et al.
Publicado: (2025)
por: Ji, Yihong, et al.
Publicado: (2025)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
por: Keskar, Maitrayee, et al.
Publicado: (2025)
por: Keskar, Maitrayee, et al.
Publicado: (2025)
Hand-Object Interaction Pretraining from Videos
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations
por: Huang, Wei-Jin, et al.
Publicado: (2026)
por: Huang, Wei-Jin, et al.
Publicado: (2026)
Learning Through Retrospection: Improving Trajectory Prediction for Automated Driving with Error Feedback
por: Hagedorn, Steffen, et al.
Publicado: (2025)
por: Hagedorn, Steffen, et al.
Publicado: (2025)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
por: Zhang, Erhang, et al.
Publicado: (2025)
por: Zhang, Erhang, et al.
Publicado: (2025)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
por: Chen, Hongjin, et al.
Publicado: (2026)
por: Chen, Hongjin, et al.
Publicado: (2026)
Scene Informer: Anchor-based Occlusion Inference and Trajectory Prediction in Partially Observable Environments
por: Lange, Bernard, et al.
Publicado: (2023)
por: Lange, Bernard, et al.
Publicado: (2023)
OPENTOUCH: Bringing Full-Hand Touch to Real-World Interaction
por: Song, Yuxin Ray, et al.
Publicado: (2025)
por: Song, Yuxin Ray, et al.
Publicado: (2025)
VISTA: A Vision and Intent-Aware Social Attention Framework for Multi-Agent Trajectory Prediction
por: Martins, Stephane Da Silva, et al.
Publicado: (2025)
por: Martins, Stephane Da Silva, et al.
Publicado: (2025)
EgoTraj-Bench: Towards Robust Trajectory Prediction Under Ego-view Noisy Observations
por: Liu, Jiayi, et al.
Publicado: (2025)
por: Liu, Jiayi, et al.
Publicado: (2025)
Towards Predicting Any Human Trajectory In Context
por: Fujii, Ryo, et al.
Publicado: (2025)
por: Fujii, Ryo, et al.
Publicado: (2025)
PhotoBot: Reference-Guided Interactive Photography via Natural Language
por: Limoyo, Oliver, et al.
Publicado: (2024)
por: Limoyo, Oliver, et al.
Publicado: (2024)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
por: Xiong, Chuyan, et al.
Publicado: (2024)
por: Xiong, Chuyan, et al.
Publicado: (2024)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
por: Wang, Ziyi, et al.
Publicado: (2025)
por: Wang, Ziyi, et al.
Publicado: (2025)
World Models for Learning Dexterous Hand-Object Interactions from Human Videos
por: Goswami, Raktim Gautam, et al.
Publicado: (2025)
por: Goswami, Raktim Gautam, et al.
Publicado: (2025)
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
por: Zhu, He, et al.
Publicado: (2025)
por: Zhu, He, et al.
Publicado: (2025)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
por: Chen, Yandu, et al.
Publicado: (2025)
por: Chen, Yandu, et al.
Publicado: (2025)
PhysHanDI: Physics-Based Reconstruction of Hand-Deformable Object Interactions
por: Lee, Jihyun, et al.
Publicado: (2026)
por: Lee, Jihyun, et al.
Publicado: (2026)
VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections
por: Kasaei, Hamidreza, et al.
Publicado: (2024)
por: Kasaei, Hamidreza, et al.
Publicado: (2024)
Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions
por: Li, Heng, et al.
Publicado: (2024)
por: Li, Heng, et al.
Publicado: (2024)
An Efficient LiDAR-Camera Fusion Network for Multi-Class 3D Dynamic Object Detection and Trajectory Prediction
por: He, Yushen, et al.
Publicado: (2025)
por: He, Yushen, et al.
Publicado: (2025)
To Move or Not to Move: Constraint-based Planning Enables Zero-Shot Generalization for Interactive Navigation
por: Vashisth, Apoorva, et al.
Publicado: (2026)
por: Vashisth, Apoorva, et al.
Publicado: (2026)
Gaze-Guided 3D Hand Motion Prediction for Detecting Intent in Egocentric Grasping Tasks
por: He, Yufei, et al.
Publicado: (2025)
por: He, Yufei, et al.
Publicado: (2025)
ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction
por: Wang, Qineng, et al.
Publicado: (2025)
por: Wang, Qineng, et al.
Publicado: (2025)
FunGraph: Functionality Aware 3D Scene Graphs for Language-Prompted Scene Interaction
por: Rotondi, Dennis, et al.
Publicado: (2025)
por: Rotondi, Dennis, et al.
Publicado: (2025)
Conditional Unscented Autoencoders for Trajectory Prediction
por: Janjoš, Faris, et al.
Publicado: (2023)
por: Janjoš, Faris, et al.
Publicado: (2023)
Li-ViP3D++: Query-Gated Deformable Camera-LiDAR Fusion for End-to-End Perception and Trajectory Prediction
por: Halinkovic, Matej, et al.
Publicado: (2026)
por: Halinkovic, Matej, et al.
Publicado: (2026)
SynHLMA:Synthesizing Hand Language Manipulation for Articulated Object with Discrete Human Object Interaction Representation
por: zhi, Wang, et al.
Publicado: (2025)
por: zhi, Wang, et al.
Publicado: (2025)
Toward Reliable AR-Guided Surgical Navigation: Interactive Deformation Modeling with Data-Driven Biomechanics and Prompts
por: Han, Zheng, et al.
Publicado: (2025)
por: Han, Zheng, et al.
Publicado: (2025)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
por: Zeng, Huajian, et al.
Publicado: (2026)
por: Zeng, Huajian, et al.
Publicado: (2026)
Progressive Pretext Task Learning for Human Trajectory Prediction
por: Lin, Xiaotong, et al.
Publicado: (2024)
por: Lin, Xiaotong, et al.
Publicado: (2024)
Work Zones challenge VLM Trajectory Planning: Toward Mitigation and Robust Autonomous Driving
por: Liao, Yifan, et al.
Publicado: (2025)
por: Liao, Yifan, et al.
Publicado: (2025)
Ejemplares similares
-
Risk-aware Trajectory Prediction by Incorporating Spatio-temporal Traffic Interaction Analysis
por: Thuremella, Divya, et al.
Publicado: (2024) -
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
por: Chen, Mingfei, et al.
Publicado: (2025) -
Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition
por: Wen, Yuhang, et al.
Publicado: (2023) -
Learning the Pedestrian-Vehicle Interaction for Pedestrian Trajectory Prediction
por: Zhang, Chi, et al.
Publicado: (2022) -
MFSeg: Efficient Multi-frame 3D Semantic Segmentation
por: Huang, Chengjie, et al.
Publicado: (2025)