Towards Generative Predictive Display for Vision-Based Teleoperation: A Zero-Shot Benchmark of Off-the-Shelf Video Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Khalil, Aws, Kwon, Jaerock |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CARIL: Confidence-Aware Regression in Imitation Learning for Autonomous Driving
di: Delavari, Elahe, et al.
Pubblicazione: (2025)
di: Delavari, Elahe, et al.
Pubblicazione: (2025)
Nonlinear Performance Degradation of Vision-Based Teleoperation under Network Latency
di: Khalil, Aws, et al.
Pubblicazione: (2026)
di: Khalil, Aws, et al.
Pubblicazione: (2026)
PLM-Net: Perception Latency Mitigation Network for Vision-Based Lateral Control of Autonomous Vehicles
di: Khalil, Aws, et al.
Pubblicazione: (2024)
di: Khalil, Aws, et al.
Pubblicazione: (2024)
Towards Real-Time Generation of Delay-Compensated Video Feeds for Outdoor Mobile Robot Teleoperation
di: Chakraborty, Neeloy, et al.
Pubblicazione: (2024)
di: Chakraborty, Neeloy, et al.
Pubblicazione: (2024)
DriveVA: Video Action Models are Zero-Shot Drivers
di: Liu, Mengmeng, et al.
Pubblicazione: (2026)
di: Liu, Mengmeng, et al.
Pubblicazione: (2026)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
di: Wang, Wen, et al.
Pubblicazione: (2023)
di: Wang, Wen, et al.
Pubblicazione: (2023)
Zero-Shot 3D Visual Grounding from Vision-Language Models
di: Li, Rong, et al.
Pubblicazione: (2025)
di: Li, Rong, et al.
Pubblicazione: (2025)
ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models
di: Zhang, Ying, et al.
Pubblicazione: (2025)
di: Zhang, Ying, et al.
Pubblicazione: (2025)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
di: Chen, Kehan, et al.
Pubblicazione: (2024)
di: Chen, Kehan, et al.
Pubblicazione: (2024)
AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System
di: Qin, Yuzhe, et al.
Pubblicazione: (2023)
di: Qin, Yuzhe, et al.
Pubblicazione: (2023)
Zero4D: Training-Free 4D Video Generation From Single Video Using Off-the-Shelf Video Diffusion
di: Park, Jangho, et al.
Pubblicazione: (2025)
di: Park, Jangho, et al.
Pubblicazione: (2025)
High-Speed Vision Improves Zero-Shot Semantic Understanding of Human Actions
di: Cao, Yongpeng, et al.
Pubblicazione: (2026)
di: Cao, Yongpeng, et al.
Pubblicazione: (2026)
Improving Zero-Shot ObjectNav with Generative Communication
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2024)
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2024)
Zero-Splat TeleAssist: A Zero-Shot Pose Estimation Framework for Semantic Teleoperation
di: Dokania, Srijan, et al.
Pubblicazione: (2025)
di: Dokania, Srijan, et al.
Pubblicazione: (2025)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
di: Batra, Sumeet, et al.
Pubblicazione: (2024)
di: Batra, Sumeet, et al.
Pubblicazione: (2024)
Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation
di: Zheng, Wanrong, et al.
Pubblicazione: (2026)
di: Zheng, Wanrong, et al.
Pubblicazione: (2026)
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
NDST: Neural Driving Style Transfer for Human-Like Vision-Based Autonomous Driving
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
di: Chen, Jiahe, et al.
Pubblicazione: (2026)
di: Chen, Jiahe, et al.
Pubblicazione: (2026)
ZeroSCD: Zero-Shot Street Scene Change Detection
di: Kannan, Shyam Sundar, et al.
Pubblicazione: (2024)
di: Kannan, Shyam Sundar, et al.
Pubblicazione: (2024)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
di: Zhang, Erhang, et al.
Pubblicazione: (2025)
di: Zhang, Erhang, et al.
Pubblicazione: (2025)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
Towards Zero-Shot Point Cloud Registration Across Diverse Scales, Scenes, and Sensor Setups
di: Lim, Hyungtae, et al.
Pubblicazione: (2026)
di: Lim, Hyungtae, et al.
Pubblicazione: (2026)
NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
di: Li, Hongyu, et al.
Pubblicazione: (2025)
di: Li, Hongyu, et al.
Pubblicazione: (2025)
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
di: Kuang, Yuxuan, et al.
Pubblicazione: (2024)
di: Kuang, Yuxuan, et al.
Pubblicazione: (2024)
ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
di: Iwase, Shun, et al.
Pubblicazione: (2025)
di: Iwase, Shun, et al.
Pubblicazione: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
di: Lyu, Kailin, et al.
Pubblicazione: (2026)
di: Lyu, Kailin, et al.
Pubblicazione: (2026)
Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology
di: Wang, Xiangyu, et al.
Pubblicazione: (2024)
di: Wang, Xiangyu, et al.
Pubblicazione: (2024)
Dream2Real: Zero-Shot 3D Object Rearrangement with Vision-Language Models
di: Kapelyukh, Ivan, et al.
Pubblicazione: (2023)
di: Kapelyukh, Ivan, et al.
Pubblicazione: (2023)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning
di: Zhang, Zixin, et al.
Pubblicazione: (2025)
di: Zhang, Zixin, et al.
Pubblicazione: (2025)
VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation
di: Chen, Hanzhi, et al.
Pubblicazione: (2025)
di: Chen, Hanzhi, et al.
Pubblicazione: (2025)
NavigateDiff: Visual Predictors are Zero-Shot Navigation Assistants
di: Qin, Yiran, et al.
Pubblicazione: (2025)
di: Qin, Yiran, et al.
Pubblicazione: (2025)
TeleOpBench: A Simulator-Centric Benchmark for Dual-Arm Dexterous Teleoperation
di: Li, Hangyu, et al.
Pubblicazione: (2025)
di: Li, Hangyu, et al.
Pubblicazione: (2025)
WildOcc: A Benchmark for Off-Road 3D Semantic Occupancy Prediction
di: Zhai, Heng, et al.
Pubblicazione: (2024)
di: Zhai, Heng, et al.
Pubblicazione: (2024)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
di: Garcia, Ricardo, et al.
Pubblicazione: (2024)
di: Garcia, Ricardo, et al.
Pubblicazione: (2024)
VBR: A Vision Benchmark in Rome
di: Brizi, Leonardo, et al.
Pubblicazione: (2024)
di: Brizi, Leonardo, et al.
Pubblicazione: (2024)
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
di: Huynh, Cuong, et al.
Pubblicazione: (2026)
di: Huynh, Cuong, et al.
Pubblicazione: (2026)
Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning
di: Ding, Runyu, et al.
Pubblicazione: (2024)
di: Ding, Runyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CARIL: Confidence-Aware Regression in Imitation Learning for Autonomous Driving
di: Delavari, Elahe, et al.
Pubblicazione: (2025) -
Nonlinear Performance Degradation of Vision-Based Teleoperation under Network Latency
di: Khalil, Aws, et al.
Pubblicazione: (2026) -
PLM-Net: Perception Latency Mitigation Network for Vision-Based Lateral Control of Autonomous Vehicles
di: Khalil, Aws, et al.
Pubblicazione: (2024) -
Towards Real-Time Generation of Delay-Compensated Video Feeds for Outdoor Mobile Robot Teleoperation
di: Chakraborty, Neeloy, et al.
Pubblicazione: (2024) -
DriveVA: Video Action Models are Zero-Shot Drivers
di: Liu, Mengmeng, et al.
Pubblicazione: (2026)