Video Generation Models in Robotics -- Applications, Research Challenges, Future Directions
Fuente:
arXiv
Saved in:
| Main Authors: | Mei, Zhiting, Yin, Tenny, Shorinwa, Ola, Badithela, Apurva, Zheng, Zhonghe, Bruno, Joseph, Bland, Madison, Zha, Lihan, Hancock, Asher, Fisac, Jaime Fernández, Dames, Philip, Majumdar, Anirudha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PlayWorld: Learning Robot World Models from Autonomous Play
by: Yin, Tenny, et al.
Published: (2026)
by: Yin, Tenny, et al.
Published: (2026)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
by: Yin, Tenny, et al.
Published: (2025)
by: Yin, Tenny, et al.
Published: (2025)
Reliable and Scalable Robot Policy Evaluation with Imperfect Simulators
by: Badithela, Apurva, et al.
Published: (2025)
by: Badithela, Apurva, et al.
Published: (2025)
Actions as Language: Fine-Tuning VLMs into VLAs Without Catastrophic Forgetting
by: Hancock, Asher J., et al.
Published: (2025)
by: Hancock, Asher J., et al.
Published: (2025)
LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer
by: Zha, Lihan, et al.
Published: (2026)
by: Zha, Lihan, et al.
Published: (2026)
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
by: Shorinwa, Ola, et al.
Published: (2025)
by: Shorinwa, Ola, et al.
Published: (2025)
Is Your Imitation Learning Policy Better than Mine? Policy Comparison with Near-Optimal Stopping
by: Snyder, David, et al.
Published: (2025)
by: Snyder, David, et al.
Published: (2025)
Beyond Binary Success: Sample-Efficient and Statistically Rigorous Robot Policy Comparison
by: Snyder, David, et al.
Published: (2026)
by: Snyder, David, et al.
Published: (2026)
Guiding Data Collection via Factored Scaling Curves
by: Zha, Lihan, et al.
Published: (2025)
by: Zha, Lihan, et al.
Published: (2025)
Run-time Observation Interventions Make Vision-Language-Action Models More Visually Robust
by: Hancock, Asher J., et al.
Published: (2024)
by: Hancock, Asher J., et al.
Published: (2024)
Perceive With Confidence: Statistical Safety Assurances for Navigation with Learning-Based Perception
by: Mei, Zhiting, et al.
Published: (2024)
by: Mei, Zhiting, et al.
Published: (2024)
Distributed Conjugate Gradient Method via Conjugate Direction Tracking
by: Shorinwa, Ola, et al.
Published: (2023)
by: Shorinwa, Ola, et al.
Published: (2023)
Splat-Nav: Safe Real-Time Robot Navigation in Gaussian Splatting Maps
by: Chen, Timothy, et al.
Published: (2024)
by: Chen, Timothy, et al.
Published: (2024)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
by: Li, Huiqiong, et al.
Published: (2026)
by: Li, Huiqiong, et al.
Published: (2026)
Distributed Quasi-Newton Method for Multi-Agent Optimization
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
Deceptive Risk Minimization: Out-of-Distribution Generalization by Deceiving Distribution Shift Detectors
by: Majumdar, Anirudha
Published: (2025)
by: Majumdar, Anirudha
Published: (2025)
Distributed Optimization Methods for Multi-Robot Systems: Part I -- A Tutorial
by: Shorinwa, Ola, et al.
Published: (2023)
by: Shorinwa, Ola, et al.
Published: (2023)
Distributed Optimization Methods for Multi-Robot Systems: Part II -- A Survey
by: Shorinwa, Ola, et al.
Published: (2023)
by: Shorinwa, Ola, et al.
Published: (2023)
From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails
by: Pandya, Ravi, et al.
Published: (2025)
by: Pandya, Ravi, et al.
Published: (2025)
SciFi-Benchmark: Leveraging Science Fiction To Improve Robot Behavior
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
DisCo: Distributed Contact-Rich Trajectory Optimization for Forceful Multi-Robot Collaboration
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
FAST-Splat: Fast, Ambiguity-Free Semantics Transfer in Gaussian Splatting
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
by: Feng, Bowen, et al.
Published: (2025)
by: Feng, Bowen, et al.
Published: (2025)
Safe, Out-of-Distribution-Adaptive MPC with Conformalized Neural Network Ensembles
by: Contreras, Jose Leopoldo, et al.
Published: (2024)
by: Contreras, Jose Leopoldo, et al.
Published: (2024)
Introspective Planning: Aligning Robots' Uncertainty with Inherent Task Ambiguity
by: Liang, Kaiqu, et al.
Published: (2024)
by: Liang, Kaiqu, et al.
Published: (2024)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
A Kinetic Theory of Encounter-Based Information Propagation in Multi-Robot Systems
by: Srivastava, Alkesh K., et al.
Published: (2026)
by: Srivastava, Alkesh K., et al.
Published: (2026)
Community level research on the potential for social enterprise development in three targeted communities in the Nelson Mandela Metropole
by: Ricardo Dames
Published: (2010)
by: Ricardo Dames
Published: (2010)
SINGER: An Onboard Generalist Vision-Language Navigation Policy for Drones
by: Adang, Maximilian, et al.
Published: (2025)
by: Adang, Maximilian, et al.
Published: (2025)
Risk-Calibrated Human-Robot Interaction via Set-Valued Intent Prediction
by: Lidard, Justin, et al.
Published: (2024)
by: Lidard, Justin, et al.
Published: (2024)
Towards Predicting Collective Performance in Multi-Robot Teams
by: Xin, Pujie, et al.
Published: (2024)
by: Xin, Pujie, et al.
Published: (2024)
Online Resynthesis of High-Level Collaborative Tasks for Robots with Changing Capabilities
by: Fang, Amy, et al.
Published: (2024)
by: Fang, Amy, et al.
Published: (2024)
Consistency in Language Models: Current Landscape, Challenges, and Future Directions
by: Novikova, Jekaterina, et al.
Published: (2025)
by: Novikova, Jekaterina, et al.
Published: (2025)
μ‐Transcranial Alternating Current Stimulation Induces Phasic Entrainment and Plastic Facilitation of Corticospinal Excitability
by: Asher Geffen, et al.
Published: (2025)
by: Asher Geffen, et al.
Published: (2025)
Distributed Multi-Robot Multi-Target Simultaneous Search and Tracking in an Unknown Non-convex Environment
by: Chen, Jun, et al.
Published: (2025)
by: Chen, Jun, et al.
Published: (2025)
Similar Items
-
PlayWorld: Learning Robot World Models from Autonomous Play
by: Yin, Tenny, et al.
Published: (2026) -
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025) -
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions
by: Shorinwa, Ola, et al.
Published: (2024)