World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mei, Zhiting, Yin, Tenny, Baker, Micah, Shorinwa, Ola, Majumdar, Anirudha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
von: Shorinwa, Ola, et al.
Veröffentlicht: (2025)
von: Shorinwa, Ola, et al.
Veröffentlicht: (2025)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
von: Wu, Tao, et al.
Veröffentlicht: (2025)
von: Wu, Tao, et al.
Veröffentlicht: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
PlayWorld: Learning Robot World Models from Autonomous Play
von: Yin, Tenny, et al.
Veröffentlicht: (2026)
von: Yin, Tenny, et al.
Veröffentlicht: (2026)
Video Generation Models in Robotics -- Applications, Research Challenges, Future Directions
von: Mei, Zhiting, et al.
Veröffentlicht: (2026)
von: Mei, Zhiting, et al.
Veröffentlicht: (2026)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
von: Feng, Bowen, et al.
Veröffentlicht: (2025)
von: Feng, Bowen, et al.
Veröffentlicht: (2025)
Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding
von: Wu, Xianjin, et al.
Veröffentlicht: (2026)
von: Wu, Xianjin, et al.
Veröffentlicht: (2026)
Objects in Generated Videos Are Slower Than They Appear: Models Suffer Sub-Earth Gravity and Don't Know Galileo's Principle...for now
von: Thozhiyoor, Varun Varma, et al.
Veröffentlicht: (2025)
von: Thozhiyoor, Varun Varma, et al.
Veröffentlicht: (2025)
Show, Don't Tell: Detecting Novel Objects by Watching Human Videos
von: Akl, James, et al.
Veröffentlicht: (2026)
von: Akl, James, et al.
Veröffentlicht: (2026)
Splat-MOVER: Multi-Stage, Open-Vocabulary Robotic Manipulation via Editable Gaussian Splatting
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
Don't Let Your Robot be Harmful: Responsible Robotic Manipulation via Safety-as-Policy
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
"Don't forget to put the milk back!" Dataset for Enabling Embodied Agents to Detect Anomalous Situations
von: Mullen Jr, James F., et al.
Veröffentlicht: (2024)
von: Mullen Jr, James F., et al.
Veröffentlicht: (2024)
Do You Know Where Your Camera Is? View-Invariant Policy Learning with Camera Conditioning
von: Jiang, Tianchong, et al.
Veröffentlicht: (2025)
von: Jiang, Tianchong, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
KnowVal: A Knowledge-Augmented and Value-Guided Autonomous Driving System
von: Xia, Zhongyu, et al.
Veröffentlicht: (2025)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2025)
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
Generating Robot Constitutions & Benchmarks for Semantic Safety
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
Rethinking Video Generation Model for the Embodied World
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
Learning Camera Movement Control from Real-World Drone Videos
von: Hou, Yunzhong, et al.
Veröffentlicht: (2024)
von: Hou, Yunzhong, et al.
Veröffentlicht: (2024)
Camera Calibration via Circular Patterns: A Comprehensive Framework with Detection Uncertainty and Unbiased Projection Model
von: Song, Chaehyeon, et al.
Veröffentlicht: (2025)
von: Song, Chaehyeon, et al.
Veröffentlicht: (2025)
Query2Uncertainty: Robust Uncertainty Quantification and Calibration for 3D Object Detection under Distribution Shift
von: Beemelmanns, Till, et al.
Veröffentlicht: (2026)
von: Beemelmanns, Till, et al.
Veröffentlicht: (2026)
iVideoGPT: Interactive VideoGPTs are Scalable World Models
von: Wu, Jialong, et al.
Veröffentlicht: (2024)
von: Wu, Jialong, et al.
Veröffentlicht: (2024)
Critiques of World Models
von: Xing, Eric, et al.
Veröffentlicht: (2025)
von: Xing, Eric, et al.
Veröffentlicht: (2025)
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
FAST-Splat: Fast, Ambiguity-Free Semantics Transfer in Gaussian Splatting
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
Causal World Modeling for Robot Control
von: Li, Lin, et al.
Veröffentlicht: (2026)
von: Li, Lin, et al.
Veröffentlicht: (2026)
Dreamitate: Real-World Visuomotor Policy Learning via Video Generation
von: Liang, Junbang, et al.
Veröffentlicht: (2024)
von: Liang, Junbang, et al.
Veröffentlicht: (2024)
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
von: Sikar, Daniel, et al.
Veröffentlicht: (2025)
von: Sikar, Daniel, et al.
Veröffentlicht: (2025)
DDP-WM: Disentangled Dynamics Prediction for Efficient World Models
von: Yin, Shicheng, et al.
Veröffentlicht: (2026)
von: Yin, Shicheng, et al.
Veröffentlicht: (2026)
Perceive With Confidence: Statistical Safety Assurances for Navigation with Learning-Based Perception
von: Mei, Zhiting, et al.
Veröffentlicht: (2024)
von: Mei, Zhiting, et al.
Veröffentlicht: (2024)
How Much Do Large Language Models Know about Human Motion? A Case Study in 3D Avatar Control
von: Li, Kunhang, et al.
Veröffentlicht: (2025)
von: Li, Kunhang, et al.
Veröffentlicht: (2025)
World Guidance: World Modeling in Condition Space for Action Generation
von: Su, Yue, et al.
Veröffentlicht: (2026)
von: Su, Yue, et al.
Veröffentlicht: (2026)
World Knowledge from AI Image Generation for Robot Control
von: Krumme, Jonas, et al.
Veröffentlicht: (2025)
von: Krumme, Jonas, et al.
Veröffentlicht: (2025)
RobotSeg: A Model and Dataset for Segmenting Robots in Image and Video
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
von: Mei, Zhiting, et al.
Veröffentlicht: (2025) -
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
von: Mei, Zhiting, et al.
Veröffentlicht: (2025) -
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025) -
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025) -
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
von: Shorinwa, Ola, et al.
Veröffentlicht: (2025)