Scalable Policy Evaluation with Video World Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tseng, Wei-Cheng, Gu, Jinwei, Zhang, Qinsheng, Mao, Hanzi, Liu, Ming-Yu, Shkurti, Florian, Yen-Chen, Lin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaussian Splatting Visual MPC for Granular Media Manipulation
by: Tseng, Wei-Cheng, et al.
Published: (2024)
by: Tseng, Wei-Cheng, et al.
Published: (2024)
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
by: Kim, Moo Jin, et al.
Published: (2026)
by: Kim, Moo Jin, et al.
Published: (2026)
What Do You Need for Compositional Generalization in Diffusion Planning?
by: Clark, Quentin, et al.
Published: (2025)
by: Clark, Quentin, et al.
Published: (2025)
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
by: Goli, Hossein, et al.
Published: (2025)
by: Goli, Hossein, et al.
Published: (2025)
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
Does Unpredictability Influence Driving Behavior?
by: Samavi, Sepehr, et al.
Published: (2023)
by: Samavi, Sepehr, et al.
Published: (2023)
RISE: Self-Improving Robot Policy with Compositional World Model
by: Yang, Jiazhi, et al.
Published: (2026)
by: Yang, Jiazhi, et al.
Published: (2026)
SICNav: Safe and Interactive Crowd Navigation using Model Predictive Control and Bilevel Optimization
by: Samavi, Sepehr, et al.
Published: (2023)
by: Samavi, Sepehr, et al.
Published: (2023)
DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion
by: Fan, Yahao, et al.
Published: (2025)
by: Fan, Yahao, et al.
Published: (2025)
Continual Model-Based Reinforcement Learning with Hypernetworks
by: Huang, Yizhou, et al.
Published: (2020)
by: Huang, Yizhou, et al.
Published: (2020)
Deploying SICNav in the Field: Safe and Interactive Crowd Navigation using MPC and Bilevel Optimization
by: Samavi, Sepehr, et al.
Published: (2025)
by: Samavi, Sepehr, et al.
Published: (2025)
Field Testing of a Stochastic Planner for ASV Navigation Using Satellite Images
by: Huang, Philip, et al.
Published: (2023)
by: Huang, Philip, et al.
Published: (2023)
SAFE: Multitask Failure Detection for Vision-Language-Action Models
by: Gu, Qiao, et al.
Published: (2025)
by: Gu, Qiao, et al.
Published: (2025)
Diversity You Can Actually Measure: A Fast, Model-Free Diversity Metric for Robotics Datasets
by: Sirigiri, Sreevardhan, et al.
Published: (2026)
by: Sirigiri, Sreevardhan, et al.
Published: (2026)
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
by: Zheng, Kaiwen, et al.
Published: (2024)
by: Zheng, Kaiwen, et al.
Published: (2024)
Robot Detection System 3: LRF groups and Coordinate System
by: Lin, Jinwei
Published: (2024)
by: Lin, Jinwei
Published: (2024)
Robot Detection System 2: Design of Sensor System
by: Lin, Jinwei
Published: (2024)
by: Lin, Jinwei
Published: (2024)
SICNav-Diffusion: Safe and Interactive Crowd Navigation with Diffusion Trajectory Predictions
by: Samavi, Sepehr, et al.
Published: (2025)
by: Samavi, Sepehr, et al.
Published: (2025)
World Simulation with Video Foundation Models for Physical AI
by: NVIDIA, et al.
Published: (2025)
by: NVIDIA, et al.
Published: (2025)
Generating Transferable Adversarial Simulation Scenarios for Self-Driving via Neural Rendering
by: Abeysirigoonawardena, Yasasa, et al.
Published: (2023)
by: Abeysirigoonawardena, Yasasa, et al.
Published: (2023)
LongBench: Evaluating Robotic Manipulation Policies on Real-World Long-Horizon Tasks
by: Chen, Xueyao, et al.
Published: (2026)
by: Chen, Xueyao, et al.
Published: (2026)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
Automated Planning Domain Inference for Task and Motion Planning
by: Huang, Jinbang, et al.
Published: (2024)
by: Huang, Jinbang, et al.
Published: (2024)
WorldGym: World Model as An Environment for Policy Evaluation
by: Quevedo, Julian, et al.
Published: (2025)
by: Quevedo, Julian, et al.
Published: (2025)
Learning Primitive Embodied World Models: Towards Scalable Robotic Learning
by: Sun, Qiao, et al.
Published: (2025)
by: Sun, Qiao, et al.
Published: (2025)
Robot Detection System 1: Front-Following
by: Lin, Jinwei
Published: (2024)
by: Lin, Jinwei
Published: (2024)
World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy
by: Liu, Xiaokang, et al.
Published: (2026)
by: Liu, Xiaokang, et al.
Published: (2026)
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
by: Gao, Shenyuan, et al.
Published: (2026)
by: Gao, Shenyuan, et al.
Published: (2026)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
by: Jain, Arhan, et al.
Published: (2025)
by: Jain, Arhan, et al.
Published: (2025)
Expert Composer Policy: Scalable Skill Repertoire for Quadruped Robots
by: Christmann, Guilherme, et al.
Published: (2024)
by: Christmann, Guilherme, et al.
Published: (2024)
Scaling World Model for Hierarchical Manipulation Policies
by: Long, Qian, et al.
Published: (2026)
by: Long, Qian, et al.
Published: (2026)
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
by: Jiang, Zhennan, et al.
Published: (2026)
by: Jiang, Zhennan, et al.
Published: (2026)
iVideoGPT: Interactive VideoGPTs are Scalable World Models
by: Wu, Jialong, et al.
Published: (2024)
by: Wu, Jialong, et al.
Published: (2024)
UniGraspTransformer: Simplified Policy Distillation for Scalable Dexterous Robotic Grasping
by: Wang, Wenbo, et al.
Published: (2024)
by: Wang, Wenbo, et al.
Published: (2024)
WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems
by: Wang, Yuchen, et al.
Published: (2026)
by: Wang, Yuchen, et al.
Published: (2026)
ISS Policy : Scalable Diffusion Policy with Implicit Scene Supervision
by: Xia, Wenlong, et al.
Published: (2025)
by: Xia, Wenlong, et al.
Published: (2025)
RaSCL: Radar to Satellite Crossview Localization
by: Abdullai, Blerim, et al.
Published: (2025)
by: Abdullai, Blerim, et al.
Published: (2025)
Cosmos-H-Surgical: Learning Surgical Robot Policies from Videos via World Modeling
by: He, Yufan, et al.
Published: (2025)
by: He, Yufan, et al.
Published: (2025)
AtomVLA: Scalable Post-Training for Robotic Manipulation via Predictive Latent World Models
by: Sun, Xiaoquan, et al.
Published: (2026)
by: Sun, Xiaoquan, et al.
Published: (2026)
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
by: Hu, Yucheng, et al.
Published: (2024)
by: Hu, Yucheng, et al.
Published: (2024)
Similar Items
-
Gaussian Splatting Visual MPC for Granular Media Manipulation
by: Tseng, Wei-Cheng, et al.
Published: (2024) -
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
by: Kim, Moo Jin, et al.
Published: (2026) -
What Do You Need for Compositional Generalization in Diffusion Planning?
by: Clark, Quentin, et al.
Published: (2025) -
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
by: Goli, Hossein, et al.
Published: (2025) -
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
by: Li, Yaxuan, et al.
Published: (2026)