SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
Fuente:
arXiv
Saved in:
| Main Authors: | He, Ziheng, Chen, Yixiang, Yang, Ning, Wu, Zhanqian, Ma, Qisen, Xu, Yuan, Yang, Jiabing, Li, Peiyan, Wu, Xiangnan, Wang, Xiaofeng, Zhu, Zheng, Liu, Jing, Liu, Nianfeng, Huang, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
by: Chen, Yixiang, et al.
Published: (2026)
by: Chen, Yixiang, et al.
Published: (2026)
EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation
by: Xu, Yuan, et al.
Published: (2025)
by: Xu, Yuan, et al.
Published: (2025)
UAOR: Uncertainty-aware Observation Reinjection for Vision-Language-Action Models
by: Yang, Jiabing, et al.
Published: (2026)
by: Yang, Jiabing, et al.
Published: (2026)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow
by: Chen, Yixiang, et al.
Published: (2025)
by: Chen, Yixiang, et al.
Published: (2025)
BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models
by: Li, Peiyan, et al.
Published: (2025)
by: Li, Peiyan, et al.
Published: (2025)
FloorPlan-VLN: A New Paradigm for Floor Plan Guided Vision-Language Navigation
by: Chen, Kehan, et al.
Published: (2026)
by: Chen, Kehan, et al.
Published: (2026)
AIR-Embodied: An Efficient Active 3DGS-based Interaction and Reconstruction Framework with Embodied Large Language Model
by: Qi, Zhenghao, et al.
Published: (2024)
by: Qi, Zhenghao, et al.
Published: (2024)
Imaginative World Modeling with Scene Graphs for Embodied Agent Navigation
by: Hu, Yue, et al.
Published: (2025)
by: Hu, Yue, et al.
Published: (2025)
VERM: Leveraging Foundation Models to Create a Virtual Eye for Efficient 3D Robotic Manipulation
by: Chen, Yixiang, et al.
Published: (2025)
by: Chen, Yixiang, et al.
Published: (2025)
Enhanced Multi-Robot SLAM System with Cross-Validation Matching and Exponential Threshold Keyframe Selection
by: He, Ang, et al.
Published: (2024)
by: He, Ang, et al.
Published: (2024)
EgoFSD: Ego-Centric Fully Sparse Paradigm with Uncertainty Denoising and Iterative Refinement for Efficient End-to-End Self-Driving
by: Su, Haisheng, et al.
Published: (2024)
by: Su, Haisheng, et al.
Published: (2024)
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
by: Wang, Boyuan, et al.
Published: (2025)
by: Wang, Boyuan, et al.
Published: (2025)
EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence
by: Wang, Xinjie, et al.
Published: (2025)
by: Wang, Xinjie, et al.
Published: (2025)
Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models
by: Xu, Haiweng, et al.
Published: (2026)
by: Xu, Haiweng, et al.
Published: (2026)
World-Gymnast: Training Robots with Reinforcement Learning in a World Model
by: Sharma, Ansh Kumar, et al.
Published: (2026)
by: Sharma, Ansh Kumar, et al.
Published: (2026)
Detecting Non-Optimal Decisions of Embodied Agents via Diversity-Guided Metamorphic Testing
by: Wu, Wenzhao, et al.
Published: (2025)
by: Wu, Wenzhao, et al.
Published: (2025)
Leveraging Large Language Model for Heterogeneous Ad Hoc Teamwork Collaboration
by: Liu, Xinzhu, et al.
Published: (2024)
by: Liu, Xinzhu, et al.
Published: (2024)
Astra: Efficient Transformer Architecture and Contrastive Dynamics Learning for Embodied Instruction Following
by: Ma, Yueen, et al.
Published: (2024)
by: Ma, Yueen, et al.
Published: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
by: Bo, Zitong, et al.
Published: (2025)
by: Bo, Zitong, et al.
Published: (2025)
MS-Mapping: Multi-session LiDAR Mapping with Wasserstein-based Keyframe Selection
by: Hu, Xiangcheng, et al.
Published: (2024)
by: Hu, Xiangcheng, et al.
Published: (2024)
An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation
by: Li, Dongjiang, et al.
Published: (2025)
by: Li, Dongjiang, et al.
Published: (2025)
Retrieval-Augmented Embodied Agents
by: Zhu, Yichen, et al.
Published: (2024)
by: Zhu, Yichen, et al.
Published: (2024)
Fast ECoT: Efficient Embodied Chain-of-Thought via Thoughts Reuse
by: Duan, Zhekai, et al.
Published: (2025)
by: Duan, Zhekai, et al.
Published: (2025)
Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction
by: Wang, Xianhao, et al.
Published: (2026)
by: Wang, Xianhao, et al.
Published: (2026)
Design and Control of an Ultra-Slender Push-Pull Multisection Continuum Manipulator for In-Situ Inspection of Aeroengine
by: Zhong, Weiheng, et al.
Published: (2024)
by: Zhong, Weiheng, et al.
Published: (2024)
Embodied Escaping: End-to-End Reinforcement Learning for Robot Navigation in Narrow Environment
by: Zheng, Han, et al.
Published: (2025)
by: Zheng, Han, et al.
Published: (2025)
Efficient End-to-End 6-Dof Grasp Detection Framework for Edge Devices with Hierarchical Heatmaps and Feature Propagation
by: Yang, Kaiqin, et al.
Published: (2024)
by: Yang, Kaiqin, et al.
Published: (2024)
C-NAV: Towards Self-Evolving Continual Object Navigation in Open World
by: Yu, Ming-Ming, et al.
Published: (2025)
by: Yu, Ming-Ming, et al.
Published: (2025)
Modeling the Mental World for Embodied AI: A Comprehensive Review
by: Liu, Biyuan, et al.
Published: (2025)
by: Liu, Biyuan, et al.
Published: (2025)
PhysiAgent: An Embodied Agent Framework in Physical World
by: Wang, Zhihao, et al.
Published: (2025)
by: Wang, Zhihao, et al.
Published: (2025)
A Survey on Robotics with Foundation Models: toward Embodied AI
by: Xu, Zhiyuan, et al.
Published: (2024)
by: Xu, Zhiyuan, et al.
Published: (2024)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
by: GigaWorld Team, et al.
Published: (2025)
by: GigaWorld Team, et al.
Published: (2025)
Submodular Optimization for Keyframe Selection & Usage in SLAM
by: Thorne, David, et al.
Published: (2024)
by: Thorne, David, et al.
Published: (2024)
EHC-MM: Embodied Holistic Control for Mobile Manipulation
by: Wang, Jiawen, et al.
Published: (2024)
by: Wang, Jiawen, et al.
Published: (2024)
WorldGym: World Model as An Environment for Policy Evaluation
by: Quevedo, Julian, et al.
Published: (2025)
by: Quevedo, Julian, et al.
Published: (2025)
Embodied Navigation Foundation Model
by: Zhang, Jiazhao, et al.
Published: (2025)
by: Zhang, Jiazhao, et al.
Published: (2025)
DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation
by: Shan, Ziyu, et al.
Published: (2026)
by: Shan, Ziyu, et al.
Published: (2026)
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
by: Chen, Yipeng, et al.
Published: (2026)
by: Chen, Yipeng, et al.
Published: (2026)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
Similar Items
-
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
by: Chen, Yixiang, et al.
Published: (2026) -
EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation
by: Xu, Yuan, et al.
Published: (2025) -
UAOR: Uncertainty-aware Observation Reinjection for Vision-Language-Action Models
by: Yang, Jiabing, et al.
Published: (2026) -
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026) -
EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow
by: Chen, Yixiang, et al.
Published: (2025)