VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic Routing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yixiao, Huo, Mingxiao, Liang, Zhixuan, Du, Yushi, Sun, Lingfeng, Lin, Haotian, Shang, Jinghuan, Peng, Chensheng, Bansal, Mohit, Ding, Mingyu, Tomizuka, Masayoshi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Pedestrian Trajectory Prediction through Posterior Sampling
by: Lin, Haotian, et al.
Published: (2024)
by: Lin, Haotian, et al.
Published: (2024)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
by: Shang, Jinghuan, et al.
Published: (2024)
by: Shang, Jinghuan, et al.
Published: (2024)
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
by: Liang, Zhixuan, et al.
Published: (2023)
by: Liang, Zhixuan, et al.
Published: (2023)
Composition Vision-Language Understanding via Segment and Depth Anything Model
by: Huo, Mingxiao, et al.
Published: (2024)
by: Huo, Mingxiao, et al.
Published: (2024)
DexHandDiff: Interaction-aware Diffusion Planning for Adaptive Dexterous Manipulation
by: Liang, Zhixuan, et al.
Published: (2024)
by: Liang, Zhixuan, et al.
Published: (2024)
Nonparametric Inverse Dynamic Models for Multimodal Interactive Robots
by: Haninger, Kevin, et al.
Published: (2019)
by: Haninger, Kevin, et al.
Published: (2019)
Physics-Aware Robotic Palletization with Online Masking Inference
by: Zhang, Tianqi, et al.
Published: (2025)
by: Zhang, Tianqi, et al.
Published: (2025)
RT-GS: Gaussian Splatting with Reflection and Transmittance Primitives
by: Zeng, Kunnong, et al.
Published: (2026)
by: Zeng, Kunnong, et al.
Published: (2026)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
by: Yuan, Puzhen, et al.
Published: (2025)
by: Yuan, Puzhen, et al.
Published: (2025)
Language-Driven Policy Distillation for Cooperative Driving in Multi-Agent Reinforcement Learning
by: Liu, Jiaqi, et al.
Published: (2024)
by: Liu, Jiaqi, et al.
Published: (2024)
Optimizing Diffusion Models for Joint Trajectory Prediction and Controllable Generation
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
CompGS: Unleashing 2D Compositionality for Compositional Text-to-3D via Dynamically Optimizing 3D Gaussians
by: Ge, Chongjian, et al.
Published: (2024)
by: Ge, Chongjian, et al.
Published: (2024)
Q-SLAM: Quadric Representations for Monocular SLAM
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
X-Drive: Cross-modality consistent multi-sensor data synthesis for driving scenarios
by: Xie, Yichen, et al.
Published: (2024)
by: Xie, Yichen, et al.
Published: (2024)
Depth-aware Volume Attention for Texture-less Stereo Matching
by: Zhao, Tong, et al.
Published: (2024)
by: Zhao, Tong, et al.
Published: (2024)
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
by: Guo, Dingkun, et al.
Published: (2024)
by: Guo, Dingkun, et al.
Published: (2024)
Imagined Potential Games: A Framework for Simulating, Learning and Evaluating Interactive Behaviors
by: Sun, Lingfeng, et al.
Published: (2024)
by: Sun, Lingfeng, et al.
Published: (2024)
Rethinking Image-to-3D Generation with Sparse Queries: Efficiency, Capacity, and Input-View Bias
by: Xu, Zhiyuan, et al.
Published: (2026)
by: Xu, Zhiyuan, et al.
Published: (2026)
DeSiRe-GS: 4D Street Gaussians for Static-Dynamic Decomposition and Surface Reconstruction for Urban Driving Scenes
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
DADP: Domain Adaptive Diffusion Policy
by: Wang, Pengcheng, et al.
Published: (2026)
by: Wang, Pengcheng, et al.
Published: (2026)
Interleave-VLA: Enhancing Robot Manipulation with Interleaved Image-Text Instructions
by: Fan, Cunxin, et al.
Published: (2025)
by: Fan, Cunxin, et al.
Published: (2025)
Adaptive Linear Path Model-Based Diffusion
by: Shimizu, Yutaka, et al.
Published: (2026)
by: Shimizu, Yutaka, et al.
Published: (2026)
Algebraic Control: Complete Stable Inversion with Necessary and Sufficient Conditions
by: Kürkçü, Burak, et al.
Published: (2024)
by: Kürkçü, Burak, et al.
Published: (2024)
Leveraging Extrinsic Dexterity for Occluded Grasping on Grasp Constraining Walls
by: Kobashi, Keita, et al.
Published: (2025)
by: Kobashi, Keita, et al.
Published: (2025)
Bisimulation metric for Model Predictive Control
by: Shimizu, Yutaka, et al.
Published: (2024)
by: Shimizu, Yutaka, et al.
Published: (2024)
DexH2R: Task-oriented Dexterous Manipulation from Human to Robots
by: Zhao, Shuqi, et al.
Published: (2024)
by: Zhao, Shuqi, et al.
Published: (2024)
P2 Explore: Efficient Exploration in Unknown Cluttered Environment with Floor Plan Prediction
by: Song, Kun, et al.
Published: (2024)
by: Song, Kun, et al.
Published: (2024)
DrPlanner: Diagnosis and Repair of Motion Planners for Automated Vehicles Using Large Language Models
by: Lin, Yuanfei, et al.
Published: (2024)
by: Lin, Yuanfei, et al.
Published: (2024)
RoadBEV: Road Surface Reconstruction in Bird's Eye View
by: Zhao, Tong, et al.
Published: (2024)
by: Zhao, Tong, et al.
Published: (2024)
URoPE: Universal Relative Position Embedding across Geometric Spaces
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
by: Tang, Weiliang, et al.
Published: (2025)
by: Tang, Weiliang, et al.
Published: (2025)
Adaptive Energy Regularization for Autonomous Gait Transition and Energy-Efficient Quadruped Locomotion
by: Liang, Boyuan, et al.
Published: (2024)
by: Liang, Boyuan, et al.
Published: (2024)
Fractional-order Modeling for Nonlinear Soft Actuators via Particle Swarm Optimization
by: Yang, Wu-Te, et al.
Published: (2025)
by: Yang, Wu-Te, et al.
Published: (2025)
EC-DIT: Scaling Diffusion Transformers with Adaptive Expert-Choice Routing
by: Sun, Haotian, et al.
Published: (2024)
by: Sun, Haotian, et al.
Published: (2024)
DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors
by: Wang, Pengcheng, et al.
Published: (2026)
by: Wang, Pengcheng, et al.
Published: (2026)
Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills
by: Chen, Justin Chih-Yao, et al.
Published: (2025)
by: Chen, Justin Chih-Yao, et al.
Published: (2025)
Index-Aligned Query Distillation for Transformer-based Incremental Object Detection
by: Ma, Mingxiao, et al.
Published: (2025)
by: Ma, Mingxiao, et al.
Published: (2025)
Similar Items
-
Joint Pedestrian Trajectory Prediction through Posterior Sampling
by: Lin, Haotian, et al.
Published: (2024) -
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
by: Wang, Yixiao, et al.
Published: (2024) -
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
by: Shang, Jinghuan, et al.
Published: (2024) -
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
by: Liang, Zhixuan, et al.
Published: (2023) -
Composition Vision-Language Understanding via Segment and Depth Anything Model
by: Huo, Mingxiao, et al.
Published: (2024)