Enhancing Vision-Based Policies with Omni-View and Cross-Modality Knowledge Distillation for Mobile Robots
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Kai, Zhao, Shiyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Visuomotor Policy for Multi-Robot Laser Tag Game
von: Li, Kai, et al.
Veröffentlicht: (2026)
von: Li, Kai, et al.
Veröffentlicht: (2026)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
Omni Differential Drive for Simultaneous Reconfiguration and Omnidirectional Mobility of Wheeled Robots
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation
von: Huang, Haoran, et al.
Veröffentlicht: (2026)
von: Huang, Haoran, et al.
Veröffentlicht: (2026)
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
von: Song, Yunzhou, et al.
Veröffentlicht: (2026)
von: Song, Yunzhou, et al.
Veröffentlicht: (2026)
LMD-PGN: Cross-Modal Knowledge Distillation from First-Person-View Images to Third-Person-View BEV Maps for Universal Point Goal Navigation
von: Uemura, Riku, et al.
Veröffentlicht: (2024)
von: Uemura, Riku, et al.
Veröffentlicht: (2024)
OmniD: Generalizable Robot Manipulation Policy via Image-Based BEV Representation
von: Mao, Jilei, et al.
Veröffentlicht: (2025)
von: Mao, Jilei, et al.
Veröffentlicht: (2025)
ODD: Omni Differential Drive for Simultaneous Reconfiguration and Omnidirectional Mobility of Wheeled Robots
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
CRKD: Enhanced Camera-Radar Object Detection with Cross-modality Knowledge Distillation
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
Learning Multi-Modal Trajectory Policies for Data-Efficient Robotic Manipulation
von: Chen, Zijia, et al.
Veröffentlicht: (2026)
von: Chen, Zijia, et al.
Veröffentlicht: (2026)
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
Teleoperated Omni-directional Dual Arm Mobile Manipulation Robotic System with Shared Control for Retail Store
von: Lima, Rolif, et al.
Veröffentlicht: (2026)
von: Lima, Rolif, et al.
Veröffentlicht: (2026)
COMPASS: Cross-embodiment Mobility Policy via Residual RL and Skill Synthesis
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
Design and Benchmarking of A Multi-Modality Sensor for Robotic Manipulation with GAN-Based Cross-Modality Interpretation
von: Zhang, Dandan, et al.
Veröffentlicht: (2025)
von: Zhang, Dandan, et al.
Veröffentlicht: (2025)
Collective Behavior Clone with Visual Attention via Neural Interaction Graph Prediction
von: Li, Kai, et al.
Veröffentlicht: (2025)
von: Li, Kai, et al.
Veröffentlicht: (2025)
LCMF: Lightweight Cross-Modality Mambaformer for Embodied Robotics VQA
von: Kang, Zeyi, et al.
Veröffentlicht: (2025)
von: Kang, Zeyi, et al.
Veröffentlicht: (2025)
VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation
von: Zhai, Xuanran, et al.
Veröffentlicht: (2025)
von: Zhai, Xuanran, et al.
Veröffentlicht: (2025)
Efficient Camera Pose Augmentation for View Generalization in Robotic Policy Learning
von: Wang, Sen, et al.
Veröffentlicht: (2026)
von: Wang, Sen, et al.
Veröffentlicht: (2026)
GUIDES: Guidance Using Instructor-Distilled Embeddings for Pre-trained Robot Policy Enhancement
von: Gao, Minquan, et al.
Veröffentlicht: (2025)
von: Gao, Minquan, et al.
Veröffentlicht: (2025)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
Online Robot Navigation and Manipulation with Distilled Vision-Language Models
von: Liu, Kangcheng
Veröffentlicht: (2024)
von: Liu, Kangcheng
Veröffentlicht: (2024)
Mobile Robotic Multi-View Photometric Stereo
von: Kumar, Suryansh
Veröffentlicht: (2025)
von: Kumar, Suryansh
Veröffentlicht: (2025)
Real‐Time Multi‐View Flower Counting With a Ground Mobile Robot
von: Daniel Petti, et al.
Veröffentlicht: (2025)
von: Daniel Petti, et al.
Veröffentlicht: (2025)
Enhanced View Planning for Robotic Harvesting: Tackling Occlusions with Imitation Learning
von: Li, Lun, et al.
Veröffentlicht: (2025)
von: Li, Lun, et al.
Veröffentlicht: (2025)
SARO: Space-Aware Robot System for Terrain Crossing via Vision-Language Model
von: Zhu, Shaoting, et al.
Veröffentlicht: (2024)
von: Zhu, Shaoting, et al.
Veröffentlicht: (2024)
SAMP: Spatial Anchor-based Motion Policy for Collision-Aware Robotic Manipulators
von: Chen, Kai, et al.
Veröffentlicht: (2025)
von: Chen, Kai, et al.
Veröffentlicht: (2025)
ACROSS: A Deformation-Based Cross-Modal Representation for Robotic Tactile Perception
von: Amri, Wadhah Zai El, et al.
Veröffentlicht: (2024)
von: Amri, Wadhah Zai El, et al.
Veröffentlicht: (2024)
SKOOTR: A SKating, Omni-Oriented, Tripedal Robot
von: Hung, Adam Joshua, et al.
Veröffentlicht: (2024)
von: Hung, Adam Joshua, et al.
Veröffentlicht: (2024)
Cortical Policy: A Dual-Stream View Transformer for Robotic Manipulation
von: Zhang, Xuening, et al.
Veröffentlicht: (2026)
von: Zhang, Xuening, et al.
Veröffentlicht: (2026)
Hybrid Consistency Policy: Decoupling Multi-Modal Diversity and Real-Time Efficiency in Robotic Manipulation
von: Zhao, Qianyou, et al.
Veröffentlicht: (2025)
von: Zhao, Qianyou, et al.
Veröffentlicht: (2025)
DVDP: An End-to-End Policy for Mobile Robot Visual Docking with RGB-D Perception
von: Min, Haohan, et al.
Veröffentlicht: (2025)
von: Min, Haohan, et al.
Veröffentlicht: (2025)
MiniTac: An Ultra-Compact 8 mm Vision-Based Tactile Sensor for Enhanced Palpation in Robot-Assisted Minimally Invasive Surgery
von: Li, Wanlin, et al.
Veröffentlicht: (2024)
von: Li, Wanlin, et al.
Veröffentlicht: (2024)
OmniVIC: A Self-Improving Variable Impedance Controller with Vision-Language In-Context Learning for Safe Robotic Manipulation
von: Zhang, Heng, et al.
Veröffentlicht: (2025)
von: Zhang, Heng, et al.
Veröffentlicht: (2025)
Modality-Augmented Fine-Tuning of Foundation Robot Policies for Cross-Embodiment Manipulation on GR1 and G1
von: Park, Junsung, et al.
Veröffentlicht: (2025)
von: Park, Junsung, et al.
Veröffentlicht: (2025)
OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation
von: Zheng, Yuhang, et al.
Veröffentlicht: (2026)
von: Zheng, Yuhang, et al.
Veröffentlicht: (2026)
UniGraspTransformer: Simplified Policy Distillation for Scalable Dexterous Robotic Grasping
von: Wang, Wenbo, et al.
Veröffentlicht: (2024)
von: Wang, Wenbo, et al.
Veröffentlicht: (2024)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
von: Xu, Charles, et al.
Veröffentlicht: (2024)
von: Xu, Charles, et al.
Veröffentlicht: (2024)
A Sonar-Visual Dataset for Cross-Modal Underwater Robot Perception
von: Chen, Weitung, et al.
Veröffentlicht: (2026)
von: Chen, Weitung, et al.
Veröffentlicht: (2026)
Driving Beyond Privilege: Distilling Dense-Reward Knowledge into Sparse-Reward Policies
von: Khanzada, Feeza Khan, et al.
Veröffentlicht: (2025)
von: Khanzada, Feeza Khan, et al.
Veröffentlicht: (2025)
MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots
von: Huang, Ting, et al.
Veröffentlicht: (2025)
von: Huang, Ting, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Visuomotor Policy for Multi-Robot Laser Tag Game
von: Li, Kai, et al.
Veröffentlicht: (2026) -
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025) -
Omni Differential Drive for Simultaneous Reconfiguration and Omnidirectional Mobility of Wheeled Robots
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024) -
Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation
von: Huang, Haoran, et al.
Veröffentlicht: (2026) -
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
von: Song, Yunzhou, et al.
Veröffentlicht: (2026)