VLM-Vac: Enhancing Smart Vacuums through VLM Knowledge Distillation and Language-Guided Experience Replay
Fuente:
arXiv
Saved in:
| Main Authors: | Mirjalili, Reihaneh, Krawez, Michael, Walter, Florian, Burgard, Wolfram |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lan-grasp: Using Large Language Models for Semantic Object Grasping and Placement
by: Mirjalili, Reihaneh, et al.
Published: (2023)
by: Mirjalili, Reihaneh, et al.
Published: (2023)
Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
by: Mirjalili, Reihaneh, et al.
Published: (2025)
by: Mirjalili, Reihaneh, et al.
Published: (2025)
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025)
by: Blei, Yannik, et al.
Published: (2025)
Leveraging Foundation Models for Enhancing Robot Perception and Action
by: Mirjalili, Reihaneh
Published: (2025)
by: Mirjalili, Reihaneh
Published: (2025)
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
VLAgents: A Policy Server for Efficient VLA Inference
by: Jülg, Tobias, et al.
Published: (2026)
by: Jülg, Tobias, et al.
Published: (2026)
CloudTrack: Scalable UAV Tracking with Cloud Semantics
by: Blei, Yannik, et al.
Published: (2024)
by: Blei, Yannik, et al.
Published: (2024)
FlowNav: Combining Flow Matching and Depth Priors for Efficient Navigation
by: Gode, Samiran, et al.
Published: (2024)
by: Gode, Samiran, et al.
Published: (2024)
VLM-Guided Experience Replay
by: Sharony, Elad, et al.
Published: (2026)
by: Sharony, Elad, et al.
Published: (2026)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
by: Krack, Pierre, et al.
Published: (2026)
by: Krack, Pierre, et al.
Published: (2026)
VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving
by: Long, Keke, et al.
Published: (2024)
by: Long, Keke, et al.
Published: (2024)
SwarmVLM: VLM-Guided Impedance Control for Autonomous Navigation of Heterogeneous Robots in Dynamic Warehousing
by: Zafar, Malaika, et al.
Published: (2025)
by: Zafar, Malaika, et al.
Published: (2025)
Towards Logic-Aware Manipulation: A Knowledge Primitive for VLM-Based Assistants in Smart Manufacturing
by: Chen, Suchang, et al.
Published: (2025)
by: Chen, Suchang, et al.
Published: (2025)
VLM-UDMC: VLM-Enhanced Unified Decision-Making and Motion Control for Urban Autonomous Driving
by: Liu, Haichao, et al.
Published: (2025)
by: Liu, Haichao, et al.
Published: (2025)
GoalVLM: VLM-driven Object Goal Navigation for Multi-Agent System
by: James, MoniJesu, et al.
Published: (2026)
by: James, MoniJesu, et al.
Published: (2026)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
by: Mahmoud, Yara, et al.
Published: (2026)
by: Mahmoud, Yara, et al.
Published: (2026)
uPLAM: Robust Panoptic Localization and Mapping Leveraging Perception Uncertainties
by: Sirohi, Kshitij, et al.
Published: (2024)
by: Sirohi, Kshitij, et al.
Published: (2024)
VLM-SFD: VLM-Assisted Siamese Flow Diffusion Framework for Dual-Arm Cooperative Manipulation
by: Chen, Jiaming, et al.
Published: (2025)
by: Chen, Jiaming, et al.
Published: (2025)
VLM-TDP: VLM-guided Trajectory-conditioned Diffusion Policy for Robust Long-Horizon Manipulation
by: Huang, Kefeng, et al.
Published: (2025)
by: Huang, Kefeng, et al.
Published: (2025)
UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms
by: Guo, Xueyang, et al.
Published: (2025)
by: Guo, Xueyang, et al.
Published: (2025)
VLM-RRT: Vision Language Model Guided RRT Search for Autonomous UAV Navigation
by: Ye, Jianlin, et al.
Published: (2025)
by: Ye, Jianlin, et al.
Published: (2025)
ReplanVLM: Replanning Robotic Tasks with Visual Language Models
by: Mei, Aoran, et al.
Published: (2024)
by: Mei, Aoran, et al.
Published: (2024)
ConPoSe: LLM-Guided Contact Point Selection for Scalable Cooperative Object Pushing
by: Steinkrüger, Noah, et al.
Published: (2025)
by: Steinkrüger, Noah, et al.
Published: (2025)
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
Swim2Real: VLM-Guided System Identification for Sim-to-Real Transfer
by: Qiu, Kevin, et al.
Published: (2026)
by: Qiu, Kevin, et al.
Published: (2026)
TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving
by: Li, Jiaxiang, et al.
Published: (2026)
by: Li, Jiaxiang, et al.
Published: (2026)
HapticVLM: VLM-Driven Texture Recognition Aimed at Intelligent Haptic Interaction
by: Khan, Muhammad Haris, et al.
Published: (2025)
by: Khan, Muhammad Haris, et al.
Published: (2025)
Enabling Dynamic Tracking in Vision-Language-Action Models via Time-Discrete and Time-Continuous Velocity Feedforward
by: Hechtl, Johannes, et al.
Published: (2026)
by: Hechtl, Johannes, et al.
Published: (2026)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
by: Guo, Ziang, et al.
Published: (2024)
by: Guo, Ziang, et al.
Published: (2024)
Agent-Agnostic Centralized Training for Decentralized Multi-Agent Cooperative Driving
by: Yan, Shengchao, et al.
Published: (2024)
by: Yan, Shengchao, et al.
Published: (2024)
VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models
by: Song, Daeun, et al.
Published: (2024)
by: Song, Daeun, et al.
Published: (2024)
CoDriveVLM: VLM-Enhanced Urban Cooperative Dispatching and Motion Planning for Future Autonomous Mobility on Demand Systems
by: Liu, Haichao, et al.
Published: (2025)
by: Liu, Haichao, et al.
Published: (2025)
A3VLM: Actionable Articulation-Aware Vision Language Model
by: Huang, Siyuan, et al.
Published: (2024)
by: Huang, Siyuan, et al.
Published: (2024)
Language as Cost: Proactive Hazard Mapping using VLM for Robot Navigation
by: Oh, Mintaek, et al.
Published: (2025)
by: Oh, Mintaek, et al.
Published: (2025)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Task-Aware Bimanual Affordance Prediction via VLM-Guided Semantic-Geometric Reasoning
by: Hahne, Fabian, et al.
Published: (2026)
by: Hahne, Fabian, et al.
Published: (2026)
iFlyBot-VLM Technical Report
by: Nie, Xin, et al.
Published: (2025)
by: Nie, Xin, et al.
Published: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
by: Xu, Runsen, et al.
Published: (2024)
by: Xu, Runsen, et al.
Published: (2024)
AERMANI-VLM: Structured Prompting and Reasoning for Aerial Manipulation with Vision Language Models
by: Mishra, Sarthak, et al.
Published: (2025)
by: Mishra, Sarthak, et al.
Published: (2025)
LocoVLM: Grounding Vision and Language for Adapting Versatile Legged Locomotion Policies
by: Nahrendra, I Made Aswin, et al.
Published: (2026)
by: Nahrendra, I Made Aswin, et al.
Published: (2026)
Similar Items
-
Lan-grasp: Using Large Language Models for Semantic Object Grasping and Placement
by: Mirjalili, Reihaneh, et al.
Published: (2023) -
Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
by: Mirjalili, Reihaneh, et al.
Published: (2025) -
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025) -
Leveraging Foundation Models for Enhancing Robot Perception and Action
by: Mirjalili, Reihaneh
Published: (2025) -
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025)