ViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Elnoor, Mohamed, Weerakoon, Kasun, Seneviratne, Gershom, Liang, Jing, Rajagopal, Vignesh, Manocha, Dinesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
BehAV: Behavioral Rule Guided Autonomy Using VLMs for Robot Navigation in Outdoor Scenes
von: Weerakoon, Kasun, et al.
Veröffentlicht: (2024)
von: Weerakoon, Kasun, et al.
Veröffentlicht: (2024)
AMCO: Adaptive Multimodal Coupling of Vision and Proprioception for Quadruped Robot Navigation in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
DR. Nav: Semantic-Geometric Representations for Proactive Dead-End Recovery and Navigation
von: Rajagopal, Vignesh, et al.
Veröffentlicht: (2025)
von: Rajagopal, Vignesh, et al.
Veröffentlicht: (2025)
CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2024)
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2024)
ProNav: Proprioceptive Traversability Estimation for Legged Robot Navigation in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2023)
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2023)
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
TOPGN: Real-time Transparent Obstacle Detection using Lidar Point Cloud Intensity for Autonomous Robot Navigation
von: Weerakoon, Kasun, et al.
Veröffentlicht: (2024)
von: Weerakoon, Kasun, et al.
Veröffentlicht: (2024)
CoNVOI: Context-aware Navigation using Vision Language Models in Outdoor and Indoor Environments
von: Sathyamoorthy, Adarsh Jagan, et al.
Veröffentlicht: (2024)
von: Sathyamoorthy, Adarsh Jagan, et al.
Veröffentlicht: (2024)
MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding
von: Liang, Jing, et al.
Veröffentlicht: (2025)
von: Liang, Jing, et al.
Veröffentlicht: (2025)
PhysGS: Bayesian-Inferred Gaussian Splatting for Physical Property Estimation
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
Splatblox: Traversability-Aware Gaussian Splatting for Outdoor Robot Navigation
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
AutoSpatial: Visual-Language Reasoning for Social Robot Navigation through Efficient Spatial Reasoning Learning
von: Kong, Yangzhe, et al.
Veröffentlicht: (2025)
von: Kong, Yangzhe, et al.
Veröffentlicht: (2025)
VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models
von: Song, Daeun, et al.
Veröffentlicht: (2024)
von: Song, Daeun, et al.
Veröffentlicht: (2024)
MTG: Mapless Trajectory Generator with Traversability Coverage for Outdoor Navigation
von: Liang, Jing, et al.
Veröffentlicht: (2023)
von: Liang, Jing, et al.
Veröffentlicht: (2023)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
CHOP: Counterfactual Human Preference Labels Improve Obstacle Avoidance in Visuomotor Navigation Policies
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2026)
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2026)
Social-LLaVA: Enhancing Robot Navigation through Human-Language Reasoning in Social Spaces
von: Payandeh, Amirreza, et al.
Veröffentlicht: (2024)
von: Payandeh, Amirreza, et al.
Veröffentlicht: (2024)
VL-TGS: Trajectory Generation and Selection using Vision Language Models in Mapless Outdoor Environments
von: Song, Daeun, et al.
Veröffentlicht: (2024)
von: Song, Daeun, et al.
Veröffentlicht: (2024)
DMCA: Dense Multi-agent Navigation using Attention and Communication
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2022)
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2022)
DTG : Diffusion-based Trajectory Generation for Mapless Global Navigation
von: Liang, Jing, et al.
Veröffentlicht: (2024)
von: Liang, Jing, et al.
Veröffentlicht: (2024)
LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
von: Ding, Hongyu, et al.
Veröffentlicht: (2025)
von: Ding, Hongyu, et al.
Veröffentlicht: (2025)
Online Robot Navigation and Manipulation with Distilled Vision-Language Models
von: Liu, Kangcheng
Veröffentlicht: (2024)
von: Liu, Kangcheng
Veröffentlicht: (2024)
VLPG-Nav: Object Navigation Using Visual Language Pose Graph and Object Localization Probability Maps
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2024)
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2024)
VLM-Based Advanced Rider Assistance System for Motorcycle Safety
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2026)
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2026)
ROVER: Regulator-Driven Robust Temporal Verification of Black-Box Robot Policies
von: Sakano, Kristy, et al.
Veröffentlicht: (2025)
von: Sakano, Kristy, et al.
Veröffentlicht: (2025)
Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation
von: Ding, Hongyu, et al.
Veröffentlicht: (2026)
von: Ding, Hongyu, et al.
Veröffentlicht: (2026)
Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement
von: Hoover, Montana, et al.
Veröffentlicht: (2026)
von: Hoover, Montana, et al.
Veröffentlicht: (2026)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025)
von: Tang, Nan, et al.
Veröffentlicht: (2025)
LAMP: Implicit Language Map for Robot Navigation
von: Lee, Sibaek, et al.
Veröffentlicht: (2026)
von: Lee, Sibaek, et al.
Veröffentlicht: (2026)
Towards Adaptive Environment Generation for Training Embodied Agents
von: Yeo, Teresa, et al.
Veröffentlicht: (2026)
von: Yeo, Teresa, et al.
Veröffentlicht: (2026)
Probing Prompt Design for Socially Compliant Robot Navigation with Vision Language Models
von: Xiao, Ling, et al.
Veröffentlicht: (2026)
von: Xiao, Ling, et al.
Veröffentlicht: (2026)
SABER: A Stealthy Agentic Black-Box Attack Framework for Vision-Language-Action Models
von: Wu, Xiyang, et al.
Veröffentlicht: (2026)
von: Wu, Xiyang, et al.
Veröffentlicht: (2026)
GaussianSSC: Triplane-Guided Directional Gaussian Fields for 3D Semantic Completion
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
Personalized Embodied Navigation for Portable Object Finding
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
LBAP: Improved Uncertainty Alignment of LLM Planners using Bayesian Inference
von: Mullen Jr., James F., et al.
Veröffentlicht: (2024)
von: Mullen Jr., James F., et al.
Veröffentlicht: (2024)
AdaNav: Adaptive Reasoning with Uncertainty for Vision-Language Navigation
von: Ding, Xin, et al.
Veröffentlicht: (2025)
von: Ding, Xin, et al.
Veröffentlicht: (2025)
GND: Global Navigation Dataset with Multi-Modal Perception and Multi-Category Traversability in Outdoor Campus Environments
von: Liang, Jing, et al.
Veröffentlicht: (2024)
von: Liang, Jing, et al.
Veröffentlicht: (2024)
MemCtrl: Using MLLMs as Active Memory Controllers on Embodied Agents
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2026)
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2026)
MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2025)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024) -
BehAV: Behavioral Rule Guided Autonomy Using VLMs for Robot Navigation in Outdoor Scenes
von: Weerakoon, Kasun, et al.
Veröffentlicht: (2024) -
AMCO: Adaptive Multimodal Coupling of Vision and Proprioception for Quadruped Robot Navigation in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024) -
DR. Nav: Semantic-Geometric Representations for Proactive Dead-End Recovery and Navigation
von: Rajagopal, Vignesh, et al.
Veröffentlicht: (2025) -
CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2024)