BehAV: Behavioral Rule Guided Autonomy Using VLMs for Robot Navigation in Outdoor Scenes
Fuente:
arXiv
Guardado en:
| Autores principales: | Weerakoon, Kasun, Elnoor, Mohamed, Seneviratne, Gershom, Rajagopal, Vignesh, Arul, Senthil Hariharan, Liang, Jing, Jaffar, Mohamed Khalid M, Manocha, Dinesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2024)
por: Elnoor, Mohamed, et al.
Publicado: (2024)
ViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation
por: Elnoor, Mohamed, et al.
Publicado: (2025)
por: Elnoor, Mohamed, et al.
Publicado: (2025)
CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
por: Seneviratne, Gershom, et al.
Publicado: (2024)
por: Seneviratne, Gershom, et al.
Publicado: (2024)
AMCO: Adaptive Multimodal Coupling of Vision and Proprioception for Quadruped Robot Navigation in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2024)
por: Elnoor, Mohamed, et al.
Publicado: (2024)
ProNav: Proprioceptive Traversability Estimation for Legged Robot Navigation in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2023)
por: Elnoor, Mohamed, et al.
Publicado: (2023)
DR. Nav: Semantic-Geometric Representations for Proactive Dead-End Recovery and Navigation
por: Rajagopal, Vignesh, et al.
Publicado: (2025)
por: Rajagopal, Vignesh, et al.
Publicado: (2025)
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
por: Seneviratne, Gershom, et al.
Publicado: (2025)
por: Seneviratne, Gershom, et al.
Publicado: (2025)
TOPGN: Real-time Transparent Obstacle Detection using Lidar Point Cloud Intensity for Autonomous Robot Navigation
por: Weerakoon, Kasun, et al.
Publicado: (2024)
por: Weerakoon, Kasun, et al.
Publicado: (2024)
CoNVOI: Context-aware Navigation using Vision Language Models in Outdoor and Indoor Environments
por: Sathyamoorthy, Adarsh Jagan, et al.
Publicado: (2024)
por: Sathyamoorthy, Adarsh Jagan, et al.
Publicado: (2024)
DMCA: Dense Multi-agent Navigation using Attention and Communication
por: Arul, Senthil Hariharan, et al.
Publicado: (2022)
por: Arul, Senthil Hariharan, et al.
Publicado: (2022)
Splatblox: Traversability-Aware Gaussian Splatting for Outdoor Robot Navigation
por: Chopra, Samarth, et al.
Publicado: (2025)
por: Chopra, Samarth, et al.
Publicado: (2025)
MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding
por: Liang, Jing, et al.
Publicado: (2025)
por: Liang, Jing, et al.
Publicado: (2025)
MTG: Mapless Trajectory Generator with Traversability Coverage for Outdoor Navigation
por: Liang, Jing, et al.
Publicado: (2023)
por: Liang, Jing, et al.
Publicado: (2023)
PhysGS: Bayesian-Inferred Gaussian Splatting for Physical Property Estimation
por: Chopra, Samarth, et al.
Publicado: (2025)
por: Chopra, Samarth, et al.
Publicado: (2025)
VLPG-Nav: Object Navigation Using Visual Language Pose Graph and Object Localization Probability Maps
por: Arul, Senthil Hariharan, et al.
Publicado: (2024)
por: Arul, Senthil Hariharan, et al.
Publicado: (2024)
CHOP: Counterfactual Human Preference Labels Improve Obstacle Avoidance in Visuomotor Navigation Policies
por: Seneviratne, Gershom, et al.
Publicado: (2026)
por: Seneviratne, Gershom, et al.
Publicado: (2026)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
por: Patel, Bhrij, et al.
Publicado: (2023)
por: Patel, Bhrij, et al.
Publicado: (2023)
AGL-NET: Aerial-Ground Cross-Modal Global Localization with Varying Scales
por: Guan, Tianrui, et al.
Publicado: (2024)
por: Guan, Tianrui, et al.
Publicado: (2024)
PiP-X: Online feedback motion planning/replanning in dynamic environments using invariant funnels
por: Jaffar, Mohamed Khalid M, et al.
Publicado: (2022)
por: Jaffar, Mohamed Khalid M, et al.
Publicado: (2022)
Listen2Scene: Interactive material-aware binaural sound propagation for reconstructed 3D scenes
por: Ratnarajah, Anton, et al.
Publicado: (2023)
por: Ratnarajah, Anton, et al.
Publicado: (2023)
EM-GANSim: Real-time and Accurate EM Simulation Using Conditional GANs for 3D Indoor Scenes
por: Wang, Ruichen, et al.
Publicado: (2024)
por: Wang, Ruichen, et al.
Publicado: (2024)
VL-TGS: Trajectory Generation and Selection using Vision Language Models in Mapless Outdoor Environments
por: Song, Daeun, et al.
Publicado: (2024)
por: Song, Daeun, et al.
Publicado: (2024)
AV-RIR: Audio-Visual Room Impulse Response Estimation
por: Ratnarajah, Anton, et al.
Publicado: (2023)
por: Ratnarajah, Anton, et al.
Publicado: (2023)
QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs
por: Denipitiyage, Dishanika, et al.
Publicado: (2026)
por: Denipitiyage, Dishanika, et al.
Publicado: (2026)
BehAVE: Behaviour Alignment of Video Game Encodings
por: Rašajski, Nemanja, et al.
Publicado: (2024)
por: Rašajski, Nemanja, et al.
Publicado: (2024)
VLM-Based Advanced Rider Assistance System for Motorcycle Safety
por: Elnoor, Mohamed, et al.
Publicado: (2026)
por: Elnoor, Mohamed, et al.
Publicado: (2026)
BoMuDANet: Unsupervised Adaptation for Visual Scene Understanding in Unstructured Driving Environments
por: Kothandaraman, Divya, et al.
Publicado: (2020)
por: Kothandaraman, Divya, et al.
Publicado: (2020)
GND: Global Navigation Dataset with Multi-Modal Perception and Multi-Category Traversability in Outdoor Campus Environments
por: Liang, Jing, et al.
Publicado: (2024)
por: Liang, Jing, et al.
Publicado: (2024)
Personalized Embodied Navigation for Portable Object Finding
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2024)
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2024)
Uncovering the Representation Geometry of Minimal Cores in Overcomplete Reasoning Traces
por: Chowdhury, Sanjoy, et al.
Publicado: (2026)
por: Chowdhury, Sanjoy, et al.
Publicado: (2026)
PACE: Data-Driven Virtual Agent Interaction in Dense and Cluttered Environments
por: Mullen, James, et al.
Publicado: (2023)
por: Mullen, James, et al.
Publicado: (2023)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
por: Lee, Yonghan, et al.
Publicado: (2026)
por: Lee, Yonghan, et al.
Publicado: (2026)
DTG : Diffusion-based Trajectory Generation for Mapless Global Navigation
por: Liang, Jing, et al.
Publicado: (2024)
por: Liang, Jing, et al.
Publicado: (2024)
Empowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
por: Zhou, Wentao, et al.
Publicado: (2025)
por: Zhou, Wentao, et al.
Publicado: (2025)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
por: Mullen Jr, James F., et al.
Publicado: (2022)
por: Mullen Jr, James F., et al.
Publicado: (2022)
Can an Embodied Agent Find Your "Cat-shaped Mug"? LLM-Guided Exploration for Zero-Shot Object Navigation
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2023)
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2023)
MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks
por: Chowdhury, Sanjoy, et al.
Publicado: (2025)
por: Chowdhury, Sanjoy, et al.
Publicado: (2025)
SocialNav-SUB: Benchmarking VLMs for Scene Understanding in Social Robot Navigation
por: Munje, Michael J., et al.
Publicado: (2025)
por: Munje, Michael J., et al.
Publicado: (2025)
The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible
por: Lovén, Lauri, et al.
Publicado: (2026)
por: Lovén, Lauri, et al.
Publicado: (2026)
MemCtrl: Using MLLMs as Active Memory Controllers on Embodied Agents
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2026)
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2026)
Ejemplares similares
-
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2024) -
ViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation
por: Elnoor, Mohamed, et al.
Publicado: (2025) -
CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
por: Seneviratne, Gershom, et al.
Publicado: (2024) -
AMCO: Adaptive Multimodal Coupling of Vision and Proprioception for Quadruped Robot Navigation in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2024) -
ProNav: Proprioceptive Traversability Estimation for Legged Robot Navigation in Outdoor Environments
por: Elnoor, Mohamed, et al.
Publicado: (2023)