OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Danyang, Yang, Zenghui, Qi, Guangpeng, Pang, Songtao, Shang, Guangyong, Ma, Qiang, Yang, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
von: Agia, Christopher, et al.
Veröffentlicht: (2024)
von: Agia, Christopher, et al.
Veröffentlicht: (2024)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
von: Chen, Junting, et al.
Veröffentlicht: (2024)
von: Chen, Junting, et al.
Veröffentlicht: (2024)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
von: Bonial, Claire, et al.
Veröffentlicht: (2024)
von: Bonial, Claire, et al.
Veröffentlicht: (2024)
A Surveillance Based Interactive Robot
von: Kavimandan, Kshitij, et al.
Veröffentlicht: (2025)
von: Kavimandan, Kshitij, et al.
Veröffentlicht: (2025)
Autonomous Navigation and Collision Avoidance for Mobile Robots: Classification and Review
von: de Carvalho, Marcus Vinicius Leal, et al.
Veröffentlicht: (2024)
von: de Carvalho, Marcus Vinicius Leal, et al.
Veröffentlicht: (2024)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
von: Deichler, Anna, et al.
Veröffentlicht: (2025)
von: Deichler, Anna, et al.
Veröffentlicht: (2025)
Industrial Robot Motion Planning with GPUs: Integration of cuRobo for Extended DOF Systems
von: Abuelsamen, Luai, et al.
Veröffentlicht: (2025)
von: Abuelsamen, Luai, et al.
Veröffentlicht: (2025)
Deployment-Time Reliability of Learned Robot Policies
von: Agia, Christopher
Veröffentlicht: (2026)
von: Agia, Christopher
Veröffentlicht: (2026)
CleanMAP: Distilling Multimodal LLMs for Confidence-Driven Crowdsourced HD Map Updates
von: Shaw, Ankit Kumar, et al.
Veröffentlicht: (2025)
von: Shaw, Ankit Kumar, et al.
Veröffentlicht: (2025)
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
von: Lukin, Stephanie M., et al.
Veröffentlicht: (2024)
von: Lukin, Stephanie M., et al.
Veröffentlicht: (2024)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation
von: Chen, Junting, et al.
Veröffentlicht: (2024)
von: Chen, Junting, et al.
Veröffentlicht: (2024)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
von: Hu, X., et al.
Veröffentlicht: (2025)
von: Hu, X., et al.
Veröffentlicht: (2025)
Who Sees What? Structured Thought-Action Sequences for Epistemic Reasoning in LLMs
von: Annese, Luca, et al.
Veröffentlicht: (2025)
von: Annese, Luca, et al.
Veröffentlicht: (2025)
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
von: Karlsson, Robin, et al.
Veröffentlicht: (2023)
von: Karlsson, Robin, et al.
Veröffentlicht: (2023)
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
von: Lemke, Oliver, et al.
Veröffentlicht: (2024)
von: Lemke, Oliver, et al.
Veröffentlicht: (2024)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
von: Yang, Shan
Veröffentlicht: (2026)
von: Yang, Shan
Veröffentlicht: (2026)
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
von: Chen, Junting, et al.
Veröffentlicht: (2025)
von: Chen, Junting, et al.
Veröffentlicht: (2025)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2023)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2023)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
Context-Dependent Affordance Computation in Vision-Language Models
von: Farzulla, Murad
Veröffentlicht: (2026)
von: Farzulla, Murad
Veröffentlicht: (2026)
PerspAct: Enhancing LLM Situated Collaboration Skills through Perspective Taking and Active Vision
von: Patania, Sabrina, et al.
Veröffentlicht: (2025)
von: Patania, Sabrina, et al.
Veröffentlicht: (2025)
CWM: Contrastive World Models for Action Feasibility Learning in Embodied Agent Pipelines
von: Banerjee, Chayan
Veröffentlicht: (2026)
von: Banerjee, Chayan
Veröffentlicht: (2026)
Conversations with Andrea: Visitors' Opinions on Android Robots in a Museum
von: Heisler, Marcel, et al.
Veröffentlicht: (2025)
von: Heisler, Marcel, et al.
Veröffentlicht: (2025)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
von: Romero, Angel, et al.
Veröffentlicht: (2025)
von: Romero, Angel, et al.
Veröffentlicht: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
von: Riva, Paolo, et al.
Veröffentlicht: (2026)
von: Riva, Paolo, et al.
Veröffentlicht: (2026)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
von: Hu, Pan
Veröffentlicht: (2025)
von: Hu, Pan
Veröffentlicht: (2025)
PRISM-TopoMap: Online Topological Mapping with Place Recognition and Scan Matching
von: Muravyev, Kirill, et al.
Veröffentlicht: (2024)
von: Muravyev, Kirill, et al.
Veröffentlicht: (2024)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
von: Aida, Adriana, et al.
Veröffentlicht: (2026)
von: Aida, Adriana, et al.
Veröffentlicht: (2026)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
VLA Foundry: A Unified Framework for Training Vision-Language-Action Models
von: Mercat, Jean, et al.
Veröffentlicht: (2026)
von: Mercat, Jean, et al.
Veröffentlicht: (2026)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
von: Babu, Abhijith, et al.
Veröffentlicht: (2026)
von: Babu, Abhijith, et al.
Veröffentlicht: (2026)
A Segmented Robot Grasping Perception Neural Network for Edge AI
von: Bröcheler, Casper, et al.
Veröffentlicht: (2025)
von: Bröcheler, Casper, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024) -
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025) -
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024) -
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
von: Agia, Christopher, et al.
Veröffentlicht: (2024) -
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
von: Chen, Junting, et al.
Veröffentlicht: (2024)