EfficientEQA: An Efficient Approach to Open-Vocabulary Embodied Question Answering
Fuente:
arXiv
Guardado en:
| Autores principales: | Cheng, Kai, Li, Zhengyuan, Sun, Xingpeng, Min, Byung-Cheol, Bedi, Amrit Singh, Bera, Aniket |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
NoisyEQA: Benchmarking Embodied Question Answering Against Noisy Queries
por: Wu, Tao, et al.
Publicado: (2024)
por: Wu, Tao, et al.
Publicado: (2024)
TrustNavGPT: Modeling Uncertainty to Improve Trustworthiness of Audio-Guided LLM-Based Robot Navigation
por: Sun, Xingpeng, et al.
Publicado: (2024)
por: Sun, Xingpeng, et al.
Publicado: (2024)
Personalized Embodied Navigation for Portable Object Finding
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2024)
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2024)
Beyond Text: Utilizing Vocal Cues to Improve Decision Making in LLMs for Robot Navigation Tasks
por: Sun, Xingpeng, et al.
Publicado: (2024)
por: Sun, Xingpeng, et al.
Publicado: (2024)
GraphEQA: Using 3D Semantic Scene Graphs for Real-time Embodied Question Answering
por: Saxena, Saumya, et al.
Publicado: (2024)
por: Saxena, Saumya, et al.
Publicado: (2024)
FAST-EQA: Efficient Embodied Question Answering with Global and Local Region Relevancy
por: Zhang, Haochen, et al.
Publicado: (2026)
por: Zhang, Haochen, et al.
Publicado: (2026)
Explore until Confident: Efficient Exploration for Embodied Question Answering
por: Ren, Allen Z., et al.
Publicado: (2024)
por: Ren, Allen Z., et al.
Publicado: (2024)
Map-based Modular Approach for Zero-shot Embodied Question Answering
por: Sakamoto, Koya, et al.
Publicado: (2024)
por: Sakamoto, Koya, et al.
Publicado: (2024)
Semantic Layering in Room Segmentation via LLMs
por: Kim, Taehyeon, et al.
Publicado: (2024)
por: Kim, Taehyeon, et al.
Publicado: (2024)
IndustryEQA: Pushing the Frontiers of Embodied Question Answering in Industrial Scenarios
por: Li, Yifan, et al.
Publicado: (2025)
por: Li, Yifan, et al.
Publicado: (2025)
ZeroSCD: Zero-Shot Street Scene Change Detection
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
PlaceFormer: Transformer-based Visual Place Recognition using Multi-Scale Patch Selection and Fusion
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
Visual Environment-Interactive Planning for Embodied Complex-Question Answering
por: Lan, Ning, et al.
Publicado: (2025)
por: Lan, Ning, et al.
Publicado: (2025)
FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction
por: Jiang, Zeyu, et al.
Publicado: (2026)
por: Jiang, Zeyu, et al.
Publicado: (2026)
Hypergraph-based Coordinated Task Allocation and Socially-aware Navigation for Multi-Robot Systems
por: Wang, Weizheng, et al.
Publicado: (2024)
por: Wang, Weizheng, et al.
Publicado: (2024)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
por: Yin, Tenny, et al.
Publicado: (2025)
por: Yin, Tenny, et al.
Publicado: (2025)
Text-guided Generation of Efficient Personalized Inspection Plans
por: Sun, Xingpeng, et al.
Publicado: (2025)
por: Sun, Xingpeng, et al.
Publicado: (2025)
ConEQsA: Concurrent and Asynchronous Embodied Questions Scheduling and Answering
por: Wang, Haisheng, et al.
Publicado: (2025)
por: Wang, Haisheng, et al.
Publicado: (2025)
Prune-Then-Plan: Step-Level Calibration for Stable Frontier Exploration in Embodied Question Answering
por: Frahm, Noah, et al.
Publicado: (2025)
por: Frahm, Noah, et al.
Publicado: (2025)
Scalable Multi-Robot Informative Path Planning for Target Mapping via Deep Reinforcement Learning
por: Vashisth, Apoorva, et al.
Publicado: (2024)
por: Vashisth, Apoorva, et al.
Publicado: (2024)
Unifying Large Language Model and Deep Reinforcement Learning for Human-in-Loop Interactive Socially-aware Navigation
por: Wang, Weizheng, et al.
Publicado: (2024)
por: Wang, Weizheng, et al.
Publicado: (2024)
A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM
por: Han, ByungOk, et al.
Publicado: (2024)
por: Han, ByungOk, et al.
Publicado: (2024)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
por: Wu, Yanmin, et al.
Publicado: (2024)
por: Wu, Yanmin, et al.
Publicado: (2024)
Open-Vocabulary Online Semantic Mapping for SLAM
por: Martins, Tomas Berriel, et al.
Publicado: (2024)
por: Martins, Tomas Berriel, et al.
Publicado: (2024)
LOVON: Legged Open-Vocabulary Object Navigator
por: Peng, Daojie, et al.
Publicado: (2025)
por: Peng, Daojie, et al.
Publicado: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
por: He, Ziheng, et al.
Publicado: (2026)
por: He, Ziheng, et al.
Publicado: (2026)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
por: Wu, Tao, et al.
Publicado: (2025)
por: Wu, Tao, et al.
Publicado: (2025)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
por: Kong, Lingdong, et al.
Publicado: (2024)
por: Kong, Lingdong, et al.
Publicado: (2024)
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
por: Chahe, Amirhosein, et al.
Publicado: (2025)
por: Chahe, Amirhosein, et al.
Publicado: (2025)
WildOS: Open-Vocabulary Object Search in the Wild
por: Shah, Hardik, et al.
Publicado: (2026)
por: Shah, Hardik, et al.
Publicado: (2026)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
HomeRobot: Open-Vocabulary Mobile Manipulation
por: Yenamandra, Sriram, et al.
Publicado: (2023)
por: Yenamandra, Sriram, et al.
Publicado: (2023)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
por: Goetting, Dylan, et al.
Publicado: (2024)
por: Goetting, Dylan, et al.
Publicado: (2024)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
por: Pham, Phu, et al.
Publicado: (2024)
por: Pham, Phu, et al.
Publicado: (2024)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
por: Ishaq, Ayesha, et al.
Publicado: (2024)
por: Ishaq, Ayesha, et al.
Publicado: (2024)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
por: Jiang, Haochen, et al.
Publicado: (2024)
por: Jiang, Haochen, et al.
Publicado: (2024)
Configurable Embodied Data Generation for Class-Agnostic RGB-D Video Segmentation
por: Opipari, Anthony, et al.
Publicado: (2024)
por: Opipari, Anthony, et al.
Publicado: (2024)
Prompter: Utilizing Large Language Model Prompting for a Data Efficient Embodied Instruction Following
por: Inoue, Yuki, et al.
Publicado: (2022)
por: Inoue, Yuki, et al.
Publicado: (2022)
Wanderland: Geometrically Grounded Simulation for Open-World Embodied AI
por: Liu, Xinhao, et al.
Publicado: (2025)
por: Liu, Xinhao, et al.
Publicado: (2025)
Kinematify: Open-Vocabulary Synthesis of High-DoF Articulated Objects
por: Wang, Jiawei, et al.
Publicado: (2025)
por: Wang, Jiawei, et al.
Publicado: (2025)
Ejemplares similares
-
NoisyEQA: Benchmarking Embodied Question Answering Against Noisy Queries
por: Wu, Tao, et al.
Publicado: (2024) -
TrustNavGPT: Modeling Uncertainty to Improve Trustworthiness of Audio-Guided LLM-Based Robot Navigation
por: Sun, Xingpeng, et al.
Publicado: (2024) -
Personalized Embodied Navigation for Portable Object Finding
por: Dorbala, Vishnu Sashank, et al.
Publicado: (2024) -
Beyond Text: Utilizing Vocal Cues to Improve Decision Making in LLMs for Robot Navigation Tasks
por: Sun, Xingpeng, et al.
Publicado: (2024) -
GraphEQA: Using 3D Semantic Scene Graphs for Real-time Embodied Question Answering
por: Saxena, Saumya, et al.
Publicado: (2024)