Visual Environment-Interactive Planning for Embodied Complex-Question Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lan, Ning, Ou, Baoshan, Xie, Xuemei, Shi, Guangming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Map-based Modular Approach for Zero-shot Embodied Question Answering
von: Sakamoto, Koya, et al.
Veröffentlicht: (2024)
von: Sakamoto, Koya, et al.
Veröffentlicht: (2024)
Prune-Then-Plan: Step-Level Calibration for Stable Frontier Exploration in Embodied Question Answering
von: Frahm, Noah, et al.
Veröffentlicht: (2025)
von: Frahm, Noah, et al.
Veröffentlicht: (2025)
Parse Graph-Based Visual-Language Interaction for Human Pose Estimation
von: Liu, Shibang, et al.
Veröffentlicht: (2025)
von: Liu, Shibang, et al.
Veröffentlicht: (2025)
ConEQsA: Concurrent and Asynchronous Embodied Questions Scheduling and Answering
von: Wang, Haisheng, et al.
Veröffentlicht: (2025)
von: Wang, Haisheng, et al.
Veröffentlicht: (2025)
EfficientEQA: An Efficient Approach to Open-Vocabulary Embodied Question Answering
von: Cheng, Kai, et al.
Veröffentlicht: (2024)
von: Cheng, Kai, et al.
Veröffentlicht: (2024)
Explore until Confident: Efficient Exploration for Embodied Question Answering
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
LingoQA: Visual Question Answering for Autonomous Driving
von: Marcu, Ana-Maria, et al.
Veröffentlicht: (2023)
von: Marcu, Ana-Maria, et al.
Veröffentlicht: (2023)
Look, Zoom, Understand: The Robotic Eyeball for Embodied Perception
von: Yang, Jiashu, et al.
Veröffentlicht: (2025)
von: Yang, Jiashu, et al.
Veröffentlicht: (2025)
EmbodiedBrain: Expanding Performance Boundaries of Task Planning for Embodied Intelligence
von: Zou, Ding, et al.
Veröffentlicht: (2025)
von: Zou, Ding, et al.
Veröffentlicht: (2025)
Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question-Localized Answering in Robotic Surgery
von: Bai, Long, et al.
Veröffentlicht: (2024)
von: Bai, Long, et al.
Veröffentlicht: (2024)
NoisyEQA: Benchmarking Embodied Question Answering Against Noisy Queries
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
FloNa: Floor Plan Guided Embodied Visual Navigation
von: Li, Jiaxin, et al.
Veröffentlicht: (2024)
von: Li, Jiaxin, et al.
Veröffentlicht: (2024)
TrackVLA: Embodied Visual Tracking in the Wild
von: Wang, Shaoan, et al.
Veröffentlicht: (2025)
von: Wang, Shaoan, et al.
Veröffentlicht: (2025)
GraphEQA: Using 3D Semantic Scene Graphs for Real-time Embodied Question Answering
von: Saxena, Saumya, et al.
Veröffentlicht: (2024)
von: Saxena, Saumya, et al.
Veröffentlicht: (2024)
EnerVerse-AC: Envisioning Embodied Environments with Action Condition
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
RoboAgent: Chaining Basic Capabilities for Embodied Task Planning
von: Xu, Peiran, et al.
Veröffentlicht: (2026)
von: Xu, Peiran, et al.
Veröffentlicht: (2026)
Closed Loop Interactive Embodied Reasoning for Robot Manipulation
von: Nazarczuk, Michal, et al.
Veröffentlicht: (2024)
von: Nazarczuk, Michal, et al.
Veröffentlicht: (2024)
Think Proprioceptively: Embodied Visual Reasoning for VLA Manipulation
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
von: Wu, Tao, et al.
Veröffentlicht: (2025)
von: Wu, Tao, et al.
Veröffentlicht: (2025)
Embodied Domain Adaptation for Object Detection
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
SVLL: Staged Vision-Language Learning for Physically Grounded Embodied Task Planning
von: Yang, Yuyuan, et al.
Veröffentlicht: (2026)
von: Yang, Yuyuan, et al.
Veröffentlicht: (2026)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models
von: Ling, Yiran, et al.
Veröffentlicht: (2026)
von: Ling, Yiran, et al.
Veröffentlicht: (2026)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
von: Liu, Enguang, et al.
Veröffentlicht: (2025)
von: Liu, Enguang, et al.
Veröffentlicht: (2025)
VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic Memory
von: Wang, Shaoan, et al.
Veröffentlicht: (2026)
von: Wang, Shaoan, et al.
Veröffentlicht: (2026)
KB-DMGen: Knowledge-Based Global Guidance and Dynamic Pose Masking for Human Image Generation
von: Liu, Shibang, et al.
Veröffentlicht: (2025)
von: Liu, Shibang, et al.
Veröffentlicht: (2025)
ACMo: Attribute Controllable Motion Generation
von: Wei, Mingjie, et al.
Veröffentlicht: (2025)
von: Wei, Mingjie, et al.
Veröffentlicht: (2025)
XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments
von: Qian, Kangan, et al.
Veröffentlicht: (2026)
von: Qian, Kangan, et al.
Veröffentlicht: (2026)
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2025)
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2025)
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
von: Xiao, Junbin, et al.
Veröffentlicht: (2026)
von: Xiao, Junbin, et al.
Veröffentlicht: (2026)
MomaGraph: State-Aware Unified Scene Graphs with Vision-Language Model for Embodied Task Planning
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
PC-NeRF: Parent-Child Neural Radiance Fields Using Sparse LiDAR Frames in Autonomous Driving Environments
von: Hu, Xiuzhong, et al.
Veröffentlicht: (2024)
von: Hu, Xiuzhong, et al.
Veröffentlicht: (2024)
AoE: Always-on Egocentric Human Video Collection for Embodied AI
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces
von: Luo, Gen, et al.
Veröffentlicht: (2025)
von: Luo, Gen, et al.
Veröffentlicht: (2025)
Expand Your SCOPE: Semantic Cognition over Potential-Based Exploration for Embodied Visual Navigation
von: Wang, Ningnan, et al.
Veröffentlicht: (2025)
von: Wang, Ningnan, et al.
Veröffentlicht: (2025)
D3D-VLP: Dynamic 3D Vision-Language-Planning Model for Embodied Grounding and Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
Embodied4C: Measuring What Matters for Embodied Vision-Language Navigation
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence
von: Wang, Xinjie, et al.
Veröffentlicht: (2025)
von: Wang, Xinjie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Map-based Modular Approach for Zero-shot Embodied Question Answering
von: Sakamoto, Koya, et al.
Veröffentlicht: (2024) -
Prune-Then-Plan: Step-Level Calibration for Stable Frontier Exploration in Embodied Question Answering
von: Frahm, Noah, et al.
Veröffentlicht: (2025) -
Parse Graph-Based Visual-Language Interaction for Human Pose Estimation
von: Liu, Shibang, et al.
Veröffentlicht: (2025) -
ConEQsA: Concurrent and Asynchronous Embodied Questions Scheduling and Answering
von: Wang, Haisheng, et al.
Veröffentlicht: (2025) -
EfficientEQA: An Efficient Approach to Open-Vocabulary Embodied Question Answering
von: Cheng, Kai, et al.
Veröffentlicht: (2024)