I Speak and You Find: Robust 3D Visual Grounding with Noisy and Ambiguous Speech Inputs

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Qi, Yu, Gu, Lipeng, Chen, Honghua, Nan, Liangliang, Wei, Mingqiang
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!