Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian
Fuente:
arXiv
Salvato in:
| Autori principali: | Chahe, Amirhosein, Zhou, Lifeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
di: Chahe, Amirhosein, et al.
Pubblicazione: (2025)
di: Chahe, Amirhosein, et al.
Pubblicazione: (2025)
Dynamic Adversarial Attacks on Autonomous Driving Systems
di: Chahe, Amirhosein, et al.
Pubblicazione: (2023)
di: Chahe, Amirhosein, et al.
Pubblicazione: (2023)
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
di: Werby, Abdelrhman, et al.
Pubblicazione: (2024)
di: Werby, Abdelrhman, et al.
Pubblicazione: (2024)
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
di: Chen, Runnan, et al.
Pubblicazione: (2024)
di: Chen, Runnan, et al.
Pubblicazione: (2024)
DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes
di: Hou, Qirui, et al.
Pubblicazione: (2025)
di: Hou, Qirui, et al.
Pubblicazione: (2025)
Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
OpenLex3D: A Tiered Evaluation Benchmark for Open-Vocabulary 3D Scene Representations
di: Kassab, Christina, et al.
Pubblicazione: (2025)
di: Kassab, Christina, et al.
Pubblicazione: (2025)
Policy-Guided World Model Planning for Language-Conditioned Visual Navigation
di: Chahe, Amirhosein, et al.
Pubblicazione: (2026)
di: Chahe, Amirhosein, et al.
Pubblicazione: (2026)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
di: Jiang, Haochen, et al.
Pubblicazione: (2024)
di: Jiang, Haochen, et al.
Pubblicazione: (2024)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
di: Wu, Yanmin, et al.
Pubblicazione: (2024)
di: Wu, Yanmin, et al.
Pubblicazione: (2024)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
di: Salzmann, Tim, et al.
Pubblicazione: (2024)
di: Salzmann, Tim, et al.
Pubblicazione: (2024)
Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models
di: Bharadwaj, Sagar, et al.
Pubblicazione: (2026)
di: Bharadwaj, Sagar, et al.
Pubblicazione: (2026)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
di: Chu, Hengshuo, et al.
Pubblicazione: (2025)
di: Chu, Hengshuo, et al.
Pubblicazione: (2025)
The 2nd Place Solution from the 3D Semantic Segmentation Track in the 2024 Waymo Open Dataset Challenge
di: Wu, Qing
Pubblicazione: (2025)
di: Wu, Qing
Pubblicazione: (2025)
GrabS: Generative Embodied Agent for 3D Object Segmentation without Scene Supervision
di: Zhang, Zihui, et al.
Pubblicazione: (2025)
di: Zhang, Zihui, et al.
Pubblicazione: (2025)
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
di: Hu, Xinggang, et al.
Pubblicazione: (2026)
di: Hu, Xinggang, et al.
Pubblicazione: (2026)
3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
di: Ke, Tsung-Wei, et al.
Pubblicazione: (2024)
di: Ke, Tsung-Wei, et al.
Pubblicazione: (2024)
SceneWeaver: All-in-One 3D Scene Synthesis with an Extensible and Self-Reflective Agent
di: Yang, Yandan, et al.
Pubblicazione: (2025)
di: Yang, Yandan, et al.
Pubblicazione: (2025)
HQ-OV3D: A High Box Quality Open-World 3D Detection Framework based on Diffision Model
di: Liu, Qi, et al.
Pubblicazione: (2025)
di: Liu, Qi, et al.
Pubblicazione: (2025)
RAPiD-Seg: Range-Aware Pointwise Distance Distribution Networks for 3D LiDAR Segmentation
di: Li, Li, et al.
Pubblicazione: (2024)
di: Li, Li, et al.
Pubblicazione: (2024)
SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
di: Jia, Baoxiong, et al.
Pubblicazione: (2024)
di: Jia, Baoxiong, et al.
Pubblicazione: (2024)
VIGS SLAM: IMU-based Large-Scale 3D Gaussian Splatting SLAM
di: Pak, Gyuhyeon, et al.
Pubblicazione: (2025)
di: Pak, Gyuhyeon, et al.
Pubblicazione: (2025)
PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI
di: Yang, Yandan, et al.
Pubblicazione: (2024)
di: Yang, Yandan, et al.
Pubblicazione: (2024)
EvObj: Learning Evolving Object-centric Representations for 3D Instance Segmentation without Scene Supervision
di: Chen, Jiahao, et al.
Pubblicazione: (2026)
di: Chen, Jiahao, et al.
Pubblicazione: (2026)
OSN: Infinite Representations of Dynamic 3D Scenes from Monocular Videos
di: Song, Ziyang, et al.
Pubblicazione: (2024)
di: Song, Ziyang, et al.
Pubblicazione: (2024)
SpaCeFormer: Fast Proposal-Free Open-Vocabulary 3D Instance Segmentation
di: Choy, Chris, et al.
Pubblicazione: (2026)
di: Choy, Chris, et al.
Pubblicazione: (2026)
Do Open-Vocabulary Detectors Transfer to Aerial Imagery? A Comparative Evaluation
di: Tsourveloudis, Christos
Pubblicazione: (2026)
di: Tsourveloudis, Christos
Pubblicazione: (2026)
OpenSUN3D: 1st Workshop Challenge on Open-Vocabulary 3D Scene Understanding
di: Engelmann, Francis, et al.
Pubblicazione: (2024)
di: Engelmann, Francis, et al.
Pubblicazione: (2024)
SceneFoundry: Generating Interactive Infinite 3D Worlds
di: Chen, ChunTeng, et al.
Pubblicazione: (2026)
di: Chen, ChunTeng, et al.
Pubblicazione: (2026)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
di: Ishaq, Ayesha, et al.
Pubblicazione: (2024)
di: Ishaq, Ayesha, et al.
Pubblicazione: (2024)
ODIN: A Single Model for 2D and 3D Segmentation
di: Jain, Ayush, et al.
Pubblicazione: (2024)
di: Jain, Ayush, et al.
Pubblicazione: (2024)
Contrastive Gaussian Clustering: Weakly Supervised 3D Scene Segmentation
di: Silva, Myrna C., et al.
Pubblicazione: (2024)
di: Silva, Myrna C., et al.
Pubblicazione: (2024)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
di: Li, Chengmeng, et al.
Pubblicazione: (2025)
di: Li, Chengmeng, et al.
Pubblicazione: (2025)
RoboLayout: Differentiable 3D Scene Generation for Embodied Agents
di: Shamsaddinlou, Ali
Pubblicazione: (2026)
di: Shamsaddinlou, Ali
Pubblicazione: (2026)
D$^3$Fields: Dynamic 3D Descriptor Fields for Zero-Shot Generalizable Rearrangement
di: Wang, Yixuan, et al.
Pubblicazione: (2023)
di: Wang, Yixuan, et al.
Pubblicazione: (2023)
Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling
di: Qiu, Xiaowen, et al.
Pubblicazione: (2025)
di: Qiu, Xiaowen, et al.
Pubblicazione: (2025)
Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding
di: Man, Yunze, et al.
Pubblicazione: (2024)
di: Man, Yunze, et al.
Pubblicazione: (2024)
Adapt3R: Adaptive 3D Scene Representation for Domain Transfer in Imitation Learning
di: Wilcox, Albert, et al.
Pubblicazione: (2025)
di: Wilcox, Albert, et al.
Pubblicazione: (2025)
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
di: Ze, Yanjie, et al.
Pubblicazione: (2024)
di: Ze, Yanjie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
di: Chahe, Amirhosein, et al.
Pubblicazione: (2025) -
Dynamic Adversarial Attacks on Autonomous Driving Systems
di: Chahe, Amirhosein, et al.
Pubblicazione: (2023) -
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
di: Werby, Abdelrhman, et al.
Pubblicazione: (2024) -
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
di: Chen, Runnan, et al.
Pubblicazione: (2024) -
DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes
di: Hou, Qirui, et al.
Pubblicazione: (2025)