Z3D: Zero-Shot 3D Visual Grounding from Images
Fuente:
arXiv
Saved in:
| Main Authors: | Drozdov, Nikita, Lemeshko, Andrey, Gavrilov, Nikita, Konushin, Anton, Rukhovich, Danila, Kolodiazhnyi, Maksim |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
by: Lemeshko, Andrey, et al.
Published: (2025)
by: Lemeshko, Andrey, et al.
Published: (2025)
TUN3D: Towards Real-World Scene Understanding from Unposed Images
by: Konushin, Anton, et al.
Published: (2025)
by: Konushin, Anton, et al.
Published: (2025)
UniDet3D: Multi-dataset Indoor 3D Object Detection
by: Kolodiazhnyi, Maksim, et al.
Published: (2024)
by: Kolodiazhnyi, Maksim, et al.
Published: (2024)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
Learning KAN-based Implicit Neural Representations for Deformable Image Registration
by: Drozdov, Nikita, et al.
Published: (2025)
by: Drozdov, Nikita, et al.
Published: (2025)
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding
by: Yin, Yufei, et al.
Published: (2026)
by: Yin, Yufei, et al.
Published: (2026)
MiCADangelo: Fine-Grained Reconstruction of Constrained CAD Models from 3D Scans
by: Karadeniz, Ahmet Serdar, et al.
Published: (2025)
by: Karadeniz, Ahmet Serdar, et al.
Published: (2025)
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
by: Yuan, Qihao, et al.
Published: (2024)
by: Yuan, Qihao, et al.
Published: (2024)
Zero-Shot 3D Visual Grounding from Vision-Language Models
by: Li, Rong, et al.
Published: (2025)
by: Li, Rong, et al.
Published: (2025)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)
by: Li, Rong, et al.
Published: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
by: Liao, Liwei, et al.
Published: (2025)
by: Liao, Liwei, et al.
Published: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
by: Xu, Runsen, et al.
Published: (2024)
by: Xu, Runsen, et al.
Published: (2024)
Visual Implicit Geometry Transformer for Autonomous Driving
by: Shirokov, Arsenii, et al.
Published: (2026)
by: Shirokov, Arsenii, et al.
Published: (2026)
ORIGEN: Zero-Shot 3D Orientation Grounding in Text-to-Image Generation
by: Min, Yunhong, et al.
Published: (2025)
by: Min, Yunhong, et al.
Published: (2025)
Enabling Training-Free Text-Based Remote Sensing Segmentation
by: Sosa, Jose, et al.
Published: (2026)
by: Sosa, Jose, et al.
Published: (2026)
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
by: Lin, Yuxuan, et al.
Published: (2025)
by: Lin, Yuxuan, et al.
Published: (2025)
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
by: Huynh, Cuong, et al.
Published: (2026)
by: Huynh, Cuong, et al.
Published: (2026)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
by: Mamedov, Timur, et al.
Published: (2026)
by: Mamedov, Timur, et al.
Published: (2026)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
by: Yuan, Zhihao, et al.
Published: (2023)
by: Yuan, Zhihao, et al.
Published: (2023)
DynaMix: Generalizable Person Re-identification via Dynamic Relabeling and Mixed Data Sampling
by: Mamedov, Timur, et al.
Published: (2025)
by: Mamedov, Timur, et al.
Published: (2025)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
by: Mamedov, Timur, et al.
Published: (2024)
by: Mamedov, Timur, et al.
Published: (2024)
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
by: Sun, Xuefei, et al.
Published: (2026)
by: Sun, Xuefei, et al.
Published: (2026)
Cooperative Face Liveness Detection from Optical Flow
by: Sokolov, Artem, et al.
Published: (2025)
by: Sokolov, Artem, et al.
Published: (2025)
A3D: Does Diffusion Dream about 3D Alignment?
by: Ignatyev, Savva, et al.
Published: (2024)
by: Ignatyev, Savva, et al.
Published: (2024)
pySpatial: Generating 3D Visual Programs for Zero-Shot Spatial Reasoning
by: Luo, Zhanpeng, et al.
Published: (2026)
by: Luo, Zhanpeng, et al.
Published: (2026)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
by: Wang, Haibo, et al.
Published: (2026)
by: Wang, Haibo, et al.
Published: (2026)
Grounding Descriptions in Images informs Zero-Shot Visual Recognition
by: Halbe, Shaunak, et al.
Published: (2024)
by: Halbe, Shaunak, et al.
Published: (2024)
VGGT: Visual Geometry Grounded Transformer
by: Wang, Jianyuan, et al.
Published: (2025)
by: Wang, Jianyuan, et al.
Published: (2025)
Articulate3D: Zero-Shot Text-Driven 3D Object Posing
by: Deb, Oishi, et al.
Published: (2025)
by: Deb, Oishi, et al.
Published: (2025)
CAD-Recode: Reverse Engineering CAD Code from Point Clouds
by: Rukhovich, Danila, et al.
Published: (2024)
by: Rukhovich, Danila, et al.
Published: (2024)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
by: T, Mukund Varma, et al.
Published: (2024)
by: T, Mukund Varma, et al.
Published: (2024)
SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding
by: Jin, Zhao, et al.
Published: (2025)
by: Jin, Zhao, et al.
Published: (2025)
ChangingGrounding: 3D Visual Grounding in Changing Scenes
by: Hu, Miao, et al.
Published: (2025)
by: Hu, Miao, et al.
Published: (2025)
SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Instance Segmentation
by: Xu, Mutian, et al.
Published: (2023)
by: Xu, Mutian, et al.
Published: (2023)
Zero-Shot Reconstruction of Animatable 3D Avatars with Cloth Dynamics from a Single Image
by: Kwon, Joohyun, et al.
Published: (2026)
by: Kwon, Joohyun, et al.
Published: (2026)
3D-PointZshotS: Geometry-Aware 3D Point Cloud Zero-Shot Semantic Segmentation Narrowing the Visual-Semantic Gap
by: Yang, Minmin, et al.
Published: (2025)
by: Yang, Minmin, et al.
Published: (2025)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
by: Yang, Anqi Joyce, et al.
Published: (2026)
by: Yang, Anqi Joyce, et al.
Published: (2026)
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
by: Lin, Jiawen, et al.
Published: (2025)
by: Lin, Jiawen, et al.
Published: (2025)
Similar Items
-
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
by: Lemeshko, Andrey, et al.
Published: (2025) -
TUN3D: Towards Real-World Scene Understanding from Unposed Images
by: Konushin, Anton, et al.
Published: (2025) -
UniDet3D: Multi-dataset Indoor 3D Object Detection
by: Kolodiazhnyi, Maksim, et al.
Published: (2024) -
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025) -
Learning KAN-based Implicit Neural Representations for Deformable Image Registration
by: Drozdov, Nikita, et al.
Published: (2025)