Z3D: Zero-Shot 3D Visual Grounding from Images
Fuente:
arXiv
Guardado en:
| Autores principales: | Drozdov, Nikita, Lemeshko, Andrey, Gavrilov, Nikita, Konushin, Anton, Rukhovich, Danila, Kolodiazhnyi, Maksim |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
por: Lemeshko, Andrey, et al.
Publicado: (2025)
por: Lemeshko, Andrey, et al.
Publicado: (2025)
TUN3D: Towards Real-World Scene Understanding from Unposed Images
por: Konushin, Anton, et al.
Publicado: (2025)
por: Konushin, Anton, et al.
Publicado: (2025)
UniDet3D: Multi-dataset Indoor 3D Object Detection
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2024)
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2024)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2025)
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2025)
Learning KAN-based Implicit Neural Representations for Deformable Image Registration
por: Drozdov, Nikita, et al.
Publicado: (2025)
por: Drozdov, Nikita, et al.
Publicado: (2025)
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
por: Skorokhodov, Vsevolod, et al.
Publicado: (2025)
por: Skorokhodov, Vsevolod, et al.
Publicado: (2025)
Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding
por: Yin, Yufei, et al.
Publicado: (2026)
por: Yin, Yufei, et al.
Publicado: (2026)
MiCADangelo: Fine-Grained Reconstruction of Constrained CAD Models from 3D Scans
por: Karadeniz, Ahmet Serdar, et al.
Publicado: (2025)
por: Karadeniz, Ahmet Serdar, et al.
Publicado: (2025)
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
por: Yuan, Qihao, et al.
Publicado: (2024)
por: Yuan, Qihao, et al.
Publicado: (2024)
Zero-Shot 3D Visual Grounding from Vision-Language Models
por: Li, Rong, et al.
Publicado: (2025)
por: Li, Rong, et al.
Publicado: (2025)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
por: Li, Rong, et al.
Publicado: (2024)
por: Li, Rong, et al.
Publicado: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
por: Liao, Liwei, et al.
Publicado: (2025)
por: Liao, Liwei, et al.
Publicado: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
por: Xu, Runsen, et al.
Publicado: (2024)
por: Xu, Runsen, et al.
Publicado: (2024)
Visual Implicit Geometry Transformer for Autonomous Driving
por: Shirokov, Arsenii, et al.
Publicado: (2026)
por: Shirokov, Arsenii, et al.
Publicado: (2026)
ORIGEN: Zero-Shot 3D Orientation Grounding in Text-to-Image Generation
por: Min, Yunhong, et al.
Publicado: (2025)
por: Min, Yunhong, et al.
Publicado: (2025)
Enabling Training-Free Text-Based Remote Sensing Segmentation
por: Sosa, Jose, et al.
Publicado: (2026)
por: Sosa, Jose, et al.
Publicado: (2026)
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
por: Sosa, Jose, et al.
Publicado: (2025)
por: Sosa, Jose, et al.
Publicado: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
por: Lin, Yuxuan, et al.
Publicado: (2025)
por: Lin, Yuxuan, et al.
Publicado: (2025)
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
por: Huynh, Cuong, et al.
Publicado: (2026)
por: Huynh, Cuong, et al.
Publicado: (2026)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
por: Mamedov, Timur, et al.
Publicado: (2026)
por: Mamedov, Timur, et al.
Publicado: (2026)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
por: Yuan, Zhihao, et al.
Publicado: (2023)
por: Yuan, Zhihao, et al.
Publicado: (2023)
DynaMix: Generalizable Person Re-identification via Dynamic Relabeling and Mixed Data Sampling
por: Mamedov, Timur, et al.
Publicado: (2025)
por: Mamedov, Timur, et al.
Publicado: (2025)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
por: Mamedov, Timur, et al.
Publicado: (2024)
por: Mamedov, Timur, et al.
Publicado: (2024)
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
por: Sun, Xuefei, et al.
Publicado: (2026)
por: Sun, Xuefei, et al.
Publicado: (2026)
Cooperative Face Liveness Detection from Optical Flow
por: Sokolov, Artem, et al.
Publicado: (2025)
por: Sokolov, Artem, et al.
Publicado: (2025)
A3D: Does Diffusion Dream about 3D Alignment?
por: Ignatyev, Savva, et al.
Publicado: (2024)
por: Ignatyev, Savva, et al.
Publicado: (2024)
pySpatial: Generating 3D Visual Programs for Zero-Shot Spatial Reasoning
por: Luo, Zhanpeng, et al.
Publicado: (2026)
por: Luo, Zhanpeng, et al.
Publicado: (2026)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
por: Wang, Haibo, et al.
Publicado: (2026)
por: Wang, Haibo, et al.
Publicado: (2026)
Grounding Descriptions in Images informs Zero-Shot Visual Recognition
por: Halbe, Shaunak, et al.
Publicado: (2024)
por: Halbe, Shaunak, et al.
Publicado: (2024)
VGGT: Visual Geometry Grounded Transformer
por: Wang, Jianyuan, et al.
Publicado: (2025)
por: Wang, Jianyuan, et al.
Publicado: (2025)
Articulate3D: Zero-Shot Text-Driven 3D Object Posing
por: Deb, Oishi, et al.
Publicado: (2025)
por: Deb, Oishi, et al.
Publicado: (2025)
CAD-Recode: Reverse Engineering CAD Code from Point Clouds
por: Rukhovich, Danila, et al.
Publicado: (2024)
por: Rukhovich, Danila, et al.
Publicado: (2024)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
por: T, Mukund Varma, et al.
Publicado: (2024)
por: T, Mukund Varma, et al.
Publicado: (2024)
SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding
por: Jin, Zhao, et al.
Publicado: (2025)
por: Jin, Zhao, et al.
Publicado: (2025)
ChangingGrounding: 3D Visual Grounding in Changing Scenes
por: Hu, Miao, et al.
Publicado: (2025)
por: Hu, Miao, et al.
Publicado: (2025)
SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Instance Segmentation
por: Xu, Mutian, et al.
Publicado: (2023)
por: Xu, Mutian, et al.
Publicado: (2023)
Zero-Shot Reconstruction of Animatable 3D Avatars with Cloth Dynamics from a Single Image
por: Kwon, Joohyun, et al.
Publicado: (2026)
por: Kwon, Joohyun, et al.
Publicado: (2026)
3D-PointZshotS: Geometry-Aware 3D Point Cloud Zero-Shot Semantic Segmentation Narrowing the Visual-Semantic Gap
por: Yang, Minmin, et al.
Publicado: (2025)
por: Yang, Minmin, et al.
Publicado: (2025)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
por: Yang, Anqi Joyce, et al.
Publicado: (2026)
por: Yang, Anqi Joyce, et al.
Publicado: (2026)
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
por: Lin, Jiawen, et al.
Publicado: (2025)
por: Lin, Jiawen, et al.
Publicado: (2025)
Ejemplares similares
-
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
por: Lemeshko, Andrey, et al.
Publicado: (2025) -
TUN3D: Towards Real-World Scene Understanding from Unposed Images
por: Konushin, Anton, et al.
Publicado: (2025) -
UniDet3D: Multi-dataset Indoor 3D Object Detection
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2024) -
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
por: Kolodiazhnyi, Maksim, et al.
Publicado: (2025) -
Learning KAN-based Implicit Neural Representations for Deformable Image Registration
por: Drozdov, Nikita, et al.
Publicado: (2025)