Zoo3D: Zero-Shot 3D Object Detection at Scene Level
Fuente:
arXiv
Saved in:
| Main Authors: | Lemeshko, Andrey, Gabdullin, Bulat, Drozdov, Nikita, Konushin, Anton, Rukhovich, Danila, Kolodiazhnyi, Maksim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Z3D: Zero-Shot 3D Visual Grounding from Images
by: Drozdov, Nikita, et al.
Published: (2026)
by: Drozdov, Nikita, et al.
Published: (2026)
TUN3D: Towards Real-World Scene Understanding from Unposed Images
by: Konushin, Anton, et al.
Published: (2025)
by: Konushin, Anton, et al.
Published: (2025)
UniDet3D: Multi-dataset Indoor 3D Object Detection
by: Kolodiazhnyi, Maksim, et al.
Published: (2024)
by: Kolodiazhnyi, Maksim, et al.
Published: (2024)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task
by: Gabdullin, Bulat, et al.
Published: (2024)
by: Gabdullin, Bulat, et al.
Published: (2024)
Improving analytical color and texture similarity estimation methods for dataset-agnostic person reidentification
by: Gabdullin, Nikita
Published: (2024)
by: Gabdullin, Nikita
Published: (2024)
The effects of Hessian eigenvalue spectral density type on the applicability of Hessian analysis to generalization capability assessment of neural networks
by: Gabdullin, Nikita
Published: (2025)
by: Gabdullin, Nikita
Published: (2025)
Investigating generalization capabilities of neural networks by means of loss landscapes and Hessian analysis
by: Gabdullin, Nikita
Published: (2024)
by: Gabdullin, Nikita
Published: (2024)
MiCADangelo: Fine-Grained Reconstruction of Constrained CAD Models from 3D Scans
by: Karadeniz, Ahmet Serdar, et al.
Published: (2025)
by: Karadeniz, Ahmet Serdar, et al.
Published: (2025)
Using predefined vector systems as latent space configuration for neural network supervised training on data with arbitrarily large number of classes
by: Gabdullin, Nikita
Published: (2025)
by: Gabdullin, Nikita
Published: (2025)
Exploring possible vector systems for faster training of neural networks with preconfigured latent spaces
by: Gabdullin, Nikita
Published: (2025)
by: Gabdullin, Nikita
Published: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
by: Lin, Yuxuan, et al.
Published: (2025)
by: Lin, Yuxuan, et al.
Published: (2025)
Articulate3D: Zero-Shot Text-Driven 3D Object Posing
by: Deb, Oishi, et al.
Published: (2025)
by: Deb, Oishi, et al.
Published: (2025)
Cooperative Face Liveness Detection from Optical Flow
by: Sokolov, Artem, et al.
Published: (2025)
by: Sokolov, Artem, et al.
Published: (2025)
VIZOR: Viewpoint-Invariant Zero-Shot Scene Graph Generation for 3D Scene Reasoning
by: Madhavaram, Vivek, et al.
Published: (2026)
by: Madhavaram, Vivek, et al.
Published: (2026)
Accelerate 3D Object Detection Models via Zero-Shot Attention Key Pruning
by: Xu, Lizhen, et al.
Published: (2025)
by: Xu, Lizhen, et al.
Published: (2025)
Enabling Training-Free Text-Based Remote Sensing Segmentation
by: Sosa, Jose, et al.
Published: (2026)
by: Sosa, Jose, et al.
Published: (2026)
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
TransBridge: Boost 3D Object Detection by Scene-Level Completion with Transformer Decoder
by: Meng, Qinghao, et al.
Published: (2025)
by: Meng, Qinghao, et al.
Published: (2025)
Zero-Shot Multi-Object Scene Completion
by: Iwase, Shun, et al.
Published: (2024)
by: Iwase, Shun, et al.
Published: (2024)
Using predefined vector systems to speed up neural network multimillion class classification
by: Gabdullin, Nikita, et al.
Published: (2026)
by: Gabdullin, Nikita, et al.
Published: (2026)
Weak-to-Strong 3D Object Detection with X-Ray Distillation
by: Gambashidze, Alexander, et al.
Published: (2024)
by: Gambashidze, Alexander, et al.
Published: (2024)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
by: Gong, Bingchen, et al.
Published: (2024)
by: Gong, Bingchen, et al.
Published: (2024)
SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model
by: Zhang, Dingyuan, et al.
Published: (2023)
by: Zhang, Dingyuan, et al.
Published: (2023)
FAWN: Floor-And-Walls Normal Regularization for Direct Neural TSDF Reconstruction
by: Sokolova, Anna, et al.
Published: (2024)
by: Sokolova, Anna, et al.
Published: (2024)
Zero-Shot Scene Change Detection
by: Cho, Kyusik, et al.
Published: (2024)
by: Cho, Kyusik, et al.
Published: (2024)
DynaMix: Generalizable Person Re-identification via Dynamic Relabeling and Mixed Data Sampling
by: Mamedov, Timur, et al.
Published: (2025)
by: Mamedov, Timur, et al.
Published: (2025)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
by: Mamedov, Timur, et al.
Published: (2024)
by: Mamedov, Timur, et al.
Published: (2024)
3D Roadway Scene Object Detection with LIDARs in Snowfall Conditions
by: Farhani, Ghazal, et al.
Published: (2025)
by: Farhani, Ghazal, et al.
Published: (2025)
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
by: Sun, Xuefei, et al.
Published: (2026)
by: Sun, Xuefei, et al.
Published: (2026)
A3D: Does Diffusion Dream about 3D Alignment?
by: Ignatyev, Savva, et al.
Published: (2024)
by: Ignatyev, Savva, et al.
Published: (2024)
ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing
by: Tang, Xiang, et al.
Published: (2025)
by: Tang, Xiang, et al.
Published: (2025)
Approaching Outside: Scaling Unsupervised 3D Object Detection from 2D Scene
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
Descrip3D: Enhancing Large Language Model-based 3D Scene Understanding with Object-Level Text Descriptions
by: Xue, Jintang, et al.
Published: (2025)
by: Xue, Jintang, et al.
Published: (2025)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
by: Yang, Anqi Joyce, et al.
Published: (2026)
by: Yang, Anqi Joyce, et al.
Published: (2026)
IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object Detection
by: Yin, Junbo, et al.
Published: (2024)
by: Yin, Junbo, et al.
Published: (2024)
Few-Shot Learning in Video and 3D Object Detection: A Survey
by: Ferdaus, Md Meftahul, et al.
Published: (2025)
by: Ferdaus, Md Meftahul, et al.
Published: (2025)
Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments
by: Zhu, Yun, et al.
Published: (2026)
by: Zhu, Yun, et al.
Published: (2026)
ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment
by: Dong, Mingyu, et al.
Published: (2026)
by: Dong, Mingyu, et al.
Published: (2026)
Similar Items
-
Z3D: Zero-Shot 3D Visual Grounding from Images
by: Drozdov, Nikita, et al.
Published: (2026) -
TUN3D: Towards Real-World Scene Understanding from Unposed Images
by: Konushin, Anton, et al.
Published: (2025) -
UniDet3D: Multi-dataset Indoor 3D Object Detection
by: Kolodiazhnyi, Maksim, et al.
Published: (2024) -
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025) -
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task
by: Gabdullin, Bulat, et al.
Published: (2024)