Zero-Shot Multi-Object Scene Completion
Fuente:
arXiv
Saved in:
| Main Authors: | Iwase, Shun, Liu, Katherine, Guizilini, Vitor, Gaidon, Adrien, Kitani, Kris, Ambrus, Rares, Zakharov, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
by: Iwase, Shun, et al.
Published: (2025)
by: Iwase, Shun, et al.
Published: (2025)
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
by: Zakharov, Sergey, et al.
Published: (2024)
by: Zakharov, Sergey, et al.
Published: (2024)
OmniShape: Zero-Shot Multi-Hypothesis Shape and Pose Estimation in the Real World
by: Liu, Katherine, et al.
Published: (2025)
by: Liu, Katherine, et al.
Published: (2025)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
$SE(3)$ Equivariant Ray Embeddings for Implicit Multi-View Depth Estimation
by: Xu, Yinshuang, et al.
Published: (2024)
by: Xu, Yinshuang, et al.
Published: (2024)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
by: Guizilini, Vitor, et al.
Published: (2024)
by: Guizilini, Vitor, et al.
Published: (2024)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
by: Guizilini, Vitor, et al.
Published: (2025)
by: Guizilini, Vitor, et al.
Published: (2025)
Incorporating dense metric depth into neural 3D representations for view synthesis and relighting
by: Chaudhury, Arkadeep Narayan, et al.
Published: (2024)
by: Chaudhury, Arkadeep Narayan, et al.
Published: (2024)
View-Invariant Policy Learning via Zero-Shot Novel View Synthesis
by: Tian, Stephen, et al.
Published: (2024)
by: Tian, Stephen, et al.
Published: (2024)
Generalizable Neural Human Renderer
by: Masuda, Mana, et al.
Published: (2024)
by: Masuda, Mana, et al.
Published: (2024)
GTR: Gaussian Splatting Tracking and Reconstruction of Unknown Objects Based on Appearance and Geometric Complexity
by: Ikeda, Takuya, et al.
Published: (2025)
by: Ikeda, Takuya, et al.
Published: (2025)
DegustaBot: Zero-Shot Visual Preference Estimation for Personalized Multi-Object Rearrangement
by: Newman, Benjamin A., et al.
Published: (2024)
by: Newman, Benjamin A., et al.
Published: (2024)
CARTO: Category and Joint Agnostic Reconstruction of ARTiculated Objects
by: Heppert, Nick, et al.
Published: (2023)
by: Heppert, Nick, et al.
Published: (2023)
Multi-Object Tracking by Hierarchical Visual Representations
by: Cao, Jinkun, et al.
Published: (2024)
by: Cao, Jinkun, et al.
Published: (2024)
Understanding Video Transformers via Universal Concept Discovery
by: Kowal, Matthew, et al.
Published: (2024)
by: Kowal, Matthew, et al.
Published: (2024)
AnyView: Synthesizing Any Novel View in Dynamic Scenes
by: Van Hoorick, Basile, et al.
Published: (2026)
by: Van Hoorick, Basile, et al.
Published: (2026)
DiffusionNOCS: Managing Symmetry and Uncertainty in Sim2Real Multi-Modal Category-level Pose Estimation
by: Ikeda, Takuya, et al.
Published: (2024)
by: Ikeda, Takuya, et al.
Published: (2024)
Joint Diffusion for Universal Hand-Object Grasp Generation
by: Cao, Jinkun, et al.
Published: (2024)
by: Cao, Jinkun, et al.
Published: (2024)
Transcrib3D: 3D Referring Expression Resolution through Large Language Models
by: Fang, Jiading, et al.
Published: (2024)
by: Fang, Jiading, et al.
Published: (2024)
G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp Synthesis
by: Ye, Yufei, et al.
Published: (2024)
by: Ye, Yufei, et al.
Published: (2024)
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024)
by: Ran, Haoxi, et al.
Published: (2024)
REST3D: Reconstructing Physically Stable 3D Scenes from a Single Image
by: Ma, Xiaoxuan, et al.
Published: (2026)
by: Ma, Xiaoxuan, et al.
Published: (2026)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)
by: Lin, Shengjie, et al.
Published: (2025)
Espresso: High Compression For Rich Extraction From Videos for Your Vision-Language Model
by: Yu, Keunwoo Peter, et al.
Published: (2024)
by: Yu, Keunwoo Peter, et al.
Published: (2024)
Multi-Person 3D Pose Estimation from Multi-View Uncalibrated Depth Cameras
by: Li, Yu-Jhe, et al.
Published: (2024)
by: Li, Yu-Jhe, et al.
Published: (2024)
CacheFlow: Fast Human Motion Prediction by Cached Normalizing Flow
by: Maeda, Takahiro, et al.
Published: (2025)
by: Maeda, Takahiro, et al.
Published: (2025)
NovaPlan: Zero-Shot Long-Horizon Manipulation via Closed-Loop Video Language Planning
by: Fu, Jiahui, et al.
Published: (2026)
by: Fu, Jiahui, et al.
Published: (2026)
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
by: Lemeshko, Andrey, et al.
Published: (2025)
by: Lemeshko, Andrey, et al.
Published: (2025)
Zero-Shot Scene Change Detection
by: Cho, Kyusik, et al.
Published: (2024)
by: Cho, Kyusik, et al.
Published: (2024)
RaySt3R: Predicting Novel Depth Maps for Zero-Shot Object Completion
by: Duisterhof, Bardienus P., et al.
Published: (2025)
by: Duisterhof, Bardienus P., et al.
Published: (2025)
MGF: Mixed Gaussian Flow for Diverse Trajectory Prediction
by: Chen, Jiahe, et al.
Published: (2024)
by: Chen, Jiahe, et al.
Published: (2024)
ExpertAF: Expert Actionable Feedback from Video
by: Ashutosh, Kumar, et al.
Published: (2024)
by: Ashutosh, Kumar, et al.
Published: (2024)
Streaming Detection of Queried Event Start
by: Eyzaguirre, Cristobal, et al.
Published: (2024)
by: Eyzaguirre, Cristobal, et al.
Published: (2024)
SIRE: SE(3) Intrinsic Rigidity Embeddings
by: Smith, Cameron, et al.
Published: (2025)
by: Smith, Cameron, et al.
Published: (2025)
Zero-Shot Monocular Scene Flow Estimation in the Wild
by: Liang, Yiqing, et al.
Published: (2025)
by: Liang, Yiqing, et al.
Published: (2025)
Self-Supervised Geometry-Guided Initialization for Robust Monocular Visual Odometry
by: Kanai, Takayuki, et al.
Published: (2024)
by: Kanai, Takayuki, et al.
Published: (2024)
Harmony4D: A Video Dataset for In-The-Wild Close Human Interactions
by: Khirodkar, Rawal, et al.
Published: (2024)
by: Khirodkar, Rawal, et al.
Published: (2024)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
by: Jin, Qunchao, et al.
Published: (2025)
by: Jin, Qunchao, et al.
Published: (2025)
Expanding Zero-Shot Object Counting with Rich Prompts
by: Zhu, Huilin, et al.
Published: (2025)
by: Zhu, Huilin, et al.
Published: (2025)
CycleSAM: Few-Shot Surgical Scene Segmentation with Cycle- and Scene-Consistent Feature Matching
by: Murali, Aditya, et al.
Published: (2024)
by: Murali, Aditya, et al.
Published: (2024)
Similar Items
-
ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
by: Iwase, Shun, et al.
Published: (2025) -
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
by: Zakharov, Sergey, et al.
Published: (2024) -
OmniShape: Zero-Shot Multi-Hypothesis Shape and Pose Estimation in the Real World
by: Liu, Katherine, et al.
Published: (2025) -
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024) -
$SE(3)$ Equivariant Ray Embeddings for Implicit Multi-View Depth Estimation
by: Xu, Yinshuang, et al.
Published: (2024)