Hypo3D: Exploring Hypothetical Reasoning in 3D
Fuente:
arXiv
Guardado en:
| Autores principales: | Mao, Ye, Luo, Weixun, Jing, Junpeng, Qiu, Anlan, Mikolajczyk, Krystian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
POMA-3D: The Point Map Way to 3D Scene Understanding
por: Mao, Ye, et al.
Publicado: (2025)
por: Mao, Ye, et al.
Publicado: (2025)
Match Stereo Videos via Bidirectional Alignment
por: Jing, Junpeng, et al.
Publicado: (2024)
por: Jing, Junpeng, et al.
Publicado: (2024)
Contrastive Language-Colored Pointmap Pretraining for Unified 3D Scene Understanding
por: Mao, Ye, et al.
Publicado: (2026)
por: Mao, Ye, et al.
Publicado: (2026)
Stereo Any Video: Temporally Consistent Stereo Matching
por: Jing, Junpeng, et al.
Publicado: (2025)
por: Jing, Junpeng, et al.
Publicado: (2025)
Lite Any Stereo: Efficient Zero-Shot Stereo Matching
por: Jing, Junpeng, et al.
Publicado: (2025)
por: Jing, Junpeng, et al.
Publicado: (2025)
From None to All: Self-Supervised 3D Reconstruction via Novel View Synthesis
por: Huang, Ranran, et al.
Publicado: (2026)
por: Huang, Ranran, et al.
Publicado: (2026)
Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching
por: Jing, Junpeng, et al.
Publicado: (2024)
por: Jing, Junpeng, et al.
Publicado: (2024)
OpenDlign: Open-World Point Cloud Understanding with Depth-Aligned Images
por: Mao, Ye, et al.
Publicado: (2024)
por: Mao, Ye, et al.
Publicado: (2024)
No Pose at All: Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
por: Huang, Ranran, et al.
Publicado: (2025)
por: Huang, Ranran, et al.
Publicado: (2025)
SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
por: Huang, Ranran, et al.
Publicado: (2025)
por: Huang, Ranran, et al.
Publicado: (2025)
UCorr: Wire Detection and Depth Estimation for Autonomous Drones
por: Kolbeinsson, Benedikt, et al.
Publicado: (2025)
por: Kolbeinsson, Benedikt, et al.
Publicado: (2025)
Multi-Class Segmentation from Aerial Views using Recursive Noise Diffusion
por: Kolbeinsson, Benedikt, et al.
Publicado: (2022)
por: Kolbeinsson, Benedikt, et al.
Publicado: (2022)
DDOS: The Drone Depth and Obstacle Segmentation Dataset
por: Kolbeinsson, Benedikt, et al.
Publicado: (2023)
por: Kolbeinsson, Benedikt, et al.
Publicado: (2023)
Language-Based Depth Hints for Monocular Depth Estimation
por: Auty, Dylan, et al.
Publicado: (2024)
por: Auty, Dylan, et al.
Publicado: (2024)
Understanding the Role of the Projector in Knowledge Distillation
por: Miles, Roy, et al.
Publicado: (2023)
por: Miles, Roy, et al.
Publicado: (2023)
Closed Loop Interactive Embodied Reasoning for Robot Manipulation
por: Nazarczuk, Michal, et al.
Publicado: (2024)
por: Nazarczuk, Michal, et al.
Publicado: (2024)
Learning to Project for Cross-Task Knowledge Distillation
por: Auty, Dylan, et al.
Publicado: (2024)
por: Auty, Dylan, et al.
Publicado: (2024)
Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning
por: He, Qingdong, et al.
Publicado: (2025)
por: He, Qingdong, et al.
Publicado: (2025)
SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions
por: Feng, Jie, et al.
Publicado: (2026)
por: Feng, Jie, et al.
Publicado: (2026)
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
por: Wei, Zeming, et al.
Publicado: (2025)
por: Wei, Zeming, et al.
Publicado: (2025)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
por: Huang, Kuan-Chih, et al.
Publicado: (2024)
por: Huang, Kuan-Chih, et al.
Publicado: (2024)
Semi-Supervised Diversity-Aware Domain Adaptation for 3D Object detection
por: Olber, Bartłomiej, et al.
Publicado: (2025)
por: Olber, Bartłomiej, et al.
Publicado: (2025)
Pathformer3D: A 3D Scanpath Transformer for 360° Images
por: Quan, Rong, et al.
Publicado: (2024)
por: Quan, Rong, et al.
Publicado: (2024)
H3D-DGS: Exploring Heterogeneous 3D Motion Representation for Deformable 3D Gaussian Splatting
por: He, Bing, et al.
Publicado: (2024)
por: He, Bing, et al.
Publicado: (2024)
SGS-3D: High-Fidelity 3D Instance Segmentation via Reliable Semantic Mask Splitting and Growing
por: Wang, Chaolei, et al.
Publicado: (2025)
por: Wang, Chaolei, et al.
Publicado: (2025)
Omni123: Exploring 3D Native Foundation Models with Limited 3D Data by Unifying Text to 2D and 3D Generation
por: Ye, Chongjie, et al.
Publicado: (2026)
por: Ye, Chongjie, et al.
Publicado: (2026)
GCA-3D: Towards Generalized and Consistent Domain Adaptation of 3D Generators
por: Li, Hengjia, et al.
Publicado: (2024)
por: Li, Hengjia, et al.
Publicado: (2024)
Exploring Recurrent Long-term Temporal Fusion for Multi-view 3D Perception
por: Han, Chunrui, et al.
Publicado: (2023)
por: Han, Chunrui, et al.
Publicado: (2023)
3D4D: An Interactive, Editable, 4D World Model via 3D Video Generation
por: He, Yunhong, et al.
Publicado: (2025)
por: He, Yunhong, et al.
Publicado: (2025)
VCVW-3D: A Virtual Construction Vehicles and Workers Dataset with 3D Annotations
por: Ding, Yuexiong, et al.
Publicado: (2023)
por: Ding, Yuexiong, et al.
Publicado: (2023)
Story3D-Agent: Exploring 3D Storytelling Visualization with Large Language Models
por: Huang, Yuzhou, et al.
Publicado: (2024)
por: Huang, Yuzhou, et al.
Publicado: (2024)
3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding
por: Huang, Ting, et al.
Publicado: (2025)
por: Huang, Ting, et al.
Publicado: (2025)
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation
por: Huang, Jiaxin, et al.
Publicado: (2025)
por: Huang, Jiaxin, et al.
Publicado: (2025)
Brain3D: EEG-to-3D Decoding of Visual Representations via Multimodal Reasoning
por: Balloni, Emanuele, et al.
Publicado: (2026)
por: Balloni, Emanuele, et al.
Publicado: (2026)
I2V3D: Controllable image-to-video generation with 3D guidance
por: Zhang, Zhiyuan, et al.
Publicado: (2025)
por: Zhang, Zhiyuan, et al.
Publicado: (2025)
R3D-AD: Reconstruction via Diffusion for 3D Anomaly Detection
por: Zhou, Zheyuan, et al.
Publicado: (2024)
por: Zhou, Zheyuan, et al.
Publicado: (2024)
M3: 3D-Spatial MultiModal Memory
por: Zou, Xueyan, et al.
Publicado: (2025)
por: Zou, Xueyan, et al.
Publicado: (2025)
ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning
por: Zhang, Yiming, et al.
Publicado: (2026)
por: Zhang, Yiming, et al.
Publicado: (2026)
Towards 3D VR-Sketch to 3D Shape Retrieval
por: Luo, Ling, et al.
Publicado: (2022)
por: Luo, Ling, et al.
Publicado: (2022)
SYM3D: Learning Symmetric Triplanes for Better 3D-Awareness of GANs
por: Yang, Jing, et al.
Publicado: (2024)
por: Yang, Jing, et al.
Publicado: (2024)
Ejemplares similares
-
POMA-3D: The Point Map Way to 3D Scene Understanding
por: Mao, Ye, et al.
Publicado: (2025) -
Match Stereo Videos via Bidirectional Alignment
por: Jing, Junpeng, et al.
Publicado: (2024) -
Contrastive Language-Colored Pointmap Pretraining for Unified 3D Scene Understanding
por: Mao, Ye, et al.
Publicado: (2026) -
Stereo Any Video: Temporally Consistent Stereo Matching
por: Jing, Junpeng, et al.
Publicado: (2025) -
Lite Any Stereo: Efficient Zero-Shot Stereo Matching
por: Jing, Junpeng, et al.
Publicado: (2025)