Saved in:
| Main Authors: | Dai, Yuhang, Yang, Xingyi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.14048 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hash3D: Training-free Acceleration for 3D Generation
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
Test3R: Learning to Reconstruct 3D at Test Time
by: Yuan, Yuheng, et al.
Published: (2025)
by: Yuan, Yuheng, et al.
Published: (2025)
WorldWarp: Propagating 3D Geometry with Asynchronous Video Diffusion
by: Kong, Hanyang, et al.
Published: (2025)
by: Kong, Hanyang, et al.
Published: (2025)
Attention Itself Could Retrieve.RetrieveVGGT: Training-Free Long Context Streaming 3D Reconstruction via Query-Key Similarity Retrieval
by: Zou, Zichen, et al.
Published: (2026)
by: Zou, Zichen, et al.
Published: (2026)
iBA: Backdoor Attack on 3D Point Cloud via Reconstructing Itself
by: Bian, Yuhao, et al.
Published: (2024)
by: Bian, Yuhao, et al.
Published: (2024)
GRACE: Estimating Geometry-level 3D Human-Scene Contact from 2D Images
by: Wang, Chengfeng, et al.
Published: (2025)
by: Wang, Chengfeng, et al.
Published: (2025)
GTR: Improving Large 3D Reconstruction Models through Geometry and Texture Refinement
by: Zhuang, Peiye, et al.
Published: (2024)
by: Zhuang, Peiye, et al.
Published: (2024)
SSR: A Training-Free Approach for Streaming 3D Reconstruction
by: Deng, Hui, et al.
Published: (2026)
by: Deng, Hui, et al.
Published: (2026)
Kinematics-based 3D Human-Object Interaction Reconstruction from Single View
by: Chen, Yuhang, et al.
Published: (2024)
by: Chen, Yuhang, et al.
Published: (2024)
Flash Sculptor: Modular 3D Worlds from Objects
by: Hu, Yujia, et al.
Published: (2025)
by: Hu, Yujia, et al.
Published: (2025)
GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction
by: Lin, Weiquan, et al.
Published: (2026)
by: Lin, Weiquan, et al.
Published: (2026)
FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling
by: Qiu, Haonan, et al.
Published: (2023)
by: Qiu, Haonan, et al.
Published: (2023)
Gesplat: Robust Pose-Free 3D Reconstruction via Geometry-Guided Gaussian Splatting
by: Lu, Jiahui, et al.
Published: (2025)
by: Lu, Jiahui, et al.
Published: (2025)
Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding
by: Zheng, Hongpei, et al.
Published: (2025)
by: Zheng, Hongpei, et al.
Published: (2025)
MC3D-AD: A Unified Geometry-aware Reconstruction Model for Multi-category 3D Anomaly Detection
by: Cheng, Jiayi, et al.
Published: (2025)
by: Cheng, Jiayi, et al.
Published: (2025)
Guiding a Diffusion Model with a Bad Version of Itself
by: Karras, Tero, et al.
Published: (2024)
by: Karras, Tero, et al.
Published: (2024)
C4D: 4D Made from 3D through Dual Correspondences
by: Wang, Shizun, et al.
Published: (2025)
by: Wang, Shizun, et al.
Published: (2025)
Refined Geometry-guided Head Avatar Reconstruction from Monocular RGB Video
by: Park, Pilseo, et al.
Published: (2025)
by: Park, Pilseo, et al.
Published: (2025)
UAVFF3D: A Geometry-Aware Benchmark for Feed-Forward UAV 3D Reconstruction
by: Yang, Xiang, et al.
Published: (2026)
by: Yang, Xiang, et al.
Published: (2026)
CaptionQA: Is Your Caption as Useful as the Image Itself?
by: Yang, Shijia, et al.
Published: (2025)
by: Yang, Shijia, et al.
Published: (2025)
IMFine: 3D Inpainting via Geometry-guided Multi-view Refinement
by: Shi, Zhihao, et al.
Published: (2025)
by: Shi, Zhihao, et al.
Published: (2025)
LoL: Longer than Longer, Scaling Video Generation to Hour
by: Cui, Justin, et al.
Published: (2026)
by: Cui, Justin, et al.
Published: (2026)
Focus on Neighbors and Know the Whole: Towards Consistent Dense Multiview Text-to-Image Generator for 3D Creation
by: Li, Bonan, et al.
Published: (2024)
by: Li, Bonan, et al.
Published: (2024)
MoRE: 3D Visual Geometry Reconstruction Meets Mixture-of-Experts
by: Gao, Jingnan, et al.
Published: (2025)
by: Gao, Jingnan, et al.
Published: (2025)
$R^2$-Mesh: Reinforcement Learning Powered Mesh Reconstruction via Geometry and Appearance Refinement
by: Wang, Haoyang, et al.
Published: (2024)
by: Wang, Haoyang, et al.
Published: (2024)
Language Model as Visual Explainer
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
Compositional Video Generation as Flow Equalization
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting
by: Tran, Anh Thuan, et al.
Published: (2026)
by: Tran, Anh Thuan, et al.
Published: (2026)
End-to-End Spatial-Temporal Transformer for Real-time 4D HOI Reconstruction
by: Zhang, Haoyu, et al.
Published: (2026)
by: Zhang, Haoyu, et al.
Published: (2026)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
by: Shen, Qiuhong, et al.
Published: (2024)
by: Shen, Qiuhong, et al.
Published: (2024)
Multi-dimensional Preference Alignment by Conditioning Reward Itself
by: Jang, Jiho, et al.
Published: (2025)
by: Jang, Jiho, et al.
Published: (2025)
Unified Scene Representation and Reconstruction for 3D Large Language Models
by: Chu, Tao, et al.
Published: (2024)
by: Chu, Tao, et al.
Published: (2024)
1000+ FPS 4D Gaussian Splatting for Dynamic Scene Rendering
by: Yuan, Yuheng, et al.
Published: (2025)
by: Yuan, Yuheng, et al.
Published: (2025)
GeoRect4D: Geometry-Compatible Generative Rectification for Dynamic Sparse-View 3D Reconstruction
by: Wu, Zhenlong, et al.
Published: (2026)
by: Wu, Zhenlong, et al.
Published: (2026)
Geometry Cloak: Preventing TGS-based 3D Reconstruction from Copyrighted Images
by: Song, Qi, et al.
Published: (2024)
by: Song, Qi, et al.
Published: (2024)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
by: Shao, Yawen, et al.
Published: (2024)
by: Shao, Yawen, et al.
Published: (2024)
Joint Reconstruction of 3D Human and Object via Contact-Based Refinement Transformer
by: Nam, Hyeongjin, et al.
Published: (2024)
by: Nam, Hyeongjin, et al.
Published: (2024)
A Refined 3D Gaussian Representation for High-Quality Dynamic Scene Reconstruction
by: Zhang, Bin, et al.
Published: (2024)
by: Zhang, Bin, et al.
Published: (2024)
MVBoost: Boost 3D Reconstruction with Multi-View Refinement
by: Liu, Xiangyu, et al.
Published: (2024)
by: Liu, Xiangyu, et al.
Published: (2024)
Similar Items
-
Hash3D: Training-free Acceleration for 3D Generation
by: Yang, Xingyi, et al.
Published: (2024) -
Test3R: Learning to Reconstruct 3D at Test Time
by: Yuan, Yuheng, et al.
Published: (2025) -
WorldWarp: Propagating 3D Geometry with Asynchronous Video Diffusion
by: Kong, Hanyang, et al.
Published: (2025) -
Attention Itself Could Retrieve.RetrieveVGGT: Training-Free Long Context Streaming 3D Reconstruction via Query-Key Similarity Retrieval
by: Zou, Zichen, et al.
Published: (2026) -
iBA: Backdoor Attack on 3D Point Cloud via Reconstructing Itself
by: Bian, Yuhao, et al.
Published: (2024)