Joint 3D Geometry Reconstruction and Motion Generation for 4D Synthesis from a Single Image
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yanran, Wang, Ziyi, Zheng, Wenzhao, Zhu, Zheng, Zhou, Jie, Lu, Jiwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniPre3D: Unified Pre-training of 3D Point Cloud Models with Cross-Modal Gaussian Splatting
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
$\bf{D^3}$QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
OGGSplat: Open Gaussian Growing for Generalizable Reconstruction with Expanded Field-of-View
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2026)
von: Zhang, Yanran, et al.
Veröffentlicht: (2026)
Measuring 3D Spatial Geometric Consistency in Dynamic Generated Videos
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
Streaming 4D Visual Geometry Transformer
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
GlobalMamba: Global Image Serialization for Vision Mamba
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning
von: Li, Yifei, et al.
Veröffentlicht: (2025)
von: Li, Yifei, et al.
Veröffentlicht: (2025)
GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
DriveTok: 3D Driving Scene Tokenization for Unified Multi-View Reconstruction and Understanding
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
Hardness-Aware Scene Synthesis for Semi-Supervised 3D Object Detection
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
Path Choice Matters for Clear Attribution in Path Methods
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Preventing Local Pitfalls in Vector Quantization via Optimal Transport
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving
von: Fei, Xin, et al.
Veröffentlicht: (2024)
von: Fei, Xin, et al.
Veröffentlicht: (2024)
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
NeXT-IMDL: Build Benchmark for NeXT-Generation Image Manipulation Detection & Localization
von: Li, Yifei, et al.
Veröffentlicht: (2025)
von: Li, Yifei, et al.
Veröffentlicht: (2025)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding
von: Wu, Yuqi, et al.
Veröffentlicht: (2024)
von: Wu, Yuqi, et al.
Veröffentlicht: (2024)
QuadricFormer: Scene as Superquadrics for 3D Semantic Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
OmniIndoor3D: Comprehensive Indoor 3D Reconstruction
von: Wei, Xiaobao, et al.
Veröffentlicht: (2025)
von: Wei, Xiaobao, et al.
Veröffentlicht: (2025)
OccSora: 4D Occupancy Generation Models as World Simulators for Autonomous Driving
von: Wang, Lening, et al.
Veröffentlicht: (2024)
von: Wang, Lening, et al.
Veröffentlicht: (2024)
Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
SFTok: Bridging the Performance Gap in Discrete Tokenizers
von: Rao, Qihang, et al.
Veröffentlicht: (2025)
von: Rao, Qihang, et al.
Veröffentlicht: (2025)
Quantize-then-Rectify: Efficient VQ-VAE Training
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
Fast Shapley Value Estimation: A Unified Approach
von: Zhang, Borui, et al.
Veröffentlicht: (2023)
von: Zhang, Borui, et al.
Veröffentlicht: (2023)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
von: Fei, Xin, et al.
Veröffentlicht: (2024)
von: Fei, Xin, et al.
Veröffentlicht: (2024)
GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation
von: Zhang, Chubin, et al.
Veröffentlicht: (2024)
von: Zhang, Chubin, et al.
Veröffentlicht: (2024)
Terra: Explorable Native 3D World Model with Point Latents
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
Vega: Learning to Drive with Natural Language Instructions
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation
von: Xu, Xiuwei, et al.
Veröffentlicht: (2025)
von: Xu, Xiuwei, et al.
Veröffentlicht: (2025)
LiDAR-HMR: 3D Human Mesh Recovery from LiDAR
von: Fan, Bohao, et al.
Veröffentlicht: (2023)
von: Fan, Bohao, et al.
Veröffentlicht: (2023)
GeoRect4D: Geometry-Compatible Generative Rectification for Dynamic Sparse-View 3D Reconstruction
von: Wu, Zhenlong, et al.
Veröffentlicht: (2026)
von: Wu, Zhenlong, et al.
Veröffentlicht: (2026)
MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
von: Zhu, Ruijie, et al.
Veröffentlicht: (2026)
von: Zhu, Ruijie, et al.
Veröffentlicht: (2026)
GenWorld: Towards Detecting AI-generated Real-world Simulation Videos
von: Chen, Weiliang, et al.
Veröffentlicht: (2025)
von: Chen, Weiliang, et al.
Veröffentlicht: (2025)
GaussianFormer-2: Probabilistic Gaussian Superposition for Efficient 3D Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
Stag-1: Towards Realistic 4D Driving Simulation with Video Generation Model
von: Wang, Lening, et al.
Veröffentlicht: (2024)
von: Wang, Lening, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
UniPre3D: Unified Pre-training of 3D Point Cloud Models with Cross-Modal Gaussian Splatting
von: Wang, Ziyi, et al.
Veröffentlicht: (2025) -
$\bf{D^3}$QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2025) -
OGGSplat: Open Gaussian Growing for Generalizable Reconstruction with Expanded Field-of-View
von: Wang, Yanbo, et al.
Veröffentlicht: (2025) -
Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
von: Wu, Yuqi, et al.
Veröffentlicht: (2025) -
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2026)