Saved in:
| Main Authors: | Gao, Chenjian, Ding, Lihe, Han, Rui, Huang, Zhanpeng, Wang, Zibin, Xue, Tianfan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.20331 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoRA-Edit: Controllable First-Frame-Guided Video Editing via Mask-Aware LoRA Fine-Tuning
by: Gao, Chenjian, et al.
Published: (2025)
by: Gao, Chenjian, et al.
Published: (2025)
Interactive3D: Create What You Want by Interactive 3D Generation
by: Dong, Shaocong, et al.
Published: (2024)
by: Dong, Shaocong, et al.
Published: (2024)
FullPart: Generating each 3D Part at Full Resolution
by: Ding, Lihe, et al.
Published: (2025)
by: Ding, Lihe, et al.
Published: (2025)
From One to More: Contextual Part Latents for 3D Generation
by: Dong, Shaocong, et al.
Published: (2025)
by: Dong, Shaocong, et al.
Published: (2025)
DiffIER: Optimizing Diffusion Models with Iterative Error Reduction
by: Chen, Ao, et al.
Published: (2025)
by: Chen, Ao, et al.
Published: (2025)
DynamicTree: Interactive Real Tree Animation via Sparse Voxel Spectrum
by: Li, Yaokun, et al.
Published: (2025)
by: Li, Yaokun, et al.
Published: (2025)
4DSloMo: 4D Reconstruction for High Speed Scene with Asynchronous Capture
by: Chen, Yutian, et al.
Published: (2025)
by: Chen, Yutian, et al.
Published: (2025)
Shining Yourself: High-Fidelity Ornaments Virtual Try-on with Diffusion Model
by: Miao, Yingmao, et al.
Published: (2025)
by: Miao, Yingmao, et al.
Published: (2025)
EGVD: Event-Guided Video Diffusion Model for Physically Realistic Large-Motion Frame Interpolation
by: Zhang, Ziran, et al.
Published: (2025)
by: Zhang, Ziran, et al.
Published: (2025)
SCPainter: A Unified Framework for Realistic 3D Asset Insertion and Novel View Synthesis
by: Dobre, Paul, et al.
Published: (2025)
by: Dobre, Paul, et al.
Published: (2025)
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection
by: Li, Jianan, et al.
Published: (2024)
by: Li, Jianan, et al.
Published: (2024)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
by: Qian, Zezhong, et al.
Published: (2025)
by: Qian, Zezhong, et al.
Published: (2025)
HDRFlow: Real-Time HDR Video Reconstruction with Large Motions
by: Xu, Gangwei, et al.
Published: (2024)
by: Xu, Gangwei, et al.
Published: (2024)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
by: Jin, Hoiyeong, et al.
Published: (2025)
by: Jin, Hoiyeong, et al.
Published: (2025)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
by: Wang, Qinghe, et al.
Published: (2025)
by: Wang, Qinghe, et al.
Published: (2025)
PvNeXt: Rethinking Network Design and Temporal Motion for Point Cloud Video Recognition
by: Wang, Jie, et al.
Published: (2025)
by: Wang, Jie, et al.
Published: (2025)
PARE: Pruning and Adaptive Routing for Efficient Video Generation
by: Wang, Yutong, et al.
Published: (2026)
by: Wang, Yutong, et al.
Published: (2026)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
by: Chen, Yutian, et al.
Published: (2026)
by: Chen, Yutian, et al.
Published: (2026)
VDOT: Efficient Unified Video Creation via Optimal Transport Distillation
by: Wang, Yutong, et al.
Published: (2025)
by: Wang, Yutong, et al.
Published: (2025)
Target-Guided Adversarial Point Cloud Transformer Towards Recognition Against Real-world Corruptions
by: Wang, Jie, et al.
Published: (2024)
by: Wang, Jie, et al.
Published: (2024)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
by: Ljungbergh, William, et al.
Published: (2025)
by: Ljungbergh, William, et al.
Published: (2025)
GLEAM: Learning Generalizable Exploration Policy for Active Mapping in Complex 3D Indoor Scenes
by: Chen, Xiao, et al.
Published: (2025)
by: Chen, Xiao, et al.
Published: (2025)
Towards Realistic Example-based Modeling via 3D Gaussian Stitching
by: Gao, Xinyu, et al.
Published: (2024)
by: Gao, Xinyu, et al.
Published: (2024)
3DGStream: On-the-Fly Training of 3D Gaussians for Efficient Streaming of Photo-Realistic Free-Viewpoint Videos
by: Sun, Jiakai, et al.
Published: (2024)
by: Sun, Jiakai, et al.
Published: (2024)
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
by: Zhao, Chengfeng, et al.
Published: (2026)
by: Zhao, Chengfeng, et al.
Published: (2026)
HOIDiffusion: Generating Realistic 3D Hand-Object Interaction Data
by: Zhang, Mengqi, et al.
Published: (2024)
by: Zhang, Mengqi, et al.
Published: (2024)
GenNBV: Generalizable Next-Best-View Policy for Active 3D Reconstruction
by: Chen, Xiao, et al.
Published: (2024)
by: Chen, Xiao, et al.
Published: (2024)
Instant4D: 4D Gaussian Splatting in Minutes
by: Luo, Zhanpeng, et al.
Published: (2025)
by: Luo, Zhanpeng, et al.
Published: (2025)
Towards Realistic and Consistent Orbital Video Generation via 3D Foundation Priors
by: Wang, Rong, et al.
Published: (2026)
by: Wang, Rong, et al.
Published: (2026)
Bilateral Guided Radiance Field Processing
by: Wang, Yuehao, et al.
Published: (2024)
by: Wang, Yuehao, et al.
Published: (2024)
Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy
by: Xue, Yuxuan, et al.
Published: (2024)
by: Xue, Yuxuan, et al.
Published: (2024)
EffectErase: Joint Video Object Removal and Insertion for High-Quality Effect Erasing
by: Fu, Yang, et al.
Published: (2026)
by: Fu, Yang, et al.
Published: (2026)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
Learning Dynamic Collaborative Network for Semi-supervised 3D Vessel Segmentation
by: Xu, Jiao, et al.
Published: (2026)
by: Xu, Jiao, et al.
Published: (2026)
LightHarmony3D: Harmonizing Illumination and Shadows for Object Insertion in 3D Gaussian Splatting
by: Huang, Tianyu, et al.
Published: (2026)
by: Huang, Tianyu, et al.
Published: (2026)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
by: Bin, Yi, et al.
Published: (2024)
by: Bin, Yi, et al.
Published: (2024)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
by: Luo, Yawen, et al.
Published: (2026)
by: Luo, Yawen, et al.
Published: (2026)
Anything in Any Scene: Photorealistic Video Object Insertion
by: Bai, Chen, et al.
Published: (2024)
by: Bai, Chen, et al.
Published: (2024)
GenesisTex: Adapting Image Denoising Diffusion to Texture Space
by: Gao, Chenjian, et al.
Published: (2024)
by: Gao, Chenjian, et al.
Published: (2024)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
Similar Items
-
LoRA-Edit: Controllable First-Frame-Guided Video Editing via Mask-Aware LoRA Fine-Tuning
by: Gao, Chenjian, et al.
Published: (2025) -
Interactive3D: Create What You Want by Interactive 3D Generation
by: Dong, Shaocong, et al.
Published: (2024) -
FullPart: Generating each 3D Part at Full Resolution
by: Ding, Lihe, et al.
Published: (2025) -
From One to More: Contextual Part Latents for 3D Generation
by: Dong, Shaocong, et al.
Published: (2025) -
DiffIER: Optimizing Diffusion Models with Iterative Error Reduction
by: Chen, Ao, et al.
Published: (2025)