Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Shuhong, Mirzaei, Ashkan, Gilitschenski, Igor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos
by: Liang, Hanxue, et al.
Published: (2024)
by: Liang, Hanxue, et al.
Published: (2024)
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
by: Yang, Jianing, et al.
Published: (2025)
by: Yang, Jianing, et al.
Published: (2025)
Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
by: Dong, Zhao, et al.
Published: (2025)
by: Dong, Zhao, et al.
Published: (2025)
Pandora3D: A Comprehensive Framework for High-Quality 3D Shape and Texture Generation
by: Yang, Jiayu, et al.
Published: (2025)
by: Yang, Jiayu, et al.
Published: (2025)
AnyHome: Open-Vocabulary Generation of Structured and Textured 3D Homes
by: Fu, Rao, et al.
Published: (2023)
by: Fu, Rao, et al.
Published: (2023)
Meta 3D TextureGen: Fast and Consistent Texture Generation for 3D Objects
by: Bensadoun, Raphael, et al.
Published: (2024)
by: Bensadoun, Raphael, et al.
Published: (2024)
LayerGS: Decomposition and Inpainting of Layered 3D Human Avatars via 2D Gaussian Splatting
by: Xu, Yinghan, et al.
Published: (2026)
by: Xu, Yinghan, et al.
Published: (2026)
Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape
by: Wu, Rundi, et al.
Published: (2023)
by: Wu, Rundi, et al.
Published: (2023)
CTRL-D: Controllable Dynamic 3D Scene Editing with Personalized 2D Diffusion
by: He, Kai, et al.
Published: (2024)
by: He, Kai, et al.
Published: (2024)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
by: Ye, Weicai, et al.
Published: (2024)
by: Ye, Weicai, et al.
Published: (2024)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
FabricDiffusion: High-Fidelity Texture Transfer for 3D Garments Generation from In-The-Wild Clothing Images
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
Meta 3D AssetGen: Text-to-Mesh Generation with High-Quality Geometry, Texture, and PBR Materials
by: Siddiqui, Yawar, et al.
Published: (2024)
by: Siddiqui, Yawar, et al.
Published: (2024)
Learning 3D-Gaussian Simulators from RGB Videos
by: Zhobro, Mikel, et al.
Published: (2025)
by: Zhobro, Mikel, et al.
Published: (2025)
SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes
by: Pfaff, Nicholas, et al.
Published: (2026)
by: Pfaff, Nicholas, et al.
Published: (2026)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
by: Shriram, Jaidev, et al.
Published: (2024)
by: Shriram, Jaidev, et al.
Published: (2024)
LuxDiT: Lighting Estimation with Video Diffusion Transformer
by: Liang, Ruofan, et al.
Published: (2025)
by: Liang, Ruofan, et al.
Published: (2025)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
by: Wei, Shenxing, et al.
Published: (2025)
by: Wei, Shenxing, et al.
Published: (2025)
Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data
by: Ma, Zhiyuan, et al.
Published: (2025)
by: Ma, Zhiyuan, et al.
Published: (2025)
SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control
by: Luo, Zhengyi, et al.
Published: (2025)
by: Luo, Zhengyi, et al.
Published: (2025)
Articraft: An Agentic System for Scalable Articulated 3D Asset Generation
by: Zhou, Matt, et al.
Published: (2026)
by: Zhou, Matt, et al.
Published: (2026)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025)
by: Song, Chenxi, et al.
Published: (2025)
Edify 3D: Scalable High-Quality 3D Asset Generation
by: NVIDIA, et al.
Published: (2024)
by: NVIDIA, et al.
Published: (2024)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
by: Ray, Arijit, et al.
Published: (2024)
by: Ray, Arijit, et al.
Published: (2024)
Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning
by: Mandalika, Sriram
Published: (2025)
by: Mandalika, Sriram
Published: (2025)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Physics-Based Motion Imitation with Adversarial Differential Discriminators
by: Zhang, Ziyu, et al.
Published: (2025)
by: Zhang, Ziyu, et al.
Published: (2025)
Free-Moving Object Reconstruction and Pose Estimation with Virtual Camera
by: Shi, Haixin, et al.
Published: (2024)
by: Shi, Haixin, et al.
Published: (2024)
Neural Implicit Representation for Building Digital Twins of Unknown Articulated Objects
by: Weng, Yijia, et al.
Published: (2024)
by: Weng, Yijia, et al.
Published: (2024)
AddBiomechanics Dataset: Capturing the Physics of Human Motion at Scale
by: Werling, Keenon, et al.
Published: (2024)
by: Werling, Keenon, et al.
Published: (2024)
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
by: Pang, Bo, et al.
Published: (2025)
by: Pang, Bo, et al.
Published: (2025)
SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control
by: Mu, Yuxuan, et al.
Published: (2025)
by: Mu, Yuxuan, et al.
Published: (2025)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
by: Shah, Ansh, et al.
Published: (2024)
by: Shah, Ansh, et al.
Published: (2024)
ImDy: Human Inverse Dynamics from Imitated Observations
by: Liu, Xinpeng, et al.
Published: (2024)
by: Liu, Xinpeng, et al.
Published: (2024)
Leveraging Foundation Models To learn the shape of semi-fluid deformable objects
by: Assal, Omar El, et al.
Published: (2024)
by: Assal, Omar El, et al.
Published: (2024)
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
by: Raji, Fadlullah, et al.
Published: (2026)
by: Raji, Fadlullah, et al.
Published: (2026)
SeqTex: Generate Mesh Textures in Video Sequence
by: Yuan, Ze, et al.
Published: (2025)
by: Yuan, Ze, et al.
Published: (2025)
TEXGen: a Generative Diffusion Model for Mesh Textures
by: Yu, Xin, et al.
Published: (2024)
by: Yu, Xin, et al.
Published: (2024)
Similar Items
-
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
by: Zheng, Shuhong, et al.
Published: (2026) -
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026) -
Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos
by: Liang, Hanxue, et al.
Published: (2024) -
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
by: Yang, Jianing, et al.
Published: (2025) -
Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
by: Dong, Zhao, et al.
Published: (2025)