Diff-3DCap: Shape Captioning with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Shu, Zhenyu, Wen, Jiawei, Li, Shiyang, Xin, Shiqing, Liu, Ligang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts
by: Shu, Zhenyu, et al.
Published: (2025)
by: Shu, Zhenyu, et al.
Published: (2025)
StrucADT: Generating Structure-controlled 3D Point Clouds with Adjacency Diffusion Transformer
by: Shu, Zhenyu, et al.
Published: (2025)
by: Shu, Zhenyu, et al.
Published: (2025)
DFG-PCN: Point Cloud Completion with Degree-Flexible Point Graph
by: Shu, Zhenyu, et al.
Published: (2025)
by: Shu, Zhenyu, et al.
Published: (2025)
TetraDiffusion: Tetrahedral Diffusion Models for 3D Shape Generation
by: Kalischek, Nikolai, et al.
Published: (2022)
by: Kalischek, Nikolai, et al.
Published: (2022)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
by: Guo, Yuwei, et al.
Published: (2023)
by: Guo, Yuwei, et al.
Published: (2023)
ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
Ghost on the Shell: An Expressive Representation of General 3D Shapes
by: Liu, Zhen, et al.
Published: (2023)
by: Liu, Zhen, et al.
Published: (2023)
DiffUHaul: A Training-Free Method for Object Dragging in Images
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
Diverse Part Synthesis for 3D Shape Creation
by: Guan, Yanran, et al.
Published: (2024)
by: Guan, Yanran, et al.
Published: (2024)
VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models
by: Han, Junlin, et al.
Published: (2024)
by: Han, Junlin, et al.
Published: (2024)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
by: Liu, Shengqi, et al.
Published: (2024)
by: Liu, Shengqi, et al.
Published: (2024)
Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D Rewards
by: Liu, Qingming, et al.
Published: (2025)
by: Liu, Qingming, et al.
Published: (2025)
LoST: Level of Semantics Tokenization for 3D Shapes
by: Dutt, Niladri Shekhar, et al.
Published: (2026)
by: Dutt, Niladri Shekhar, et al.
Published: (2026)
Enhancing Diffusion-based Point Cloud Generation with Smoothness Constraint
by: Li, Yukun, et al.
Published: (2024)
by: Li, Yukun, et al.
Published: (2024)
Modeling Aesthetic Preferences in 3D Shapes: A Large-Scale Paired Comparison Study Across Object Categories
by: Dev, Kapil
Published: (2025)
by: Dev, Kapil
Published: (2025)
LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR
by: Fekri, Pedram, et al.
Published: (2026)
by: Fekri, Pedram, et al.
Published: (2026)
MatCLIP: Light- and Shape-Insensitive Assignment of PBR Material Models
by: Birsak, Michael, et al.
Published: (2025)
by: Birsak, Michael, et al.
Published: (2025)
SENS: Part-Aware Sketch-based Implicit Neural Shape Modeling
by: Binninger, Alexandre, et al.
Published: (2023)
by: Binninger, Alexandre, et al.
Published: (2023)
DiffBody: Diffusion-based Pose and Shape Editing of Human Images
by: Okuyama, Yuta, et al.
Published: (2024)
by: Okuyama, Yuta, et al.
Published: (2024)
Diffusion 3D Features (Diff3F): Decorating Untextured Shapes with Distilled Semantic Features
by: Dutt, Niladri Shekhar, et al.
Published: (2023)
by: Dutt, Niladri Shekhar, et al.
Published: (2023)
Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models
by: Ge, Songwei, et al.
Published: (2023)
by: Ge, Songwei, et al.
Published: (2023)
Flexible Motion In-betweening with Diffusion Models
by: Cohan, Setareh, et al.
Published: (2024)
by: Cohan, Setareh, et al.
Published: (2024)
Distilling Diffusion Models into Conditional GANs
by: Kang, Minguk, et al.
Published: (2024)
by: Kang, Minguk, et al.
Published: (2024)
Recovering 3D Shapes from Ultra-Fast Motion-Blurred Images
by: Yu, Fei, et al.
Published: (2026)
by: Yu, Fei, et al.
Published: (2026)
Interpreting the Weight Space of Customized Diffusion Models
by: Dravid, Amil, et al.
Published: (2024)
by: Dravid, Amil, et al.
Published: (2024)
DiffH2O: Diffusion-Based Synthesis of Hand-Object Interactions from Textual Descriptions
by: Christen, Sammy, et al.
Published: (2024)
by: Christen, Sammy, et al.
Published: (2024)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
by: Gandikota, Rohit, et al.
Published: (2025)
by: Gandikota, Rohit, et al.
Published: (2025)
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
by: Cui, Jiahao, et al.
Published: (2024)
by: Cui, Jiahao, et al.
Published: (2024)
DMesh++: An Efficient Differentiable Mesh for Complex Shapes
by: Son, Sanghyun, et al.
Published: (2024)
by: Son, Sanghyun, et al.
Published: (2024)
DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion
by: Huang, Yukun, et al.
Published: (2024)
by: Huang, Yukun, et al.
Published: (2024)
Curved Diffusion: A Generative Model With Optical Geometry Control
by: Voynov, Andrey, et al.
Published: (2023)
by: Voynov, Andrey, et al.
Published: (2023)
Single Mesh Diffusion Models with Field Latents for Texture Generation
by: Mitchel, Thomas W., et al.
Published: (2023)
by: Mitchel, Thomas W., et al.
Published: (2023)
Enhancing Image Layout Control with Loss-Guided Diffusion Models
by: Patel, Zakaria, et al.
Published: (2024)
by: Patel, Zakaria, et al.
Published: (2024)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
by: Avrahami, Omri, et al.
Published: (2023)
by: Avrahami, Omri, et al.
Published: (2023)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
by: Wei, Yuanxin, et al.
Published: (2025)
by: Wei, Yuanxin, et al.
Published: (2025)
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
by: Engelhardt, Andreas, et al.
Published: (2025)
by: Engelhardt, Andreas, et al.
Published: (2025)
DreamTime: An Improved Optimization Strategy for Diffusion-Guided 3D Generation
by: Huang, Yukun, et al.
Published: (2023)
by: Huang, Yukun, et al.
Published: (2023)
Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models
by: Laria, Héctor, et al.
Published: (2025)
by: Laria, Héctor, et al.
Published: (2025)
Your Latent Mask is Wrong: Pixel-Equivalent Latent Compositing for Diffusion Models
by: Bradbury, Rowan, et al.
Published: (2025)
by: Bradbury, Rowan, et al.
Published: (2025)
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
by: Xing, Xiaoyan, et al.
Published: (2024)
by: Xing, Xiaoyan, et al.
Published: (2024)
Similar Items
-
GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts
by: Shu, Zhenyu, et al.
Published: (2025) -
StrucADT: Generating Structure-controlled 3D Point Clouds with Adjacency Diffusion Transformer
by: Shu, Zhenyu, et al.
Published: (2025) -
DFG-PCN: Point Cloud Completion with Degree-Flexible Point Graph
by: Shu, Zhenyu, et al.
Published: (2025) -
TetraDiffusion: Tetrahedral Diffusion Models for 3D Shape Generation
by: Kalischek, Nikolai, et al.
Published: (2022) -
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
by: Guo, Yuwei, et al.
Published: (2023)