Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zhenwei, Wang, Tengfei, He, Zexin, Hancke, Gerhard, Liu, Ziwei, Lau, Rynson W. H. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
by: Qu, Zefan, et al.
Published: (2025)
by: Qu, Zefan, et al.
Published: (2025)
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
by: Liu, Yuhao, et al.
Published: (2025)
by: Liu, Yuhao, et al.
Published: (2025)
Boosting Weakly-Supervised Referring Image Segmentation via Progressive Comprehension
by: Yang, Zaiquan, et al.
Published: (2024)
by: Yang, Zaiquan, et al.
Published: (2024)
Color Shift Estimation-and-Correction for Image Enhancement
by: Li, Yiyu, et al.
Published: (2024)
by: Li, Yiyu, et al.
Published: (2024)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
by: Huang, Tianyu, et al.
Published: (2025)
by: Huang, Tianyu, et al.
Published: (2025)
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
by: Yang, Zaiquan, et al.
Published: (2025)
by: Yang, Zaiquan, et al.
Published: (2025)
LuSh-NeRF: Lighting up and Sharpening NeRFs for Low-light Scenes
by: Qu, Zefan, et al.
Published: (2024)
by: Qu, Zefan, et al.
Published: (2024)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
SeHDR: Single-Exposure HDR Novel View Synthesis via 3D Gaussian Bracketing
by: Li, Yiyu, et al.
Published: (2025)
by: Li, Yiyu, et al.
Published: (2025)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding
by: Zhao, Youjun, et al.
Published: (2024)
by: Zhao, Youjun, et al.
Published: (2024)
Neural LightRig: Unlocking Accurate Object Normal and Material Estimation with Multi-Light Diffusion
by: He, Zexin, et al.
Published: (2024)
by: He, Zexin, et al.
Published: (2024)
3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors
by: Hong, Fangzhou, et al.
Published: (2024)
by: Hong, Fangzhou, et al.
Published: (2024)
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
by: Huang, Tianyu, et al.
Published: (2024)
by: Huang, Tianyu, et al.
Published: (2024)
X-Dreamer: Creating High-quality 3D Content by Bridging the Domain Gap Between Text-to-2D and Text-to-3D Generation
by: Ma, Yiwei, et al.
Published: (2023)
by: Ma, Yiwei, et al.
Published: (2023)
SIC3D: Style Image Conditioned Text-to-3D Gaussian Splatting Generation
by: He, Ming, et al.
Published: (2026)
by: He, Ming, et al.
Published: (2026)
LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation
by: Tang, Jiaxiang, et al.
Published: (2024)
by: Tang, Jiaxiang, et al.
Published: (2024)
Text-Image Conditioned 3D Generation
by: Cen, Jiazhong, et al.
Published: (2026)
by: Cen, Jiazhong, et al.
Published: (2026)
Recasting Regional Lighting for Shadow Removal
by: Liu, Yuhao, et al.
Published: (2024)
by: Liu, Yuhao, et al.
Published: (2024)
Conditional Text-to-Image Generation with Reference Guidance
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface Detection
by: Lin, Jiaying, et al.
Published: (2022)
by: Lin, Jiaying, et al.
Published: (2022)
ComboVerse: Compositional 3D Assets Creation Using Spatially-Aware Diffusion Guidance
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
3DGen-Bench: Comprehensive Benchmark Suite for 3D Generative Models
by: Zhang, Yuhan, et al.
Published: (2025)
by: Zhang, Yuhan, et al.
Published: (2025)
RefSTAR: Blind Facial Image Restoration with Reference Selection, Transfer, and Reconstruction
by: Yin, Zhicun, et al.
Published: (2025)
by: Yin, Zhicun, et al.
Published: (2025)
DreamComposer++: Empowering Diffusion Models with Multi-View Conditions for 3D Content Generation
by: Yang, Yunhan, et al.
Published: (2025)
by: Yang, Yunhan, et al.
Published: (2025)
MDeRainNet: An Efficient Macro-pixel Image Rain Removal Network
by: Yan, Tao, et al.
Published: (2024)
by: Yan, Tao, et al.
Published: (2024)
PI3D: Efficient Text-to-3D Generation with Pseudo-Image Diffusion
by: Liu, Ying-Tian, et al.
Published: (2023)
by: Liu, Ying-Tian, et al.
Published: (2023)
Delving into Dark Regions for Robust Shadow Detection
by: Guan, Huankang, et al.
Published: (2024)
by: Guan, Huankang, et al.
Published: (2024)
Inverse Rendering of Glossy Objects via the Neural Plenoptic Function and Radiance Fields
by: Wang, Haoyuan, et al.
Published: (2024)
by: Wang, Haoyuan, et al.
Published: (2024)
AID: Attention Interpolation of Text-to-Image Diffusion
by: He, Qiyuan, et al.
Published: (2024)
by: He, Qiyuan, et al.
Published: (2024)
Diff-Plugin: Revitalizing Details for Diffusion-based Low-level Tasks
by: Liu, Yuhao, et al.
Published: (2024)
by: Liu, Yuhao, et al.
Published: (2024)
Omni6D: Large-Vocabulary 3D Object Dataset for Category-Level 6D Object Pose Estimation
by: Zhang, Mengchen, et al.
Published: (2024)
by: Zhang, Mengchen, et al.
Published: (2024)
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
by: Zhou, Yun, et al.
Published: (2025)
by: Zhou, Yun, et al.
Published: (2025)
MirrorMamba: Towards Scalable and Robust Mirror Detection in Videos
by: Song, Rui, et al.
Published: (2025)
by: Song, Rui, et al.
Published: (2025)
SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
Revisiting the Integration of Convolution and Attention for Vision Backbone
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation
by: Li, Zhiqi, et al.
Published: (2025)
by: Li, Zhiqi, et al.
Published: (2025)
Similar Items
-
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
by: Wang, Zhenwei, et al.
Published: (2024) -
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
by: Qu, Zefan, et al.
Published: (2025) -
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
by: Liu, Yuhao, et al.
Published: (2025) -
Boosting Weakly-Supervised Referring Image Segmentation via Progressive Comprehension
by: Yang, Zaiquan, et al.
Published: (2024) -
Color Shift Estimation-and-Correction for Image Enhancement
by: Li, Yiyu, et al.
Published: (2024)