Focus on Neighbors and Know the Whole: Towards Consistent Dense Multiview Text-to-Image Generator for 3D Creation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Bonan, Zhang, Zicheng, Yang, Xingyi, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hash3D: Training-free Acceleration for 3D Generation
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Compositional Video Generation as Flow Equalization
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Language Model as Visual Explainer
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Top-Down Compression: Revisit Efficient Vision Token Projection for Visual Instruction Tuning
von: li, Bonan, et al.
Veröffentlicht: (2025)
von: li, Bonan, et al.
Veröffentlicht: (2025)
Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition
von: Zhou, Bangbang, et al.
Veröffentlicht: (2024)
von: Zhou, Bangbang, et al.
Veröffentlicht: (2024)
Flash Sculptor: Modular 3D Worlds from Objects
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
C4D: 4D Made from 3D through Dual Correspondences
von: Wang, Shizun, et al.
Veröffentlicht: (2025)
von: Wang, Shizun, et al.
Veröffentlicht: (2025)
Control and Realism: Best of Both Worlds in Layout-to-Image without Training
von: Li, Bonan, et al.
Veröffentlicht: (2025)
von: Li, Bonan, et al.
Veröffentlicht: (2025)
Test3R: Learning to Reconstruct 3D at Test Time
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
Efficient Gaussian Splatting for Monocular Dynamic Scene Rendering via Sparse Time-Variant Attribute Modeling
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
Unsegment Anything by Simulating Deformation
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
Relation Rectification in Diffusion Model
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
1000+ FPS 4D Gaussian Splatting for Dynamic Scene Rendering
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
WorldWarp: Propagating 3D Geometry with Asynchronous Video Diffusion
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
Neural Metamorphosis
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Vista3D: Unravel the 3D Darkside of a Single Image
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
Adaptively Clustering Neighbor Elements for Image-Text Generation
von: Wang, Zihua, et al.
Veröffentlicht: (2023)
von: Wang, Zihua, et al.
Veröffentlicht: (2023)
Image Editing As Programs with Diffusion Models
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
Video-Infinity: Distributed Long Video Generation
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
Few-shot Implicit Function Generation via Equivariance
von: Huang, Suizhi, et al.
Veröffentlicht: (2025)
von: Huang, Suizhi, et al.
Veröffentlicht: (2025)
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
Kolmogorov-Arnold Transformer
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Semi-supervised Dense Keypoints Using Unlabeled Multiview Images
von: Yu, Zhixuan, et al.
Veröffentlicht: (2021)
von: Yu, Zhixuan, et al.
Veröffentlicht: (2021)
Evaluating Multiview Object Consistency in Humans and Image Models
von: Bonnen, Tyler, et al.
Veröffentlicht: (2024)
von: Bonnen, Tyler, et al.
Veröffentlicht: (2024)
DragGaussian: Enabling Drag-style Manipulation on 3D Gaussian Representation
von: Shen, Sitian, et al.
Veröffentlicht: (2024)
von: Shen, Sitian, et al.
Veröffentlicht: (2024)
GFlow: Recovering 4D World from Monocular Video
von: Wang, Shizun, et al.
Veröffentlicht: (2024)
von: Wang, Shizun, et al.
Veröffentlicht: (2024)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
von: Zhu, Shenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Shenhao, et al.
Veröffentlicht: (2024)
ConTEXTure: Consistent Multiview Images to Texture
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2024)
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2024)
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation
von: Li, Niantong, et al.
Veröffentlicht: (2026)
von: Li, Niantong, et al.
Veröffentlicht: (2026)
MVD$^2$: Efficient Multiview 3D Reconstruction for Multiview Diffusion
von: Zheng, Xin-Yang, et al.
Veröffentlicht: (2024)
von: Zheng, Xin-Yang, et al.
Veröffentlicht: (2024)
ConDense: Consistent 2D/3D Pre-training for Dense and Sparse Features from Multi-View Images
von: Zhang, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoshuai, et al.
Veröffentlicht: (2024)
CMD: Controllable Multiview Diffusion for 3D Editing and Progressive Generation
von: Li, Peng, et al.
Veröffentlicht: (2025)
von: Li, Peng, et al.
Veröffentlicht: (2025)
Progressive3D: Progressively Local Editing for Text-to-3D Content Creation with Complex Semantic Prompts
von: Cheng, Xinhua, et al.
Veröffentlicht: (2023)
von: Cheng, Xinhua, et al.
Veröffentlicht: (2023)
TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2025)
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2025)
Dragen3D: Multiview Geometry Consistent 3D Gaussian Generation with Drag-Based Control
von: Yan, Jinbo, et al.
Veröffentlicht: (2025)
von: Yan, Jinbo, et al.
Veröffentlicht: (2025)
StyDeSty: Min-Max Stylization and Destylization for Single Domain Generalization
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
MOC-3D: Manifold-Order Consistency for Text-to-3D Generation
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
ViewDiff: 3D-Consistent Image Generation with Text-to-Image Models
von: Höllein, Lukas, et al.
Veröffentlicht: (2024)
von: Höllein, Lukas, et al.
Veröffentlicht: (2024)
Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
von: Xiao, Lingao, et al.
Veröffentlicht: (2025)
von: Xiao, Lingao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hash3D: Training-free Acceleration for 3D Generation
von: Yang, Xingyi, et al.
Veröffentlicht: (2024) -
Compositional Video Generation as Flow Equalization
von: Yang, Xingyi, et al.
Veröffentlicht: (2024) -
Language Model as Visual Explainer
von: Yang, Xingyi, et al.
Veröffentlicht: (2024) -
Top-Down Compression: Revisit Efficient Vision Token Projection for Visual Instruction Tuning
von: li, Bonan, et al.
Veröffentlicht: (2025) -
Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition
von: Zhou, Bangbang, et al.
Veröffentlicht: (2024)