ControlDreamer: Blending Geometry and Style in Text-to-3D
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Oh, Yeongtak, Choi, Jooyoung, Kim, Yongsung, Park, Minjun, Shin, Chaehun, Yoon, Sungroh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Style-Friendly SNR Sampler for Style-Driven Generation
von: Choi, Jooyoung, et al.
Veröffentlicht: (2024)
von: Choi, Jooyoung, et al.
Veröffentlicht: (2024)
Improving Geometry in Sparse-View 3DGS via Reprojection-based DoF Separation
von: Kim, Yongsung, et al.
Veröffentlicht: (2024)
von: Kim, Yongsung, et al.
Veröffentlicht: (2024)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025)
von: Park, Sangha, et al.
Veröffentlicht: (2025)
Improving Diffusion-Based Generative Models via Approximated Optimal Transport
von: Kim, Daegyu, et al.
Veröffentlicht: (2024)
von: Kim, Daegyu, et al.
Veröffentlicht: (2024)
Disentangled Motion Modeling for Video Frame Interpolation
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
von: Shin, Chaehun, et al.
Veröffentlicht: (2025)
von: Shin, Chaehun, et al.
Veröffentlicht: (2025)
Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2024)
RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
von: Oh, Yeongtak, et al.
Veröffentlicht: (2025)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2025)
Diagnosing and Correcting Concept Omission in Multimodal Diffusion Transformers
von: Baek, Kanghyun, et al.
Veröffentlicht: (2026)
von: Baek, Kanghyun, et al.
Veröffentlicht: (2026)
DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
Contextualized Visual Personalization in Vision-Language Models
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
HeSS: Head Sensitivity Score for Sparsity Redistribution in VGGT
von: Kim, Yongsung, et al.
Veröffentlicht: (2026)
von: Kim, Yongsung, et al.
Veröffentlicht: (2026)
STAG: Structural Test-time Alignment of Gradients for Online Adaptation
von: Shin, Juhyeon, et al.
Veröffentlicht: (2024)
von: Shin, Juhyeon, et al.
Veröffentlicht: (2024)
Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
von: Kang, Minjun, et al.
Veröffentlicht: (2025)
von: Kang, Minjun, et al.
Veröffentlicht: (2025)
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
von: Song, Jaewoo, et al.
Veröffentlicht: (2025)
Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
von: Kang, Minjun, et al.
Veröffentlicht: (2026)
von: Kang, Minjun, et al.
Veröffentlicht: (2026)
JointDreamer: Ensuring Geometry Consistency and Text Congruence in Text-to-3D Generation via Joint Score Distillation
von: Jiang, Chenhan, et al.
Veröffentlicht: (2024)
von: Jiang, Chenhan, et al.
Veröffentlicht: (2024)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
CKNN: Cleansed k-Nearest Neighbor for Unsupervised Video Anomaly Detection
von: Yi, Jihun, et al.
Veröffentlicht: (2024)
von: Yi, Jihun, et al.
Veröffentlicht: (2024)
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
von: Kim, Jin Hyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jin Hyeon, et al.
Veröffentlicht: (2026)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
ClotheDreamer: Text-Guided Garment Generation with 3D Gaussians
von: Liu, Yufei, et al.
Veröffentlicht: (2024)
von: Liu, Yufei, et al.
Veröffentlicht: (2024)
Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
StyleBlend: Enhancing Style-Specific Content Creation in Text-to-Image Diffusion Models
von: Chen, Zichong, et al.
Veröffentlicht: (2025)
von: Chen, Zichong, et al.
Veröffentlicht: (2025)
VividDreamer: Towards High-Fidelity and Efficient Text-to-3D Generation
von: Chen, Zixuan, et al.
Veröffentlicht: (2024)
von: Chen, Zixuan, et al.
Veröffentlicht: (2024)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
von: Jiang, Lutao, et al.
Veröffentlicht: (2024)
von: Jiang, Lutao, et al.
Veröffentlicht: (2024)
RecDreamer: Consistent Text-to-3D Generation via Uniform Score Distillation
von: Zheng, Chenxi, et al.
Veröffentlicht: (2025)
von: Zheng, Chenxi, et al.
Veröffentlicht: (2025)
VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation
von: Zhuo, Wenjie, et al.
Veröffentlicht: (2024)
von: Zhuo, Wenjie, et al.
Veröffentlicht: (2024)
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
von: Qu, Zefan, et al.
Veröffentlicht: (2025)
von: Qu, Zefan, et al.
Veröffentlicht: (2025)
Normality Addition via Normality Detection in Industrial Image Anomaly Detection Models
von: Yi, Jihun, et al.
Veröffentlicht: (2024)
von: Yi, Jihun, et al.
Veröffentlicht: (2024)
3D-SceneDreamer: Text-Driven 3D-Consistent Scene Generation
von: Zhang, Frank, et al.
Veröffentlicht: (2024)
von: Zhang, Frank, et al.
Veröffentlicht: (2024)
PlacidDreamer: Advancing Harmony in Text-to-3D Generation
von: Huang, Shuo, et al.
Veröffentlicht: (2024)
von: Huang, Shuo, et al.
Veröffentlicht: (2024)
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
von: Park, Sangha, et al.
Veröffentlicht: (2025)
von: Park, Sangha, et al.
Veröffentlicht: (2025)
SteinDreamer: Variance Reduction for Text-to-3D Score Distillation via Stein Identity
von: Wang, Peihao, et al.
Veröffentlicht: (2023)
von: Wang, Peihao, et al.
Veröffentlicht: (2023)
FlowDreamer: Exploring High Fidelity Text-to-3D Generation via Rectified Flow
von: Li, Hangyu, et al.
Veröffentlicht: (2024)
von: Li, Hangyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Style-Friendly SNR Sampler for Style-Driven Generation
von: Choi, Jooyoung, et al.
Veröffentlicht: (2024) -
Improving Geometry in Sparse-View 3DGS via Reprojection-based DoF Separation
von: Kim, Yongsung, et al.
Veröffentlicht: (2024) -
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
von: Shin, Chaehun, et al.
Veröffentlicht: (2024) -
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025) -
Improving Diffusion-Based Generative Models via Approximated Optimal Transport
von: Kim, Daegyu, et al.
Veröffentlicht: (2024)