Retrieval-Augmented Score Distillation for Text-to-3D Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Seo, Junyoung, Hong, Susung, Jang, Wooseok, Kim, Inès Hyeonsu, Kwak, Minseop, Lee, Doyup, Kim, Seungryong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
by: Seo, Junyoung, et al.
Published: (2023)
by: Seo, Junyoung, et al.
Published: (2023)
Geometry-Aware Score Distillation via 3D Consistent Noising and Gradient Consistency Modeling
by: Kwak, Min-Seop, et al.
Published: (2024)
by: Kwak, Min-Seop, et al.
Published: (2024)
DirecT2V: Large Language Models are Frame-Level Directors for Zero-Shot Text-to-Video Generation
by: Hong, Susung, et al.
Published: (2023)
by: Hong, Susung, et al.
Published: (2023)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025)
by: Kwon, Minkyung, et al.
Published: (2025)
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
Domain Generalization Using Large Pretrained Models with Mixture-of-Adapters
by: Lee, Gyuseong, et al.
Published: (2023)
by: Lee, Gyuseong, et al.
Published: (2023)
ControlFace: Harnessing Facial Parametric Control for Face Rigging
by: Jang, Wooseok, et al.
Published: (2024)
by: Jang, Wooseok, et al.
Published: (2024)
WorldKV: Efficient World Memory with World Retrieval and Compression
by: Yi, Jung, et al.
Published: (2026)
by: Yi, Jung, et al.
Published: (2026)
Where and How to Perturb: On the Design of Perturbation Guidance in Diffusion and Flow Models
by: Ahn, Donghoon, et al.
Published: (2025)
by: Ahn, Donghoon, et al.
Published: (2025)
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
by: Lee, JoungBin, et al.
Published: (2025)
by: Lee, JoungBin, et al.
Published: (2025)
V-Warper: Appearance-Consistent Video Diffusion Personalization via Value Warping
by: Lee, Hyunkoo, et al.
Published: (2025)
by: Lee, Hyunkoo, et al.
Published: (2025)
Exploring Temporally-Aware Features for Point Tracking
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
D$^2$USt3R: Enhancing 3D Reconstruction for Dynamic Scenes
by: Han, Jisang, et al.
Published: (2025)
by: Han, Jisang, et al.
Published: (2025)
Harnessing the Power of Training-Free Techniques in Text-to-2D Generation for Text-to-3D Generation via Score Distillation Sampling
by: Lee, Junhong, et al.
Published: (2025)
by: Lee, Junhong, et al.
Published: (2025)
TAG: Tangential Amplifying Guidance for Hallucination-Resistant Sampling
by: Cho, Hyunmin, et al.
Published: (2025)
by: Cho, Hyunmin, et al.
Published: (2025)
Diffusion Model for Dense Matching
by: Nam, Jisu, et al.
Published: (2023)
by: Nam, Jisu, et al.
Published: (2023)
Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression
by: Yi, Jung, et al.
Published: (2025)
by: Yi, Jung, et al.
Published: (2025)
COIN: Confidence Score-Guided Distillation for Annotation-Free Cell Segmentation
by: Jo, Sanghyun, et al.
Published: (2025)
by: Jo, Sanghyun, et al.
Published: (2025)
AnthroTAP: Learning Point Tracking with Real-World Motion
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
Emergent Outlier View Rejection in Visual Geometry Grounded Transformers
by: Han, Jisang, et al.
Published: (2025)
by: Han, Jisang, et al.
Published: (2025)
Uncertainty-aware Semantic Mapping in Off-road Environments with Dempster-Shafer Theory of Evidence
by: Kim, Junyoung, et al.
Published: (2024)
by: Kim, Junyoung, et al.
Published: (2024)
MV-TAP: Tracking Any Point in Multi-View Videos
by: Koo, Jahyeok, et al.
Published: (2025)
by: Koo, Jahyeok, et al.
Published: (2025)
Effective Rank Analysis and Regularization for Enhanced 3D Gaussian Splatting
by: Hyung, Junha, et al.
Published: (2024)
by: Hyung, Junha, et al.
Published: (2024)
Connecting Consistency Distillation to Score Distillation for Text-to-3D Generation
by: Li, Zongrui, et al.
Published: (2024)
by: Li, Zongrui, et al.
Published: (2024)
DivCon-NeRF: Diverse and Consistent Ray Augmentation for Few-Shot NeRF
by: Lee, Ingyun, et al.
Published: (2025)
by: Lee, Ingyun, et al.
Published: (2025)
NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image
by: Jeong, Yoonwoo, et al.
Published: (2023)
by: Jeong, Yoonwoo, et al.
Published: (2023)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
by: Kim, Chaehyun, et al.
Published: (2025)
by: Kim, Chaehyun, et al.
Published: (2025)
APPLE: Attribute-Preserving Pseudo-Labeling for Diffusion-Based Face Swapping
by: Kang, Jiwon, et al.
Published: (2026)
by: Kang, Jiwon, et al.
Published: (2026)
Perturb-and-Revise: Flexible 3D Editing with Generative Trajectories
by: Hong, Susung, et al.
Published: (2024)
by: Hong, Susung, et al.
Published: (2024)
Hybrid Video Diffusion Models with 2D Triplane and 3D Wavelet Representation
by: Kim, Kihong, et al.
Published: (2024)
by: Kim, Kihong, et al.
Published: (2024)
Smoothed Energy Guidance: Guiding Diffusion Models with Reduced Energy Curvature of Attention
by: Hong, Susung
Published: (2024)
by: Hong, Susung
Published: (2024)
Robust 3D-Masked Part-level Editing in 3D Gaussian Splatting with Regularized Score Distillation Sampling
by: Kim, Hayeon, et al.
Published: (2025)
by: Kim, Hayeon, et al.
Published: (2025)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
Semantic Score Distillation Sampling for Compositional Text-to-3D Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Repurposing Geometric Foundation Models for Multi-view Diffusion
by: Jang, Wooseok, et al.
Published: (2026)
by: Jang, Wooseok, et al.
Published: (2026)
Talk3D: High-Fidelity Talking Portrait Synthesis via Personalized 3D Generative Prior
by: Ko, Jaehoon, et al.
Published: (2024)
by: Ko, Jaehoon, et al.
Published: (2024)
MonoSAOD: Monocular 3D Object Detection with Sparsely Annotated Label
by: Jung, Junyoung, et al.
Published: (2026)
by: Jung, Junyoung, et al.
Published: (2026)
GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping
by: Seo, Junyoung, et al.
Published: (2024)
by: Seo, Junyoung, et al.
Published: (2024)
Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors
by: Jeon, Subin, et al.
Published: (2025)
by: Jeon, Subin, et al.
Published: (2025)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
by: Hyung, Junha, et al.
Published: (2024)
by: Hyung, Junha, et al.
Published: (2024)
Similar Items
-
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
by: Seo, Junyoung, et al.
Published: (2023) -
Geometry-Aware Score Distillation via 3D Consistent Noising and Gradient Consistency Modeling
by: Kwak, Min-Seop, et al.
Published: (2024) -
DirecT2V: Large Language Models are Frame-Level Directors for Zero-Shot Text-to-Video Generation
by: Hong, Susung, et al.
Published: (2023) -
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025) -
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
by: Kim, Inès Hyeonsu, et al.
Published: (2024)