DECOR:Decomposition and Projection of Text Embeddings for Text-to-Image Customization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jang, Geonhui, Kim, Jin-Hwa, Park, Yong-Hyun, Kim, Junho, Lee, Gayoung, Jeong, Yonghyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Direct Unlearning Optimization for Robust and Safe Text-to-Image Models
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
StyleKeeper: Prevent Content Leakage using Negative Visual Query Guidance
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
Visual Style Prompting with Swapping Self-Attention
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2024)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2024)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion
von: Kim, Jiwon, et al.
Veröffentlicht: (2025)
von: Kim, Jiwon, et al.
Veröffentlicht: (2025)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
Learning to Customize Text-to-Image Diffusion In Diverse Context
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
Enhancing Creative Generation on Stable Diffusion-based Models
von: Han, Jiyeon, et al.
Veröffentlicht: (2025)
von: Han, Jiyeon, et al.
Veröffentlicht: (2025)
Advancing Cross-Domain Generalizability in Face Anti-Spoofing: Insights, Design, and Metrics
von: Kim, Hyojin, et al.
Veröffentlicht: (2024)
von: Kim, Hyojin, et al.
Veröffentlicht: (2024)
DECOR: Deep Embedding Clustering with Orientation Robustness
von: Jothiraj, Fiona Victoria Stanley, et al.
Veröffentlicht: (2025)
von: Jothiraj, Fiona Victoria Stanley, et al.
Veröffentlicht: (2025)
Selectively Informative Description can Reduce Undesired Embedding Entanglements in Text-to-Image Personalization
von: Kim, Jimyeong, et al.
Veröffentlicht: (2024)
von: Kim, Jimyeong, et al.
Veröffentlicht: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
ParTY: Part-Guidance for Expressive Text-to-Motion Synthesis
von: Heo, KunHo, et al.
Veröffentlicht: (2026)
von: Heo, KunHo, et al.
Veröffentlicht: (2026)
Aligned Novel View Image and Geometry Synthesis via Cross-modal Attention Instillation
von: Kwak, Min-Seop, et al.
Veröffentlicht: (2025)
von: Kwak, Min-Seop, et al.
Veröffentlicht: (2025)
Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis
von: Lee, Jonghyun, et al.
Veröffentlicht: (2024)
von: Lee, Jonghyun, et al.
Veröffentlicht: (2024)
HP-GAN: Harnessing pretrained networks for GAN improvement with FakeTwins and discriminator consistency
von: Son, Geonhui, et al.
Veröffentlicht: (2026)
von: Son, Geonhui, et al.
Veröffentlicht: (2026)
Unified Diffusion Transformer for High-fidelity Text-Aware Image Restoration
von: Kim, Jin Hyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jin Hyeon, et al.
Veröffentlicht: (2025)
MINDiff: Mask-Integrated Negative Attention for Controlling Overfitting in Text-to-Image Personalization
von: Jeong, Seulgi, et al.
Veröffentlicht: (2025)
von: Jeong, Seulgi, et al.
Veröffentlicht: (2025)
Grounding World Simulation Models in a Real-World Metropolis
von: Seo, Junyoung, et al.
Veröffentlicht: (2026)
von: Seo, Junyoung, et al.
Veröffentlicht: (2026)
Tuning-Free Image Customization with Image and Text Guidance
von: Li, Pengzhi, et al.
Veröffentlicht: (2024)
von: Li, Pengzhi, et al.
Veröffentlicht: (2024)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Geometry-Aware Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2026)
von: Lee, Junho, et al.
Veröffentlicht: (2026)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
von: Kumari, Nupur, et al.
Veröffentlicht: (2024)
von: Kumari, Nupur, et al.
Veröffentlicht: (2024)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift
von: Kim, Gihoon, et al.
Veröffentlicht: (2025)
von: Kim, Gihoon, et al.
Veröffentlicht: (2025)
One-Shot Structure-Aware Stylized Image Synthesis
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
Text-Aware Image Restoration with Diffusion Models
von: Min, Jaewon, et al.
Veröffentlicht: (2025)
von: Min, Jaewon, et al.
Veröffentlicht: (2025)
Is There a Better Source Distribution than Gaussian? Exploring Source Distributions for Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2025)
von: Lee, Junho, et al.
Veröffentlicht: (2025)
Group-wise Scaling and Orthogonal Decomposition for Domain-Invariant Feature Extraction in Face Anti-Spoofing
von: Jung, Seungjin, et al.
Veröffentlicht: (2025)
von: Jung, Seungjin, et al.
Veröffentlicht: (2025)
Calligrapher: Freestyle Text Image Customization
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
Fully Geometric Panoramic Localization
von: Kim, Junho, et al.
Veröffentlicht: (2024)
von: Kim, Junho, et al.
Veröffentlicht: (2024)
DragText: Rethinking Text Embedding in Point-based Image Editing
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
Text Embedding is Not All You Need: Attention Control for Text-to-Image Semantic Alignment with Text Self-Attention Maps
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Direct Unlearning Optimization for Robust and Safe Text-to-Image Models
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024) -
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
von: Lee, Gayoung, et al.
Veröffentlicht: (2025) -
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024) -
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
von: Seo, Junyoung, et al.
Veröffentlicht: (2023) -
StyleKeeper: Prevent Content Leakage using Negative Visual Query Guidance
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)