Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Seo, Hoigi, Jeong, Wongi, Seo, Jae-sun, Chun, Se Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Personalization of Quantized Diffusion Model without Backpropagation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
von: Jeong, Wongi, et al.
Veröffentlicht: (2026)
von: Jeong, Wongi, et al.
Veröffentlicht: (2026)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2026)
von: Seo, Hoigi, et al.
Veröffentlicht: (2026)
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
von: Jeong, Wongi, et al.
Veröffentlicht: (2025)
von: Jeong, Wongi, et al.
Veröffentlicht: (2025)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2024)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding
von: Jang, Ji Ha, et al.
Veröffentlicht: (2024)
von: Jang, Ji Ha, et al.
Veröffentlicht: (2024)
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
von: Kim, Heechang, et al.
Veröffentlicht: (2025)
von: Kim, Heechang, et al.
Veröffentlicht: (2025)
Localized Concept Erasure for Text-to-Image Diffusion Models Using Training-Free Gated Low-Rank Adaptation
von: Lee, Byung Hyun, et al.
Veröffentlicht: (2025)
von: Lee, Byung Hyun, et al.
Veröffentlicht: (2025)
Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models
von: Kim, Hyungjin, et al.
Veröffentlicht: (2025)
von: Kim, Hyungjin, et al.
Veröffentlicht: (2025)
DragText: Rethinking Text Embedding in Point-based Image Editing
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules
von: Bang, Junseo, et al.
Veröffentlicht: (2026)
von: Bang, Junseo, et al.
Veröffentlicht: (2026)
Visual Words Meet BM25: Sparse Auto-Encoder Visual Word Scoring for Image Retrieval
von: Han, Donghoon, et al.
Veröffentlicht: (2026)
von: Han, Donghoon, et al.
Veröffentlicht: (2026)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
M2-Encoder: Advancing Bilingual Image-Text Understanding by Large-scale Efficient Pretraining
von: Guo, Qingpei, et al.
Veröffentlicht: (2024)
von: Guo, Qingpei, et al.
Veröffentlicht: (2024)
TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment
von: Yuan, Jiquan, et al.
Veröffentlicht: (2024)
von: Yuan, Jiquan, et al.
Veröffentlicht: (2024)
Contribution-based Low-Rank Adaptation with Pre-training Model for Real Image Restoration
von: Park, Donwon, et al.
Veröffentlicht: (2024)
von: Park, Donwon, et al.
Veröffentlicht: (2024)
TextCraftor: Your Text Encoder Can be Image Quality Controller
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
Text-guided 3D Human Motion Generation with Keyframe-based Parallel Skip Transformer
von: Geng, Zichen, et al.
Veröffentlicht: (2024)
von: Geng, Zichen, et al.
Veröffentlicht: (2024)
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
Prompt Decoupling for Text-to-Image Person Re-identification
von: Li, Weihao, et al.
Veröffentlicht: (2024)
von: Li, Weihao, et al.
Veröffentlicht: (2024)
Continual Multiple Instance Learning with Enhanced Localization for Histopathological Whole Slide Image Analysis
von: Lee, Byung Hyun, et al.
Veröffentlicht: (2025)
von: Lee, Byung Hyun, et al.
Veröffentlicht: (2025)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
KRAST: Knowledge-Augmented Robotic Action Recognition with Structured Text for Vision-Language Models
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
von: Wu, Mingrui, et al.
Veröffentlicht: (2025)
von: Wu, Mingrui, et al.
Veröffentlicht: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
The Era of Foundation Models in Medical Imaging is Approaching : A Scoping Review of the Clinical Value of Large-Scale Generative AI Applications in Radiology
von: Seo, Inwoo, et al.
Veröffentlicht: (2024)
von: Seo, Inwoo, et al.
Veröffentlicht: (2024)
Representations of Text and Images Align From Layer One
von: Wybitul, Evžen, et al.
Veröffentlicht: (2026)
von: Wybitul, Evžen, et al.
Veröffentlicht: (2026)
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
von: Zhang, Huixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Huixuan, et al.
Veröffentlicht: (2025)
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
HARIVO: Harnessing Text-to-Image Models for Video Generation
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
KOALA: Empirical Lessons Toward Memory-Efficient and Fast Diffusion Models for Text-to-Image Synthesis
von: Lee, Youngwan, et al.
Veröffentlicht: (2023)
von: Lee, Youngwan, et al.
Veröffentlicht: (2023)
TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis
von: Rahman, Kazi Mahathir, et al.
Veröffentlicht: (2025)
von: Rahman, Kazi Mahathir, et al.
Veröffentlicht: (2025)
FlipConcept: Tuning-Free Multi-Concept Personalization for Text-to-Image Generation
von: Woo, Young Beom, et al.
Veröffentlicht: (2025)
von: Woo, Young Beom, et al.
Veröffentlicht: (2025)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
von: Zhou, Linhan, et al.
Veröffentlicht: (2025)
von: Zhou, Linhan, et al.
Veröffentlicht: (2025)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
von: Kwak, Jaehyun, et al.
Veröffentlicht: (2025)
von: Kwak, Jaehyun, et al.
Veröffentlicht: (2025)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
von: Yu, Xinyao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Efficient Personalization of Quantized Diffusion Model without Backpropagation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025) -
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
von: Jeong, Wongi, et al.
Veröffentlicht: (2026) -
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025) -
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2026) -
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
von: Jeong, Wongi, et al.
Veröffentlicht: (2025)