Guided Score identity Distillation for Data-Free One-Step Text-to-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Mingyuan, Wang, Zhendong, Zheng, Huangjie, Huang, Hai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Score identity Distillation: Exponentially Fast Distillation of Pretrained Diffusion Models for One-Step Generation
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
Adversarial Score identity Distillation: Rapidly Surpassing the Teacher in One Step
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
Denoising Score Distillation: From Noisy Diffusion Pretraining to One-Step High-Quality Generation
di: Chen, Tianyu, et al.
Pubblicazione: (2025)
di: Chen, Tianyu, et al.
Pubblicazione: (2025)
Few-Step Diffusion via Score identity Distillation
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025)
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025)
Score Distillation of Flow Matching Models
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025)
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025)
Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
One-Step Diffusion Distillation through Score Implicit Matching
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)
di: Saxon, Michael, et al.
Pubblicazione: (2024)
di: Saxon, Michael, et al.
Pubblicazione: (2024)
Holistic Evaluation for Interleaved Text-and-Image Generation
di: Liu, Minqian, et al.
Pubblicazione: (2024)
di: Liu, Minqian, et al.
Pubblicazione: (2024)
Text-only Synthesis for Image Captioning
di: Zhou, Qing, et al.
Pubblicazione: (2024)
di: Zhou, Qing, et al.
Pubblicazione: (2024)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
di: Zhou, Shijie, et al.
Pubblicazione: (2025)
di: Zhou, Shijie, et al.
Pubblicazione: (2025)
Automatic Evaluation for Text-to-image Generation: Task-decomposed Framework, Distilled Training, and Meta-evaluation Benchmark
di: Tu, Rong-Cheng, et al.
Pubblicazione: (2024)
di: Tu, Rong-Cheng, et al.
Pubblicazione: (2024)
ComCLIP: Training-Free Compositional Image and Text Matching
di: Jiang, Kenan, et al.
Pubblicazione: (2022)
di: Jiang, Kenan, et al.
Pubblicazione: (2022)
Enhancing and Accelerating Diffusion-Based Inverse Problem Solving through Measurements Optimization
di: Chen, Tianyu, et al.
Pubblicazione: (2024)
di: Chen, Tianyu, et al.
Pubblicazione: (2024)
VideoScore2: Think before You Score in Generative Video Evaluation
di: He, Xuan, et al.
Pubblicazione: (2025)
di: He, Xuan, et al.
Pubblicazione: (2025)
How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
VEGA: Learning Interleaved Image-Text Comprehension in Vision-Language Large Models
di: Zhou, Chenyu, et al.
Pubblicazione: (2024)
di: Zhou, Chenyu, et al.
Pubblicazione: (2024)
Text or Image? What is More Important in Cross-Domain Generalization Capabilities of Hate Meme Detection Models?
di: Aggarwal, Piush, et al.
Pubblicazione: (2024)
di: Aggarwal, Piush, et al.
Pubblicazione: (2024)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation
di: Chen, Wenting, et al.
Pubblicazione: (2023)
di: Chen, Wenting, et al.
Pubblicazione: (2023)
Text-to-3D Generation using Jensen-Shannon Score Distillation
di: Do, Khoi, et al.
Pubblicazione: (2025)
di: Do, Khoi, et al.
Pubblicazione: (2025)
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
di: Zhang, Leixin, et al.
Pubblicazione: (2024)
di: Zhang, Leixin, et al.
Pubblicazione: (2024)
Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads
di: Kou, Siqi, et al.
Pubblicazione: (2024)
di: Kou, Siqi, et al.
Pubblicazione: (2024)
RusCode: Russian Cultural Code Benchmark for Text-to-Image Generation
di: Vasilev, Viacheslav, et al.
Pubblicazione: (2025)
di: Vasilev, Viacheslav, et al.
Pubblicazione: (2025)
Accurate Scene Text Recognition with Efficient Model Scaling and Cloze Self-Distillation
di: Maracani, Andrea, et al.
Pubblicazione: (2025)
di: Maracani, Andrea, et al.
Pubblicazione: (2025)
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
di: Liu, Delong, et al.
Pubblicazione: (2023)
di: Liu, Delong, et al.
Pubblicazione: (2023)
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
di: Du, Yang, et al.
Pubblicazione: (2025)
di: Du, Yang, et al.
Pubblicazione: (2025)
Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
CAPEEN: Image Captioning with Early Exits and Knowledge Distillation
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text
di: Zhang, Junzhe, et al.
Pubblicazione: (2025)
di: Zhang, Junzhe, et al.
Pubblicazione: (2025)
TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
di: Feng, Weixi, et al.
Pubblicazione: (2024)
di: Feng, Weixi, et al.
Pubblicazione: (2024)
StarVector: Generating Scalable Vector Graphics Code from Images and Text
di: Rodriguez, Juan A., et al.
Pubblicazione: (2023)
di: Rodriguez, Juan A., et al.
Pubblicazione: (2023)
CoMat: Aligning Text-to-Image Diffusion Model with Image-to-Text Concept Matching
di: Jiang, Dongzhi, et al.
Pubblicazione: (2024)
di: Jiang, Dongzhi, et al.
Pubblicazione: (2024)
Discriminative Probing and Tuning for Text-to-Image Generation
di: Qu, Leigang, et al.
Pubblicazione: (2024)
di: Qu, Leigang, et al.
Pubblicazione: (2024)
Text-Printed Image: Bridging the Image-Text Modality Gap for Text-centric Training of Large Vision-Language Models
di: Yamabe, Shojiro, et al.
Pubblicazione: (2025)
di: Yamabe, Shojiro, et al.
Pubblicazione: (2025)
DreamPolish: Domain Score Distillation With Progressive Geometry Generation
di: Cheng, Yean, et al.
Pubblicazione: (2024)
di: Cheng, Yean, et al.
Pubblicazione: (2024)
Boosting Medical Image-based Cancer Detection via Text-guided Supervision from Reports
di: Guo, Guangyu, et al.
Pubblicazione: (2024)
di: Guo, Guangyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Score identity Distillation: Exponentially Fast Distillation of Pretrained Diffusion Models for One-Step Generation
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024) -
Adversarial Score identity Distillation: Rapidly Surpassing the Teacher in One Step
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024) -
Denoising Score Distillation: From Noisy Diffusion Pretraining to One-Step High-Quality Generation
di: Chen, Tianyu, et al.
Pubblicazione: (2025) -
Few-Step Diffusion via Score identity Distillation
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025) -
Score Distillation of Flow Matching Models
di: Zhou, Mingyuan, et al.
Pubblicazione: (2025)