Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Haoyuan, Xia, Bo, Chang, Yongzhe, Wang, Xueqian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping
von: Sun, Haoyuan, et al.
Veröffentlicht: (2026)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2026)
MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment
von: Gao, Zhiting, et al.
Veröffentlicht: (2025)
von: Gao, Zhiting, et al.
Veröffentlicht: (2025)
Learning Multi-dimensional Human Preference for Text-to-Image Generation
von: Zhang, Sixian, et al.
Veröffentlicht: (2024)
von: Zhang, Sixian, et al.
Veröffentlicht: (2024)
Fast Prompt Alignment for Text-to-Image Generation
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
Scalable Ranked Preference Optimization for Text-to-Image Generation
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
AGFSync: Leveraging AI-Generated Feedback for Preference Optimization in Text-to-Image Generation
von: An, Jingkun, et al.
Veröffentlicht: (2024)
von: An, Jingkun, et al.
Veröffentlicht: (2024)
Unleashing the Potential of Large Language Models for Text-to-Image Generation through Autoregressive Representation Alignment
von: Xie, Xing, et al.
Veröffentlicht: (2025)
von: Xie, Xing, et al.
Veröffentlicht: (2025)
Instant Preference Alignment for Text-to-Image Diffusion Models
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
HCMA: Hierarchical Cross-model Alignment for Grounded Text-to-Image Generation
von: Wang, Hang, et al.
Veröffentlicht: (2025)
von: Wang, Hang, et al.
Veröffentlicht: (2025)
Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
BideDPO: Conditional Image Generation with Simultaneous Text and Condition Alignment
von: Zhou, Dewei, et al.
Veröffentlicht: (2025)
von: Zhou, Dewei, et al.
Veröffentlicht: (2025)
FocusDiff: Advancing Fine-Grained Text-Image Alignment for Autoregressive Visual Generation through RL
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
DETONATE: A Benchmark for Text-to-Image Alignment and Kernelized Direct Preference Optimization
von: Prasad, Renjith, et al.
Veröffentlicht: (2025)
von: Prasad, Renjith, et al.
Veröffentlicht: (2025)
Instruction-augmented Multimodal Alignment for Image-Text and Element Matching
von: Yue, Xinli, et al.
Veröffentlicht: (2025)
von: Yue, Xinli, et al.
Veröffentlicht: (2025)
OSPO: Object-Centric Self-Improving Preference Optimization for Text-to-Image Generation
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
PopAlign: Population-Level Alignment for Fair Text-to-Image Generation
von: Li, Shufan, et al.
Veröffentlicht: (2024)
von: Li, Shufan, et al.
Veröffentlicht: (2024)
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
von: Zhang, Bingqing, et al.
Veröffentlicht: (2024)
von: Zhang, Bingqing, et al.
Veröffentlicht: (2024)
High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance
von: Gao, Danyi
Veröffentlicht: (2025)
von: Gao, Danyi
Veröffentlicht: (2025)
What Makes a Good Generated Image? Investigating Human and Multimodal LLM Image Preference Alignment
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think
von: Chen, Liang, et al.
Veröffentlicht: (2025)
von: Chen, Liang, et al.
Veröffentlicht: (2025)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
von: Song, Xinyang, et al.
Veröffentlicht: (2025)
von: Song, Xinyang, et al.
Veröffentlicht: (2025)
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025)
von: Ba, Ying, et al.
Veröffentlicht: (2025)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
Continual Learning for Image Captioning through Improved Image-Text Alignment
von: Taetz, Bertram, et al.
Veröffentlicht: (2025)
von: Taetz, Bertram, et al.
Veröffentlicht: (2025)
Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization
von: Liu, Zhuohan, et al.
Veröffentlicht: (2026)
von: Liu, Zhuohan, et al.
Veröffentlicht: (2026)
Exploring Motion-Language Alignment for Text-driven Motion Generation
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
CritiFusion: Semantic Critique and Spectral Alignment for Faithful Text-to-Image Generation
von: Chen, ZhenQi, et al.
Veröffentlicht: (2025)
von: Chen, ZhenQi, et al.
Veröffentlicht: (2025)
Hierarchical Vision-Language Alignment for Text-to-Image Generation via Diffusion Models
von: Johnson, Emily, et al.
Veröffentlicht: (2025)
von: Johnson, Emily, et al.
Veröffentlicht: (2025)
Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation
von: Zhang, Wenchao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenchao, et al.
Veröffentlicht: (2025)
Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
HuViDPO:Enhancing Video Generation through Direct Preference Optimization for Human-Centric Alignment
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization
von: Wu, Xianfeng, et al.
Veröffentlicht: (2025)
von: Wu, Xianfeng, et al.
Veröffentlicht: (2025)
Cross Paradigm Representation and Alignment Transformer for Image Deraining
von: Zou, Shun, et al.
Veröffentlicht: (2025)
von: Zou, Shun, et al.
Veröffentlicht: (2025)
Generalized Small Object Detection:A Point-Prompted Paradigm and Benchmark
von: Zhu, Haoran, et al.
Veröffentlicht: (2026)
von: Zhu, Haoran, et al.
Veröffentlicht: (2026)
Anomaly-Preference Image Generation
von: Wang, Fuyun, et al.
Veröffentlicht: (2026)
von: Wang, Fuyun, et al.
Veröffentlicht: (2026)
CSGO: Content-Style Composition in Text-to-Image Generation
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
Text2Traffic: A Text-to-Image Generation and Editing Method for Traffic Scenes
von: Lv, Feng, et al.
Veröffentlicht: (2025)
von: Lv, Feng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
von: Luo, Yifu, et al.
Veröffentlicht: (2025) -
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025) -
Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping
von: Sun, Haoyuan, et al.
Veröffentlicht: (2026) -
MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment
von: Gao, Zhiting, et al.
Veröffentlicht: (2025) -
Learning Multi-dimensional Human Preference for Text-to-Image Generation
von: Zhang, Sixian, et al.
Veröffentlicht: (2024)