Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ba, Ying, Zhang, Tianyu, Bai, Yalong, Mo, Wenyi, Liang, Tao, Su, Bing, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pareto-Guided Optimal Transport for Multi-Reward Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2026)
von: Ba, Ying, et al.
Veröffentlicht: (2026)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
StyleInject: Parameter Efficient Tuning of Text-to-Image Diffusion Models
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
REVEALER: Reinforcement-Guided Visual Reasoning for Element-Level Text-Image Alignment Evaluation
von: Shi, Fulin, et al.
Veröffentlicht: (2025)
von: Shi, Fulin, et al.
Veröffentlicht: (2025)
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
Spatio-Temporal Branching for Motion Prediction using Motion Increments
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation
von: Zhang, Wenchao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenchao, et al.
Veröffentlicht: (2025)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
von: Cui, Tianyu, et al.
Veröffentlicht: (2025)
von: Cui, Tianyu, et al.
Veröffentlicht: (2025)
Personalized Safety Alignment for Text-to-Image Diffusion Models
von: Lei, Yu, et al.
Veröffentlicht: (2025)
von: Lei, Yu, et al.
Veröffentlicht: (2025)
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think
von: Chen, Liang, et al.
Veröffentlicht: (2025)
von: Chen, Liang, et al.
Veröffentlicht: (2025)
High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance
von: Gao, Danyi
Veröffentlicht: (2025)
von: Gao, Danyi
Veröffentlicht: (2025)
GEA: Generation-Enhanced Alignment for Text-to-Image Person Retrieval
von: Zou, Hao, et al.
Veröffentlicht: (2025)
von: Zou, Hao, et al.
Veröffentlicht: (2025)
Beyond Pixels: Text Enhances Generalization in Real-World Image Restoration
von: Sun, Haoze, et al.
Veröffentlicht: (2024)
von: Sun, Haoze, et al.
Veröffentlicht: (2024)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
Reward Incremental Learning in Text-to-Image Generation
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
Personalized Reward Modeling for Text-to-Image Generation
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
Explainable AI-Generated Image Detection RewardBench
von: Yang, Michael, et al.
Veröffentlicht: (2025)
von: Yang, Michael, et al.
Veröffentlicht: (2025)
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation
von: Feng, Xuelu, et al.
Veröffentlicht: (2025)
von: Feng, Xuelu, et al.
Veröffentlicht: (2025)
Cycle Consistency as Reward: Learning Image-Text Alignment without Human Preferences
von: Bahng, Hyojin, et al.
Veröffentlicht: (2025)
von: Bahng, Hyojin, et al.
Veröffentlicht: (2025)
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
von: Agarwal, Aishwarya, et al.
Veröffentlicht: (2024)
von: Agarwal, Aishwarya, et al.
Veröffentlicht: (2024)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
von: Shen, Guibao, et al.
Veröffentlicht: (2024)
von: Shen, Guibao, et al.
Veröffentlicht: (2024)
Fast Prompt Alignment for Text-to-Image Generation
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
DreamVideo: High-Fidelity Image-to-Video Generation with Image Retention and Text Guidance
von: Wang, Cong, et al.
Veröffentlicht: (2023)
von: Wang, Cong, et al.
Veröffentlicht: (2023)
IFAdapter: Instance Feature Control for Grounded Text-to-Image Generation
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
STAR: Scale-wise Text-conditioned AutoRegressive image generation
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
Enhancing Spatial Understanding in Image Generation via Reward Modeling
von: Tang, Zhenyu, et al.
Veröffentlicht: (2026)
von: Tang, Zhenyu, et al.
Veröffentlicht: (2026)
InstructEngine: Instruction-driven Text-to-Image Alignment
von: Lu, Xingyu, et al.
Veröffentlicht: (2025)
von: Lu, Xingyu, et al.
Veröffentlicht: (2025)
The Image as Its Own Reward: Reinforcement Learning with Adversarial Reward for Image Generation
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
Towards RGB-NIR Cross-modality Image Registration and Beyond
von: Li, Huadong, et al.
Veröffentlicht: (2024)
von: Li, Huadong, et al.
Veröffentlicht: (2024)
More Than Generation: Unifying Generation and Depth Estimation via Text-to-Image Diffusion Models
von: Lin, Hongkai, et al.
Veröffentlicht: (2025)
von: Lin, Hongkai, et al.
Veröffentlicht: (2025)
MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Pareto-Guided Optimal Transport for Multi-Reward Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2026) -
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025) -
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)