Cycle Consistency as Reward: Learning Image-Text Alignment without Human Preferences
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bahng, Hyojin, Chan, Caroline, Durand, Fredo, Isola, Phillip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DreamReward: Text-to-3D Generation with Human Preference
von: Ye, Junliang, et al.
Veröffentlicht: (2024)
von: Ye, Junliang, et al.
Veröffentlicht: (2024)
DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data
von: Fu, Stephanie, et al.
Veröffentlicht: (2023)
von: Fu, Stephanie, et al.
Veröffentlicht: (2023)
Reward Incremental Learning in Text-to-Image Generation
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
CycleNet: Rethinking Cycle Consistency in Text-Guided Diffusion for Image Manipulation
von: Xu, Sihan, et al.
Veröffentlicht: (2023)
von: Xu, Sihan, et al.
Veröffentlicht: (2023)
When Does Perceptual Alignment Benefit Vision Representations?
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models
von: Wang, Austin, et al.
Veröffentlicht: (2026)
von: Wang, Austin, et al.
Veröffentlicht: (2026)
Words That Make Language Models Perceive
von: Wang, Sophie L., et al.
Veröffentlicht: (2025)
von: Wang, Sophie L., et al.
Veröffentlicht: (2025)
Adaptive Length Image Tokenization via Recurrent Allocation
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
Single-pass Adaptive Image Tokenization for Minimum Program Search
von: Duggal, Shivam, et al.
Veröffentlicht: (2025)
von: Duggal, Shivam, et al.
Veröffentlicht: (2025)
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
von: Gupta, Sharut, et al.
Veröffentlicht: (2025)
von: Gupta, Sharut, et al.
Veröffentlicht: (2025)
Personalized Representation from Personalized Generation
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
Attention Guided Alignment in Efficient Vision-Language Models
von: Mahajan, Shweta, et al.
Veröffentlicht: (2025)
von: Mahajan, Shweta, et al.
Veröffentlicht: (2025)
Information Theoretic Text-to-Image Alignment
von: Wang, Chao, et al.
Veröffentlicht: (2024)
von: Wang, Chao, et al.
Veröffentlicht: (2024)
UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward
von: Cheng, Yufeng, et al.
Veröffentlicht: (2025)
von: Cheng, Yufeng, et al.
Veröffentlicht: (2025)
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
Training Neural Networks from Scratch with Parallel Low-Rank Adapters
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
Consistent View Alignment Improves Foundation Models for 3D Medical Image Segmentation
von: Vaish, Puru, et al.
Veröffentlicht: (2025)
von: Vaish, Puru, et al.
Veröffentlicht: (2025)
Exploring Text-to-Motion Generation with Human Preference
von: Sheng, Jenny, et al.
Veröffentlicht: (2024)
von: Sheng, Jenny, et al.
Veröffentlicht: (2024)
Learning to Rank Caption Chains for Video-Text Alignment
von: Blume, Ansel, et al.
Veröffentlicht: (2026)
von: Blume, Ansel, et al.
Veröffentlicht: (2026)
Towards Better Alignment: Training Diffusion Models with Reinforcement Learning Against Sparse Rewards
von: Hu, Zijing, et al.
Veröffentlicht: (2025)
von: Hu, Zijing, et al.
Veröffentlicht: (2025)
RFMI: Estimating Mutual Information on Rectified Flow for Text-to-Image Alignment
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
Text Embedding is Not All You Need: Attention Control for Text-to-Image Semantic Alignment with Text Self-Attention Maps
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
Activation Reward Models for Few-Shot Model Alignment
von: Chai, Tianning, et al.
Veröffentlicht: (2025)
von: Chai, Tianning, et al.
Veröffentlicht: (2025)
Efficiency without Compromise: CLIP-aided Text-to-Image GANs with Increased Diversity
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
The Platonic Representation Hypothesis
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
von: Avrahami, Omri, et al.
Veröffentlicht: (2023)
von: Avrahami, Omri, et al.
Veröffentlicht: (2023)
Learning an Image Editing Model without Image Editing Pairs
von: Kumari, Nupur, et al.
Veröffentlicht: (2025)
von: Kumari, Nupur, et al.
Veröffentlicht: (2025)
Divergence Minimization Preference Optimization for Diffusion Model Alignment
von: Li, Binxu, et al.
Veröffentlicht: (2025)
von: Li, Binxu, et al.
Veröffentlicht: (2025)
Parallel Tempering Initial Sampling in Inference-Time Reward Alignment
von: Oh, Myeongjun, et al.
Veröffentlicht: (2026)
von: Oh, Myeongjun, et al.
Veröffentlicht: (2026)
Subject-driven Text-to-Image Generation via Preference-based Reinforcement Learning
von: Miao, Yanting, et al.
Veröffentlicht: (2024)
von: Miao, Yanting, et al.
Veröffentlicht: (2024)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
von: Chatterjee, Abhiroop, et al.
Veröffentlicht: (2025)
von: Chatterjee, Abhiroop, et al.
Veröffentlicht: (2025)
Text-centric Alignment for Multi-Modality Learning
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
Fine-Grained GRPO for Precise Preference Alignment in Flow Models
von: Zhou, Yujie, et al.
Veröffentlicht: (2025)
von: Zhou, Yujie, et al.
Veröffentlicht: (2025)
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium
von: Hu, Jiaming, et al.
Veröffentlicht: (2026)
von: Hu, Jiaming, et al.
Veröffentlicht: (2026)
Test-time Alignment of Diffusion Models without Reward Over-optimization
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DreamReward: Text-to-3D Generation with Human Preference
von: Ye, Junliang, et al.
Veröffentlicht: (2024) -
DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data
von: Fu, Stephanie, et al.
Veröffentlicht: (2023) -
Reward Incremental Learning in Text-to-Image Generation
von: Wang, Maorong, et al.
Veröffentlicht: (2024) -
CycleNet: Rethinking Cycle Consistency in Text-Guided Diffusion for Image Manipulation
von: Xu, Sihan, et al.
Veröffentlicht: (2023) -
When Does Perceptual Alignment Benefit Vision Representations?
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)