PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ping, Bowen, Jia, Chengyou, Luo, Minnan, Xia, Changliang, Shen, Xin, Dang, Zhuohang, Qian, Hangwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
von: Xia, Changliang, et al.
Veröffentlicht: (2025)
von: Xia, Changliang, et al.
Veröffentlicht: (2025)
Why Settle for One? Text-to-ImageSet Generation and Evaluation
von: Jia, Chengyou, et al.
Veröffentlicht: (2025)
von: Jia, Chengyou, et al.
Veröffentlicht: (2025)
Flow-Factory: A Unified Framework for Reinforcement Learning in Flow-Matching Models
von: Ping, Bowen, et al.
Veröffentlicht: (2026)
von: Ping, Bowen, et al.
Veröffentlicht: (2026)
AutoGPS: Automated Geometry Problem Solving via Multimodal Formalization and Deductive Reasoning
von: Ping, Bowen, et al.
Veröffentlicht: (2025)
von: Ping, Bowen, et al.
Veröffentlicht: (2025)
From Ideal to Real: Unified and Data-Efficient Dense Prediction for Real-World Scenarios
von: Xia, Changliang, et al.
Veröffentlicht: (2025)
von: Xia, Changliang, et al.
Veröffentlicht: (2025)
Multi-Modal Dataset Distillation in the Wild
von: Dang, Zhuohang, et al.
Veröffentlicht: (2025)
von: Dang, Zhuohang, et al.
Veröffentlicht: (2025)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
PSDiff: Diffusion Model for Person Search with Iterative and Collaborative Refinement
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
Disentangled Representation Learning with Transmitted Information Bottleneck
von: Dang, Zhuohang, et al.
Veröffentlicht: (2023)
von: Dang, Zhuohang, et al.
Veröffentlicht: (2023)
AgentStore: Scalable Integration of Heterogeneous Agents As Specialized Generalist Computer Assistant
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
PaCo-FR: Patch-Pixel Aligned End-to-End Codebook Learning for Facial Representation Pre-training
von: Xie, Yin, et al.
Veröffentlicht: (2025)
von: Xie, Yin, et al.
Veröffentlicht: (2025)
Disentangled Noisy Correspondence Learning
von: Dang, Zhuohang, et al.
Veröffentlicht: (2024)
von: Dang, Zhuohang, et al.
Veröffentlicht: (2024)
PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation
von: Cao, Haofan, et al.
Veröffentlicht: (2026)
von: Cao, Haofan, et al.
Veröffentlicht: (2026)
PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards
von: Wang, Shulei, et al.
Veröffentlicht: (2025)
von: Wang, Shulei, et al.
Veröffentlicht: (2025)
CoFFT: Chain of Foresight-Focus Thought for Visual Language Models
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning
von: Hao, Heng, et al.
Veröffentlicht: (2025)
von: Hao, Heng, et al.
Veröffentlicht: (2025)
PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling
von: Jian, Ai, et al.
Veröffentlicht: (2025)
von: Jian, Ai, et al.
Veröffentlicht: (2025)
Chart-RL: Generalized Chart Comprehension via Reinforcement Learning with Verifiable Rewards
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
Evaluating Reward Model Generalization via Pairwise Maximum Discrepancy Competitions
von: Luo, Shunyang, et al.
Veröffentlicht: (2026)
von: Luo, Shunyang, et al.
Veröffentlicht: (2026)
Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
DyCoRM: Dynamic Criterion-Aware Reward Modeling for Text-to-Image Generation
von: Qian, Jiaying, et al.
Veröffentlicht: (2026)
von: Qian, Jiaying, et al.
Veröffentlicht: (2026)
RL-I2IT: Image-to-Image Translation with Deep Reinforcement Learning
von: Hu, Jing, et al.
Veröffentlicht: (2023)
von: Hu, Jing, et al.
Veröffentlicht: (2023)
ARIADNE: Agentic Reward-Informed Adaptive Decision Exploration via Blackboard-Driven MCTS for Competitive Program Generation
von: Wei, Minnan, et al.
Veröffentlicht: (2026)
von: Wei, Minnan, et al.
Veröffentlicht: (2026)
A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization
von: Xu, Wenyuan, et al.
Veröffentlicht: (2025)
von: Xu, Wenyuan, et al.
Veröffentlicht: (2025)
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation
von: Feng, Xuelu, et al.
Veröffentlicht: (2025)
von: Feng, Xuelu, et al.
Veröffentlicht: (2025)
Pairwise Calibrated Rewards for Pluralistic Alignment
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
FreRA: A Frequency-Refined Augmentation for Contrastive Learning on Time Series Classification
von: Tian, Tian, et al.
Veröffentlicht: (2025)
von: Tian, Tian, et al.
Veröffentlicht: (2025)
Exploring the Effectiveness and Interpretability of Texts in LLM-based Time Series Models
von: Sun, Zhengke, et al.
Veröffentlicht: (2025)
von: Sun, Zhengke, et al.
Veröffentlicht: (2025)
The Image as Its Own Reward: Reinforcement Learning with Adversarial Reward for Image Generation
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
EditScore: Unlocking Online RL for Image Editing via High-Fidelity Reward Modeling
von: Luo, Xin, et al.
Veröffentlicht: (2025)
von: Luo, Xin, et al.
Veröffentlicht: (2025)
DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization
von: Deng, Mengyi, et al.
Veröffentlicht: (2026)
von: Deng, Mengyi, et al.
Veröffentlicht: (2026)
UNEX-RL: Reinforcing Long-Term Rewards in Multi-Stage Recommender Systems with UNidirectional EXecution
von: Zhang, Gengrui, et al.
Veröffentlicht: (2024)
von: Zhang, Gengrui, et al.
Veröffentlicht: (2024)
ACE-RL: Adaptive Constraint-Enhanced Reward for Long-form Generation Reinforcement Learning
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
ELO-Rated Sequence Rewards: Advancing Reinforcement Learning Models
von: Ju, Qi, et al.
Veröffentlicht: (2024)
von: Ju, Qi, et al.
Veröffentlicht: (2024)
DiPaCo: Distributed Path Composition
von: Douillard, Arthur, et al.
Veröffentlicht: (2024)
von: Douillard, Arthur, et al.
Veröffentlicht: (2024)
Reinforcing Consistency in Video MLLMs with Structured Rewards
von: Quan, Yihao, et al.
Veröffentlicht: (2026)
von: Quan, Yihao, et al.
Veröffentlicht: (2026)
Enhancing the Outcome Reward-based RL Training of MLLMs with Self-Consistency Sampling
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
Revista CoPaLa
Veröffentlicht: (2020)
Veröffentlicht: (2020)
Ähnliche Einträge
-
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
von: Jia, Chengyou, et al.
Veröffentlicht: (2024) -
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
von: Xia, Changliang, et al.
Veröffentlicht: (2025) -
Why Settle for One? Text-to-ImageSet Generation and Evaluation
von: Jia, Chengyou, et al.
Veröffentlicht: (2025) -
Flow-Factory: A Unified Framework for Reinforcement Learning in Flow-Matching Models
von: Ping, Bowen, et al.
Veröffentlicht: (2026) -
AutoGPS: Automated Geometry Problem Solving via Multimodal Formalization and Deductive Reasoning
von: Ping, Bowen, et al.
Veröffentlicht: (2025)