Parrot: Pareto-optimal Multi-Reward Reinforcement Learning Framework for Text-to-Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Seung Hyun, Li, Yinxiao, Ke, Junjie, Yoo, Innfarn, Zhang, Han, Yu, Jiahui, Wang, Qifei, Deng, Fei, Entis, Glenn, He, Junfeng, Li, Gang, Kim, Sangpil, Essa, Irfan, Yang, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Calibrated Multi-Preference Optimization for Aligning Diffusion Models
by: Lee, Kyungmin, et al.
Published: (2025)
by: Lee, Kyungmin, et al.
Published: (2025)
Cropper: Vision-Language Model for Image Cropping through In-Context Learning
by: Lee, Seung Hyun, et al.
Published: (2024)
by: Lee, Seung Hyun, et al.
Published: (2024)
Optical Diffusion Models for Image Generation
by: Oguz, Ilker, et al.
Published: (2024)
by: Oguz, Ilker, et al.
Published: (2024)
PRDP: Proximal Reward Difference Prediction for Large-Scale Reward Finetuning of Diffusion Models
by: Deng, Fei, et al.
Published: (2024)
by: Deng, Fei, et al.
Published: (2024)
Predicting Depressive Symptoms through Emotion Pairs within Asian American Families
by: Youm, Sangpil, et al.
Published: (2026)
by: Youm, Sangpil, et al.
Published: (2026)
Beyond Sparse Rewards: Enhancing Reinforcement Learning with Language Model Critique in Text Generation
by: Cao, Meng, et al.
Published: (2024)
by: Cao, Meng, et al.
Published: (2024)
HierSum: A Global and Local Attention Mechanism for Video Summarization
by: Beedu, Apoorva, et al.
Published: (2025)
by: Beedu, Apoorva, et al.
Published: (2025)
On the Efficacy of Text-Based Input Modalities for Action Anticipation
by: Beedu, Apoorva, et al.
Published: (2024)
by: Beedu, Apoorva, et al.
Published: (2024)
CATSplat: Context-Aware Transformer with Spatial Guidance for Generalizable 3D Gaussian Splatting from A Single-View Image
by: Roh, Wonseok, et al.
Published: (2024)
by: Roh, Wonseok, et al.
Published: (2024)
Parrot Captions Teach CLIP to Spot Text
by: Lin, Yiqi, et al.
Published: (2023)
by: Lin, Yiqi, et al.
Published: (2023)
Beyond Stereotypes: Exploring How Minority College Students Experience Stigma on Reddit
by: Han, Chaeeun, et al.
Published: (2025)
by: Han, Chaeeun, et al.
Published: (2025)
Information Seeking and Communication among International Students on Reddit
by: Han, Chaeeun, et al.
Published: (2024)
by: Han, Chaeeun, et al.
Published: (2024)
MoCHA: Denoising Caption Supervision for Motion-Text Retrieval
by: Warner, Nikolai, et al.
Published: (2026)
by: Warner, Nikolai, et al.
Published: (2026)
Pareto Inverse Reinforcement Learning for Diverse Expert Policy Generation
by: Kim, Woo Kyung, et al.
Published: (2024)
by: Kim, Woo Kyung, et al.
Published: (2024)
Leveraging Procedural Knowledge and Task Hierarchies for Efficient Instructional Video Pre-training
by: Samel, Karan, et al.
Published: (2025)
by: Samel, Karan, et al.
Published: (2025)
SLAIM: Robust Dense Neural SLAM for Online Tracking and Mapping
by: Cartillier, Vincent, et al.
Published: (2024)
by: Cartillier, Vincent, et al.
Published: (2024)
3D Semantic MapNet: Building Maps for Multi-Object Re-Identification in 3D
by: Cartillier, Vincent, et al.
Published: (2024)
by: Cartillier, Vincent, et al.
Published: (2024)
Focus-N-Fix: Region-Aware Fine-Tuning for Text-to-Image Generation
by: Xing, Xiaoying, et al.
Published: (2025)
by: Xing, Xiaoying, et al.
Published: (2025)
Agentic Reinforcement Learning with Implicit Step Rewards
by: Liu, Xiaoqian, et al.
Published: (2025)
by: Liu, Xiaoqian, et al.
Published: (2025)
Balancing Rewards in Text Summarization: Multi-Objective Reinforcement Learning via HyperVolume Optimization
by: Song, Junjie, et al.
Published: (2025)
by: Song, Junjie, et al.
Published: (2025)
Minimax optimal transfer learning for high-dimensional additive regression
by: Moon, Seung Hyun
Published: (2025)
by: Moon, Seung Hyun
Published: (2025)
HALO: Human-Aligned End-to-end Image Retargeting with Layered Transformations
by: Xu, Yiran, et al.
Published: (2025)
by: Xu, Yiran, et al.
Published: (2025)
The Stochastic Parrot on LLM's Shoulder: A Summative Assessment of Physical Concept Understanding
by: Yu, Mo, et al.
Published: (2025)
by: Yu, Mo, et al.
Published: (2025)
Mamba Fusion: Learning Actions Through Questioning
by: Dong, Zhikang, et al.
Published: (2024)
by: Dong, Zhikang, et al.
Published: (2024)
Exploring Efficient Foundational Multi-modal Models for Video Summarization
by: Samel, Karan, et al.
Published: (2024)
by: Samel, Karan, et al.
Published: (2024)
Dr. Carl Parrot
by: NA
Published: (1911)
by: NA
Published: (1911)
Fashion-VDM: Video Diffusion Model for Virtual Try-On
by: Karras, Johanna, et al.
Published: (2024)
by: Karras, Johanna, et al.
Published: (2024)
Parrot: Multilingual Visual Instruction Tuning
by: Sun, Hai-Long, et al.
Published: (2024)
by: Sun, Hai-Long, et al.
Published: (2024)
Comment on “The Effects of Financing Green and Brown Sectors: What Do Theories and Evidence Say?”
by: Qifei Zhu
Published: (2025)
by: Qifei Zhu
Published: (2025)
Chiplet Actuary: A Quantitative Cost Model and Multi-Chiplet Architecture Exploration
by: Feng, Yinxiao, et al.
Published: (2022)
by: Feng, Yinxiao, et al.
Published: (2022)
Switch-Less Dragonfly on Wafers: A Scalable Interconnection Architecture based on Wafer-Scale Integration
by: Feng, Yinxiao, et al.
Published: (2024)
by: Feng, Yinxiao, et al.
Published: (2024)
Pareto optimal proxy metrics
by: Zito, Alessandro, et al.
Published: (2023)
by: Zito, Alessandro, et al.
Published: (2023)
Pareto Set Learning for Multi-Objective Reinforcement Learning
by: Liu, Erlong, et al.
Published: (2025)
by: Liu, Erlong, et al.
Published: (2025)
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
MapReduce LoRA: Advancing the Pareto Front in Multi-Preference Optimization for Generative Models
by: Chen, Chieh-Yun, et al.
Published: (2025)
by: Chen, Chieh-Yun, et al.
Published: (2025)
Real-Time ESFP: Estimating, Smoothing, Filtering, and Pose-Mapping
by: Cui, Qifei, et al.
Published: (2025)
by: Cui, Qifei, et al.
Published: (2025)
Deep Pareto Reinforcement Learning for Multi-Objective Recommender Systems
by: Li, Pan, et al.
Published: (2024)
by: Li, Pan, et al.
Published: (2024)
RedParrot: Accelerating NL-to-DSL for Business Analytics via Query Semantic Caching
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
Reward-Agnostic Prompt Optimization for Text-to-Image Diffusion Models
by: Kim, Semin, et al.
Published: (2025)
by: Kim, Semin, et al.
Published: (2025)
Pareto-Guided Optimal Transport for Multi-Reward Alignment
by: Ba, Ying, et al.
Published: (2026)
by: Ba, Ying, et al.
Published: (2026)
Similar Items
-
Calibrated Multi-Preference Optimization for Aligning Diffusion Models
by: Lee, Kyungmin, et al.
Published: (2025) -
Cropper: Vision-Language Model for Image Cropping through In-Context Learning
by: Lee, Seung Hyun, et al.
Published: (2024) -
Optical Diffusion Models for Image Generation
by: Oguz, Ilker, et al.
Published: (2024) -
PRDP: Proximal Reward Difference Prediction for Large-Scale Reward Finetuning of Diffusion Models
by: Deng, Fei, et al.
Published: (2024) -
Predicting Depressive Symptoms through Emotion Pairs within Asian American Families
by: Youm, Sangpil, et al.
Published: (2026)