Pareto-Guided Optimal Transport for Multi-Reward Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ba, Ying, Zhang, Tianyu, Zhou, Mohan, Bai, Yalong, Mo, Wenyi, Zhang, Guiwei, Su, Bing, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025)
von: Ba, Ying, et al.
Veröffentlicht: (2025)
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Spatio-Temporal Branching for Motion Prediction using Motion Increments
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
von: Li, Jiangmeng, et al.
Veröffentlicht: (2022)
StyleInject: Parameter Efficient Tuning of Text-to-Image Diffusion Models
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
von: Zhou, Mohan, et al.
Veröffentlicht: (2024)
Pre-training CLIP against Data Poisoning with Optimal Transport-based Matching and Alignment
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
POCA: Pareto-Optimal Curriculum Alignment for Visual Text Generation
von: Fan, Yaohou, et al.
Veröffentlicht: (2026)
von: Fan, Yaohou, et al.
Veröffentlicht: (2026)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
Parrot: Pareto-optimal Multi-Reward Reinforcement Learning Framework for Text-to-Image Generation
von: Lee, Seung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Seung Hyun, et al.
Veröffentlicht: (2024)
Rad-VLSM: A Cross-Modal Framework with Semantics-Assisted Prompting for Medical Segmentation and Diagnosis
von: Zhang, Fengyi, et al.
Veröffentlicht: (2026)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2026)
Optimizing Distributional Geometry Alignment with Optimal Transport for Generative Dataset Distillation
von: Cui, Xiao, et al.
Veröffentlicht: (2025)
von: Cui, Xiao, et al.
Veröffentlicht: (2025)
From Poses to Identity: Training-Free Person Re-Identification via Feature Centralization
von: Yuan, Chao, et al.
Veröffentlicht: (2025)
von: Yuan, Chao, et al.
Veröffentlicht: (2025)
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale
von: Tang, Zhicong, et al.
Veröffentlicht: (2026)
von: Tang, Zhicong, et al.
Veröffentlicht: (2026)
ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
ReAlign: Bilingual Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
REVEALER: Reinforcement-Guided Visual Reasoning for Element-Level Text-Image Alignment Evaluation
von: Shi, Fulin, et al.
Veröffentlicht: (2025)
von: Shi, Fulin, et al.
Veröffentlicht: (2025)
Reinforcing Few-step Generators via Reward-Tilted Distribution Matching
von: Huang, Yushi, et al.
Veröffentlicht: (2026)
von: Huang, Yushi, et al.
Veröffentlicht: (2026)
Pareto-Guided Optimization for Uncertainty-Aware Medical Image Segmentation
von: Zhang, Jinming, et al.
Veröffentlicht: (2026)
von: Zhang, Jinming, et al.
Veröffentlicht: (2026)
STAR: Scale-wise Text-conditioned AutoRegressive image generation
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
Mahalanobis Distance-based Multi-view Optimal Transport for Multi-view Crowd Localization
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
Multi-dimensional Preference Alignment by Conditioning Reward Itself
von: Jang, Jiho, et al.
Veröffentlicht: (2025)
von: Jang, Jiho, et al.
Veröffentlicht: (2025)
GRAM-MAMBA: Holistic Feature Alignment for Wireless Perception with Adaptive Low-Rank Compensation
von: Yang, Weiqi, et al.
Veröffentlicht: (2025)
von: Yang, Weiqi, et al.
Veröffentlicht: (2025)
HOT: Harmonic-Constrained Optimal Transport for Remote Photoplethysmography Domain Adaptation
von: Nguyen, Ba-Thinh, et al.
Veröffentlicht: (2026)
von: Nguyen, Ba-Thinh, et al.
Veröffentlicht: (2026)
Pareto-Enhanced Portrait Generation: Vision-Aligned Text Supervision for Alignment, Realism, and Aesthetics
von: Wang, Yunlong, et al.
Veröffentlicht: (2026)
von: Wang, Yunlong, et al.
Veröffentlicht: (2026)
Split-Fuse-Transport: Annotation-Free Saliency via Dual Clustering and Optimal Transport Alignment
von: Ramzan, Muhammad Umer, et al.
Veröffentlicht: (2025)
von: Ramzan, Muhammad Umer, et al.
Veröffentlicht: (2025)
Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport
von: Yu, Qinkai, et al.
Veröffentlicht: (2025)
von: Yu, Qinkai, et al.
Veröffentlicht: (2025)
Gradient-Guided Parameter Mask for Multi-Scenario Image Restoration Under Adverse Weather
von: Guo, Jilong, et al.
Veröffentlicht: (2024)
von: Guo, Jilong, et al.
Veröffentlicht: (2024)
Mutual Information Guided Optimal Transport for Unsupervised Visible-Infrared Person Re-identification
von: Zhang, Zhizhong, et al.
Veröffentlicht: (2024)
von: Zhang, Zhizhong, et al.
Veröffentlicht: (2024)
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
ParetoSlider: Diffusion Models Post-Training for Continuous Reward Control
von: Golan, Shelly, et al.
Veröffentlicht: (2026)
von: Golan, Shelly, et al.
Veröffentlicht: (2026)
ExpAlign: Expectation-Guided Vision-Language Alignment for Open-Vocabulary Grounding
von: Hu, Junyi, et al.
Veröffentlicht: (2026)
von: Hu, Junyi, et al.
Veröffentlicht: (2026)
SafeGRPO: Self-Rewarded Multimodal Safety Alignment via Rule-Governed Policy Optimization
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
Improving Vision-language Models with Perception-centric Process Reward Models
von: Min, Yingqian, et al.
Veröffentlicht: (2026)
von: Min, Yingqian, et al.
Veröffentlicht: (2026)
Multi-Reward as Condition for Instruction-based Image Editing
von: Gu, Xin, et al.
Veröffentlicht: (2024)
von: Gu, Xin, et al.
Veröffentlicht: (2024)
PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models
von: Li, Yifan, et al.
Veröffentlicht: (2024)
von: Li, Yifan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025) -
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025) -
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)