RFSR: Improving ISR Diffusion Models via Reward Feedback Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Xiaopeng, Lin, Qinwei, Gao, Yu, Zhong, Yujie, Feng, Chengjian, Li, Dengjie, Zhao, Zheng, Hu, Jie, Ma, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TASR: Timestep-Aware Diffusion Model for Image Super-Resolution
von: Lin, Qinwei, et al.
Veröffentlicht: (2024)
von: Lin, Qinwei, et al.
Veröffentlicht: (2024)
RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case
von: Xiao, Baihui, et al.
Veröffentlicht: (2025)
von: Xiao, Baihui, et al.
Veröffentlicht: (2025)
UniMD: Towards Unifying Moment Retrieval and Temporal Action Detection
von: Zeng, Yingsen, et al.
Veröffentlicht: (2024)
von: Zeng, Yingsen, et al.
Veröffentlicht: (2024)
InstaGen: Enhancing Object Detection by Training on Synthetic Dataset
von: Feng, Chengjian, et al.
Veröffentlicht: (2024)
von: Feng, Chengjian, et al.
Veröffentlicht: (2024)
AP-CAP: Advancing High-Quality Data Synthesis for Animal Pose Estimation via a Controllable Image Generation Pipeline
von: Wang, Lei, et al.
Veröffentlicht: (2025)
von: Wang, Lei, et al.
Veröffentlicht: (2025)
Manga Generation via Layout-controllable Diffusion
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
Matten: Video Generation with Mamba-Attention
von: Gao, Yu, et al.
Veröffentlicht: (2024)
von: Gao, Yu, et al.
Veröffentlicht: (2024)
DisTime: Distribution-based Time Representation for Video Large Language Models
von: Zeng, Yingsen, et al.
Veröffentlicht: (2025)
von: Zeng, Yingsen, et al.
Veröffentlicht: (2025)
LinVT: Empower Your Image-level Large Language Model to Understand Videos
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
von: Huang, Zhijian, et al.
Veröffentlicht: (2024)
von: Huang, Zhijian, et al.
Veröffentlicht: (2024)
InstructVEdit: A Holistic Approach for Instructional Video Editing
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
DiffusionReward: Enhancing Blind Face Restoration through Reward Feedback Learning
von: Wu, Bin, et al.
Veröffentlicht: (2025)
von: Wu, Bin, et al.
Veröffentlicht: (2025)
Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis
von: Lin, Tianwei, et al.
Veröffentlicht: (2026)
von: Lin, Tianwei, et al.
Veröffentlicht: (2026)
Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
MRStyle: A Unified Framework for Color Style Transfer with Multi-Modality Reference
von: Huang, Jiancheng, et al.
Veröffentlicht: (2024)
von: Huang, Jiancheng, et al.
Veröffentlicht: (2024)
RoboUniView: Visual-Language Model with Unified View Representation for Robotic Manipulation
von: Liu, Fanfan, et al.
Veröffentlicht: (2024)
von: Liu, Fanfan, et al.
Veröffentlicht: (2024)
Advancing Visual Large Language Model for Multi-granular Versatile Perception
von: Xiang, Wentao, et al.
Veröffentlicht: (2025)
von: Xiang, Wentao, et al.
Veröffentlicht: (2025)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
von: Ge, Wenhang, et al.
Veröffentlicht: (2026)
von: Ge, Wenhang, et al.
Veröffentlicht: (2026)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
Monocular Gaussian SLAM with Language Extended Loop Closure
von: Lan, Tian, et al.
Veröffentlicht: (2024)
von: Lan, Tian, et al.
Veröffentlicht: (2024)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
UniFL: Improve Latent Diffusion Model via Unified Feedback Learning
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
High-quality Image Dehazing with Diffusion Model
von: Yu, Hu, et al.
Veröffentlicht: (2023)
von: Yu, Hu, et al.
Veröffentlicht: (2023)
HiMix: Reducing Computational Complexity in Large Vision-Language Models
von: Zhang, Xuange, et al.
Veröffentlicht: (2025)
von: Zhang, Xuange, et al.
Veröffentlicht: (2025)
LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories
von: Liang, Zhanhao, et al.
Veröffentlicht: (2026)
von: Liang, Zhanhao, et al.
Veröffentlicht: (2026)
DIffSteISR: Harnessing Diffusion Prior for Superior Real-world Stereo Image Super-Resolution
von: Zhou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Zhou, Yuanbo, et al.
Veröffentlicht: (2024)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
von: Huang, Tao, et al.
Veröffentlicht: (2023)
von: Huang, Tao, et al.
Veröffentlicht: (2023)
Gramformer: Learning Crowd Counting via Graph-Modulated Transformer
von: Lin, Hui, et al.
Veröffentlicht: (2024)
von: Lin, Hui, et al.
Veröffentlicht: (2024)
MindBench: A Comprehensive Benchmark for Mind Map Structure Recognition and Analysis
von: Chen, Lei, et al.
Veröffentlicht: (2024)
von: Chen, Lei, et al.
Veröffentlicht: (2024)
X-SAM: From Segment Anything to Any Segmentation
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
Beyond Reward Margin: Rethinking and Resolving Likelihood Displacement in Diffusion Models via Video Generation
von: Xu, Ruojun, et al.
Veröffentlicht: (2025)
von: Xu, Ruojun, et al.
Veröffentlicht: (2025)
ParaUni: Enhance Generation in Unified Multimodal Model with Reinforcement-driven Hierarchical Parallel Information Interaction
von: Tan, Jiangtong, et al.
Veröffentlicht: (2025)
von: Tan, Jiangtong, et al.
Veröffentlicht: (2025)
Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward
von: Xu, Xiaogang, et al.
Veröffentlicht: (2025)
von: Xu, Xiaogang, et al.
Veröffentlicht: (2025)
Personalized Cross-Modal Emotional Correlation Learning for Speech-Preserving Facial Expression Manipulation
von: Chen, Tianshui, et al.
Veröffentlicht: (2026)
von: Chen, Tianshui, et al.
Veröffentlicht: (2026)
ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO
von: Ahn, Daechul, et al.
Veröffentlicht: (2024)
von: Ahn, Daechul, et al.
Veröffentlicht: (2024)
ProxT2I: Efficient Reward-Guided Text-to-Image Generation via Proximal Diffusion
von: Fang, Zhenghan, et al.
Veröffentlicht: (2025)
von: Fang, Zhenghan, et al.
Veröffentlicht: (2025)
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
Diffusion-Classifier Synergy: Reward-Aligned Learning via Mutual Boosting Loop for FSCIL
von: Wu, Ruitao, et al.
Veröffentlicht: (2025)
von: Wu, Ruitao, et al.
Veröffentlicht: (2025)
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
von: Liu, Gongye, et al.
Veröffentlicht: (2026)
von: Liu, Gongye, et al.
Veröffentlicht: (2026)
HyperSeg: Towards Universal Visual Segmentation with Large Language Model
von: Wei, Cong, et al.
Veröffentlicht: (2024)
von: Wei, Cong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TASR: Timestep-Aware Diffusion Model for Image Super-Resolution
von: Lin, Qinwei, et al.
Veröffentlicht: (2024) -
RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case
von: Xiao, Baihui, et al.
Veröffentlicht: (2025) -
UniMD: Towards Unifying Moment Retrieval and Temporal Action Detection
von: Zeng, Yingsen, et al.
Veröffentlicht: (2024) -
InstaGen: Enhancing Object Detection by Training on Synthetic Dataset
von: Feng, Chengjian, et al.
Veröffentlicht: (2024) -
AP-CAP: Advancing High-Quality Data Synthesis for Animal Pose Estimation via a Controllable Image Generation Pipeline
von: Wang, Lei, et al.
Veröffentlicht: (2025)