Fake it till You Make it: Reward Modeling as Discriminative Prediction
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Runtao, Zhan, Jiahao, He, Yingqing, Wei, Chen, Yuille, Alan, Chen, Qifeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
por: Liu, Runtao, et al.
Publicado: (2024)
por: Liu, Runtao, et al.
Publicado: (2024)
ModelGrow: Continual Text-to-Video Pre-training with Model Expansion and Language Understanding Enhancement
por: Rao, Zhefan, et al.
Publicado: (2024)
por: Rao, Zhefan, et al.
Publicado: (2024)
Latent Guard: a Safety Framework for Text-to-image Generation
por: Liu, Runtao, et al.
Publicado: (2024)
por: Liu, Runtao, et al.
Publicado: (2024)
LongVideoAgent: Multi-Agent Reasoning with Long Videos
por: Liu, Runtao, et al.
Publicado: (2025)
por: Liu, Runtao, et al.
Publicado: (2025)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
por: Liu, Runtao, et al.
Publicado: (2024)
por: Liu, Runtao, et al.
Publicado: (2024)
Robust-R1: Degradation-Aware Reasoning for Robust Visual Understanding
por: Tang, Jiaqi, et al.
Publicado: (2025)
por: Tang, Jiaqi, et al.
Publicado: (2025)
Leveraging AI Predicted and Expert Revised Annotations in Interactive Segmentation: Continual Tuning or Full Training?
por: Zhang, Tiezheng, et al.
Publicado: (2024)
por: Zhang, Tiezheng, et al.
Publicado: (2024)
Dynamic Analysis and Adaptive Discriminator for Fake News Detection
por: Su, Xinqi, et al.
Publicado: (2024)
por: Su, Xinqi, et al.
Publicado: (2024)
Fake It till You Make It: Curricular Dynamic Forgery Augmentations towards General Deepfake Detection
por: Lin, Yuzhen, et al.
Publicado: (2024)
por: Lin, Yuzhen, et al.
Publicado: (2024)
LLMs Meet Multimodal Generation and Editing: A Survey
por: He, Yingqing, et al.
Publicado: (2024)
por: He, Yingqing, et al.
Publicado: (2024)
Shared LoRA Subspaces for almost Strict Continual Learning
por: Kaushik, Prakhar, et al.
Publicado: (2026)
por: Kaushik, Prakhar, et al.
Publicado: (2026)
The Universal Weight Subspace Hypothesis
por: Kaushik, Prakhar, et al.
Publicado: (2025)
por: Kaushik, Prakhar, et al.
Publicado: (2025)
SpA2V: Harnessing Spatial Auditory Cues for Audio-driven Spatially-aware Video Generation
por: Pham, Kien T., et al.
Publicado: (2025)
por: Pham, Kien T., et al.
Publicado: (2025)
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation
por: Zhang, Zihui, et al.
Publicado: (2026)
por: Zhang, Zihui, et al.
Publicado: (2026)
Ancestral Mamba: Enhancing Selective Discriminant Space Model with Online Visual Prototype Learning for Efficient and Robust Discriminant Approach
por: Qin, Jiahao, et al.
Publicado: (2025)
por: Qin, Jiahao, et al.
Publicado: (2025)
A Bayesian Approach to OOD Robustness in Image Classification
por: Kaushik, Prakhar, et al.
Publicado: (2024)
por: Kaushik, Prakhar, et al.
Publicado: (2024)
Differences That Matter: Auditing Models for Capability Gap Discovery and Rectification
por: Liu, Qihao, et al.
Publicado: (2025)
por: Liu, Qihao, et al.
Publicado: (2025)
GENFIG1: Visual Summaries of Scholarly Work as a Challenge for Vision-Language Models
por: Guan, Yaohan, et al.
Publicado: (2026)
por: Guan, Yaohan, et al.
Publicado: (2026)
Fake It Till You Make It: Using Synthetic Data and Domain Knowledge for Improved Text-Based Learning for LGE Detection
por: Jacob, Athira J, et al.
Publicado: (2025)
por: Jacob, Athira J, et al.
Publicado: (2025)
Zoom and Shift are All You Need
por: Qin, Jiahao
Publicado: (2024)
por: Qin, Jiahao
Publicado: (2024)
GARDO: Reinforcing Diffusion Models without Reward Hacking
por: He, Haoran, et al.
Publicado: (2025)
por: He, Haoran, et al.
Publicado: (2025)
MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation
por: Xie, Jiahao, et al.
Publicado: (2023)
por: Xie, Jiahao, et al.
Publicado: (2023)
Open Your Eyes: Vision Enhances Message Passing Neural Networks in Link Prediction
por: Wei, Yanbin, et al.
Publicado: (2025)
por: Wei, Yanbin, et al.
Publicado: (2025)
Can You Count to Nine? A Human Evaluation Benchmark for Counting Limits in Modern Text-to-Video Models
por: Guo, Xuyang, et al.
Publicado: (2025)
por: Guo, Xuyang, et al.
Publicado: (2025)
RewardFlow: Generate Images by Optimizing What You Reward
por: Susladkar, Onkar, et al.
Publicado: (2026)
por: Susladkar, Onkar, et al.
Publicado: (2026)
Spatial457: A Diagnostic Benchmark for 6D Spatial Reasoning of Large Multimodal Models
por: Wang, Xingrui, et al.
Publicado: (2025)
por: Wang, Xingrui, et al.
Publicado: (2025)
The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
por: Ai, Wei, et al.
Publicado: (2026)
por: Ai, Wei, et al.
Publicado: (2026)
Towards High-Order Mean Flow Generative Models: Feasibility, Expressivity, and Provably Efficient Criteria
por: Cao, Yang, et al.
Publicado: (2025)
por: Cao, Yang, et al.
Publicado: (2025)
Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose Estimation
por: Kaushik, Prakhar, et al.
Publicado: (2024)
por: Kaushik, Prakhar, et al.
Publicado: (2024)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
por: Chin, Zhi-Yi, et al.
Publicado: (2023)
por: Chin, Zhi-Yi, et al.
Publicado: (2023)
Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering
por: Wang, Xingrui, et al.
Publicado: (2024)
por: Wang, Xingrui, et al.
Publicado: (2024)
Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model
por: Yang, Kai, et al.
Publicado: (2023)
por: Yang, Kai, et al.
Publicado: (2023)
FuRL: Visual-Language Models as Fuzzy Rewards for Reinforcement Learning
por: Fu, Yuwei, et al.
Publicado: (2024)
por: Fu, Yuwei, et al.
Publicado: (2024)
Intention-aware Denoising Diffusion Model for Trajectory Prediction
por: Liu, Chen, et al.
Publicado: (2024)
por: Liu, Chen, et al.
Publicado: (2024)
Online Reward-Weighted Fine-Tuning of Flow Matching with Wasserstein Regularization
por: Fan, Jiajun, et al.
Publicado: (2025)
por: Fan, Jiajun, et al.
Publicado: (2025)
Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models
por: Lee, Jeongjae, et al.
Publicado: (2026)
por: Lee, Jeongjae, et al.
Publicado: (2026)
Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction
por: Liu, Yu, et al.
Publicado: (2025)
por: Liu, Yu, et al.
Publicado: (2025)
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers
por: Yang, Chenglin, et al.
Publicado: (2023)
por: Yang, Chenglin, et al.
Publicado: (2023)
Perception-R1: Advancing Multimodal Reasoning Capabilities of MLLMs via Visual Perception Reward
por: Xiao, Tong, et al.
Publicado: (2025)
por: Xiao, Tong, et al.
Publicado: (2025)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
por: Zhang, Zhihong, et al.
Publicado: (2025)
por: Zhang, Zhihong, et al.
Publicado: (2025)
Ejemplares similares
-
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
por: Liu, Runtao, et al.
Publicado: (2024) -
ModelGrow: Continual Text-to-Video Pre-training with Model Expansion and Language Understanding Enhancement
por: Rao, Zhefan, et al.
Publicado: (2024) -
Latent Guard: a Safety Framework for Text-to-image Generation
por: Liu, Runtao, et al.
Publicado: (2024) -
LongVideoAgent: Multi-Agent Reasoning with Long Videos
por: Liu, Runtao, et al.
Publicado: (2025) -
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
por: Liu, Runtao, et al.
Publicado: (2024)