Self-Corrected Image Generation with Explainable Latent Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Yinyi, Gokhale, Hrishikesh, Savvides, Marios, Wang, Jindong, He, Shengfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LatentUMM: Dual Latent Alignment for Unified Multimodal Models
von: Luo, Yinyi, et al.
Veröffentlicht: (2026)
von: Luo, Yinyi, et al.
Veröffentlicht: (2026)
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
von: Chen, Hao, et al.
Veröffentlicht: (2022)
von: Chen, Hao, et al.
Veröffentlicht: (2022)
STELAR-VISION: Self-Topology-Aware Efficient Learning for Aligned Reasoning in Vision
von: Li, Chen, et al.
Veröffentlicht: (2025)
von: Li, Chen, et al.
Veröffentlicht: (2025)
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
von: Chen, Hao, et al.
Veröffentlicht: (2022)
von: Chen, Hao, et al.
Veröffentlicht: (2022)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
von: Liu, Gongye, et al.
Veröffentlicht: (2026)
von: Liu, Gongye, et al.
Veröffentlicht: (2026)
Reward Guided Latent Consistency Distillation
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
CCCaption: Dual-Reward Reinforcement Learning for Complete and Correct Image Captioning
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
RewardFlow: Generate Images by Optimizing What You Reward
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing
von: Wang, Weixing, et al.
Veröffentlicht: (2025)
von: Wang, Weixing, et al.
Veröffentlicht: (2025)
Personalized Reward Modeling for Text-to-Image Generation
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
Image Tokenizer Needs Post-Training
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
Low-light Pedestrian Detection in Visible and Infrared Image Feeds: Issues and Challenges
von: Akilan, Thangarajah, et al.
Veröffentlicht: (2023)
von: Akilan, Thangarajah, et al.
Veröffentlicht: (2023)
SOLAR: Scalable Optimization of Large-scale Architecture for Reasoning
von: Li, Chen, et al.
Veröffentlicht: (2025)
von: Li, Chen, et al.
Veröffentlicht: (2025)
Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
LaRE$^2$: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection
von: Luo, Yunpeng, et al.
Veröffentlicht: (2024)
von: Luo, Yunpeng, et al.
Veröffentlicht: (2024)
A Survey on Responsible Generative AI: What to Generate and What Not
von: Gu, Jindong
Veröffentlicht: (2024)
von: Gu, Jindong
Veröffentlicht: (2024)
OT-ALD: Aligning Latent Distributions with Optimal Transport for Accelerated Image-to-Image Translation
von: Wang, Zhanpeng, et al.
Veröffentlicht: (2025)
von: Wang, Zhanpeng, et al.
Veröffentlicht: (2025)
Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation
von: Chen, Wenting, et al.
Veröffentlicht: (2023)
von: Chen, Wenting, et al.
Veröffentlicht: (2023)
Latent Guard: a Safety Framework for Text-to-image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
Latent Modulated Function for Computational Optimal Continuous Image Representation
von: He, Zongyao, et al.
Veröffentlicht: (2024)
von: He, Zongyao, et al.
Veröffentlicht: (2024)
MonoCon: A general framework for learning ultra-compact high-fidelity representations using monotonicity constraints
von: Gokhale, Shreyas
Veröffentlicht: (2025)
von: Gokhale, Shreyas
Veröffentlicht: (2025)
Self-Explainable Affordance Learning with Embodied Caption
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2024)
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2024)
Spatial-Aware Latent Initialization for Controllable Image Generation
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
Latent Expression Generation for Referring Image Segmentation and Grounding
von: Yu, Seonghoon, et al.
Veröffentlicht: (2025)
von: Yu, Seonghoon, et al.
Veröffentlicht: (2025)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
von: Wei, Yibing, et al.
Veröffentlicht: (2024)
von: Wei, Yibing, et al.
Veröffentlicht: (2024)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
von: Trang, Bailey, et al.
Veröffentlicht: (2025)
von: Trang, Bailey, et al.
Veröffentlicht: (2025)
CAR: Contrast-Agnostic Deformable Medical Image Registration with Contrast-Invariant Latent Regularization
von: Wang, Yinsong, et al.
Veröffentlicht: (2024)
von: Wang, Yinsong, et al.
Veröffentlicht: (2024)
Adaptive Clinical-Aware Latent Diffusion for Multimodal Brain Image Generation and Missing Modality Imputation
von: Zhou, Rong, et al.
Veröffentlicht: (2026)
von: Zhou, Rong, et al.
Veröffentlicht: (2026)
OISD: On-Policy Internal Self-Distillation of Language Models
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor
von: Chen, Jiali, et al.
Veröffentlicht: (2024)
von: Chen, Jiali, et al.
Veröffentlicht: (2024)
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
MICA: Towards Explainable Skin Lesion Diagnosis via Multi-Level Image-Concept Alignment
von: Bie, Yequan, et al.
Veröffentlicht: (2024)
von: Bie, Yequan, et al.
Veröffentlicht: (2024)
Latent Action Control for Reasoning-Guided Unified Image Generation
von: Zhai, Fuxiang, et al.
Veröffentlicht: (2026)
von: Zhai, Fuxiang, et al.
Veröffentlicht: (2026)
DAVID-XR1: Detecting AI-Generated Videos with Explainable Reasoning
von: Gao, Yifeng, et al.
Veröffentlicht: (2025)
von: Gao, Yifeng, et al.
Veröffentlicht: (2025)
Self-Correcting Self-Consuming Loops for Generative Model Training
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection
von: Cao, Huangsen, et al.
Veröffentlicht: (2025)
von: Cao, Huangsen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LatentUMM: Dual Latent Alignment for Unified Multimodal Models
von: Luo, Yinyi, et al.
Veröffentlicht: (2026) -
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
von: Qiu, Kai, et al.
Veröffentlicht: (2025) -
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
von: Chen, Hao, et al.
Veröffentlicht: (2022) -
STELAR-VISION: Self-Topology-Aware Efficient Learning for Aligned Reasoning in Vision
von: Li, Chen, et al.
Veröffentlicht: (2025) -
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
von: Chen, Hao, et al.
Veröffentlicht: (2022)