Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Seungwook, Cho, Minsu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Harnessing the Power of Training-Free Techniques in Text-to-2D Generation for Text-to-3D Generation via Score Distillation Sampling
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
3D Geometric Shape Assembly via Efficient Point Cloud Matching
von: Lee, Nahyuk, et al.
Veröffentlicht: (2024)
von: Lee, Nahyuk, et al.
Veröffentlicht: (2024)
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
FreeAction: Training-Free Techniques for Enhanced Fidelity of Trajectory-to-Video Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
IRIS: Intrinsic Reward Image Synthesis
von: Chen, Yihang, et al.
Veröffentlicht: (2025)
von: Chen, Yihang, et al.
Veröffentlicht: (2025)
Generic Event Boundary Detection via Denoising Diffusion
von: Hwang, Jaejun, et al.
Veröffentlicht: (2025)
von: Hwang, Jaejun, et al.
Veröffentlicht: (2025)
Personalized Reward Modeling for Text-to-Image Generation
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
RapidMV: Leveraging Spatio-Angular Representations for Efficient and Consistent Text-to-Multi-View Synthesis
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
Maestro: Self-Improving Text-to-Image Generation via Agent Orchestration
von: Wan, Xingchen, et al.
Veröffentlicht: (2025)
von: Wan, Xingchen, et al.
Veröffentlicht: (2025)
Learning SO(3)-Invariant Semantic Correspondence via Local Shape Transform
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
Self-Corrected Image Generation with Explainable Latent Rewards
von: Luo, Yinyi, et al.
Veröffentlicht: (2026)
von: Luo, Yinyi, et al.
Veröffentlicht: (2026)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
DiffExp: Efficient Exploration in Reward Fine-tuning for Text-to-Image Diffusion Models
von: Chae, Daewon, et al.
Veröffentlicht: (2025)
von: Chae, Daewon, et al.
Veröffentlicht: (2025)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning
von: Guo, Xiaojun, et al.
Veröffentlicht: (2025)
von: Guo, Xiaojun, et al.
Veröffentlicht: (2025)
Few-Shot Pattern Detection via Template Matching and Regression
von: Jo, Eunchan, et al.
Veröffentlicht: (2025)
von: Jo, Eunchan, et al.
Veröffentlicht: (2025)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
Image Clustering Conditioned on Text Criteria
von: Kwon, Sehyun, et al.
Veröffentlicht: (2023)
von: Kwon, Sehyun, et al.
Veröffentlicht: (2023)
Addressing Diverging Training Costs using BEVRestore for High-resolution Bird's Eye View Map Construction
von: Kim, Minsu, et al.
Veröffentlicht: (2024)
von: Kim, Minsu, et al.
Veröffentlicht: (2024)
RewardFlow: Generate Images by Optimizing What You Reward
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
von: Wu, Xiaoshi, et al.
Veröffentlicht: (2024)
von: Wu, Xiaoshi, et al.
Veröffentlicht: (2024)
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
Clustering-based Image-Text Graph Matching for Domain Generalization
von: Park, Nokyung, et al.
Veröffentlicht: (2023)
von: Park, Nokyung, et al.
Veröffentlicht: (2023)
Elucidating Optimal Reward-Diversity Tradeoffs in Text-to-Image Diffusion Models
von: Jena, Rohit, et al.
Veröffentlicht: (2024)
von: Jena, Rohit, et al.
Veröffentlicht: (2024)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
Improving Physical Object State Representation in Text-to-Image Generative Systems
von: Chen, Tianle, et al.
Veröffentlicht: (2025)
von: Chen, Tianle, et al.
Veröffentlicht: (2025)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
Reflexive Guidance: Improving OoDD in Vision-Language Models via Self-Guided Image-Adaptive Concept Generation
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
Exploring Intrinsic Properties of Medical Images for Self-Supervised Binary Semantic Segmentation
von: Singh, Pranav, et al.
Veröffentlicht: (2024)
von: Singh, Pranav, et al.
Veröffentlicht: (2024)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
von: Chae, Daewon, et al.
Veröffentlicht: (2023)
von: Chae, Daewon, et al.
Veröffentlicht: (2023)
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
Combinative Matching for Geometric Shape Assembly
von: Lee, Nahyuk, et al.
Veröffentlicht: (2025)
von: Lee, Nahyuk, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Harnessing the Power of Training-Free Techniques in Text-to-2D Generation for Text-to-3D Generation via Score Distillation Sampling
von: Lee, Junhong, et al.
Veröffentlicht: (2025) -
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
von: Kim, Seungwook, et al.
Veröffentlicht: (2025) -
3D Geometric Shape Assembly via Efficient Point Cloud Matching
von: Lee, Nahyuk, et al.
Veröffentlicht: (2024) -
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2024) -
FreeAction: Training-Free Techniques for Enhanced Fidelity of Trajectory-to-Video Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)