Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Guiyu, Gao, Huan-ang, Jiang, Zijian, Zhao, Hao, Zheng, Zhedong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RIGI: Rectifying Image-to-3D Generation Inconsistency via Uncertainty-aware Learning
di: Wang, Jiacheng, et al.
Pubblicazione: (2024)
di: Wang, Jiacheng, et al.
Pubblicazione: (2024)
FairDiff: Fair Segmentation with Point-Image Diffusion
di: Li, Wenyi, et al.
Pubblicazione: (2024)
di: Li, Wenyi, et al.
Pubblicazione: (2024)
Diffusion-based Visual Anagram as Multi-task Learning
di: Xu, Zhiyuan, et al.
Pubblicazione: (2024)
di: Xu, Zhiyuan, et al.
Pubblicazione: (2024)
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024)
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024)
Harnessing Uncertainty-aware Bounding Boxes for Unsupervised 3D Object Detection
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024)
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024)
Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models
di: Zhang, Ruiyang, et al.
Pubblicazione: (2025)
di: Zhang, Ruiyang, et al.
Pubblicazione: (2025)
Dual-frame Fluid Motion Estimation with Test-time Optimization and Zero-divergence Loss
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
DriveCtrl: Conditioned Sim-to-Real Driving Video Generation
di: Zhao, Haonan, et al.
Pubblicazione: (2026)
di: Zhao, Haonan, et al.
Pubblicazione: (2026)
AnomalyLMM: Bridging Generative Knowledge and Discriminative Retrieval for Text-Based Person Anomaly Search
di: Ju, Hao, et al.
Pubblicazione: (2025)
di: Ju, Hao, et al.
Pubblicazione: (2025)
Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization
di: Chen, Yiyang, et al.
Pubblicazione: (2022)
di: Chen, Yiyang, et al.
Pubblicazione: (2022)
Uncertainty-Aware Trajectory Prediction: A Unified Framework Harnessing Positional and Semantic Uncertainties
di: Sun, Jintao, et al.
Pubblicazione: (2026)
di: Sun, Jintao, et al.
Pubblicazione: (2026)
Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation
di: Chen, Mu, et al.
Pubblicazione: (2023)
di: Chen, Mu, et al.
Pubblicazione: (2023)
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
di: Gao, Huan-ang, et al.
Pubblicazione: (2024)
di: Gao, Huan-ang, et al.
Pubblicazione: (2024)
Latency-aware Road Anomaly Segmentation in Videos: A Photorealistic Dataset and New Metrics
di: Tian, Beiwen, et al.
Pubblicazione: (2024)
di: Tian, Beiwen, et al.
Pubblicazione: (2024)
StepNet: Spatial-temporal Part-aware Network for Isolated Sign Language Recognition
di: Shen, Xiaolong, et al.
Pubblicazione: (2022)
di: Shen, Xiaolong, et al.
Pubblicazione: (2022)
Harnessing Weak Pair Uncertainty for Text-based Person Search
di: Sun, Jintao, et al.
Pubblicazione: (2026)
di: Sun, Jintao, et al.
Pubblicazione: (2026)
RGM: Reconstructing High-fidelity 3D Car Assets with Relightable 3D-GS Generative Model from a Single Image
di: Chen, Xiaoxue, et al.
Pubblicazione: (2024)
di: Chen, Xiaoxue, et al.
Pubblicazione: (2024)
Training-Free Model Merging for Multi-target Domain Adaptation
di: Li, Wenyi, et al.
Pubblicazione: (2024)
di: Li, Wenyi, et al.
Pubblicazione: (2024)
EmoCtrl: Controllable Emotional Image Content Generation
di: Yang, Jingyuan, et al.
Pubblicazione: (2025)
di: Yang, Jingyuan, et al.
Pubblicazione: (2025)
Scale-adaptive UAV Geo-localization via Height-aware Partition Learning
di: Chen, Quan, et al.
Pubblicazione: (2024)
di: Chen, Quan, et al.
Pubblicazione: (2024)
Collaborative Group: Composed Image Retrieval via Consensus Learning from Noisy Annotations
di: Zhang, Xu, et al.
Pubblicazione: (2023)
di: Zhang, Xu, et al.
Pubblicazione: (2023)
CameraCtrl II: Dynamic Scene Exploration via Camera-controlled Video Diffusion Models
di: He, Hao, et al.
Pubblicazione: (2025)
di: He, Hao, et al.
Pubblicazione: (2025)
P-MapNet: Far-seeing Map Generator Enhanced by both SDMap and HDMap Priors
di: Jiang, Zhou, et al.
Pubblicazione: (2024)
di: Jiang, Zhou, et al.
Pubblicazione: (2024)
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
di: Gao, Mingju, et al.
Pubblicazione: (2025)
di: Gao, Mingju, et al.
Pubblicazione: (2025)
SGOR: Outlier Removal by Leveraging Semantic and Geometric Information for Robust Point Cloud Registration
di: Zhao, Guiyu, et al.
Pubblicazione: (2024)
di: Zhao, Guiyu, et al.
Pubblicazione: (2024)
The Image as Its Own Reward: Reinforcement Learning with Adversarial Reward for Image Generation
di: Mao, Weijia, et al.
Pubblicazione: (2025)
di: Mao, Weijia, et al.
Pubblicazione: (2025)
FB-4D: Spatial-Temporal Coherent Dynamic 3D Content Generation with Feature Banks
di: Li, Jinwei, et al.
Pubblicazione: (2025)
di: Li, Jinwei, et al.
Pubblicazione: (2025)
Delving into Mapping Uncertainty for Mapless Trajectory Prediction
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
Instilling Multi-round Thinking to Text-guided Image Generation
di: Zeng, Lidong, et al.
Pubblicazione: (2024)
di: Zeng, Lidong, et al.
Pubblicazione: (2024)
SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
di: Zhang, Ruiyang, et al.
Pubblicazione: (2026)
di: Zhang, Ruiyang, et al.
Pubblicazione: (2026)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
Robust and Generalizable GNN Fine-Tuning via Uncertainty-aware Adapter Learning
di: Jiang, Bo, et al.
Pubblicazione: (2025)
di: Jiang, Bo, et al.
Pubblicazione: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
di: He, Hao, et al.
Pubblicazione: (2024)
di: He, Hao, et al.
Pubblicazione: (2024)
GrndCtrl: Grounding World Models via Self-Supervised Reward Alignment
di: He, Haoyang, et al.
Pubblicazione: (2025)
di: He, Haoyang, et al.
Pubblicazione: (2025)
Progressive Correspondence Regenerator for Robust 3D Registration
di: Zhao, Guiyu, et al.
Pubblicazione: (2025)
di: Zhao, Guiyu, et al.
Pubblicazione: (2025)
Alias-free 4D Gaussian Splatting
di: Chen, Zilong, et al.
Pubblicazione: (2025)
di: Chen, Zilong, et al.
Pubblicazione: (2025)
Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
Cross-PCR: A Robust Cross-Source Point Cloud Registration Framework
di: Zhao, Guiyu, et al.
Pubblicazione: (2024)
di: Zhao, Guiyu, et al.
Pubblicazione: (2024)
SafeCtrl: Region-Based Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
di: Zhang, Lingyun, et al.
Pubblicazione: (2025)
di: Zhang, Lingyun, et al.
Pubblicazione: (2025)
DynamiCtrl: Rethinking the Basic Structure and the Role of Text for High-quality Human Image Animation
di: Zhao, Haoyu, et al.
Pubblicazione: (2025)
di: Zhao, Haoyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
RIGI: Rectifying Image-to-3D Generation Inconsistency via Uncertainty-aware Learning
di: Wang, Jiacheng, et al.
Pubblicazione: (2024) -
FairDiff: Fair Segmentation with Point-Image Diffusion
di: Li, Wenyi, et al.
Pubblicazione: (2024) -
Diffusion-based Visual Anagram as Multi-task Learning
di: Xu, Zhiyuan, et al.
Pubblicazione: (2024) -
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024) -
Harnessing Uncertainty-aware Bounding Boxes for Unsupervised 3D Object Detection
di: Zhang, Ruiyang, et al.
Pubblicazione: (2024)