Initialization is Half the Battle: Generating Diverse Images from a Guidance Potential Posterior
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Xiang, Liu, Dianbo, Kawaguchi, Kenji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models
von: Li, Xiang, et al.
Veröffentlicht: (2023)
von: Li, Xiang, et al.
Veröffentlicht: (2023)
RSGen: Enhancing Layout-Driven Remote Sensing Image Generation with Diverse Edge Guidance
von: Hou, Xianbao, et al.
Veröffentlicht: (2026)
von: Hou, Xianbao, et al.
Veröffentlicht: (2026)
Spatial-Aware Latent Initialization for Controllable Image Generation
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
DiverseGRPO: Mitigating Mode Collapse in Image Generation via Diversity-Aware GRPO
von: Liu, Henglin, et al.
Veröffentlicht: (2025)
von: Liu, Henglin, et al.
Veröffentlicht: (2025)
Training-free Guidance in Text-to-Video Generation via Multimodal Planning and Structured Noise Initialization
von: Li, Jialu, et al.
Veröffentlicht: (2025)
von: Li, Jialu, et al.
Veröffentlicht: (2025)
Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance
von: Wu, Song, et al.
Veröffentlicht: (2026)
von: Wu, Song, et al.
Veröffentlicht: (2026)
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models
von: Saini, Harshvardhan, et al.
Veröffentlicht: (2026)
von: Saini, Harshvardhan, et al.
Veröffentlicht: (2026)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation
von: Luo, Yan, et al.
Veröffentlicht: (2026)
von: Luo, Yan, et al.
Veröffentlicht: (2026)
Guidance Matters: Rethinking the Evaluation Pitfall for Text-to-Image Generation
von: Xie, Dian, et al.
Veröffentlicht: (2026)
von: Xie, Dian, et al.
Veröffentlicht: (2026)
Bringing Diversity from Diffusion Models to Semantic-Guided Face Asset Generation
von: Cai, Yunxuan, et al.
Veröffentlicht: (2025)
von: Cai, Yunxuan, et al.
Veröffentlicht: (2025)
Hybrid Global-Local Representation with Augmented Spatial Guidance for Zero-Shot Referring Image Segmentation
von: Liu, Ting, et al.
Veröffentlicht: (2025)
von: Liu, Ting, et al.
Veröffentlicht: (2025)
PostEdit: Posterior Sampling for Efficient Zero-Shot Image Editing
von: Tian, Feng, et al.
Veröffentlicht: (2024)
von: Tian, Feng, et al.
Veröffentlicht: (2024)
Evolve to Inspire: Novelty Search for Diverse Image Generation
von: Inch, Alex, et al.
Veröffentlicht: (2025)
von: Inch, Alex, et al.
Veröffentlicht: (2025)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
von: Trang, Bailey, et al.
Veröffentlicht: (2025)
von: Trang, Bailey, et al.
Veröffentlicht: (2025)
BrainDreamer: Reasoning-Coherent and Controllable Image Generation from EEG Brain Signals via Language Guidance
von: Wang, Ling, et al.
Veröffentlicht: (2024)
von: Wang, Ling, et al.
Veröffentlicht: (2024)
Dual-Channel Attention Guidance for Training-Free Image Editing Control in Diffusion Transformers
von: Li, Guandong
Veröffentlicht: (2026)
von: Li, Guandong
Veröffentlicht: (2026)
ClothHMR: 3D Mesh Recovery of Humans in Diverse Clothing from Single Image
von: Gao, Yunqi, et al.
Veröffentlicht: (2025)
von: Gao, Yunqi, et al.
Veröffentlicht: (2025)
CannyEdit: Selective Canny Control and Dual-Prompt Guidance for Training-Free Image Editing
von: Xie, Weiyan, et al.
Veröffentlicht: (2025)
von: Xie, Weiyan, et al.
Veröffentlicht: (2025)
Semantic Guidance Tuning for Text-To-Image Diffusion Models
von: Kang, Hyun, et al.
Veröffentlicht: (2023)
von: Kang, Hyun, et al.
Veröffentlicht: (2023)
MCIE: Multimodal LLM-Driven Complex Instruction Image Editing with Spatial Guidance
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
ConVQG: Contrastive Visual Question Generation with Multimodal Guidance
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
Laplacian Score Sharpening for Mitigating Hallucination in Diffusion Models
von: C, Barath Chandran., et al.
Veröffentlicht: (2025)
von: C, Barath Chandran., et al.
Veröffentlicht: (2025)
LD-RPS: Zero-Shot Unified Image Restoration via Latent Diffusion Recurrent Posterior Sampling
von: Li, Huaqiu, et al.
Veröffentlicht: (2025)
von: Li, Huaqiu, et al.
Veröffentlicht: (2025)
Resource-Efficient Motion Control for Video Generation via Dynamic Mask Guidance
von: Feng, Sicong, et al.
Veröffentlicht: (2025)
von: Feng, Sicong, et al.
Veröffentlicht: (2025)
HMAFlow: Learning More Accurate Optical Flow via Hierarchical Motion Field Alignment
von: Ma, Dianbo, et al.
Veröffentlicht: (2024)
von: Ma, Dianbo, et al.
Veröffentlicht: (2024)
Fostering Video Reasoning via Next-Event Prediction
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
Geometry-Correct Diffusion Posterior Sampling with Denoiser-Pullback Curvature Guidance and Manifold-Aligned Damping
von: Shin, Seunghyeok, et al.
Veröffentlicht: (2026)
von: Shin, Seunghyeok, et al.
Veröffentlicht: (2026)
Make It Efficient: Dynamic Sparse Attention for Autoregressive Image Generation
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
von: Lingenberg, Tobias, et al.
Veröffentlicht: (2024)
von: Lingenberg, Tobias, et al.
Veröffentlicht: (2024)
Designing and Generating Diverse, Equitable Face Image Datasets for Face Verification Tasks
von: Baltsou, Georgia, et al.
Veröffentlicht: (2025)
von: Baltsou, Georgia, et al.
Veröffentlicht: (2025)
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
von: Molino, Daniele, et al.
Veröffentlicht: (2026)
von: Molino, Daniele, et al.
Veröffentlicht: (2026)
FastInit: Fast Noise Initialization for Temporally Consistent Video Generation
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
Curriculum Group Policy Optimization: Adaptive Sampling for Unleashing the Potential of Text-to-Image Generation
von: Li, Baoteng, et al.
Veröffentlicht: (2026)
von: Li, Baoteng, et al.
Veröffentlicht: (2026)
Reflexive Guidance: Improving OoDD in Vision-Language Models via Self-Guided Image-Adaptive Concept Generation
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
Stabilizing Diffusion Posterior Sampling by Noise--Frequency Continuation
von: Tian, Feng, et al.
Veröffentlicht: (2026)
von: Tian, Feng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models
von: Li, Xiang, et al.
Veröffentlicht: (2023) -
RSGen: Enhancing Layout-Driven Remote Sensing Image Generation with Diverse Edge Guidance
von: Hou, Xianbao, et al.
Veröffentlicht: (2026) -
Spatial-Aware Latent Initialization for Controllable Image Generation
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024) -
DiverseGRPO: Mitigating Mode Collapse in Image Generation via Diversity-Aware GRPO
von: Liu, Henglin, et al.
Veröffentlicht: (2025) -
Training-free Guidance in Text-to-Video Generation via Multimodal Planning and Structured Noise Initialization
von: Li, Jialu, et al.
Veröffentlicht: (2025)