Prompt Refinement with Image Pivot for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhan, Jingtao, Ai, Qingyao, Liu, Yiqun, Pan, Yingwei, Yao, Ting, Mao, Jiaxin, Ma, Shaoping, Mei, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Capability-aware Prompt Reformulation Learning for Text-to-Image Generation
von: Zhan, Jingtao, et al.
Veröffentlicht: (2024)
von: Zhan, Jingtao, et al.
Veröffentlicht: (2024)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
FreeInpaint: Tuning-free Prompt Alignment and Visual Rationality Enhancement in Image Inpainting
von: Gong, Chao, et al.
Veröffentlicht: (2025)
von: Gong, Chao, et al.
Veröffentlicht: (2025)
Evaluating Intelligence via Trial and Error
von: Zhan, Jingtao, et al.
Veröffentlicht: (2025)
von: Zhan, Jingtao, et al.
Veröffentlicht: (2025)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
Creatively Upscaling Images with Global-Regional Priors
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
MotionPro: A Precise Motion Controller for Image-to-Video Generation
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
Visual Autoregressive Modeling for Instruction-Guided Image Editing
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
TRIP: Temporal Residual Learning with Image Noise Prior for Image-to-Video Diffusion Models
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2024)
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024)
von: Yao, Ting, et al.
Veröffentlicht: (2024)
DreamVAR: Taming Reinforced Visual Autoregressive Model for High-Fidelity Subject-Driven Image Generation
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
DreamJourney: Perpetual View Generation with Video Diffusion Models
von: Pan, Bo, et al.
Veröffentlicht: (2025)
von: Pan, Bo, et al.
Veröffentlicht: (2025)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
Iterative Prompt Refinement for Safer Text-to-Image Generation
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
Improving Text-guided Object Inpainting with Semantic Pre-inpainting
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
G-Refine: A General Quality Refiner for Text-to-Image Generation
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
EDITOR: Effective and Interpretable Prompt Inversion for Text-to-Image Diffusion Models
von: Li, Mingzhe, et al.
Veröffentlicht: (2025)
von: Li, Mingzhe, et al.
Veröffentlicht: (2025)
IPVTON: Image-based 3D Virtual Try-on with Image Prompt Adapter
von: Zhong, Xiaojing, et al.
Veröffentlicht: (2025)
von: Zhong, Xiaojing, et al.
Veröffentlicht: (2025)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
von: Jing, Zonglei, et al.
Veröffentlicht: (2025)
von: Jing, Zonglei, et al.
Veröffentlicht: (2025)
Learning from Mistakes: Iterative Prompt Relabeling for Text-to-Image Diffusion Model Training
von: Chen, Xinyan, et al.
Veröffentlicht: (2023)
von: Chen, Xinyan, et al.
Veröffentlicht: (2023)
Training-Free Text-to-Image Compositional Food Generation via Prompt Grafting
von: Pan, Xinyue, et al.
Veröffentlicht: (2026)
von: Pan, Xinyue, et al.
Veröffentlicht: (2026)
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
von: Feng, Chun-Mei, et al.
Veröffentlicht: (2025)
von: Feng, Chun-Mei, et al.
Veröffentlicht: (2025)
Tailored Visions: Enhancing Text-to-Image Generation with Personalized Prompt Rewriting
von: Chen, Zijie, et al.
Veröffentlicht: (2023)
von: Chen, Zijie, et al.
Veröffentlicht: (2023)
Boosting Diffusion Models with Moving Average Sampling in Frequency Domain
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
Improving Virtual Try-On with Garment-focused Diffusion Models
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
RPiAE: A Representation-Pivoted Autoencoder Enhancing Both Image Generation and Editing
von: Gong, Yue, et al.
Veröffentlicht: (2026)
von: Gong, Yue, et al.
Veröffentlicht: (2026)
Unified Prompt Attack Against Text-to-Image Generation Models
von: Peng, Duo, et al.
Veröffentlicht: (2025)
von: Peng, Duo, et al.
Veröffentlicht: (2025)
ProGEO: Generating Prompts through Image-Text Contrastive Learning for Visual Geo-localization
von: Mao, Chen, et al.
Veröffentlicht: (2024)
von: Mao, Chen, et al.
Veröffentlicht: (2024)
Optimizing Prompts for Text-to-Image Generation
von: Hao, Yaru, et al.
Veröffentlicht: (2022)
von: Hao, Yaru, et al.
Veröffentlicht: (2022)
Setting the Stage: Text-Driven Scene-Consistent Image Generation
von: Xie, Cong, et al.
Veröffentlicht: (2025)
von: Xie, Cong, et al.
Veröffentlicht: (2025)
Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image
von: Zhao, Qi, et al.
Veröffentlicht: (2025)
von: Zhao, Qi, et al.
Veröffentlicht: (2025)
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Capability-aware Prompt Reformulation Learning for Text-to-Image Generation
von: Zhan, Jingtao, et al.
Veröffentlicht: (2024) -
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024) -
FreeInpaint: Tuning-free Prompt Alignment and Visual Rationality Enhancement in Image Inpainting
von: Gong, Chao, et al.
Veröffentlicht: (2025) -
Evaluating Intelligence via Trial and Error
von: Zhan, Jingtao, et al.
Veröffentlicht: (2025) -
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)