I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
Fuente:
arXiv
Saved in:
| Main Authors: | Fanelli, Nicola, Vessio, Gennaro, Castellano, Giovanna |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ArtSeek: Deep artwork understanding via multimodal in-context reasoning and late interaction retrieval
by: Fanelli, Nicola, et al.
Published: (2025)
by: Fanelli, Nicola, et al.
Published: (2025)
Art2Mus: Bridging Visual Arts and Music through Cross-Modal Generation
by: Rinaldi, Ivan, et al.
Published: (2024)
by: Rinaldi, Ivan, et al.
Published: (2024)
Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts
by: De Marinis, Pasquale, et al.
Published: (2024)
by: De Marinis, Pasquale, et al.
Published: (2024)
Take a Peek: Efficient Encoder Adaptation for Few-Shot Semantic Segmentation via LoRA
by: De Marinis, Pasquale, et al.
Published: (2025)
by: De Marinis, Pasquale, et al.
Published: (2025)
Art2Mus: Artwork-to-Music Generation via Visual Conditioning and Large-Scale Cross-Modal Alignment
by: Rinaldi, Ivan, et al.
Published: (2026)
by: Rinaldi, Ivan, et al.
Published: (2026)
RoWeeder: Unsupervised Weed Mapping through Crop-Row Detection
by: De Marinis, Pasquale, et al.
Published: (2024)
by: De Marinis, Pasquale, et al.
Published: (2024)
Matching-Based Few-Shot Semantic Segmentation Models Are Interpretable by Design
by: De Marinis, Pasquale, et al.
Published: (2025)
by: De Marinis, Pasquale, et al.
Published: (2025)
GuidPaint: Class-Guided Image Inpainting with Diffusion Models
by: Wang, Qimin, et al.
Published: (2025)
by: Wang, Qimin, et al.
Published: (2025)
DistillFSS: Synthesizing Few-Shot Knowledge into a Lightweight Segmentation Model
by: De Marinis, Pasquale, et al.
Published: (2025)
by: De Marinis, Pasquale, et al.
Published: (2025)
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
by: Jin, Qixuan, et al.
Published: (2024)
by: Jin, Qixuan, et al.
Published: (2024)
MTADiffusion: Mask Text Alignment Diffusion Model for Object Inpainting
by: Huang, Jun, et al.
Published: (2025)
by: Huang, Jun, et al.
Published: (2025)
Text Image Inpainting via Global Structure-Guided Diffusion Models
by: Zhu, Shipeng, et al.
Published: (2024)
by: Zhu, Shipeng, et al.
Published: (2024)
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
by: Manukyan, Hayk, et al.
Published: (2023)
by: Manukyan, Hayk, et al.
Published: (2023)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
by: Jiang, Longtao, et al.
Published: (2025)
by: Jiang, Longtao, et al.
Published: (2025)
HarmonPaint: Harmonized Training-Free Diffusion Inpainting
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
PrefPaint: Aligning Image Inpainting Diffusion Model with Human Preference
by: Liu, Kendong, et al.
Published: (2024)
by: Liu, Kendong, et al.
Published: (2024)
DreamCom: Finetuning Text-guided Inpainting Model for Image Composition
by: Lu, Lingxiao, et al.
Published: (2023)
by: Lu, Lingxiao, et al.
Published: (2023)
Explainable offline automatic signature verifier to support forensic handwriting examiners
by: Diaz, Moises, et al.
Published: (2024)
by: Diaz, Moises, et al.
Published: (2024)
A New Chinese Landscape Paintings Generation Model based on Stable Diffusion using DreamBooth
by: Gu, Yujia, et al.
Published: (2024)
by: Gu, Yujia, et al.
Published: (2024)
TD-Paint: Faster Diffusion Inpainting Through Time Aware Pixel Conditioning
by: Mayet, Tsiry, et al.
Published: (2024)
by: Mayet, Tsiry, et al.
Published: (2024)
PromptFlare: Prompt-Generalized Defense via Cross-Attention Decoy in Diffusion-Based Inpainting
by: Na, Hohyun, et al.
Published: (2025)
by: Na, Hohyun, et al.
Published: (2025)
Diffree: Text-Guided Shape Free Object Inpainting with Diffusion Model
by: Zhao, Lirui, et al.
Published: (2024)
by: Zhao, Lirui, et al.
Published: (2024)
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts
by: Wan, Zhen, et al.
Published: (2024)
by: Wan, Zhen, et al.
Published: (2024)
Flow-Guided Diffusion for Video Inpainting
by: Gu, Bohai, et al.
Published: (2023)
by: Gu, Bohai, et al.
Published: (2023)
VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
by: Kuzyk, Oleh, et al.
Published: (2025)
by: Kuzyk, Oleh, et al.
Published: (2025)
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
by: Yu, Yongsheng, et al.
Published: (2025)
by: Yu, Yongsheng, et al.
Published: (2025)
One Stone with Two Birds: A Null-Text-Null Frequency-Aware Diffusion Models for Text-Guided Image Inpainting
by: Liu, Haipeng, et al.
Published: (2025)
by: Liu, Haipeng, et al.
Published: (2025)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
by: Yan, Weicai, et al.
Published: (2025)
by: Yan, Weicai, et al.
Published: (2025)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
by: Zhang, Jiaxu, et al.
Published: (2025)
by: Zhang, Jiaxu, et al.
Published: (2025)
DreamLayer: Simultaneous Multi-Layer Generation via Diffusion Mode
by: Huang, Junjia, et al.
Published: (2025)
by: Huang, Junjia, et al.
Published: (2025)
RePaintGS: Reference-Guided Gaussian Splatting for Realistic and View-Consistent 3D Scene Inpainting
by: Seo, Ji Hyun, et al.
Published: (2025)
by: Seo, Ji Hyun, et al.
Published: (2025)
InstaInpaint: Instant 3D-Scene Inpainting with Masked Large Reconstruction Model
by: You, Junqi, et al.
Published: (2025)
by: You, Junqi, et al.
Published: (2025)
PatternPaint: Practical Layout Pattern Generation Using Diffusion-Based Inpainting
by: Zhou, Guanglei, et al.
Published: (2024)
by: Zhou, Guanglei, et al.
Published: (2024)
Reference-Guided Diffusion Inpainting For Multimodal Counterfactual Generation
by: Buburuzan, Alexandru
Published: (2025)
by: Buburuzan, Alexandru
Published: (2025)
Caption Generation for Dongba Paintings via Prompt Learning and Semantic Fusion
by: Qian, Shuangwu, et al.
Published: (2026)
by: Qian, Shuangwu, et al.
Published: (2026)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
by: Liu, Baolin, et al.
Published: (2023)
by: Liu, Baolin, et al.
Published: (2023)
Promptable Game Models: Text-Guided Game Simulation via Masked Diffusion Models
by: Menapace, Willi, et al.
Published: (2023)
by: Menapace, Willi, et al.
Published: (2023)
SafePaint: Anti-forensic Image Inpainting with Domain Adaptation
by: Chen, Dunyun, et al.
Published: (2024)
by: Chen, Dunyun, et al.
Published: (2024)
Dynamically enhanced static handwriting representation for Parkinson's disease detection
by: Diaz, Moises, et al.
Published: (2024)
by: Diaz, Moises, et al.
Published: (2024)
DreamDrone: Text-to-Image Diffusion Models are Zero-shot Perpetual View Generators
by: Kong, Hanyang, et al.
Published: (2023)
by: Kong, Hanyang, et al.
Published: (2023)
Similar Items
-
ArtSeek: Deep artwork understanding via multimodal in-context reasoning and late interaction retrieval
by: Fanelli, Nicola, et al.
Published: (2025) -
Art2Mus: Bridging Visual Arts and Music through Cross-Modal Generation
by: Rinaldi, Ivan, et al.
Published: (2024) -
Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts
by: De Marinis, Pasquale, et al.
Published: (2024) -
Take a Peek: Efficient Encoder Adaptation for Few-Shot Semantic Segmentation via LoRA
by: De Marinis, Pasquale, et al.
Published: (2025) -
Art2Mus: Artwork-to-Music Generation via Visual Conditioning and Large-Scale Cross-Modal Alignment
by: Rinaldi, Ivan, et al.
Published: (2026)