MCGM: Mask Conditional Text-to-Image Generative Model
Fuente:
arXiv
Guardado en:
| Autores principales: | Skaik, Rami, Rossi, Leonardo, Fontanini, Tomaso, Prati, Andrea |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Semantic Image Synthesis via Class-Adaptive Cross-Attention
por: Fontanini, Tomaso, et al.
Publicado: (2023)
por: Fontanini, Tomaso, et al.
Publicado: (2023)
MARS: Paying more attention to visual attributes for text-based person search
por: Ergasti, Alex, et al.
Publicado: (2024)
por: Ergasti, Alex, et al.
Publicado: (2024)
Swin2-MoSE: A New Single Image Super-Resolution Model for Remote Sensing
por: Rossi, Leonardo, et al.
Publicado: (2024)
por: Rossi, Leonardo, et al.
Publicado: (2024)
Memory-augmented Online Video Anomaly Detection
por: Rossi, Leonardo, et al.
Publicado: (2023)
por: Rossi, Leonardo, et al.
Publicado: (2023)
WaveMAE: Wavelet decomposition Masked Auto-Encoder for Remote Sensing
por: Bernuzzi, Vittorio, et al.
Publicado: (2025)
por: Bernuzzi, Vittorio, et al.
Publicado: (2025)
Adversarial Identity Injection for Semantic Face Image Synthesis
por: Tarollo, Giuseppe, et al.
Publicado: (2024)
por: Tarollo, Giuseppe, et al.
Publicado: (2024)
Mamba-ST: State Space Model for Efficient Style Transfer
por: Botti, Filippo, et al.
Publicado: (2024)
por: Botti, Filippo, et al.
Publicado: (2024)
Controllable Face Synthesis with Semantic Latent Diffusion Models
por: Ergasti, Alex, et al.
Publicado: (2024)
por: Ergasti, Alex, et al.
Publicado: (2024)
SISMA: Semantic Face Image Synthesis with Mamba
por: Botti, Filippo, et al.
Publicado: (2025)
por: Botti, Filippo, et al.
Publicado: (2025)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
por: Zheng, Zirui, et al.
Publicado: (2025)
por: Zheng, Zirui, et al.
Publicado: (2025)
U-Shape Mamba: State Space Model for faster diffusion
por: Ergasti, Alex, et al.
Publicado: (2025)
por: Ergasti, Alex, et al.
Publicado: (2025)
Progressive Image Restoration via Text-Conditioned Video Generation
por: Kang, Peng, et al.
Publicado: (2025)
por: Kang, Peng, et al.
Publicado: (2025)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
por: Wang, Ziming, et al.
Publicado: (2023)
por: Wang, Ziming, et al.
Publicado: (2023)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
por: Li, Yiheng, et al.
Publicado: (2026)
por: Li, Yiheng, et al.
Publicado: (2026)
$^R$FLAV: Rolling Flow matching for infinite Audio Video generation
por: Ergasti, Alex, et al.
Publicado: (2025)
por: Ergasti, Alex, et al.
Publicado: (2025)
Image Clustering Conditioned on Text Criteria
por: Kwon, Sehyun, et al.
Publicado: (2023)
por: Kwon, Sehyun, et al.
Publicado: (2023)
LOTS of Fashion! Multi-Conditioning for Image Generation via Sketch-Text Pairing
por: Girella, Federico, et al.
Publicado: (2025)
por: Girella, Federico, et al.
Publicado: (2025)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
por: Kang, Wonjun, et al.
Publicado: (2025)
por: Kang, Wonjun, et al.
Publicado: (2025)
Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning
por: Zhang, Yu, et al.
Publicado: (2025)
por: Zhang, Yu, et al.
Publicado: (2025)
MacDiff: Unified Skeleton Modeling with Masked Conditional Diffusion
por: Wu, Lehong, et al.
Publicado: (2024)
por: Wu, Lehong, et al.
Publicado: (2024)
Generating Satellite Imagery Data for Wildfire Detection through Mask-Conditioned Generative AI
por: Martin, Valeria, et al.
Publicado: (2026)
por: Martin, Valeria, et al.
Publicado: (2026)
VidVec: Unlocking Video MLLM Embeddings for Video-Text Retrieval
por: Tzachor, Issar, et al.
Publicado: (2026)
por: Tzachor, Issar, et al.
Publicado: (2026)
Personalized Reward Modeling for Text-to-Image Generation
por: Lee, Jeongeun, et al.
Publicado: (2025)
por: Lee, Jeongeun, et al.
Publicado: (2025)
HU-based Foreground Masking for 3D Medical Masked Image Modeling
por: Lee, Jin, et al.
Publicado: (2025)
por: Lee, Jin, et al.
Publicado: (2025)
Patch-enhanced Mask Encoder Prompt Image Generation
por: Xu, Shusong, et al.
Publicado: (2024)
por: Xu, Shusong, et al.
Publicado: (2024)
Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models
por: Kim, Hyungjin, et al.
Publicado: (2025)
por: Kim, Hyungjin, et al.
Publicado: (2025)
Conditional Diffusion Model for Longitudinal Medical Image Generation
por: Dao, Duy-Phuong, et al.
Publicado: (2024)
por: Dao, Duy-Phuong, et al.
Publicado: (2024)
CGI: Identifying Conditional Generative Models with Example Images
por: Zhou, Zhi, et al.
Publicado: (2025)
por: Zhou, Zhi, et al.
Publicado: (2025)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
por: Sampaio, Georgia Gabriela, et al.
Publicado: (2024)
por: Sampaio, Georgia Gabriela, et al.
Publicado: (2024)
MedSegFactory: Text-Guided Generation of Medical Image-Mask Pairs
por: Mao, Jiawei, et al.
Publicado: (2025)
por: Mao, Jiawei, et al.
Publicado: (2025)
HARIVO: Harnessing Text-to-Image Models for Video Generation
por: Kwon, Mingi, et al.
Publicado: (2024)
por: Kwon, Mingi, et al.
Publicado: (2024)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
por: Vice, Jordan, et al.
Publicado: (2024)
por: Vice, Jordan, et al.
Publicado: (2024)
Interactive Visual Assessment for Text-to-Image Generation Models
por: Mi, Xiaoyue, et al.
Publicado: (2024)
por: Mi, Xiaoyue, et al.
Publicado: (2024)
Self-Balanced R-CNN for Instance Segmentation
por: Rossi, Leonardo, et al.
Publicado: (2024)
por: Rossi, Leonardo, et al.
Publicado: (2024)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
por: Zhao, Yu, et al.
Publicado: (2024)
por: Zhao, Yu, et al.
Publicado: (2024)
Promptable Game Models: Text-Guided Game Simulation via Masked Diffusion Models
por: Menapace, Willi, et al.
Publicado: (2023)
por: Menapace, Willi, et al.
Publicado: (2023)
Model-based Cleaning of the QUILT-1M Pathology Dataset for Text-Conditional Image Synthesis
por: Aubreville, Marc, et al.
Publicado: (2024)
por: Aubreville, Marc, et al.
Publicado: (2024)
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
por: Li, Meiling, et al.
Publicado: (2024)
por: Li, Meiling, et al.
Publicado: (2024)
AEMIM: Adversarial Examples Meet Masked Image Modeling
por: Xiang, Wenzhao, et al.
Publicado: (2024)
por: Xiang, Wenzhao, et al.
Publicado: (2024)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
por: Chen, Muxi, et al.
Publicado: (2024)
por: Chen, Muxi, et al.
Publicado: (2024)
Ejemplares similares
-
Semantic Image Synthesis via Class-Adaptive Cross-Attention
por: Fontanini, Tomaso, et al.
Publicado: (2023) -
MARS: Paying more attention to visual attributes for text-based person search
por: Ergasti, Alex, et al.
Publicado: (2024) -
Swin2-MoSE: A New Single Image Super-Resolution Model for Remote Sensing
por: Rossi, Leonardo, et al.
Publicado: (2024) -
Memory-augmented Online Video Anomaly Detection
por: Rossi, Leonardo, et al.
Publicado: (2023) -
WaveMAE: Wavelet decomposition Masked Auto-Encoder for Remote Sensing
por: Bernuzzi, Vittorio, et al.
Publicado: (2025)