Evaluating Attribute Confusion in Fashion Text-to-Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Ziyue, Girella, Federico, Wang, Yiming, Talon, Davide |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LOTS of Fashion! Multi-Conditioning for Image Generation via Sketch-Text Pairing
por: Girella, Federico, et al.
Publicado: (2025)
por: Girella, Federico, et al.
Publicado: (2025)
Multi-Level Conditioning by Pairing Localized Text and Sketch for Fashion Image Generation
por: Liu, Ziyue, et al.
Publicado: (2026)
por: Liu, Ziyue, et al.
Publicado: (2026)
Seeing the Abstract: Translating the Abstract Language for Vision Language Models
por: Talon, Davide, et al.
Publicado: (2025)
por: Talon, Davide, et al.
Publicado: (2025)
One VLM to Keep it Learning: Generation and Balancing for Data-free Continual Visual Question Answering
por: Das, Deepayan, et al.
Publicado: (2024)
por: Das, Deepayan, et al.
Publicado: (2024)
Leveraging Latent Diffusion Models for Training-Free In-Distribution Data Augmentation for Surface Defect Detection
por: Girella, Federico, et al.
Publicado: (2024)
por: Girella, Federico, et al.
Publicado: (2024)
FashionPose: Text to Pose to Relight Image Generation for Personalized Fashion Visualization
por: Shi, Chuancheng, et al.
Publicado: (2025)
por: Shi, Chuancheng, et al.
Publicado: (2025)
Training-Free Personalization via Retrieval and Reasoning on Fingerprints
por: Das, Deepayan, et al.
Publicado: (2025)
por: Das, Deepayan, et al.
Publicado: (2025)
FashionComposer: Compositional Fashion Image Generation
por: Ji, Sihui, et al.
Publicado: (2024)
por: Ji, Sihui, et al.
Publicado: (2024)
Object-Attribute Binding in Text-to-Image Generation: Evaluation and Control
por: Trusca, Maria Mihaela, et al.
Publicado: (2024)
por: Trusca, Maria Mihaela, et al.
Publicado: (2024)
Fashion-RAG: Multimodal Fashion Image Editing via Retrieval-Augmented Generation
por: Sanguigni, Fulvio, et al.
Publicado: (2025)
por: Sanguigni, Fulvio, et al.
Publicado: (2025)
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
por: Huang, Jiale, et al.
Publicado: (2024)
por: Huang, Jiale, et al.
Publicado: (2024)
FashionMAC: Deformation-Free Fashion Image Generation with Fine-Grained Model Appearance Customization
por: Zhang, Rong, et al.
Publicado: (2025)
por: Zhang, Rong, et al.
Publicado: (2025)
ObjectAdd: Adding Objects into Image via a Training-Free Diffusion Modification Fashion
por: Zhang, Ziyue, et al.
Publicado: (2024)
por: Zhang, Ziyue, et al.
Publicado: (2024)
Q-Save: Towards Scoring and Attribution for Generated Video Evaluation
por: Wu, Xiele, et al.
Publicado: (2025)
por: Wu, Xiele, et al.
Publicado: (2025)
Apply Hierarchical-Chain-of-Generation to Complex Attributes Text-to-3D Generation
por: Qin, Yiming, et al.
Publicado: (2025)
por: Qin, Yiming, et al.
Publicado: (2025)
How to Take a Memorable Picture? Empowering Users with Actionable Feedback
por: Laiti, Francesco, et al.
Publicado: (2026)
por: Laiti, Francesco, et al.
Publicado: (2026)
Multimodal-Conditioned Latent Diffusion Models for Fashion Image Editing
por: Baldrati, Alberto, et al.
Publicado: (2024)
por: Baldrati, Alberto, et al.
Publicado: (2024)
Which Model Generated This Image? A Model-Agnostic Approach for Origin Attribution
por: Liu, Fengyuan, et al.
Publicado: (2024)
por: Liu, Fengyuan, et al.
Publicado: (2024)
ConceptGuard: Continual Personalized Text-to-Image Generation with Forgetting and Confusion Mitigation
por: Guo, Zirun, et al.
Publicado: (2025)
por: Guo, Zirun, et al.
Publicado: (2025)
Resolving Multi-Condition Confusion for Finetuning-Free Personalized Image Generation
por: Huang, Qihan, et al.
Publicado: (2024)
por: Huang, Qihan, et al.
Publicado: (2024)
Diffusion-based Image Generation for In-distribution Data Augmentation in Surface Defect Detection
por: Capogrosso, Luigi, et al.
Publicado: (2024)
por: Capogrosso, Luigi, et al.
Publicado: (2024)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
por: Zhan, Zechao, et al.
Publicado: (2024)
por: Zhan, Zechao, et al.
Publicado: (2024)
ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images
por: Kong, Xianghao, et al.
Publicado: (2025)
por: Kong, Xianghao, et al.
Publicado: (2025)
Improving Compositional Attribute Binding in Text-to-Image Generative Models via Enhanced Text Embeddings
por: Zarei, Arman, et al.
Publicado: (2024)
por: Zarei, Arman, et al.
Publicado: (2024)
Quality and Quantity: Unveiling a Million High-Quality Images for Text-to-Image Synthesis in Fashion Design
por: Yu, Jia, et al.
Publicado: (2023)
por: Yu, Jia, et al.
Publicado: (2023)
Attribution as Retrieval: Model-Agnostic AI-Generated Image Attribution
por: Wang, Hongsong, et al.
Publicado: (2026)
por: Wang, Hongsong, et al.
Publicado: (2026)
Unleashing the Potential of Large Language Models for Text-to-Image Generation through Autoregressive Representation Alignment
por: Xie, Xing, et al.
Publicado: (2025)
por: Xie, Xing, et al.
Publicado: (2025)
RAIGen: Rare Attribute Identification in Text-to-Image Generative Models
por: Sreelatha, Silpa Vadakkeeveetil, et al.
Publicado: (2026)
por: Sreelatha, Silpa Vadakkeeveetil, et al.
Publicado: (2026)
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models
por: Argyrou, Georgia, et al.
Publicado: (2024)
por: Argyrou, Georgia, et al.
Publicado: (2024)
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation
por: Li, Niantong, et al.
Publicado: (2026)
por: Li, Niantong, et al.
Publicado: (2026)
Composing Object Relations and Attributes for Image-Text Matching
por: Pham, Khoi, et al.
Publicado: (2024)
por: Pham, Khoi, et al.
Publicado: (2024)
Detecting Origin Attribution for Text-to-Image Diffusion Models
por: Xu, Katherine, et al.
Publicado: (2024)
por: Xu, Katherine, et al.
Publicado: (2024)
Generic Knowledge Boosted Pre-training For Remote Sensing Images
por: Huang, Ziyue, et al.
Publicado: (2024)
por: Huang, Ziyue, et al.
Publicado: (2024)
Dynamic Inter-Class Confusion-Aware Encoder for Audio-Visual Fusion in Human Activity Recognition
por: Cong, Kaixuan, et al.
Publicado: (2025)
por: Cong, Kaixuan, et al.
Publicado: (2025)
Content-Adaptive Image Retouching Guided by Attribute-Based Text Representation
por: Zhu, Hancheng, et al.
Publicado: (2025)
por: Zhu, Hancheng, et al.
Publicado: (2025)
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition
por: He, Yu, et al.
Publicado: (2026)
por: He, Yu, et al.
Publicado: (2026)
ConfusionBench: An Expert-Validated Benchmark for Confusion Recognition and Localization in Educational Videos
por: Dong, Lu, et al.
Publicado: (2026)
por: Dong, Lu, et al.
Publicado: (2026)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
por: Cui, Tianyu, et al.
Publicado: (2025)
por: Cui, Tianyu, et al.
Publicado: (2025)
AnyText2: Visual Text Generation and Editing With Customizable Attributes
por: Tuo, Yuxiang, et al.
Publicado: (2024)
por: Tuo, Yuxiang, et al.
Publicado: (2024)
HAIFIT: Human-to-AI Fashion Image Translation
por: Jiang, Jianan, et al.
Publicado: (2024)
por: Jiang, Jianan, et al.
Publicado: (2024)
Ejemplares similares
-
LOTS of Fashion! Multi-Conditioning for Image Generation via Sketch-Text Pairing
por: Girella, Federico, et al.
Publicado: (2025) -
Multi-Level Conditioning by Pairing Localized Text and Sketch for Fashion Image Generation
por: Liu, Ziyue, et al.
Publicado: (2026) -
Seeing the Abstract: Translating the Abstract Language for Vision Language Models
por: Talon, Davide, et al.
Publicado: (2025) -
One VLM to Keep it Learning: Generation and Balancing for Data-free Continual Visual Question Answering
por: Das, Deepayan, et al.
Publicado: (2024) -
Leveraging Latent Diffusion Models for Training-Free In-Distribution Data Augmentation for Surface Defect Detection
por: Girella, Federico, et al.
Publicado: (2024)