Fine-Grained Customized Fashion Design with Image-into-Prompt benchmark and dataset from LMM
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Hui, You, Yi, Chen, Qiqi, Zhang, Bingfeng, Huang, George Q. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FashionMAC: Deformation-Free Fashion Image Generation with Fine-Grained Model Appearance Customization
di: Zhang, Rong, et al.
Pubblicazione: (2025)
di: Zhang, Rong, et al.
Pubblicazione: (2025)
IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design
di: Shen, Fei, et al.
Pubblicazione: (2025)
di: Shen, Fei, et al.
Pubblicazione: (2025)
Prompt2Fashion: An automatically generated fashion dataset
di: Argyrou, Georgia, et al.
Pubblicazione: (2024)
di: Argyrou, Georgia, et al.
Pubblicazione: (2024)
LMM4LMM: Benchmarking and Evaluating Large-multimodal Image Generation with LMMs
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
di: Huang, Jiale, et al.
Pubblicazione: (2024)
di: Huang, Jiale, et al.
Pubblicazione: (2024)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
di: Chen, Wenting, et al.
Pubblicazione: (2024)
di: Chen, Wenting, et al.
Pubblicazione: (2024)
FashionComposer: Compositional Fashion Image Generation
di: Ji, Sihui, et al.
Pubblicazione: (2024)
di: Ji, Sihui, et al.
Pubblicazione: (2024)
Language Prompt vs. Image Enhancement: Boosting Object Detection With CLIP in Hazy Environments
di: Pang, Jian, et al.
Pubblicazione: (2026)
di: Pang, Jian, et al.
Pubblicazione: (2026)
FashionLOGO: Prompting Multimodal Large Language Models for Fashion Logo Embeddings
di: Wang, Zhen, et al.
Pubblicazione: (2023)
di: Wang, Zhen, et al.
Pubblicazione: (2023)
LMM-PCQA: Assisting Point Cloud Quality Assessment with LMM
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
FashionMV: Product-Level Composed Image Retrieval with Multi-View Fashion Data
di: Yuan, Peng, et al.
Pubblicazione: (2026)
di: Yuan, Peng, et al.
Pubblicazione: (2026)
Head-wise Adaptive Rotary Positional Encoding for Fine-Grained Image Generation
di: Li, Jiaye, et al.
Pubblicazione: (2025)
di: Li, Jiaye, et al.
Pubblicazione: (2025)
FashionPose: Text to Pose to Relight Image Generation for Personalized Fashion Visualization
di: Shi, Chuancheng, et al.
Pubblicazione: (2025)
di: Shi, Chuancheng, et al.
Pubblicazione: (2025)
SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing
di: Xiao, Yicheng, et al.
Pubblicazione: (2026)
di: Xiao, Yicheng, et al.
Pubblicazione: (2026)
F-LMM: Grounding Frozen Large Multimodal Models
di: Wu, Size, et al.
Pubblicazione: (2024)
di: Wu, Size, et al.
Pubblicazione: (2024)
LMM-Regularized CLIP Embeddings for Image Classification
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization
di: Song, Quanjian, et al.
Pubblicazione: (2026)
di: Song, Quanjian, et al.
Pubblicazione: (2026)
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models
di: Argyrou, Georgia, et al.
Pubblicazione: (2024)
di: Argyrou, Georgia, et al.
Pubblicazione: (2024)
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
di: He, Jing, et al.
Pubblicazione: (2024)
di: He, Jing, et al.
Pubblicazione: (2024)
AttriPrompt: Dynamic Prompt Composition Learning for CLIP
di: Zhan, Qiqi, et al.
Pubblicazione: (2025)
di: Zhan, Qiqi, et al.
Pubblicazione: (2025)
Disturbing Image Detection Using LMM-Elicited Emotion Embeddings
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
Quality and Quantity: Unveiling a Million High-Quality Images for Text-to-Image Synthesis in Fashion Design
di: Yu, Jia, et al.
Pubblicazione: (2023)
di: Yu, Jia, et al.
Pubblicazione: (2023)
HCC-3D: Hierarchical Compensatory Compression for 98% 3D Token Reduction in Vision-Language Models
di: Zhang, Liheng, et al.
Pubblicazione: (2025)
di: Zhang, Liheng, et al.
Pubblicazione: (2025)
FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval
di: Gardères, François, et al.
Pubblicazione: (2026)
di: Gardères, François, et al.
Pubblicazione: (2026)
Adaptive Bidirectional Displacement for Semi-Supervised Medical Image Segmentation
di: Chi, Hanyang, et al.
Pubblicazione: (2024)
di: Chi, Hanyang, et al.
Pubblicazione: (2024)
Adaptive Point-Prompt Tuning: Fine-Tuning Heterogeneous Foundation Models for 3D Point Cloud Analysis
di: Li, Mengke, et al.
Pubblicazione: (2025)
di: Li, Mengke, et al.
Pubblicazione: (2025)
User-Friendly Customized Generation with Multi-Modal Prompts
di: Zhong, Linhao, et al.
Pubblicazione: (2024)
di: Zhong, Linhao, et al.
Pubblicazione: (2024)
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance
di: Xu, Haohang, et al.
Pubblicazione: (2026)
di: Xu, Haohang, et al.
Pubblicazione: (2026)
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
di: Agarwal, Aishwarya, et al.
Pubblicazione: (2024)
di: Agarwal, Aishwarya, et al.
Pubblicazione: (2024)
Double Banking on Knowledge: Customized Modulation and Prototypes for Multi-Modality Semi-supervised Medical Image Segmentation
di: Chen, Yingyu, et al.
Pubblicazione: (2024)
di: Chen, Yingyu, et al.
Pubblicazione: (2024)
HAIFIT: Human-to-AI Fashion Image Translation
di: Jiang, Jianan, et al.
Pubblicazione: (2024)
di: Jiang, Jianan, et al.
Pubblicazione: (2024)
Fine-Grained Zero-Shot Composed Image Retrieval with Complementary Visual-Semantic Integration
di: Ye, Yongcong, et al.
Pubblicazione: (2026)
di: Ye, Yongcong, et al.
Pubblicazione: (2026)
Anatomy-Aware Text-Visual Fusion with Dual-Perspective Prompts for Fine-Grained Lumbar Spine Segmentation
di: Lian, Sheng, et al.
Pubblicazione: (2025)
di: Lian, Sheng, et al.
Pubblicazione: (2025)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
di: Wu, Fan, et al.
Pubblicazione: (2025)
di: Wu, Fan, et al.
Pubblicazione: (2025)
ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images
di: Kong, Xianghao, et al.
Pubblicazione: (2025)
di: Kong, Xianghao, et al.
Pubblicazione: (2025)
A Multihead Continual Learning Framework for Fine-Grained Fashion Image Retrieval with Contrastive Learning and Exponential Moving Average Distillation
di: Xiao, Ling, et al.
Pubblicazione: (2026)
di: Xiao, Ling, et al.
Pubblicazione: (2026)
FGENet: Fine-Grained Extraction Network for Congested Crowd Counting
di: Ma, Hao-Yuan, et al.
Pubblicazione: (2024)
di: Ma, Hao-Yuan, et al.
Pubblicazione: (2024)
AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
UltraEdit: Instruction-based Fine-Grained Image Editing at Scale
di: Zhao, Haozhe, et al.
Pubblicazione: (2024)
di: Zhao, Haozhe, et al.
Pubblicazione: (2024)
Empower Vision Applications with LoRA LMM
di: Mi, Liang, et al.
Pubblicazione: (2024)
di: Mi, Liang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FashionMAC: Deformation-Free Fashion Image Generation with Fine-Grained Model Appearance Customization
di: Zhang, Rong, et al.
Pubblicazione: (2025) -
IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design
di: Shen, Fei, et al.
Pubblicazione: (2025) -
Prompt2Fashion: An automatically generated fashion dataset
di: Argyrou, Georgia, et al.
Pubblicazione: (2024) -
LMM4LMM: Benchmarking and Evaluating Large-multimodal Image Generation with LMMs
di: Wang, Jiarui, et al.
Pubblicazione: (2025) -
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
di: Huang, Jiale, et al.
Pubblicazione: (2024)