Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Dai, Dawei, Jia, Mingming, Zhou, Yinxiu, Xing, Hang, Li, Chenghang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Face-MakeUpV2: Facial Consistency Learning for Controllable Text-to-Image Generation
di: Dai, Dawei, et al.
Pubblicazione: (2025)
di: Dai, Dawei, et al.
Pubblicazione: (2025)
15M Multimodal Facial Image-Text Dataset
di: Dai, Dawei, et al.
Pubblicazione: (2024)
di: Dai, Dawei, et al.
Pubblicazione: (2024)
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
Multimodal Prompt Alignment for Facial Expression Recognition
di: Ma, Fuyan, et al.
Pubblicazione: (2025)
di: Ma, Fuyan, et al.
Pubblicazione: (2025)
Prompt Decoupling for Text-to-Image Person Re-identification
di: Li, Weihao, et al.
Pubblicazione: (2024)
di: Li, Weihao, et al.
Pubblicazione: (2024)
IdentiFace : A VGG Based Multimodal Facial Biometric System
di: Rabea, Mahmoud, et al.
Pubblicazione: (2024)
di: Rabea, Mahmoud, et al.
Pubblicazione: (2024)
Controllable Talking Face Generation by Implicit Facial Keypoints Editing
di: Zhao, Dong, et al.
Pubblicazione: (2024)
di: Zhao, Dong, et al.
Pubblicazione: (2024)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
Multivariate Diffusion Transformer with Decoupled Attention for High-Fidelity Mask-Text Collaborative Facial Generation
di: Cao, Yushe, et al.
Pubblicazione: (2025)
di: Cao, Yushe, et al.
Pubblicazione: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
Text-Driven Emotionally Continuous Talking Face Generation
di: Yang, Hao, et al.
Pubblicazione: (2026)
di: Yang, Hao, et al.
Pubblicazione: (2026)
FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing
di: Feng, Guanwen, et al.
Pubblicazione: (2025)
di: Feng, Guanwen, et al.
Pubblicazione: (2025)
Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making
di: Dai, Siyuan, et al.
Pubblicazione: (2025)
di: Dai, Siyuan, et al.
Pubblicazione: (2025)
MASTER: Multimodal Segmentation with Text Prompts
di: Liu, Fuyang, et al.
Pubblicazione: (2025)
di: Liu, Fuyang, et al.
Pubblicazione: (2025)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
di: Zhou, Linhan, et al.
Pubblicazione: (2025)
di: Zhou, Linhan, et al.
Pubblicazione: (2025)
Scale Up Composed Image Retrieval Learning via Modification Text Generation
di: Zhou, Yinan, et al.
Pubblicazione: (2025)
di: Zhou, Yinan, et al.
Pubblicazione: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
Smile on the Face, Sadness in the Eyes: Bridging the Emotion Gap with a Multimodal Dataset of Eye and Facial Behaviors
di: Liu, Kejun, et al.
Pubblicazione: (2025)
di: Liu, Kejun, et al.
Pubblicazione: (2025)
Weakly Supervised Gaussian Contrastive Grounding with Large Multimodal Models for Video Question Answering
di: Wang, Haibo, et al.
Pubblicazione: (2024)
di: Wang, Haibo, et al.
Pubblicazione: (2024)
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data
di: Berman, William, et al.
Pubblicazione: (2024)
di: Berman, William, et al.
Pubblicazione: (2024)
FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
di: Teo, Christopher T. H, et al.
Pubblicazione: (2024)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
di: Park, Sangha, et al.
Pubblicazione: (2025)
di: Park, Sangha, et al.
Pubblicazione: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
di: Sung-Bin, Kim, et al.
Pubblicazione: (2025)
di: Sung-Bin, Kim, et al.
Pubblicazione: (2025)
Spatially Covariant Image Registration with Text Prompts
di: Chen, Xiang, et al.
Pubblicazione: (2023)
di: Chen, Xiang, et al.
Pubblicazione: (2023)
AlphaFace: High Fidelity and Real-time Face Swapper Robust to Facial Pose
di: Yu, Jongmin, et al.
Pubblicazione: (2026)
di: Yu, Jongmin, et al.
Pubblicazione: (2026)
Enhance Multimodal Consistency and Coherence for Text-Image Plan Generation
di: Lu, Xiaoxin, et al.
Pubblicazione: (2025)
di: Lu, Xiaoxin, et al.
Pubblicazione: (2025)
Transferable Adversarial Facial Images for Privacy Protection
di: Li, Minghui, et al.
Pubblicazione: (2024)
di: Li, Minghui, et al.
Pubblicazione: (2024)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
di: Zhou, Shijie, et al.
Pubblicazione: (2025)
di: Zhou, Shijie, et al.
Pubblicazione: (2025)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
di: Binyamin, Lital, et al.
Pubblicazione: (2024)
di: Binyamin, Lital, et al.
Pubblicazione: (2024)
A novel Facial Recognition technique with Focusing on Masked Faces
di: Abdullah, Dana A, et al.
Pubblicazione: (2025)
di: Abdullah, Dana A, et al.
Pubblicazione: (2025)
Multimodal Foundation Models Exploit Text to Make Medical Image Predictions
di: Buckley, Thomas, et al.
Pubblicazione: (2023)
di: Buckley, Thomas, et al.
Pubblicazione: (2023)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
di: Chatterjee, Abhiroop, et al.
Pubblicazione: (2025)
di: Chatterjee, Abhiroop, et al.
Pubblicazione: (2025)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
di: Wu, Shang, et al.
Pubblicazione: (2026)
di: Wu, Shang, et al.
Pubblicazione: (2026)
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
di: Yuan, Lingzhi, et al.
Pubblicazione: (2025)
A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation
di: Feng, Yukang, et al.
Pubblicazione: (2025)
di: Feng, Yukang, et al.
Pubblicazione: (2025)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Face-MakeUpV2: Facial Consistency Learning for Controllable Text-to-Image Generation
di: Dai, Dawei, et al.
Pubblicazione: (2025) -
15M Multimodal Facial Image-Text Dataset
di: Dai, Dawei, et al.
Pubblicazione: (2024) -
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026) -
Multimodal Prompt Alignment for Facial Expression Recognition
di: Ma, Fuyan, et al.
Pubblicazione: (2025) -
Prompt Decoupling for Text-to-Image Person Re-identification
di: Li, Weihao, et al.
Pubblicazione: (2024)