FaceGemma: Enhancing Image Captioning with Facial Attributes for Portrait Images
Fuente:
arXiv
Saved in:
| Main Authors: | Haque, Naimul, Labiba, Iffat, Akter, Sadia |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
by: Cha, Hyunsoo, et al.
Published: (2025)
by: Cha, Hyunsoo, et al.
Published: (2025)
DynaSeg: A Deep Dynamic Fusion Method for Unsupervised Image Segmentation Incorporating Feature Similarity and Spatial Continuity
by: Guermazi, Boujemaa, et al.
Published: (2024)
by: Guermazi, Boujemaa, et al.
Published: (2024)
Maximizing Generalization: The Effect of Different Augmentation Techniques on Lightweight Vision Transformer for Bengali Character Classification
by: Chowdhury, Rafi Hassan, et al.
Published: (2026)
by: Chowdhury, Rafi Hassan, et al.
Published: (2026)
EasyPortrait -- Face Parsing and Portrait Segmentation Dataset
by: Kvanchiani, Karina, et al.
Published: (2023)
by: Kvanchiani, Karina, et al.
Published: (2023)
FaceShield: Defending Facial Image against Deepfake Threats
by: Jeong, Jaehwan, et al.
Published: (2024)
by: Jeong, Jaehwan, et al.
Published: (2024)
FaceFilterSense: A Filter-Resistant Face Recognition and Facial Attribute Analysis Framework
by: Tiwari, Shubham, et al.
Published: (2024)
by: Tiwari, Shubham, et al.
Published: (2024)
MD-Face: MoE-Enhanced Label-Free Disentangled Representation for Interactive Facial Attribute Editing
by: Cui, Xuan, et al.
Published: (2026)
by: Cui, Xuan, et al.
Published: (2026)
Aesthetic Image Captioning with Saliency Enhanced MLLMs
by: Tao, Yilin, et al.
Published: (2025)
by: Tao, Yilin, et al.
Published: (2025)
InstaFace: Identity-Preserving Facial Editing with Single Image Inference
by: Khan, MD Wahiduzzaman, et al.
Published: (2025)
by: Khan, MD Wahiduzzaman, et al.
Published: (2025)
BiSe-Unet: A Lightweight Dual-path U-Net with Attention-refined Context for Real-time Medical Image Segmentation
by: Hossain, M Iffat, et al.
Published: (2026)
by: Hossain, M Iffat, et al.
Published: (2026)
Enhancing Descriptive Captions with Visual Attributes for Multimodal Perception
by: Sun, Yanpeng, et al.
Published: (2024)
by: Sun, Yanpeng, et al.
Published: (2024)
BornoViT: A Novel Efficient Vision Transformer for Bengali Handwritten Basic Characters Classification
by: Chowdhury, Rafi Hassan, et al.
Published: (2026)
by: Chowdhury, Rafi Hassan, et al.
Published: (2026)
Evaluating Demographic Misrepresentation in Image-to-Image Portrait Editing
by: Seo, Huichan, et al.
Published: (2026)
by: Seo, Huichan, et al.
Published: (2026)
JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation
by: Cao, Xuyang, et al.
Published: (2024)
by: Cao, Xuyang, et al.
Published: (2024)
Bias Analysis for Synthetic Face Detection: A Case Study of the Impact of Facial Attributes
by: Lamsaf, Asmae, et al.
Published: (2025)
by: Lamsaf, Asmae, et al.
Published: (2025)
Hierarchical Vectorization for Portrait Images
by: Fu, Qian, et al.
Published: (2022)
by: Fu, Qian, et al.
Published: (2022)
FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs
by: Wang, Xiaoqin, et al.
Published: (2025)
by: Wang, Xiaoqin, et al.
Published: (2025)
FaceSnap: Enhanced ID-fidelity Network for Tuning-free Portrait Customization
by: Zhai, Benxiang, et al.
Published: (2026)
by: Zhai, Benxiang, et al.
Published: (2026)
GRIHA: Synthesizing 2-Dimensional Building Layouts from Images Captured using a Smart Phone
by: Goyal, Shreya, et al.
Published: (2021)
by: Goyal, Shreya, et al.
Published: (2021)
CaptionQA: Is Your Caption as Useful as the Image Itself?
by: Yang, Shijia, et al.
Published: (2025)
by: Yang, Shijia, et al.
Published: (2025)
Generating an Image From 1,000 Words: Enhancing Text-to-Image With Structured Captions
by: Gutflaish, Eyal, et al.
Published: (2025)
by: Gutflaish, Eyal, et al.
Published: (2025)
PerFace: Metric Learning in Perceptual Facial Similarity for Enhanced Face Anonymization
by: Kumagai, Haruka, et al.
Published: (2025)
by: Kumagai, Haruka, et al.
Published: (2025)
FaceMixup: Enhancing Facial Expression Recognition through Mixed Face Regularization
by: Faria, Fabio A., et al.
Published: (2024)
by: Faria, Fabio A., et al.
Published: (2024)
DynamicFace: High-Quality and Consistent Face Swapping for Image and Video using Composable 3D Facial Priors
by: Wang, Runqi, et al.
Published: (2025)
by: Wang, Runqi, et al.
Published: (2025)
SDFD: Building a Versatile Synthetic Face Image Dataset with Diverse Attributes
by: Baltsou, Georgia, et al.
Published: (2024)
by: Baltsou, Georgia, et al.
Published: (2024)
Inserting Faces inside Captions: Image Captioning with Attention Guided Merging
by: Tevissen, Yannis, et al.
Published: (2024)
by: Tevissen, Yannis, et al.
Published: (2024)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
by: Möller, Lucas, et al.
Published: (2024)
by: Möller, Lucas, et al.
Published: (2024)
CaptionSmiths: Flexibly Controlling Language Pattern in Image Captioning
by: Saito, Kuniaki, et al.
Published: (2025)
by: Saito, Kuniaki, et al.
Published: (2025)
Image Generation from Image Captioning -- Invertible Approach
by: Menon, Nandakishore S, et al.
Published: (2024)
by: Menon, Nandakishore S, et al.
Published: (2024)
Generating Attribution Reports for Manipulated Facial Images: A Dataset and Baseline
by: Lian, Jingchun, et al.
Published: (2024)
by: Lian, Jingchun, et al.
Published: (2024)
Q-Bench-Portrait: Benchmarking Multimodal Large Language Models on Portrait Image Quality Perception
by: Wu, Sijing, et al.
Published: (2026)
by: Wu, Sijing, et al.
Published: (2026)
Enhancing Perception of Key Changes in Remote Sensing Image Change Captioning
by: Yang, Cong, et al.
Published: (2024)
by: Yang, Cong, et al.
Published: (2024)
Multi-Attribute guided Thermal Face Image Translation based on Latent Diffusion Model
by: Cai, Mingshu, et al.
Published: (2025)
by: Cai, Mingshu, et al.
Published: (2025)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
by: Dai, Dawei, et al.
Published: (2025)
by: Dai, Dawei, et al.
Published: (2025)
GNN-ViTCap: GNN-Enhanced Multiple Instance Learning with Vision Transformers for Whole Slide Image Classification and Captioning
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning
by: Zhang, Lin, et al.
Published: (2025)
by: Zhang, Lin, et al.
Published: (2025)
MagicStyle: Portrait Stylization Based on Reference Image
by: Deng, Zhaoli, et al.
Published: (2024)
by: Deng, Zhaoli, et al.
Published: (2024)
Revealing Unintentional Information Leakage in Low-Dimensional Facial Portrait Representations
by: Anderson, Kathleen, et al.
Published: (2025)
by: Anderson, Kathleen, et al.
Published: (2025)
What Makes for Good Image Captions?
by: Chen, Delong, et al.
Published: (2024)
by: Chen, Delong, et al.
Published: (2024)
Benchmarking and Improving Detail Image Caption
by: Dong, Hongyuan, et al.
Published: (2024)
by: Dong, Hongyuan, et al.
Published: (2024)
Similar Items
-
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
by: Cha, Hyunsoo, et al.
Published: (2025) -
DynaSeg: A Deep Dynamic Fusion Method for Unsupervised Image Segmentation Incorporating Feature Similarity and Spatial Continuity
by: Guermazi, Boujemaa, et al.
Published: (2024) -
Maximizing Generalization: The Effect of Different Augmentation Techniques on Lightweight Vision Transformer for Bengali Character Classification
by: Chowdhury, Rafi Hassan, et al.
Published: (2026) -
EasyPortrait -- Face Parsing and Portrait Segmentation Dataset
by: Kvanchiani, Karina, et al.
Published: (2023) -
FaceShield: Defending Facial Image against Deepfake Threats
by: Jeong, Jaehwan, et al.
Published: (2024)