FaceLLM: A Multimodal Large Language Model for Face Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Shahreza, Hatef Otroshi, Marcel, Sébastien |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Multimodal Large Language Models for Face Recognition
by: Shahreza, Hatef Otroshi, et al.
Published: (2025)
by: Shahreza, Hatef Otroshi, et al.
Published: (2025)
Demographic Fairness in Multimodal LLMs: A Benchmark of Gender and Ethnicity Bias in Face Verification
by: Öztürk, Ünsal, et al.
Published: (2026)
by: Öztürk, Ünsal, et al.
Published: (2026)
Evaluating Multimodal Large Language Models for Heterogeneous Face Recognition
by: Shahreza, Hatef Otroshi, et al.
Published: (2026)
by: Shahreza, Hatef Otroshi, et al.
Published: (2026)
HyperFace: Generating Synthetic Face Recognition Datasets by Exploring Face Embedding Hypersphere
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
Face Reconstruction from Face Embeddings using Adapter to a Face Foundation Model
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
Unveiling Synthetic Faces: How Synthetic Datasets Can Expose Real Identities
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)
Synthetic Face Datasets Generation via Latent Space Exploration from Brownian Identity Diffusion
by: Geissbühler, David, et al.
Published: (2024)
by: Geissbühler, David, et al.
Published: (2024)
EdgeFace: Efficient Face Recognition Model for Edge Devices
by: George, Anjith, et al.
Published: (2023)
by: George, Anjith, et al.
Published: (2023)
Exploring ChatGPT for Face Presentation Attack Detection in Zero and Few-Shot in-Context Learning
by: Komaty, Alain, et al.
Published: (2025)
by: Komaty, Alain, et al.
Published: (2025)
ChatGPT and biometrics: an assessment of face recognition, gender detection, and age estimation capabilities
by: Hassanpour, Ahmad, et al.
Published: (2024)
by: Hassanpour, Ahmad, et al.
Published: (2024)
Approximating Optimal Morphing Attacks using Template Inversion
by: Colbois, Laurent, et al.
Published: (2024)
by: Colbois, Laurent, et al.
Published: (2024)
Model Pairing Using Embedding Translation for Backdoor Attack Detection on Open-Set Classification Tasks
by: Unnervik, Alexander, et al.
Published: (2024)
by: Unnervik, Alexander, et al.
Published: (2024)
ArtFace: Towards Historical Portrait Face Identification via Model Adaptation
by: Poh, Francois, et al.
Published: (2025)
by: Poh, Francois, et al.
Published: (2025)
Face-Human-Bench: A Comprehensive Benchmark of Face and Human Understanding for Multi-modal Assistants
by: Qin, Lixiong, et al.
Published: (2025)
by: Qin, Lixiong, et al.
Published: (2025)
More Distinctively Black and Feminine Faces Lead to Increased Stereotyping in Vision-Language Models
by: Lee, Messi H. J., et al.
Published: (2024)
by: Lee, Messi H. J., et al.
Published: (2024)
PointLLM: Empowering Large Language Models to Understand Point Clouds
by: Xu, Runsen, et al.
Published: (2023)
by: Xu, Runsen, et al.
Published: (2023)
II-Bench: An Image Implication Understanding Benchmark for Multimodal Large Language Models
by: Liu, Ziqiang, et al.
Published: (2024)
by: Liu, Ziqiang, et al.
Published: (2024)
Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos
by: Chen, Haodong, et al.
Published: (2026)
by: Chen, Haodong, et al.
Published: (2026)
TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
by: Ren, Shuhuai, et al.
Published: (2023)
by: Ren, Shuhuai, et al.
Published: (2023)
Benchmarking Egocentric Clinical Intent Understanding Capability for Medical Multimodal Large Language Models
by: Liu, Shaonan, et al.
Published: (2026)
by: Liu, Shaonan, et al.
Published: (2026)
Model Composition for Multimodal Large Language Models
by: Chen, Chi, et al.
Published: (2024)
by: Chen, Chi, et al.
Published: (2024)
A Survey on Agentic Multimodal Large Language Models
by: Yao, Huanjin, et al.
Published: (2025)
by: Yao, Huanjin, et al.
Published: (2025)
A Survey on Benchmarks of Multimodal Large Language Models
by: Li, Jian, et al.
Published: (2024)
by: Li, Jian, et al.
Published: (2024)
A Survey on Evaluation of Multimodal Large Language Models
by: Huang, Jiaxing, et al.
Published: (2024)
by: Huang, Jiaxing, et al.
Published: (2024)
On Pre-training of Multimodal Language Models Customized for Chart Understanding
by: Fan, Wan-Cyuan, et al.
Published: (2024)
by: Fan, Wan-Cyuan, et al.
Published: (2024)
Humor in Pixels: Benchmarking Large Multimodal Models Understanding of Online Comics
by: Ryan, Yuriel, et al.
Published: (2025)
by: Ryan, Yuriel, et al.
Published: (2025)
StreetviewLLM: Extracting Geographic Information Using a Chain-of-Thought Multimodal Large Language Model
by: Li, Zongrong, et al.
Published: (2024)
by: Li, Zongrong, et al.
Published: (2024)
Towards Understanding Graphical Perception in Large Multimodal Models
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Investigation of Accuracy and Bias in Face Recognition Trained with Synthetic Data
by: Korshunov, Pavel, et al.
Published: (2025)
by: Korshunov, Pavel, et al.
Published: (2025)
FFAA: Multimodal Large Language Model based Explainable Open-World Face Forgery Analysis Assistant
by: Huang, Zhengchao, et al.
Published: (2024)
by: Huang, Zhengchao, et al.
Published: (2024)
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025)
by: Zhang, Zongmeng, et al.
Published: (2025)
MLLM-CL: Continual Learning for Multimodal Large Language Models
by: Zhao, Hongbo, et al.
Published: (2025)
by: Zhao, Hongbo, et al.
Published: (2025)
Forgotten Polygons: Multimodal Large Language Models are Shape-Blind
by: Rudman, William, et al.
Published: (2025)
by: Rudman, William, et al.
Published: (2025)
Evaluating Large Language Models on Multimodal Chemistry Olympiad Exams
by: Cui, Yiming, et al.
Published: (2025)
by: Cui, Yiming, et al.
Published: (2025)
Cross-modal Information Flow in Multimodal Large Language Models
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models
by: Wang, Yuqing, et al.
Published: (2023)
by: Wang, Yuqing, et al.
Published: (2023)
BLINK: Multimodal Large Language Models Can See but Not Perceive
by: Fu, Xingyu, et al.
Published: (2024)
by: Fu, Xingyu, et al.
Published: (2024)
In-Depth and In-Breadth: Pre-training Multimodal Language Models Customized for Comprehensive Chart Understanding
by: Fan, Wan-Cyuan, et al.
Published: (2025)
by: Fan, Wan-Cyuan, et al.
Published: (2025)
S3Editor: A Sparse Semantic-Disentangled Self-Training Framework for Face Video Editing
by: Wang, Guangzhi, et al.
Published: (2024)
by: Wang, Guangzhi, et al.
Published: (2024)
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards
by: Hong, Jixiang, et al.
Published: (2025)
by: Hong, Jixiang, et al.
Published: (2025)
Similar Items
-
Benchmarking Multimodal Large Language Models for Face Recognition
by: Shahreza, Hatef Otroshi, et al.
Published: (2025) -
Demographic Fairness in Multimodal LLMs: A Benchmark of Gender and Ethnicity Bias in Face Verification
by: Öztürk, Ünsal, et al.
Published: (2026) -
Evaluating Multimodal Large Language Models for Heterogeneous Face Recognition
by: Shahreza, Hatef Otroshi, et al.
Published: (2026) -
HyperFace: Generating Synthetic Face Recognition Datasets by Exploring Face Embedding Hypersphere
by: Shahreza, Hatef Otroshi, et al.
Published: (2024) -
Face Reconstruction from Face Embeddings using Adapter to a Face Foundation Model
by: Shahreza, Hatef Otroshi, et al.
Published: (2024)