LVFace: Progressive Cluster Optimization for Large Vision Models in Face Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | You, Jinghan, Li, Shanglin, Sun, Yuanrui, Wei, Jiangchuan, Guo, Mingyu, Feng, Chao, Ran, Jiao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
di: Chen, Renjie, et al.
Pubblicazione: (2025)
di: Chen, Renjie, et al.
Pubblicazione: (2025)
ContentV: Efficient Training of Video Generation Models with Limited Compute
di: Lin, Wenfeng, et al.
Pubblicazione: (2025)
di: Lin, Wenfeng, et al.
Pubblicazione: (2025)
RepFace: Refining Closed-Set Noise with Progressive Label Correction for Face Recognition
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion
di: Wei, Jiangchuan, et al.
Pubblicazione: (2025)
di: Wei, Jiangchuan, et al.
Pubblicazione: (2025)
CascadeV: An Implementation of Wurstchen Architecture for Video Generation
di: Lin, Wenfeng, et al.
Pubblicazione: (2025)
di: Lin, Wenfeng, et al.
Pubblicazione: (2025)
FaceLiVT: Face Recognition using Linear Vision Transformer with Structural Reparameterization For Mobile Device
di: Setyawan, Novendra, et al.
Pubblicazione: (2025)
di: Setyawan, Novendra, et al.
Pubblicazione: (2025)
PVG: Progressive Vision Graph for Vision Recognition
di: Wu, Jiafu, et al.
Pubblicazione: (2023)
di: Wu, Jiafu, et al.
Pubblicazione: (2023)
Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals
di: Shanglin, Yang
Pubblicazione: (2026)
di: Shanglin, Yang
Pubblicazione: (2026)
Scalable Vision Language Model Training via High Quality Data Curation
di: Dong, Hongyuan, et al.
Pubblicazione: (2025)
di: Dong, Hongyuan, et al.
Pubblicazione: (2025)
Evaluating Multimodal Large Language Models for Heterogeneous Face Recognition
di: Shahreza, Hatef Otroshi, et al.
Pubblicazione: (2026)
di: Shahreza, Hatef Otroshi, et al.
Pubblicazione: (2026)
Boosting Multi-modal Keyphrase Prediction with Dynamic Chain-of-Thought in Vision-Language Models
di: Ma, Qihang, et al.
Pubblicazione: (2025)
di: Ma, Qihang, et al.
Pubblicazione: (2025)
Face-MLLM: A Large Face Perception Model
di: Sun, Haomiao, et al.
Pubblicazione: (2024)
di: Sun, Haomiao, et al.
Pubblicazione: (2024)
Emotion Separation and Recognition from a Facial Expression by Generating the Poker Face with Vision Transformers
di: Li, Jia, et al.
Pubblicazione: (2022)
di: Li, Jia, et al.
Pubblicazione: (2022)
CemiFace: Center-based Semi-hard Synthetic Face Generation for Face Recognition
di: Sun, Zhonglin, et al.
Pubblicazione: (2024)
di: Sun, Zhonglin, et al.
Pubblicazione: (2024)
FaceLiVTv2: An Improved Hybrid Architecture for Efficient Mobile Face Recognition
di: Setyawan, Novendra, et al.
Pubblicazione: (2026)
di: Setyawan, Novendra, et al.
Pubblicazione: (2026)
Towards Large-Scale Pose-Invariant Face Recognition Using Face Defrontalization
di: Mesec, Patrik, et al.
Pubblicazione: (2025)
di: Mesec, Patrik, et al.
Pubblicazione: (2025)
Sample Correlation for Fingerprinting Deep Face Recognition
di: Guan, Jiyang, et al.
Pubblicazione: (2024)
di: Guan, Jiyang, et al.
Pubblicazione: (2024)
A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
di: Liu, Daizong, et al.
Pubblicazione: (2024)
di: Liu, Daizong, et al.
Pubblicazione: (2024)
Rethinking Vision-Language Model in Face Forensics: Multi-Modal Interpretable Forged Face Detector
di: Guo, Xiao, et al.
Pubblicazione: (2025)
di: Guo, Xiao, et al.
Pubblicazione: (2025)
FaceLinkGen: Rethinking Identity Leakage in Privacy-Preserving Face Recognition with Identity Extraction
di: Guo, Wenqi, et al.
Pubblicazione: (2026)
di: Guo, Wenqi, et al.
Pubblicazione: (2026)
Adversarial Attacks on Both Face Recognition and Face Anti-spoofing Models
di: Zhou, Fengfan, et al.
Pubblicazione: (2024)
di: Zhou, Fengfan, et al.
Pubblicazione: (2024)
IIR-VLM: In-Context Instance-level Recognition for Large Vision-Language Models
di: Shi, Liang, et al.
Pubblicazione: (2026)
di: Shi, Liang, et al.
Pubblicazione: (2026)
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens
di: Fan, Qihang, et al.
Pubblicazione: (2024)
di: Fan, Qihang, et al.
Pubblicazione: (2024)
BiggerGait: Unlocking Gait Recognition with Layer-wise Representations from Large Vision Models
di: Ye, Dingqiang, et al.
Pubblicazione: (2025)
di: Ye, Dingqiang, et al.
Pubblicazione: (2025)
Shape and Texture Recognition in Large Vision-Language Models
di: Eppel, Sagi, et al.
Pubblicazione: (2025)
di: Eppel, Sagi, et al.
Pubblicazione: (2025)
Mitigating Information Loss under High Pruning Rates for Efficient Large Vision Language Models
di: Fu, Mingyu, et al.
Pubblicazione: (2025)
di: Fu, Mingyu, et al.
Pubblicazione: (2025)
FastFace: Fast-converging Scheduler for Large-scale Face Recognition Training with One GPU
di: Gong, Xueyuan, et al.
Pubblicazione: (2024)
di: Gong, Xueyuan, et al.
Pubblicazione: (2024)
Manipulation Facing Threats: Evaluating Physical Vulnerabilities in End-to-End Vision Language Action Models
di: Cheng, Hao, et al.
Pubblicazione: (2024)
di: Cheng, Hao, et al.
Pubblicazione: (2024)
CryptoFace: End-to-End Encrypted Face Recognition
di: Ao, Wei, et al.
Pubblicazione: (2025)
di: Ao, Wei, et al.
Pubblicazione: (2025)
Benchmarking Vision Foundation Models for Domain-Generalizable Face Anti-Spoofing
di: Feng, Mika, et al.
Pubblicazione: (2026)
di: Feng, Mika, et al.
Pubblicazione: (2026)
In-context Learning of Vision Language Models for Detection of Physical and Digital Attacks against Face Recognition Systems
di: Gonzalez-Soler, Lazaro Janier, et al.
Pubblicazione: (2025)
di: Gonzalez-Soler, Lazaro Janier, et al.
Pubblicazione: (2025)
FRoundation: Are Foundation Models Ready for Face Recognition?
di: Chettaoui, Tahar, et al.
Pubblicazione: (2024)
di: Chettaoui, Tahar, et al.
Pubblicazione: (2024)
Recognition through Reasoning: Reinforcing Image Geo-localization with Large Vision-Language Models
di: Li, Ling, et al.
Pubblicazione: (2025)
di: Li, Ling, et al.
Pubblicazione: (2025)
DH-FaceVid-1K: A Large-Scale High-Quality Dataset for Face Video Generation
di: Di, Donglin, et al.
Pubblicazione: (2024)
di: Di, Donglin, et al.
Pubblicazione: (2024)
FaceCat: Enhancing Face Recognition Security with a Unified Diffusion Model
di: Chen, Jiawei, et al.
Pubblicazione: (2024)
di: Chen, Jiawei, et al.
Pubblicazione: (2024)
GVTNet: Graph Vision Transformer For Face Super-Resolution
di: Yang, Chao, et al.
Pubblicazione: (2025)
di: Yang, Chao, et al.
Pubblicazione: (2025)
KeyPoint Relative Position Encoding for Face Recognition
di: Kim, Minchul, et al.
Pubblicazione: (2024)
di: Kim, Minchul, et al.
Pubblicazione: (2024)
FaceShield: Explainable Face Anti-Spoofing with Multimodal Large Language Models
di: Wang, Hongyang, et al.
Pubblicazione: (2025)
di: Wang, Hongyang, et al.
Pubblicazione: (2025)
Contextual Emotion Recognition using Large Vision Language Models
di: Etesam, Yasaman, et al.
Pubblicazione: (2024)
di: Etesam, Yasaman, et al.
Pubblicazione: (2024)
PSMamba: Progressive Self-supervised Vision Mamba for Plant Disease Recognition
di: Mamun, Abdullah Al, et al.
Pubblicazione: (2025)
di: Mamun, Abdullah Al, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
di: Chen, Renjie, et al.
Pubblicazione: (2025) -
ContentV: Efficient Training of Video Generation Models with Limited Compute
di: Lin, Wenfeng, et al.
Pubblicazione: (2025) -
RepFace: Refining Closed-Set Noise with Progressive Label Correction for Face Recognition
di: Zhang, Jie, et al.
Pubblicazione: (2024) -
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion
di: Wei, Jiangchuan, et al.
Pubblicazione: (2025) -
CascadeV: An Implementation of Wurstchen Architecture for Video Generation
di: Lin, Wenfeng, et al.
Pubblicazione: (2025)