Generalist Multimodal LLMs Gain Biometric Expertise via Human Salience
Fuente:
arXiv
Saved in:
| Main Authors: | Piland, Jacob, Dowling, Byron, Sweet, Christopher, Czajka, Adam |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VISER: Visually-Informed System for Enhanced Robustness in Open-Set Iris Presentation Attack Detection
by: Dowling, Byron, et al.
Published: (2026)
by: Dowling, Byron, et al.
Published: (2026)
Divisive Decisions: Improving Salience-Based Training for Generalization in Binary Classification Tasks
by: Piland, Jacob, et al.
Published: (2025)
by: Piland, Jacob, et al.
Published: (2025)
AutoSIGHT: Automatic Eye Tracking-based System for Immediate Grading of Human experTise
by: Dowling, Byron, et al.
Published: (2025)
by: Dowling, Byron, et al.
Published: (2025)
SAGE: Saliency-Guided Contrastive Embeddings
by: Crum, Colton R., et al.
Published: (2025)
by: Crum, Colton R., et al.
Published: (2025)
Grains of Saliency: Optimizing Saliency-based Training of Biometric Attack Detection Models
by: Crum, Colton R., et al.
Published: (2024)
by: Crum, Colton R., et al.
Published: (2024)
When Humans Judge Irises: Pupil Size Normalization as an Aid and Synthetic Irises as a Challenge
by: Mitcheff, Mahsa, et al.
Published: (2026)
by: Mitcheff, Mahsa, et al.
Published: (2026)
Unifying Biomedical Vision-Language Expertise: Towards a Generalist Foundation Model via Multi-CLIP Knowledge Distillation
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Saliency-Guided Training for Fingerprint Presentation Attack Detection
by: Webster, Samuel, et al.
Published: (2025)
by: Webster, Samuel, et al.
Published: (2025)
Training Better Deep Learning Models Using Human Saliency
by: Boyd, Aidan, et al.
Published: (2024)
by: Boyd, Aidan, et al.
Published: (2024)
EyeFound: A Multimodal Generalist Foundation Model for Ophthalmic Imaging
by: Shi, Danli, et al.
Published: (2024)
by: Shi, Danli, et al.
Published: (2024)
FaceSaliencyAug: Mitigating Geographic, Gender and Stereotypical Biases via Saliency-Based Data Augmentation
by: Kumar, Teerath, et al.
Published: (2024)
by: Kumar, Teerath, et al.
Published: (2024)
IdentiFace : A VGG Based Multimodal Facial Biometric System
by: Rabea, Mahmoud, et al.
Published: (2024)
by: Rabea, Mahmoud, et al.
Published: (2024)
Reflect to Inform: Boosting Multimodal Reasoning via Information-Gain-Driven Verification
by: Lv, Shuai, et al.
Published: (2026)
by: Lv, Shuai, et al.
Published: (2026)
Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs
by: Fu, Xingyu, et al.
Published: (2025)
by: Fu, Xingyu, et al.
Published: (2025)
Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency
by: Wen, Ziqi, et al.
Published: (2026)
by: Wen, Ziqi, et al.
Published: (2026)
SiNC+: Adaptive Camera-Based Vitals with Unsupervised Learning of Periodic Signals
by: Speth, Jeremy, et al.
Published: (2024)
by: Speth, Jeremy, et al.
Published: (2024)
Is Visual Realism Enough? Evaluating Gait Biometric Fidelity in Generative AI Human Animation
by: DeAndres-Tame, Ivan, et al.
Published: (2025)
by: DeAndres-Tame, Ivan, et al.
Published: (2025)
Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning
by: LASA Team, et al.
Published: (2025)
by: LASA Team, et al.
Published: (2025)
Image Generators are Generalist Vision Learners
by: Gabeur, Valentin, et al.
Published: (2026)
by: Gabeur, Valentin, et al.
Published: (2026)
Vision Generalist Model: A Survey
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Alpha Divergence Losses for Biometric Verification
by: Koutsianos, Dimitrios, et al.
Published: (2025)
by: Koutsianos, Dimitrios, et al.
Published: (2025)
Salience Adjustment for Context-Based Emotion Recognition
by: Han, Bin, et al.
Published: (2025)
by: Han, Bin, et al.
Published: (2025)
Normal-Abnormal Guided Generalist Anomaly Detection
by: Wang, Yuexin, et al.
Published: (2025)
by: Wang, Yuexin, et al.
Published: (2025)
Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward
by: Islam, Muhammad, et al.
Published: (2025)
by: Islam, Muhammad, et al.
Published: (2025)
Saliency Guided Longitudinal Medical Visual Question Answering
by: Wu, Jialin, et al.
Published: (2025)
by: Wu, Jialin, et al.
Published: (2025)
Towards Learning a Generalist Model for Embodied Navigation
by: Zheng, Duo, et al.
Published: (2023)
by: Zheng, Duo, et al.
Published: (2023)
Learning User Embeddings from Human Gaze for Personalised Saliency Prediction
by: Strohm, Florian, et al.
Published: (2024)
by: Strohm, Florian, et al.
Published: (2024)
OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks
by: Hu, Wenbo, et al.
Published: (2026)
by: Hu, Wenbo, et al.
Published: (2026)
Towards Billion-scale Multi-modal Biometric Search
by: Koner, Arka, et al.
Published: (2026)
by: Koner, Arka, et al.
Published: (2026)
Model Compression Techniques in Biometrics Applications: A Survey
by: Caldeira, Eduarda, et al.
Published: (2024)
by: Caldeira, Eduarda, et al.
Published: (2024)
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
by: Sony, Redwan, et al.
Published: (2025)
by: Sony, Redwan, et al.
Published: (2025)
Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
by: Guo, Garvin, et al.
Published: (2026)
by: Guo, Garvin, et al.
Published: (2026)
SUM: Saliency Unification through Mamba for Visual Attention Modeling
by: Hosseini, Alireza, et al.
Published: (2024)
by: Hosseini, Alireza, et al.
Published: (2024)
GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis
by: Mitra, Sirshapan, et al.
Published: (2025)
by: Mitra, Sirshapan, et al.
Published: (2025)
Flexible ViG: Learning the Self-Saliency for Flexible Object Recognition
by: Zuo, Lin, et al.
Published: (2024)
by: Zuo, Lin, et al.
Published: (2024)
Forward Learning for Gradient-based Black-box Saliency Map Generation
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
Saliency Suppressed, Semantics Surfaced: Visual Transformations in Neural Networks and the Brain
by: Opiełka, Gustaw, et al.
Published: (2024)
by: Opiełka, Gustaw, et al.
Published: (2024)
MENTOR: Human Perception-Guided Pretraining for Increased Generalization
by: Crum, Colton R., et al.
Published: (2023)
by: Crum, Colton R., et al.
Published: (2023)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025)
by: Rekimoto, Jun
Published: (2025)
What is the Visual Cognition Gap between Humans and Multimodal LLMs?
by: Cao, Xu, et al.
Published: (2024)
by: Cao, Xu, et al.
Published: (2024)
Similar Items
-
VISER: Visually-Informed System for Enhanced Robustness in Open-Set Iris Presentation Attack Detection
by: Dowling, Byron, et al.
Published: (2026) -
Divisive Decisions: Improving Salience-Based Training for Generalization in Binary Classification Tasks
by: Piland, Jacob, et al.
Published: (2025) -
AutoSIGHT: Automatic Eye Tracking-based System for Immediate Grading of Human experTise
by: Dowling, Byron, et al.
Published: (2025) -
SAGE: Saliency-Guided Contrastive Embeddings
by: Crum, Colton R., et al.
Published: (2025) -
Grains of Saliency: Optimizing Saliency-based Training of Biometric Attack Detection Models
by: Crum, Colton R., et al.
Published: (2024)