Understanding Cross Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Changqing, Qin, Huafeng, El-Yacoubi, Mounim A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hybrid Transformer for Early Alzheimer's Detection: Integration of Handwriting-Based 2D Images and 1D Signal Features
by: Gong, Changqing, et al.
Published: (2024)
by: Gong, Changqing, et al.
Published: (2024)
Neural Architecture Search based Global-local Vision Mamba for Palm-Vein Recognition
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
Adversarial Masking Contrastive Learning for vein recognition
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
Adversarial AutoMixup
by: Qin, Huafeng, et al.
Published: (2023)
by: Qin, Huafeng, et al.
Published: (2023)
SUMix: Mixup with Semantic and Uncertain Information
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
EM-DARTS: Hierarchical Differentiable Architecture Search for Eye Movement Recognition
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
Aesthetic Assessment of Chinese Handwritings Based on Vision Language Models
by: Zheng, Chen, et al.
Published: (2026)
by: Zheng, Chen, et al.
Published: (2026)
ScriptViT: Vision Transformer-Based Personalized Handwriting Generation
by: Acharya, Sajjan, et al.
Published: (2025)
by: Acharya, Sajjan, et al.
Published: (2025)
Vision-Language Model Based Handwriting Verification
by: Chauhan, Mihir, et al.
Published: (2024)
by: Chauhan, Mihir, et al.
Published: (2024)
EmMixformer: Mix transformer for eye movement recognition
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
Relax DARTS: Relaxing the Constraints of Differentiable Architecture Search for Eye Movement Recognition
by: Zhu, Hongyu, et al.
Published: (2024)
by: Zhu, Hongyu, et al.
Published: (2024)
Understanding Retrieval-Augmented Task Adaptation for Vision-Language Models
by: Ming, Yifei, et al.
Published: (2024)
by: Ming, Yifei, et al.
Published: (2024)
MsMemoryGAN: A Multi-scale Memory GAN for Palm-vein Adversarial Purification
by: Qin, Huafeng, et al.
Published: (2024)
by: Qin, Huafeng, et al.
Published: (2024)
StarLKNet: Star Mixup with Large Kernel Networks for Palm Vein Identification
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
Representing Online Handwriting for Recognition in Large Vision-Language Models
by: Fadeeva, Anastasiia, et al.
Published: (2024)
by: Fadeeva, Anastasiia, et al.
Published: (2024)
InkSight: Offline-to-Online Handwriting Conversion by Teaching Vision-Language Models to Read and Write
by: Mitrevski, Blagoj, et al.
Published: (2024)
by: Mitrevski, Blagoj, et al.
Published: (2024)
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
by: Baechler, Gilles, et al.
Published: (2024)
by: Baechler, Gilles, et al.
Published: (2024)
Cross-Task Attack: A Self-Supervision Generative Framework Based on Attention Shift
by: Zeng, Qingyuan, et al.
Published: (2024)
by: Zeng, Qingyuan, et al.
Published: (2024)
AdaRing: Towards Ultra-Light Vision-Language Adaptation via Cross-Layer Tensor Ring Decomposition
by: Huang, Ying, et al.
Published: (2025)
by: Huang, Ying, et al.
Published: (2025)
Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator
by: Qin, Luozheng, et al.
Published: (2026)
by: Qin, Luozheng, et al.
Published: (2026)
FreshMem: Brain-Inspired Frequency-Space Hybrid Memory for Streaming Video Understanding
by: Li, Kangcong, et al.
Published: (2026)
by: Li, Kangcong, et al.
Published: (2026)
Cross-Layer Vision Smoothing: Enhancing Visual Understanding via Sustained Focus on Key Objects in Large Vision-Language Models
by: Zhao, Jianfei, et al.
Published: (2025)
by: Zhao, Jianfei, et al.
Published: (2025)
Probing Vision-Language Understanding through the Visual Entailment Task: promises and pitfalls
by: Pitta, Elena, et al.
Published: (2025)
by: Pitta, Elena, et al.
Published: (2025)
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding
by: Mathew, Athul M., et al.
Published: (2025)
by: Mathew, Athul M., et al.
Published: (2025)
Polar Coordinate-Based 2D Pose Prior with Neural Distance Field
by: Gan, Qi, et al.
Published: (2025)
by: Gan, Qi, et al.
Published: (2025)
Exploring Task-Level Optimal Prompts for Visual In-Context Learning
by: Zhu, Yan, et al.
Published: (2025)
by: Zhu, Yan, et al.
Published: (2025)
Are Unified Vision-Language Models Necessary: Generalization Across Understanding and Generation
by: Zhang, Jihai, et al.
Published: (2025)
by: Zhang, Jihai, et al.
Published: (2025)
PEEK: Picking Essential frames via Efficient Knowledge distillation
by: Steunou, Killian, et al.
Published: (2026)
by: Steunou, Killian, et al.
Published: (2026)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
by: Ma, Huan, et al.
Published: (2024)
by: Ma, Huan, et al.
Published: (2024)
Chain-of-Adaptation: Surgical Vision-Language Adaptation with Reinforcement Learning
by: Li, Jiajie, et al.
Published: (2026)
by: Li, Jiajie, et al.
Published: (2026)
Self-Supervised Learning Based Handwriting Verification
by: Chauhan, Mihir, et al.
Published: (2024)
by: Chauhan, Mihir, et al.
Published: (2024)
SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
by: Wu, Xue, et al.
Published: (2026)
by: Wu, Xue, et al.
Published: (2026)
General Scene Adaptation for Vision-and-Language Navigation
by: Hong, Haodong, et al.
Published: (2025)
by: Hong, Haodong, et al.
Published: (2025)
Evolving Prompt Adaptation for Vision-Language Models
by: Zhang, Enming, et al.
Published: (2026)
by: Zhang, Enming, et al.
Published: (2026)
BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation
by: Tong, Huanyang, et al.
Published: (2026)
by: Tong, Huanyang, et al.
Published: (2026)
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
by: Kong, Fei, et al.
Published: (2025)
by: Kong, Fei, et al.
Published: (2025)
HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation
by: Lin, Tianwei, et al.
Published: (2025)
by: Lin, Tianwei, et al.
Published: (2025)
Telling Human and Machine Handwriting Apart
by: Leiva, Luis A., et al.
Published: (2026)
by: Leiva, Luis A., et al.
Published: (2026)
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition
by: Bhatia, Gagan, et al.
Published: (2024)
by: Bhatia, Gagan, et al.
Published: (2024)
When Deep Learning Fails: Limitations of Recurrent Models on Stroke-Based Handwriting for Alzheimer's Disease Detection
by: Nardone, Emanuele, et al.
Published: (2025)
by: Nardone, Emanuele, et al.
Published: (2025)
Similar Items
-
Hybrid Transformer for Early Alzheimer's Detection: Integration of Handwriting-Based 2D Images and 1D Signal Features
by: Gong, Changqing, et al.
Published: (2024) -
Neural Architecture Search based Global-local Vision Mamba for Palm-Vein Recognition
by: Qin, Huafeng, et al.
Published: (2024) -
Adversarial Masking Contrastive Learning for vein recognition
by: Qin, Huafeng, et al.
Published: (2024) -
Adversarial AutoMixup
by: Qin, Huafeng, et al.
Published: (2023) -
SUMix: Mixup with Semantic and Uncertain Information
by: Qin, Huafeng, et al.
Published: (2024)