PE-CLIP: A Parameter-Efficient Fine-Tuning of Vision Language Models for Dynamic Facial Expression Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Saadi, Ibtissam, Hadid, Abdenour, Cunningham, Douglas W., Taleb-Ahmed, Abdelmalik, Hillali, Yassin El |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Shuffle Vision Transformer: Lightweight, Fast and Efficient Recognition of Driver Facial Expression
by: Saadi, Ibtissam, et al.
Published: (2024)
by: Saadi, Ibtissam, et al.
Published: (2024)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
by: Keita, Mamadou, et al.
Published: (2025)
by: Keita, Mamadou, et al.
Published: (2025)
Bi-LORA: A Vision-Language Approach for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
FIDAVL: Fake Image Detection and Attribution using Vision-Language Model
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
Bi‐ LORA : A Vision‐Language Approach for Synthetic Image Detection
by: Mamadou Keita, et al.
Published: (2025)
by: Mamadou Keita, et al.
Published: (2025)
SPARK-IL: Spectral Retrieval-Augmented RAG for Knowledge-driven Deepfake Detection via Incremental Learning
by: Eutamene, Hessen Bougueffa, et al.
Published: (2026)
by: Eutamene, Hessen Bougueffa, et al.
Published: (2026)
Conflict-Aware Multimodal Fusion for Ambivalence and Hesitancy Recognition
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
RAVID: Retrieval-Augmented Visual Detection: A Knowledge-Driven Approach for AI-Generated Image Identification
by: Keita, Mamadou, et al.
Published: (2025)
by: Keita, Mamadou, et al.
Published: (2025)
Can Visual Mamba Improve AI-Generated Image Detection? An In-Depth Investigation
by: Keita, Mamadou, et al.
Published: (2026)
by: Keita, Mamadou, et al.
Published: (2026)
SAViL-Det: Semantic-Aware Vision-Language Model for Multi-Script Text Detection
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
Recent Advances in Medical Imaging Segmentation: A Survey
by: Bougourzi, Fares, et al.
Published: (2025)
by: Bougourzi, Fares, et al.
Published: (2025)
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
Decoding Matters: Efficient Mamba-Based Decoder with Distribution-Aware Deep Supervision for Medical Image Segmentation
by: Bougourzi, Fares, et al.
Published: (2026)
by: Bougourzi, Fares, et al.
Published: (2026)
When Geoscience Meets Generative AI and Large Language Models: Foundations, Trends, and Future Challenges
by: Hadid, Abdenour, et al.
Published: (2024)
by: Hadid, Abdenour, et al.
Published: (2024)
CSIM: A Copula-based similarity index sensitive to local changes for Image quality assessment
by: Ghazouali, Safouane El, et al.
Published: (2024)
by: Ghazouali, Safouane El, et al.
Published: (2024)
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
Some Applications envisaged for the new generation of communications networks 6G
by: Latreche, Sofiane, et al.
Published: (2025)
by: Latreche, Sofiane, et al.
Published: (2025)
Terahertz for Radar applications and Wireless Communication
by: Latreche, Sofiane, et al.
Published: (2025)
by: Latreche, Sofiane, et al.
Published: (2025)
TempoKGAT: A Novel Graph Attention Network Approach for Temporal Graph Analysis
by: Sasal, Lena, et al.
Published: (2024)
by: Sasal, Lena, et al.
Published: (2024)
Knowledge-Based Convolutional Neural Network for the Simulation and Prediction of Two-Phase Darcy Flows
by: Elabid, Zakaria, et al.
Published: (2024)
by: Elabid, Zakaria, et al.
Published: (2024)
When geoscience meets generative AI and large language models: Foundations, trends, and future challenges
by: Abdenour Hadid, et al.
Published: (2024)
by: Abdenour Hadid, et al.
Published: (2024)
Sim-CLIP: Unsupervised Siamese Adversarial Fine-Tuning for Robust and Semantically-Rich Vision-Language Models
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
by: Alyami, Sarah, et al.
Published: (2025)
by: Alyami, Sarah, et al.
Published: (2025)
D-TrAttUnet: Toward Hybrid CNN-Transformer Architecture for Generic and Subtle Segmentation in Medical Images
by: Bougourzi, Fares, et al.
Published: (2024)
by: Bougourzi, Fares, et al.
Published: (2024)
Electrostatic Force Regularization for Neural Structured Pruning
by: Ferdi, Abdesselam, et al.
Published: (2024)
by: Ferdi, Abdesselam, et al.
Published: (2024)
A$^{3}$lign-DFER: Pioneering Comprehensive Dynamic Affective Alignment for Dynamic Facial Expression Recognition with CLIP
by: Tao, Zeng, et al.
Published: (2024)
by: Tao, Zeng, et al.
Published: (2024)
Recognition of Facial Expressions Using Vision Transformer
by: Paula Ivone Rodríguez-Azar
Published: (2022)
by: Paula Ivone Rodríguez-Azar
Published: (2022)
TG-PhyNN: An Enhanced Physically-Aware Graph Neural Network framework for forecasting Spatio-Temporal Data
by: Elabid, Zakaria, et al.
Published: (2024)
by: Elabid, Zakaria, et al.
Published: (2024)
Robust Dynamic Facial Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Prompting Visual-Language Models for Dynamic Facial Expression Recognition
by: Zhao, Zengqun, et al.
Published: (2023)
by: Zhao, Zengqun, et al.
Published: (2023)
MER-CLIP: AU-Guided Vision-Language Alignment for Micro-Expression Recognition
by: Liu, Shifeng, et al.
Published: (2025)
by: Liu, Shifeng, et al.
Published: (2025)
CS3D: An Efficient Facial Expression Recognition via Event Vision
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Rethinking Attention Gated with Hybrid Dual Pyramid Transformer-CNN for Generalized Segmentation in Medical Imaging
by: Bougourzi, Fares, et al.
Published: (2024)
by: Bougourzi, Fares, et al.
Published: (2024)
Vision-Language Model Fine-Tuning via Simple Parameter-Efficient Modification
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Mixed Text Recognition with Efficient Parameter Fine-Tuning and Transformer
by: Chang, Da, et al.
Published: (2024)
by: Chang, Da, et al.
Published: (2024)
Dynamic Resolution Guidance for Facial Expression Recognition
by: Wang, Songpan, et al.
Published: (2024)
by: Wang, Songpan, et al.
Published: (2024)
SegDT: A Diffusion Transformer-Based Segmentation Model for Medical Imaging
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
Similar Items
-
Shuffle Vision Transformer: Lightweight, Fast and Efficient Recognition of Driver Facial Expression
by: Saadi, Ibtissam, et al.
Published: (2024) -
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024) -
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
by: Keita, Mamadou, et al.
Published: (2025) -
Bi-LORA: A Vision-Language Approach for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024) -
FIDAVL: Fake Image Detection and Attribution using Vision-Language Model
by: Keita, Mamadou, et al.
Published: (2024)