An Evaluation of a Visual Question Answering Strategy for Zero-shot Facial Expression Recognition in Still Images
Fuente:
arXiv
Saved in:
| Main Authors: | Castrillón-Santana, Modesto, Santana, Oliverio J, Freire-Obregón, David, Hernández-Sosa, Daniel, Lorenzo-Navarro, Javier |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Modeling Cultural Bias in Facial Expression Recognition with Adaptive Agents
by: Freire-Obregón, David, et al.
Published: (2025)
by: Freire-Obregón, David, et al.
Published: (2025)
Predicting Soccer Penalty Kick Direction Using Human Action Recognition
by: Freire-Obregón, David, et al.
Published: (2025)
by: Freire-Obregón, David, et al.
Published: (2025)
Look, Listen, and Answer: Overcoming Biases for Audio-Visual Question Answering
by: Ma, Jie, et al.
Published: (2024)
by: Ma, Jie, et al.
Published: (2024)
Biased Heritage: How Datasets Shape Models in Facial Expression Recognition
by: Dominguez-Catena, Iris, et al.
Published: (2025)
by: Dominguez-Catena, Iris, et al.
Published: (2025)
An Empirical Study for Representations of Videos in Video Question Answering via MLLMs
by: Li, Zhi, et al.
Published: (2025)
by: Li, Zhi, et al.
Published: (2025)
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
Towards Bi-Hemispheric Emotion Mapping through EEG: A Dual-Stream Neural Network Approach
by: Freire-Obregón, David, et al.
Published: (2024)
by: Freire-Obregón, David, et al.
Published: (2024)
A Large-Scale Re-identification Analysis in Sporting Scenarios: the Betrayal of Reaching a Critical Point
by: Freire-Obregón, David, et al.
Published: (2023)
by: Freire-Obregón, David, et al.
Published: (2023)
Robust Visual Question Answering: Datasets, Methods, and Future Challenges
by: Ma, Jie, et al.
Published: (2023)
by: Ma, Jie, et al.
Published: (2023)
Cinéaste: A Fine-grained Contextual Movie Question Answering Benchmark
by: Shah, Nisarg A., et al.
Published: (2025)
by: Shah, Nisarg A., et al.
Published: (2025)
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
by: Karimi, Ehsan, et al.
Published: (2025)
by: Karimi, Ehsan, et al.
Published: (2025)
Co-Speech Gesture and Facial Expression Generation for Non-Photorealistic 3D Characters
by: Omine, Taisei, et al.
Published: (2025)
by: Omine, Taisei, et al.
Published: (2025)
Step-CoT: Stepwise Visual Chain-of-Thought for Medical Visual Question Answering
by: Fan, Lin, et al.
Published: (2026)
by: Fan, Lin, et al.
Published: (2026)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
by: Yim, Wen-wai, et al.
Published: (2025)
by: Yim, Wen-wai, et al.
Published: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
Mixture of Rationale: Multi-Modal Reasoning Mixture for Visual Question Answering
by: Li, Tao, et al.
Published: (2024)
by: Li, Tao, et al.
Published: (2024)
The Quest for Visual Understanding: A Journey Through the Evolution of Visual Question Answering
by: Pandey, Anupam, et al.
Published: (2025)
by: Pandey, Anupam, et al.
Published: (2025)
SUN Team's Contribution to ABAW 2024 Competition: Audio-visual Valence-Arousal Estimation and Expression Recognition
by: Dresvyanskiy, Denis, et al.
Published: (2024)
by: Dresvyanskiy, Denis, et al.
Published: (2024)
Product Review Based on Optimized Facial Expression Detection
by: Chaugule, Vikrant, et al.
Published: (2026)
by: Chaugule, Vikrant, et al.
Published: (2026)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
by: Li, Huibin, et al.
Published: (2025)
by: Li, Huibin, et al.
Published: (2025)
NAAQA: A Neural Architecture for Acoustic Question Answering
by: Abdelnour, Jerome, et al.
Published: (2021)
by: Abdelnour, Jerome, et al.
Published: (2021)
NeuroGaze-Distill: Brain-informed Distillation and Depression-Inspired Geometric Priors for Robust Facial Emotion Recognition
by: Li, Zilin, et al.
Published: (2025)
by: Li, Zilin, et al.
Published: (2025)
Facial Emotion Recognition does not detect feeling unsafe in automated driving
by: van Elburg, Abel, et al.
Published: (2025)
by: van Elburg, Abel, et al.
Published: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
by: Wang, Yiming, et al.
Published: (2026)
by: Wang, Yiming, et al.
Published: (2026)
Explainable AI for Analyzing Person-Specific Patterns in Facial Recognition Tasks
by: Borsukiewicz, Paweł Jakub, et al.
Published: (2025)
by: Borsukiewicz, Paweł Jakub, et al.
Published: (2025)
Emo3D: Metric and Benchmarking Dataset for 3D Facial Expression Generation from Emotion Description
by: Dehghani, Mahshid, et al.
Published: (2024)
by: Dehghani, Mahshid, et al.
Published: (2024)
Tri-VQA: Triangular Reasoning Medical Visual Question Answering for Multi-Attribute Analysis
by: Fan, Lin, et al.
Published: (2024)
by: Fan, Lin, et al.
Published: (2024)
Visual Graph Question Answering with ASP and LLMs for Language Parsing
by: Bauer, Jakob Johannes, et al.
Published: (2025)
by: Bauer, Jakob Johannes, et al.
Published: (2025)
Complex Facial Expression Recognition Using Deep Knowledge Distillation of Basic Features
by: Maiden, Angus, et al.
Published: (2023)
by: Maiden, Angus, et al.
Published: (2023)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
by: Li, Xinting, et al.
Published: (2023)
by: Li, Xinting, et al.
Published: (2023)
Distinguishing Visually Similar Actions: Prompt-Guided Semantic Prototype Modulation for Few-Shot Action Recognition
by: Li, Xiaoyang, et al.
Published: (2025)
by: Li, Xiaoyang, et al.
Published: (2025)
Facial Attribute Based Text Guided Face Anonymization
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
A Comparative Analysis of Recurrent and Attention Architectures for Isolated Sign Language Recognition
by: Alishzade, Nigar, et al.
Published: (2025)
by: Alishzade, Nigar, et al.
Published: (2025)
MSPCaps: A Multi-Scale Patchify Capsule Network with Cross-Agreement Routing for Visual Recognition
by: Hu, Yudong, et al.
Published: (2025)
by: Hu, Yudong, et al.
Published: (2025)
Image-Based Leopard Seal Recognition: Approaches and Challenges in Current Automated Systems
by: Salazar, Jorge Yero, et al.
Published: (2024)
by: Salazar, Jorge Yero, et al.
Published: (2024)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
by: Panek, Vojtech, et al.
Published: (2026)
by: Panek, Vojtech, et al.
Published: (2026)
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
by: Ghosh, Archishman, et al.
Published: (2026)
by: Ghosh, Archishman, et al.
Published: (2026)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
by: Zhang, Xinyi, et al.
Published: (2026)
by: Zhang, Xinyi, et al.
Published: (2026)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
by: Li, Danyang, et al.
Published: (2025)
by: Li, Danyang, et al.
Published: (2025)
Image-based Facial Rig Inversion
by: Yang, Tianxiang, et al.
Published: (2025)
by: Yang, Tianxiang, et al.
Published: (2025)
Similar Items
-
Modeling Cultural Bias in Facial Expression Recognition with Adaptive Agents
by: Freire-Obregón, David, et al.
Published: (2025) -
Predicting Soccer Penalty Kick Direction Using Human Action Recognition
by: Freire-Obregón, David, et al.
Published: (2025) -
Look, Listen, and Answer: Overcoming Biases for Audio-Visual Question Answering
by: Ma, Jie, et al.
Published: (2024) -
Biased Heritage: How Datasets Shape Models in Facial Expression Recognition
by: Dominguez-Catena, Iris, et al.
Published: (2025) -
An Empirical Study for Representations of Videos in Video Question Answering via MLLMs
by: Li, Zhi, et al.
Published: (2025)