SimpsonsVQA: Enhancing Inquiry-Based Learning with a Tailored Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Huynh, Ngoc Dung, Bouadjenek, Mohamed Reda, Aryal, Sunil, Razzak, Imran, Hacid, Hakim |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual question answering: from early developments to recent advances -- a survey
by: Huynh, Ngoc Dung, et al.
Published: (2025)
by: Huynh, Ngoc Dung, et al.
Published: (2025)
SVLA: A Unified Speech-Vision-Language Assistant with Multimodal Reasoning and Speech Generation
by: Huynh, Ngoc Dung, et al.
Published: (2025)
by: Huynh, Ngoc Dung, et al.
Published: (2025)
Deep Learning for Sports Video Event Detection: Tasks, Datasets, Methods, and Challenges
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
Margin-bounded Confidence Scores for Out-of-Distribution Detection
by: Tamang, Lakpa D., et al.
Published: (2024)
by: Tamang, Lakpa D., et al.
Published: (2024)
Rolling Ball Optimizer: Learning by ironing out loss landscape wrinkles
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
TOTNet: Occlusion-Aware Temporal Tracking for Robust Ball Detection in Sports Videos
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
Training Machine Learning models at the Edge: A Survey
by: Khouas, Aymen Rayane, et al.
Published: (2024)
by: Khouas, Aymen Rayane, et al.
Published: (2024)
Far From Sight, Far From Mind: Inverse Distance Weighting for Graph Federated Recommendation
by: Khouas, Aymen Rayane, et al.
Published: (2025)
by: Khouas, Aymen Rayane, et al.
Published: (2025)
Data Quality in Edge Machine Learning: A State-of-the-Art Survey
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2024)
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2024)
Falcon Perception
by: Bevli, Aviraj, et al.
Published: (2026)
by: Bevli, Aviraj, et al.
Published: (2026)
DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes
by: Al-Mohannadi, Aisha, et al.
Published: (2026)
by: Al-Mohannadi, Aisha, et al.
Published: (2026)
SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models
by: Chaybouti, Sofian, et al.
Published: (2025)
by: Chaybouti, Sofian, et al.
Published: (2025)
Unmasking Gender Bias in Recommendation Systems and Enhancing Category-Aware Fairness
by: Kheya, Tahsin Alamgir, et al.
Published: (2025)
by: Kheya, Tahsin Alamgir, et al.
Published: (2025)
Retinal Lipidomics Associations as Candidate Biomarkers for Cardiovascular Health
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
Memory-Augmented Multimodal LLMs for Surgical VQA via Self-Contained Inquiry
by: Hou, Wenjun, et al.
Published: (2024)
by: Hou, Wenjun, et al.
Published: (2024)
Robust Atypical Mitosis Classification with DenseNet121: Stain-Aware Augmentation and Hybrid Loss for Domain Generalization
by: Dukre, Adinath, et al.
Published: (2025)
by: Dukre, Adinath, et al.
Published: (2025)
The Eye as a Window to Systemic Health: A Survey of Retinal Imaging from Classical Techniques to Oculomics
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
Multi-Focus Temporal Shifting for Precise Event Spotting in Sports Videos
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
Integrated Oculomics and Lipidomics Reveal Microvascular Metabolic Signatures Associated with Cardiovascular Health in a Healthy Cohort
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
DHGCN: Dynamic Hop Graph Convolution Network for Self-Supervised Point Cloud Learning
by: Jiang, Jincen, et al.
Published: (2024)
by: Jiang, Jincen, et al.
Published: (2024)
Advancing Medical Image Segmentation with Mini-Net: A Lightweight Solution Tailored for Efficient Segmentation of Medical Images
by: Javed, Syed, et al.
Published: (2024)
by: Javed, Syed, et al.
Published: (2024)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025)
by: Liu, Xiwei, et al.
Published: (2025)
Interactive Interface For Semantic Segmentation Dataset Synthesis
by: Tran, Ngoc-Do, et al.
Published: (2025)
by: Tran, Ngoc-Do, et al.
Published: (2025)
DeepChest: Dynamic Gradient-Free Task Weighting for Effective Multi-Task Learning in Chest X-ray Classification
by: Mohamed, Youssef, et al.
Published: (2025)
by: Mohamed, Youssef, et al.
Published: (2025)
Vision-Language Models Can't See the Obvious
by: Dahou, Yasser, et al.
Published: (2025)
by: Dahou, Yasser, et al.
Published: (2025)
StackOverflowVQA: Stack Overflow Visual Question Answering Dataset
by: Mirzaei, Motahhare, et al.
Published: (2024)
by: Mirzaei, Motahhare, et al.
Published: (2024)
Region Guided Attention Network for Retinal Vessel Segmentation
by: Javed, Syed, et al.
Published: (2024)
by: Javed, Syed, et al.
Published: (2024)
Disentanglement-Based Equivariant Learning for Compositional VQA
by: Du, Zhou, et al.
Published: (2026)
by: Du, Zhou, et al.
Published: (2026)
Adversary-Robust Graph-Based Learning of WSIs
by: Gheshlaghi, Saba Heidari, et al.
Published: (2024)
by: Gheshlaghi, Saba Heidari, et al.
Published: (2024)
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
TESL-Net: A Transformer-Enhanced CNN for Accurate Skin Lesion Segmentation
by: Iqbal, Shahzaib, et al.
Published: (2024)
by: Iqbal, Shahzaib, et al.
Published: (2024)
The Pursuit of Fairness in Artificial Intelligence Models: A Survey
by: Kheya, Tahsin Alamgir, et al.
Published: (2024)
by: Kheya, Tahsin Alamgir, et al.
Published: (2024)
Enhancing Document VQA Models via Retrieval-Augmented Generation
by: López, Eric, et al.
Published: (2025)
by: López, Eric, et al.
Published: (2025)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
CMSA-Net: Causal Multi-scale Aggregation with Adaptive Multi-source Reference for Video Polyp Segmentation
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
Enhanced Multi-level Features for Very High Resolution Remote Sensing Scene Classification
by: Sitaula, Chiranjibi, et al.
Published: (2023)
by: Sitaula, Chiranjibi, et al.
Published: (2023)
HDRSDR-VQA: A Subjective Video Quality Dataset for HDR and SDR Comparative Evaluation
by: Chen, Bowen, et al.
Published: (2025)
by: Chen, Bowen, et al.
Published: (2025)
The Role of AI in Early Detection of Life-Threatening Diseases: A Retinal Imaging Perspective
by: Khan, Tariq M, et al.
Published: (2025)
by: Khan, Tariq M, et al.
Published: (2025)
A Knowledge-driven Adaptive Collaboration of LLMs for Enhancing Medical Decision-making
by: Wu, Xiao, et al.
Published: (2025)
by: Wu, Xiao, et al.
Published: (2025)
MedMO: Grounding and Understanding Multimodal Large Language Model for Medical Images
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Similar Items
-
Visual question answering: from early developments to recent advances -- a survey
by: Huynh, Ngoc Dung, et al.
Published: (2025) -
SVLA: A Unified Speech-Vision-Language Assistant with Multimodal Reasoning and Speech Generation
by: Huynh, Ngoc Dung, et al.
Published: (2025) -
Deep Learning for Sports Video Event Detection: Tasks, Datasets, Methods, and Challenges
by: Xu, Hao, et al.
Published: (2025) -
Margin-bounded Confidence Scores for Out-of-Distribution Detection
by: Tamang, Lakpa D., et al.
Published: (2024) -
Rolling Ball Optimizer: Learning by ironing out loss landscape wrinkles
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)