TinyVQA: Compact Multimodal Deep Neural Network for Visual Question Answering on Resource-Constrained Devices
Fuente:
arXiv
Saved in:
| Main Authors: | Rashid, Hasib-Al, Sarkar, Argho, Gangopadhyay, Aryya, Rahnemoonfar, Maryam, Mohsenin, Tinoosh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TinyM$^2$Net-V3: Memory-Aware Compressed Multimodal Deep Neural Networks for Sustainable Edge Deployment
by: Rashid, Hasib-Al, et al.
Published: (2024)
by: Rashid, Hasib-Al, et al.
Published: (2024)
Decentralised Resource Sharing in TinyML: Wireless Bilayer Gossip Parallel SGD for Collaborative Learning
by: Bao, Ziyuan, et al.
Published: (2025)
by: Bao, Ziyuan, et al.
Published: (2025)
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
by: Karimi, Ehsan, et al.
Published: (2025)
by: Karimi, Ehsan, et al.
Published: (2025)
SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation
by: Humes, Edward, et al.
Published: (2026)
by: Humes, Edward, et al.
Published: (2026)
EdgeNavMamba: Mamba Optimized Object Detection for Energy Efficient Edge Devices
by: Aalishah, Romina, et al.
Published: (2025)
by: Aalishah, Romina, et al.
Published: (2025)
Energy-Aware FPGA Implementation of Spiking Neural Network with LIF Neurons
by: Ali, Asmer Hamid, et al.
Published: (2024)
by: Ali, Asmer Hamid, et al.
Published: (2024)
Skip-WaveNet: A Wavelet based Multi-scale Architecture to Trace Snow Layers in Radar Echograms
by: Varshney, Debvrat, et al.
Published: (2023)
by: Varshney, Debvrat, et al.
Published: (2023)
ATLASv2: LLM-Guided Adaptive Landmark Acquisition and Navigation on the Edge
by: Walczak, Mikolaj, et al.
Published: (2025)
by: Walczak, Mikolaj, et al.
Published: (2025)
MedMambaLite: Hardware-Aware Mamba for Medical Image Classification
by: Aalishah, Romina, et al.
Published: (2025)
by: Aalishah, Romina, et al.
Published: (2025)
MambaLiteSR: Image Super-Resolution with Low-Rank Mamba using Knowledge Distillation
by: Aalishah, Romina, et al.
Published: (2025)
by: Aalishah, Romina, et al.
Published: (2025)
Enabling On-Device Medical AI Assistants via Input-Driven Saliency Adaptation
by: Kallakurik, Uttej, et al.
Published: (2025)
by: Kallakurik, Uttej, et al.
Published: (2025)
FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
by: Singh, Shubhankar, et al.
Published: (2024)
by: Singh, Shubhankar, et al.
Published: (2024)
VQA-MHUG: A Gaze Dataset to Study Multimodal Neural Attention in Visual Question Answering
by: Sood, Ekta, et al.
Published: (2021)
by: Sood, Ekta, et al.
Published: (2021)
Hilbert-Augmented Reinforcement Learning for Scalable Multi-Robot Coverage and Exploration
by: Gurunathan, Tamil Selvan, et al.
Published: (2026)
by: Gurunathan, Tamil Selvan, et al.
Published: (2026)
BERT-VQA: Visual Question Answering on Plots
by: Vu, Tai, et al.
Published: (2025)
by: Vu, Tai, et al.
Published: (2025)
Visual Robustness Benchmark for Visual Question Answering (VQA)
by: Ishmam, Md Farhan, et al.
Published: (2024)
by: Ishmam, Md Farhan, et al.
Published: (2024)
Graph Neural Networks for Emulation of Finite-Element Ice Dynamics in Greenland and Antarctic Ice Sheets
by: Koo, Younghyun, et al.
Published: (2024)
by: Koo, Younghyun, et al.
Published: (2024)
Multi-branch Spatio-Temporal Graph Neural Network For Efficient Ice Layer Thickness Prediction
by: Liu, Zesheng, et al.
Published: (2024)
by: Liu, Zesheng, et al.
Published: (2024)
Hierarchical Information-sharing Convolutional Neural Network for the Prediction of Arctic Sea Ice Concentration and Velocity
by: Koo, Younghyun, et al.
Published: (2023)
by: Koo, Younghyun, et al.
Published: (2023)
Prediction of Sea Ice Velocity and Concentration in the Arctic Ocean using Physics-informed Neural Network
by: Koo, Younghyun, et al.
Published: (2025)
by: Koo, Younghyun, et al.
Published: (2025)
Learning Spatio-Temporal Patterns of Polar Ice Layers With Physics-Informed Graph Neural Network
by: Liu, Zesheng, et al.
Published: (2024)
by: Liu, Zesheng, et al.
Published: (2024)
Graph Neural Network as Computationally Efficient Emulator of Ice-sheet and Sea-level System Model (ISSM)
by: Koo, Younghyun, et al.
Published: (2024)
by: Koo, Younghyun, et al.
Published: (2024)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Think First, Assign Next (ThiFAN-VQA): A Two-stage Chain-of-Thought Framework for Post-Disaster Damage Assessment
by: Karimi, Ehsan, et al.
Published: (2025)
by: Karimi, Ehsan, et al.
Published: (2025)
K-STEMIT: Knowledge-Informed Spatio-Temporal Efficient Multi-Branch Graph Neural Network for Subsurface Stratigraphy Thickness Estimation from Radar Data
by: Liu, Zesheng, et al.
Published: (2026)
by: Liu, Zesheng, et al.
Published: (2026)
DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes
by: Al-Mohannadi, Aisha, et al.
Published: (2026)
by: Al-Mohannadi, Aisha, et al.
Published: (2026)
A Survey of Post-Quantum Cryptography Support in Cryptographic Libraries
by: Ahmed, Nadeem, et al.
Published: (2025)
by: Ahmed, Nadeem, et al.
Published: (2025)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
by: Zhang, Xiaoman, et al.
Published: (2023)
by: Zhang, Xiaoman, et al.
Published: (2023)
CommVQA: Situating Visual Question Answering in Communicative Contexts
by: Naik, Nandita Shankar, et al.
Published: (2024)
by: Naik, Nandita Shankar, et al.
Published: (2024)
VQA$^2$: Visual Question Answering for Video Quality Assessment
by: Jia, Ziheng, et al.
Published: (2024)
by: Jia, Ziheng, et al.
Published: (2024)
RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering
by: Zhang, Chengyi, et al.
Published: (2026)
by: Zhang, Chengyi, et al.
Published: (2026)
GenAI at the Edge: Comprehensive Survey on Empowering Edge Devices
by: Navardi, Mozhgan, et al.
Published: (2025)
by: Navardi, Mozhgan, et al.
Published: (2025)
AdaDocVQA: Adaptive Framework for Long Document Visual Question Answering in Low-Resource Settings
by: Li, Haoxuan, et al.
Published: (2025)
by: Li, Haoxuan, et al.
Published: (2025)
M$^3$-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering
by: Ma, Jiatong, et al.
Published: (2026)
by: Ma, Jiatong, et al.
Published: (2026)
StackOverflowVQA: Stack Overflow Visual Question Answering Dataset
by: Mirzaei, Motahhare, et al.
Published: (2024)
by: Mirzaei, Motahhare, et al.
Published: (2024)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
by: Vu, Sinh Trong, et al.
Published: (2025)
by: Vu, Sinh Trong, et al.
Published: (2025)
Invited Paper: BitMedViT: Ternary-Quantized Vision Transformer for Medical AI Assistants on the Edge
by: Walczak, Mikolaj, et al.
Published: (2025)
by: Walczak, Mikolaj, et al.
Published: (2025)
RAFT -- A Domain Adaptation Framework for RGB & LiDAR Semantic Segmentation
by: Humes, Edward, et al.
Published: (2025)
by: Humes, Edward, et al.
Published: (2025)
Efficient Federated Finetuning of Tiny Transformers with Resource-Constrained Devices
by: Pfeiffer, Kilian, et al.
Published: (2024)
by: Pfeiffer, Kilian, et al.
Published: (2024)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
Similar Items
-
TinyM$^2$Net-V3: Memory-Aware Compressed Multimodal Deep Neural Networks for Sustainable Edge Deployment
by: Rashid, Hasib-Al, et al.
Published: (2024) -
Decentralised Resource Sharing in TinyML: Wireless Bilayer Gossip Parallel SGD for Collaborative Learning
by: Bao, Ziyuan, et al.
Published: (2025) -
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
by: Karimi, Ehsan, et al.
Published: (2025) -
SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation
by: Humes, Edward, et al.
Published: (2026) -
EdgeNavMamba: Mamba Optimized Object Detection for Energy Efficient Edge Devices
by: Aalishah, Romina, et al.
Published: (2025)