Saved in:
| Main Authors: | Abir, Mushfiqur Rahman, Hosain, Md. Tanzib, Abdullah-Al-Jubair, Md., Mridha, M. F. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.16807 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
by: Hossain, Md. Mithun, et al.
Published: (2025)
by: Hossain, Md. Mithun, et al.
Published: (2025)
Pattern Recognition Tasks with Personalized Federated Learning
by: Rahman, Md. Arifur, et al.
Published: (2026)
by: Rahman, Md. Arifur, et al.
Published: (2026)
MIC: Medical Image Classification Using Chest X-ray (COVID-19 and Pneumonia) Dataset with the Help of CNN and Customized CNN
by: Fahad, Nafiz, et al.
Published: (2024)
by: Fahad, Nafiz, et al.
Published: (2024)
From Explanations to Architecture: Explainability-Driven CNN Refinement for Brain Tumor Classification in MRI
by: Gupta, Rajan Das, et al.
Published: (2025)
by: Gupta, Rajan Das, et al.
Published: (2025)
A Survey on 3D Egocentric Human Pose Estimation
by: Azam, Md Mushfiqur, et al.
Published: (2024)
by: Azam, Md Mushfiqur, et al.
Published: (2024)
ConMatFormer: A Multi-attention and Transformer Integrated ConvNext based Deep Learning Model for Enhanced Diabetic Foot Ulcer Classification
by: Rifat, Raihan Ahamed, et al.
Published: (2025)
by: Rifat, Raihan Ahamed, et al.
Published: (2025)
Soybean Disease Detection via Interpretable Hybrid CNN-GNN: Integrating MobileNetV2 and GraphSAGE with Cross-Modal Attention
by: Jahin, Md Abrar, et al.
Published: (2025)
by: Jahin, Md Abrar, et al.
Published: (2025)
AG-EgoPose: Leveraging Action-Guided Motion and Kinematic Joint Encoding for Egocentric 3D Pose Estimation
by: Azam, Md Mushfiqur, et al.
Published: (2026)
by: Azam, Md Mushfiqur, et al.
Published: (2026)
Vision Transformers for End-to-End Quark-Gluon Jet Classification from Calorimeter Images
by: Jahin, Md Abrar, et al.
Published: (2025)
by: Jahin, Md Abrar, et al.
Published: (2025)
MobilePlantViT: A Mobile-friendly Hybrid ViT for Generalized Plant Disease Image Classification
by: Tonmoy, Moshiur Rahman, et al.
Published: (2025)
by: Tonmoy, Moshiur Rahman, et al.
Published: (2025)
Interpretable Dynamic Graph Neural Networks for Small Occluded Object Detection and Tracking
by: Soudeep, Shahriar, et al.
Published: (2024)
by: Soudeep, Shahriar, et al.
Published: (2024)
Can Multi-turn Self-refined Single Agent LMs with Retrieval Solve Hard Coding Problems?
by: Hosain, Md Tanzib, et al.
Published: (2025)
by: Hosain, Md Tanzib, et al.
Published: (2025)
BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla
by: Islam, Ariful, et al.
Published: (2025)
by: Islam, Ariful, et al.
Published: (2025)
A Methodological and Structural Review of Hand Gesture Recognition Across Diverse Data Modalities
by: Shin, Jungpil, et al.
Published: (2024)
by: Shin, Jungpil, et al.
Published: (2024)
From Image to Language: A Critical Analysis of Visual Question Answering (VQA) Approaches, Challenges, and Opportunities
by: Ishmam, Md Farhan, et al.
Published: (2023)
by: Ishmam, Md Farhan, et al.
Published: (2023)
GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection
by: Mia, Md Sohag, et al.
Published: (2025)
by: Mia, Md Sohag, et al.
Published: (2025)
DyCAF-Net: Dynamic Class-Aware Fusion Network
by: Jahin, Md Abrar, et al.
Published: (2025)
by: Jahin, Md Abrar, et al.
Published: (2025)
Counting Through Occlusion: Framework for Open World Amodal Counting
by: Arib, Safaeid Hossain, et al.
Published: (2025)
by: Arib, Safaeid Hossain, et al.
Published: (2025)
MosquitoFusion: A Multiclass Dataset for Real-Time Detection of Mosquitoes, Swarms, and Breeding Sites Using Deep Learning
by: Sayeedi, Md. Faiyaz Abdullah, et al.
Published: (2024)
by: Sayeedi, Md. Faiyaz Abdullah, et al.
Published: (2024)
Less Is More? Selective Visual Attention to High-Importance Regions for Multimodal Radiology Summarization
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2026)
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2026)
Improving Fine-Grained Rice Leaf Disease Detection via Angular-Compactness Dual Loss Learning
by: Mia, Md. Rokon, et al.
Published: (2026)
by: Mia, Md. Rokon, et al.
Published: (2026)
Multimodal Programming in Computer Science with Interactive Assistance Powered by Large Language Model
by: Gupta, Rajan Das, et al.
Published: (2025)
by: Gupta, Rajan Das, et al.
Published: (2025)
MF-GCN: A Multi-Frequency Graph Convolutional Network for Tri-Modal Depression Detection Using Eye-Tracking, Facial, and Acoustic Features
by: Rahman, Sejuti, et al.
Published: (2025)
by: Rahman, Sejuti, et al.
Published: (2025)
A CNN-Based Malaria Diagnosis from Blood Cell Images with SHAP and LIME Explainability
by: Abir, Md. Ismiel Hossen, et al.
Published: (2025)
by: Abir, Md. Ismiel Hossen, et al.
Published: (2025)
Seam Carving as Feature Pooling in CNN
by: Jubair, Mohammad Imrul
Published: (2024)
by: Jubair, Mohammad Imrul
Published: (2024)
Foundation Model-Powered 3D Few-Shot Class Incremental Learning via Training-free Adaptor
by: Ahmadi, Sahar, et al.
Published: (2024)
by: Ahmadi, Sahar, et al.
Published: (2024)
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
by: Hosain, Md Tanzib, et al.
Published: (2025)
by: Hosain, Md Tanzib, et al.
Published: (2025)
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models
by: Samin, Niamul Hassan, et al.
Published: (2026)
by: Samin, Niamul Hassan, et al.
Published: (2026)
MK-UNet: Multi-kernel Lightweight CNN for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models
by: Rahman, Md Ashikur, et al.
Published: (2026)
by: Rahman, Md Ashikur, et al.
Published: (2026)
LoMix: Learnable Weighted Multi-Scale Logits Mixing for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
Edge-Native Digitization of Handwritten Marksheets: A Hybrid Heuristic-Deep Learning Framework
by: Hossain, Md. Irtiza, et al.
Published: (2025)
by: Hossain, Md. Irtiza, et al.
Published: (2025)
VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites
by: Islam, Md. Adnanul, et al.
Published: (2025)
by: Islam, Md. Adnanul, et al.
Published: (2025)
Deep Fusion Model for Brain Tumor Classification Using Fine-Grained Gradient Preservation
by: Islam, Niful, et al.
Published: (2024)
by: Islam, Niful, et al.
Published: (2024)
PULSAR: Graph based Positive Unlabeled Learning with Multi Stream Adaptive Convolutions for Parkinson's Disease Recognition
by: Alam, Md. Zarif Ul, et al.
Published: (2023)
by: Alam, Md. Zarif Ul, et al.
Published: (2023)
Detecting Unauthorized Vehicles using Deep Learning for Smart Cities: A Case Study on Bangladesh
by: Sukanto, Sudipto Das, et al.
Published: (2025)
by: Sukanto, Sudipto Das, et al.
Published: (2025)
Comparative Performance Analysis of Transformer-Based Pre-Trained Models for Detecting Keratoconus Disease
by: Ahmed, Nayeem, et al.
Published: (2024)
by: Ahmed, Nayeem, et al.
Published: (2024)
Less Is More: An Explainable AI Framework for Lightweight Malaria Classification
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
A CNN Approach to Automated Detection and Classification of Brain Tumors
by: Hasan, Md. Zahid, et al.
Published: (2025)
by: Hasan, Md. Zahid, et al.
Published: (2025)
ILASH: A Predictive Neural Architecture Search Framework for Multi-Task Applications
by: Rahman, Md Hafizur, et al.
Published: (2024)
by: Rahman, Md Hafizur, et al.
Published: (2024)
Similar Items
-
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
by: Hossain, Md. Mithun, et al.
Published: (2025) -
Pattern Recognition Tasks with Personalized Federated Learning
by: Rahman, Md. Arifur, et al.
Published: (2026) -
MIC: Medical Image Classification Using Chest X-ray (COVID-19 and Pneumonia) Dataset with the Help of CNN and Customized CNN
by: Fahad, Nafiz, et al.
Published: (2024) -
From Explanations to Architecture: Explainability-Driven CNN Refinement for Brain Tumor Classification in MRI
by: Gupta, Rajan Das, et al.
Published: (2025) -
A Survey on 3D Egocentric Human Pose Estimation
by: Azam, Md Mushfiqur, et al.
Published: (2024)