IMVB7t: A Multi-Modal Model for Food Preferences based on Artificially Produced Traits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abir, Mushfiqur Rahman, Hosain, Md. Tanzib, Abdullah-Al-Jubair, Md., Mridha, M. F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
von: Hossain, Md. Mithun, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Mithun, et al.
Veröffentlicht: (2025)
A Survey on 3D Egocentric Human Pose Estimation
von: Azam, Md Mushfiqur, et al.
Veröffentlicht: (2024)
von: Azam, Md Mushfiqur, et al.
Veröffentlicht: (2024)
MIC: Medical Image Classification Using Chest X-ray (COVID-19 and Pneumonia) Dataset with the Help of CNN and Customized CNN
von: Fahad, Nafiz, et al.
Veröffentlicht: (2024)
von: Fahad, Nafiz, et al.
Veröffentlicht: (2024)
From Explanations to Architecture: Explainability-Driven CNN Refinement for Brain Tumor Classification in MRI
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
ConMatFormer: A Multi-attention and Transformer Integrated ConvNext based Deep Learning Model for Enhanced Diabetic Foot Ulcer Classification
von: Rifat, Raihan Ahamed, et al.
Veröffentlicht: (2025)
von: Rifat, Raihan Ahamed, et al.
Veröffentlicht: (2025)
Pattern Recognition Tasks with Personalized Federated Learning
von: Rahman, Md. Arifur, et al.
Veröffentlicht: (2026)
von: Rahman, Md. Arifur, et al.
Veröffentlicht: (2026)
Soybean Disease Detection via Interpretable Hybrid CNN-GNN: Integrating MobileNetV2 and GraphSAGE with Cross-Modal Attention
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
AG-EgoPose: Leveraging Action-Guided Motion and Kinematic Joint Encoding for Egocentric 3D Pose Estimation
von: Azam, Md Mushfiqur, et al.
Veröffentlicht: (2026)
von: Azam, Md Mushfiqur, et al.
Veröffentlicht: (2026)
Vision Transformers for End-to-End Quark-Gluon Jet Classification from Calorimeter Images
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
A Methodological and Structural Review of Hand Gesture Recognition Across Diverse Data Modalities
von: Shin, Jungpil, et al.
Veröffentlicht: (2024)
von: Shin, Jungpil, et al.
Veröffentlicht: (2024)
GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection
von: Mia, Md Sohag, et al.
Veröffentlicht: (2025)
von: Mia, Md Sohag, et al.
Veröffentlicht: (2025)
Interpretable Dynamic Graph Neural Networks for Small Occluded Object Detection and Tracking
von: Soudeep, Shahriar, et al.
Veröffentlicht: (2024)
von: Soudeep, Shahriar, et al.
Veröffentlicht: (2024)
BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
MobilePlantViT: A Mobile-friendly Hybrid ViT for Generalized Plant Disease Image Classification
von: Tonmoy, Moshiur Rahman, et al.
Veröffentlicht: (2025)
von: Tonmoy, Moshiur Rahman, et al.
Veröffentlicht: (2025)
From Image to Language: A Critical Analysis of Visual Question Answering (VQA) Approaches, Challenges, and Opportunities
von: Ishmam, Md Farhan, et al.
Veröffentlicht: (2023)
von: Ishmam, Md Farhan, et al.
Veröffentlicht: (2023)
Can Multi-turn Self-refined Single Agent LMs with Retrieval Solve Hard Coding Problems?
von: Hosain, Md Tanzib, et al.
Veröffentlicht: (2025)
von: Hosain, Md Tanzib, et al.
Veröffentlicht: (2025)
Counting Through Occlusion: Framework for Open World Amodal Counting
von: Arib, Safaeid Hossain, et al.
Veröffentlicht: (2025)
von: Arib, Safaeid Hossain, et al.
Veröffentlicht: (2025)
MosquitoFusion: A Multiclass Dataset for Real-Time Detection of Mosquitoes, Swarms, and Breeding Sites Using Deep Learning
von: Sayeedi, Md. Faiyaz Abdullah, et al.
Veröffentlicht: (2024)
von: Sayeedi, Md. Faiyaz Abdullah, et al.
Veröffentlicht: (2024)
DyCAF-Net: Dynamic Class-Aware Fusion Network
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
von: Jahin, Md Abrar, et al.
Veröffentlicht: (2025)
MK-UNet: Multi-kernel Lightweight CNN for Medical Image Segmentation
von: Rahman, Md Mostafijur, et al.
Veröffentlicht: (2025)
von: Rahman, Md Mostafijur, et al.
Veröffentlicht: (2025)
Seam Carving as Feature Pooling in CNN
von: Jubair, Mohammad Imrul
Veröffentlicht: (2024)
von: Jubair, Mohammad Imrul
Veröffentlicht: (2024)
LoMix: Learnable Weighted Multi-Scale Logits Mixing for Medical Image Segmentation
von: Rahman, Md Mostafijur, et al.
Veröffentlicht: (2025)
von: Rahman, Md Mostafijur, et al.
Veröffentlicht: (2025)
MF-GCN: A Multi-Frequency Graph Convolutional Network for Tri-Modal Depression Detection Using Eye-Tracking, Facial, and Acoustic Features
von: Rahman, Sejuti, et al.
Veröffentlicht: (2025)
von: Rahman, Sejuti, et al.
Veröffentlicht: (2025)
A CNN-Based Malaria Diagnosis from Blood Cell Images with SHAP and LIME Explainability
von: Abir, Md. Ismiel Hossen, et al.
Veröffentlicht: (2025)
von: Abir, Md. Ismiel Hossen, et al.
Veröffentlicht: (2025)
Foundation Model-Powered 3D Few-Shot Class Incremental Learning via Training-free Adaptor
von: Ahmadi, Sahar, et al.
Veröffentlicht: (2024)
von: Ahmadi, Sahar, et al.
Veröffentlicht: (2024)
Edge-Native Digitization of Handwritten Marksheets: A Hybrid Heuristic-Deep Learning Framework
von: Hossain, Md. Irtiza, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Irtiza, et al.
Veröffentlicht: (2025)
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models
von: Samin, Niamul Hassan, et al.
Veröffentlicht: (2026)
von: Samin, Niamul Hassan, et al.
Veröffentlicht: (2026)
Less Is More? Selective Visual Attention to High-Importance Regions for Multimodal Radiology Summarization
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2026)
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2026)
Improving Fine-Grained Rice Leaf Disease Detection via Angular-Compactness Dual Loss Learning
von: Mia, Md. Rokon, et al.
Veröffentlicht: (2026)
von: Mia, Md. Rokon, et al.
Veröffentlicht: (2026)
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models
von: Rahman, Md Ashikur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Ashikur, et al.
Veröffentlicht: (2026)
PULSAR: Graph based Positive Unlabeled Learning with Multi Stream Adaptive Convolutions for Parkinson's Disease Recognition
von: Alam, Md. Zarif Ul, et al.
Veröffentlicht: (2023)
von: Alam, Md. Zarif Ul, et al.
Veröffentlicht: (2023)
Less Is More: An Explainable AI Framework for Lightweight Malaria Classification
von: Kafi, Md Abdullah Al, et al.
Veröffentlicht: (2025)
von: Kafi, Md Abdullah Al, et al.
Veröffentlicht: (2025)
VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites
von: Islam, Md. Adnanul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Adnanul, et al.
Veröffentlicht: (2025)
Real-Time Multi-Modal Embedded Vision Framework for Object Detection Facial Emotion Recognition and Biometric Identification on Low-Power Edge Platforms
von: Zahid, S. M. Khalid Bin, et al.
Veröffentlicht: (2026)
von: Zahid, S. M. Khalid Bin, et al.
Veröffentlicht: (2026)
Comparative Performance Analysis of Transformer-Based Pre-Trained Models for Detecting Keratoconus Disease
von: Ahmed, Nayeem, et al.
Veröffentlicht: (2024)
von: Ahmed, Nayeem, et al.
Veröffentlicht: (2024)
Deep Fusion Model for Brain Tumor Classification Using Fine-Grained Gradient Preservation
von: Islam, Niful, et al.
Veröffentlicht: (2024)
von: Islam, Niful, et al.
Veröffentlicht: (2024)
ILASH: A Predictive Neural Architecture Search Framework for Multi-Task Applications
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2024)
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2024)
Detecting Unauthorized Vehicles using Deep Learning for Smart Cities: A Case Study on Bangladesh
von: Sukanto, Sudipto Das, et al.
Veröffentlicht: (2025)
von: Sukanto, Sudipto Das, et al.
Veröffentlicht: (2025)
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
von: Hosain, Md Tanzib, et al.
Veröffentlicht: (2025)
von: Hosain, Md Tanzib, et al.
Veröffentlicht: (2025)
A Heterogeneous Two-Stream Framework for Video Action Recognition with Comparative Fusion Analysis
von: Rahaman, Md. Afzalur, et al.
Veröffentlicht: (2026)
von: Rahaman, Md. Afzalur, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
von: Hossain, Md. Mithun, et al.
Veröffentlicht: (2025) -
A Survey on 3D Egocentric Human Pose Estimation
von: Azam, Md Mushfiqur, et al.
Veröffentlicht: (2024) -
MIC: Medical Image Classification Using Chest X-ray (COVID-19 and Pneumonia) Dataset with the Help of CNN and Customized CNN
von: Fahad, Nafiz, et al.
Veröffentlicht: (2024) -
From Explanations to Architecture: Explainability-Driven CNN Refinement for Brain Tumor Classification in MRI
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025) -
ConMatFormer: A Multi-attention and Transformer Integrated ConvNext based Deep Learning Model for Enhanced Diabetic Foot Ulcer Classification
von: Rifat, Raihan Ahamed, et al.
Veröffentlicht: (2025)