An Explainable Vision-Language Model Framework with Adaptive PID-Tversky Loss for Lumbar Spinal Stenosis Diagnosis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sk., Md. Sajeebul Islam, Shawon, Md. Mehedi Hasan, Alam, Md. Golam Rabiul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FoundationalECGNet: A Lightweight Foundational Model for ECG-based Multitask Cardiac Analysis
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2025)
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2025)
Interpretable Heart Disease Prediction via a Weighted Ensemble Model: A Large-Scale Study with SHAP and Surrogate Decision Trees
von: Hasnat, Md Abrar, et al.
Veröffentlicht: (2025)
von: Hasnat, Md Abrar, et al.
Veröffentlicht: (2025)
Erase to Retain: Low Rank Adaptation Guided Selective Unlearning in Medical Segmentation Networks
von: Datta, Nirjhor, et al.
Veröffentlicht: (2025)
von: Datta, Nirjhor, et al.
Veröffentlicht: (2025)
DUA-D2C: Dynamic Uncertainty Aware Method for Overfitting Remediation in Deep Learning
von: Siddiqui, Md. Saiful Bari, et al.
Veröffentlicht: (2024)
von: Siddiqui, Md. Saiful Bari, et al.
Veröffentlicht: (2024)
Transforming Precision: A Comparative Analysis of Vision Transformers, CNNs, and Traditional ML for Knee Osteoarthritis Severity Diagnosis
von: Apon, Tasnim Sakib, et al.
Veröffentlicht: (2024)
von: Apon, Tasnim Sakib, et al.
Veröffentlicht: (2024)
A Computer Vision Based Approach for Stalking Detection Using a CNN-LSTM-MLP Hybrid Fusion Model
von: Hasan, Murad, et al.
Veröffentlicht: (2024)
von: Hasan, Murad, et al.
Veröffentlicht: (2024)
SynthEnsemble: A Fusion of CNN, Vision Transformer, and Hybrid Models for Multi-Label Chest X-Ray Classification
von: Ashraf, S. M. Nabil, et al.
Veröffentlicht: (2023)
von: Ashraf, S. M. Nabil, et al.
Veröffentlicht: (2023)
SS-DPPN: A self-supervised dual-path foundation model for the generalizable cardiac audio representation
von: Muna, Ummy Maria, et al.
Veröffentlicht: (2025)
von: Muna, Ummy Maria, et al.
Veröffentlicht: (2025)
A Two-Stage Multitask Vision-Language Framework for Explainable Crop Disease Visual Question Answering
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2026)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2026)
ElderFallGuard: Real-Time IoT and Computer Vision-Based Fall Detection System for Elderly Safety
von: Riahi, Tasrifur, et al.
Veröffentlicht: (2025)
von: Riahi, Tasrifur, et al.
Veröffentlicht: (2025)
Extreme Model Compression for Edge Vision-Language Models: Sparse Temporal Token Fusion and Adaptive Neural Compression
von: Tanvir, Md Tasnin, et al.
Veröffentlicht: (2025)
von: Tanvir, Md Tasnin, et al.
Veröffentlicht: (2025)
Beyond Perception Errors: Semantic Fixation in Large Vision-Language Models
von: Alam, Md Tanvirul
Veröffentlicht: (2026)
von: Alam, Md Tanvirul
Veröffentlicht: (2026)
HeBA: Heterogeneous Bottleneck Adapters for Robust Vision-Language Models
von: Islam, Md Jahidul
Veröffentlicht: (2026)
von: Islam, Md Jahidul
Veröffentlicht: (2026)
Attentive Dilated Convolution for Automatic Sleep Staging using Force-directed Layout
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
ReHARK: Refined Hybrid Adaptive RBF Kernels for Robust One-Shot Vision-Language Adaptation
von: Islam, Md Jahidul
Veröffentlicht: (2026)
von: Islam, Md Jahidul
Veröffentlicht: (2026)
CAST: Channel-Aware Spatial Transfer Learning with Pseudo-Image Radar for Sign Language Recognition
von: Shujon, Md. Shakhoyat Rahman, et al.
Veröffentlicht: (2026)
von: Shujon, Md. Shakhoyat Rahman, et al.
Veröffentlicht: (2026)
Vision Models for Medical Imaging: A Hybrid Approach for PCOS Detection from Ultrasound Scans
von: Hoque, Md Mahmudul, et al.
Veröffentlicht: (2026)
von: Hoque, Md Mahmudul, et al.
Veröffentlicht: (2026)
Bengali Sign Language Recognition through Hand Pose Estimation using Multi-Branch Spatial-Temporal Attention Model
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
DSVTLA: Deep Swin Vision Transformer-Based Transfer Learning Architecture for Multi-Type Cancer Histopathological Cancer Image Classification
von: Khan, Muazzem Hussain, et al.
Veröffentlicht: (2026)
von: Khan, Muazzem Hussain, et al.
Veröffentlicht: (2026)
Context-Aware Asymmetric Ensembling for Interpretable Retinopathy of Prematurity Screening via Active Query and Vascular Attention
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2026)
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2026)
Less Is More? Selective Visual Attention to High-Importance Regions for Multimodal Radiology Summarization
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2026)
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2026)
Explainable AI-Driven Detection of Human Monkeypox Using Deep Learning and Vision Transformers: A Comprehensive Analysis
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
An Explainable Agentic AI Framework for Uncertainty-Aware and Abstention-Enabled Acute Ischemic Stroke Imaging Decisions
von: Islam, Md Rashadul
Veröffentlicht: (2026)
von: Islam, Md Rashadul
Veröffentlicht: (2026)
Interpretable Gallbladder Ultrasound Diagnosis: A Lightweight Web-Mobile Software Platform with Real-Time XAI
von: Bhoyan, Fuyad Hasan, et al.
Veröffentlicht: (2025)
von: Bhoyan, Fuyad Hasan, et al.
Veröffentlicht: (2025)
Retrieval Augmented Enhanced Dual Co-Attention Framework for Target Aware Multimodal Bengali Hateful Meme Detection
von: Tanvir, Raihan, et al.
Veröffentlicht: (2026)
von: Tanvir, Raihan, et al.
Veröffentlicht: (2026)
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
A Lightweight and Explainable DenseNet-121 Framework for Grape Leaf Disease Classification
von: Haque, Md. Ehsanul, et al.
Veröffentlicht: (2026)
von: Haque, Md. Ehsanul, et al.
Veröffentlicht: (2026)
Counting Through Occlusion: Framework for Open World Amodal Counting
von: Arib, Safaeid Hossain, et al.
Veröffentlicht: (2025)
von: Arib, Safaeid Hossain, et al.
Veröffentlicht: (2025)
IKIWISI: An Interactive Visual Pattern Generator for Evaluating the Reliability of Vision-Language Models Without Ground Truth
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2025)
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2025)
JaiLIP: Jailbreaking Vision-Language Models via Loss Guided Image Perturbation
von: Mia, Md Jueal, et al.
Veröffentlicht: (2025)
von: Mia, Md Jueal, et al.
Veröffentlicht: (2025)
Vision-Based Lane Following and Traffic Sign Recognition for Resource-Constrained Autonomous Vehicles
von: Islam, Md Tanjemul, et al.
Veröffentlicht: (2026)
von: Islam, Md Tanjemul, et al.
Veröffentlicht: (2026)
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2
von: Islam, Md. Rakibul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Rakibul, et al.
Veröffentlicht: (2025)
Size and Smoothness Aware Adaptive Focal Loss for Small Tumor Segmentation
von: Islam, Md Rakibul, et al.
Veröffentlicht: (2024)
von: Islam, Md Rakibul, et al.
Veröffentlicht: (2024)
PULSAR: Graph based Positive Unlabeled Learning with Multi Stream Adaptive Convolutions for Parkinson's Disease Recognition
von: Alam, Md. Zarif Ul, et al.
Veröffentlicht: (2023)
von: Alam, Md. Zarif Ul, et al.
Veröffentlicht: (2023)
Using Computer Vision for Skin Disease Diagnosis in Bangladesh Enhancing Interpretability and Transparency in Deep Learning Models for Skin Cancer Classification
von: Islam, Rafiul, et al.
Veröffentlicht: (2025)
von: Islam, Rafiul, et al.
Veröffentlicht: (2025)
DL$^3$M: A Vision-to-Language Framework for Expert-Level Medical Reasoning through Deep Learning and Large Language Models
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
Attention Is not Everything: Efficient Alternatives for Vision
von: Kazi, Nur Mohammad, et al.
Veröffentlicht: (2026)
von: Kazi, Nur Mohammad, et al.
Veröffentlicht: (2026)
Unleashing the Power of Transfer Learning Model for Sophisticated Insect Detection: Revolutionizing Insect Classification
von: Hasan, Md. Mahmudul, et al.
Veröffentlicht: (2024)
von: Hasan, Md. Mahmudul, et al.
Veröffentlicht: (2024)
MSRANetV2: An Explainable Deep Learning Architecture for Multi-class Classification of Colorectal Histopathological Images
von: Sarkar, Ovi, et al.
Veröffentlicht: (2025)
von: Sarkar, Ovi, et al.
Veröffentlicht: (2025)
SYNAPSE-Net: A Unified Framework with Lesion-Aware Hierarchical Gating for Robust Segmentation of Heterogeneous Brain Lesions
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2025)
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FoundationalECGNet: A Lightweight Foundational Model for ECG-based Multitask Cardiac Analysis
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2025) -
Interpretable Heart Disease Prediction via a Weighted Ensemble Model: A Large-Scale Study with SHAP and Surrogate Decision Trees
von: Hasnat, Md Abrar, et al.
Veröffentlicht: (2025) -
Erase to Retain: Low Rank Adaptation Guided Selective Unlearning in Medical Segmentation Networks
von: Datta, Nirjhor, et al.
Veröffentlicht: (2025) -
DUA-D2C: Dynamic Uncertainty Aware Method for Overfitting Remediation in Deep Learning
von: Siddiqui, Md. Saiful Bari, et al.
Veröffentlicht: (2024) -
Transforming Precision: A Comparative Analysis of Vision Transformers, CNNs, and Traditional ML for Knee Osteoarthritis Severity Diagnosis
von: Apon, Tasnim Sakib, et al.
Veröffentlicht: (2024)