Less Is More? Selective Visual Attention to High-Importance Regions for Multimodal Radiology Summarization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Naznin, Mst. Fahmida Sultana, Faruq, Adnan Ibney, Rahman, Mushfiqur, Mondal, Niloy Kumar, Shawon, Md. Mehedi Hasan, Hasan, Md Rakibul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CSTRL: Context-Driven Sequential Transfer Learning for Abstractive Radiology Report Summarization
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2025)
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2025)
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
An Explainable Vision-Language Model Framework with Adaptive PID-Tversky Loss for Lumbar Spinal Stenosis Diagnosis
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2026)
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2026)
Context-Aware Asymmetric Ensembling for Interpretable Retinopathy of Prematurity Screening via Active Query and Vascular Attention
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2026)
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2026)
Attentive Dilated Convolution for Automatic Sleep Staging using Force-directed Layout
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
Privacy-Preserving Empathy Detection in Video Interactions
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
Are You Really Empathic? Evidence from Trait, State and Speaker-Perceived Empathy, and Physiological Signals
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
SS-DPPN: A self-supervised dual-path foundation model for the generalizable cardiac audio representation
von: Muna, Ummy Maria, et al.
Veröffentlicht: (2025)
von: Muna, Ummy Maria, et al.
Veröffentlicht: (2025)
FunnelNet: An End-to-End Deep Learning Framework to Monitor Digital Heart Murmur in Real-Time
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
UPLME: Uncertainty-Aware Probabilistic Language Modelling for Robust Empathy Regression
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection
von: Mia, Md Sohag, et al.
Veröffentlicht: (2025)
von: Mia, Md Sohag, et al.
Veröffentlicht: (2025)
Attention Is not Everything: Efficient Alternatives for Vision
von: Kazi, Nur Mohammad, et al.
Veröffentlicht: (2026)
von: Kazi, Nur Mohammad, et al.
Veröffentlicht: (2026)
Predictive Health Analysis in Industry 5.0: A Scientometric and Systematic Review of Motion Capture in Construction
von: Rahman, Md Hadisur, et al.
Veröffentlicht: (2024)
von: Rahman, Md Hadisur, et al.
Veröffentlicht: (2024)
Stability Criteria and Optoelectronic Properties of Mg3ZBr3 (Z = As, Sb, Bi) Perovskites for Evaluating the Performance in PIN Photo Diode
von: Mohiuddin, Md, et al.
Veröffentlicht: (2025)
von: Mohiuddin, Md, et al.
Veröffentlicht: (2025)
Hierarchical Sentiment Analysis Framework for Hate Speech Detection: Implementing Binary and Multiclass Classification Strategy
von: Naznin, Faria, et al.
Veröffentlicht: (2024)
von: Naznin, Faria, et al.
Veröffentlicht: (2024)
Dialectal Toxicity Detection: Evaluating LLM-as-a-Judge Consistency Across Language Varieties
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
Rep3Net: An Approach Exploiting Multimodal Representation for Molecular Bioactivity Prediction
von: Islam, Sabrina, et al.
Veröffentlicht: (2025)
von: Islam, Sabrina, et al.
Veröffentlicht: (2025)
Real-Time Multi-Modal Embedded Vision Framework for Object Detection Facial Emotion Recognition and Biometric Identification on Low-Power Edge Platforms
von: Zahid, S. M. Khalid Bin, et al.
Veröffentlicht: (2026)
von: Zahid, S. M. Khalid Bin, et al.
Veröffentlicht: (2026)
MADE-for-ASD: A Multi-Atlas Deep Ensemble Network for Diagnosing Autism Spectrum Disorder
von: Liu, Xuehan, et al.
Veröffentlicht: (2024)
von: Liu, Xuehan, et al.
Veröffentlicht: (2024)
A Two-Stage Multitask Vision-Language Framework for Explainable Crop Disease Visual Question Answering
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2026)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2026)
Explainable AI-Driven Detection of Human Monkeypox Using Deep Learning and Vision Transformers: A Comprehensive Analysis
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
Annotate Rhetorical Relations with INCEpTION: A Comparison with Automatic Approaches
von: Emon, Mehedi Hasan
Veröffentlicht: (2025)
von: Emon, Mehedi Hasan
Veröffentlicht: (2025)
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
A Deep Learning-based Multimodal Depth-Aware Dynamic Hand Gesture Recognition System
von: Mahmud, Hasan, et al.
Veröffentlicht: (2021)
von: Mahmud, Hasan, et al.
Veröffentlicht: (2021)
Bengali Sign Language Recognition through Hand Pose Estimation using Multi-Branch Spatial-Temporal Attention Model
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
Labels Generated by Large Language Models Help Measure People's Empathy in Vitro
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
Towards Interpretable Radiology Report Generation via Concept Bottlenecks using a Multi-Agentic RAG
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2024)
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2024)
FAARM: Firmware Attestation and Authentication Framework for Mali GPUs
von: Hasan, Md. Mehedi
Veröffentlicht: (2025)
von: Hasan, Md. Mehedi
Veröffentlicht: (2025)
Interpretable Heart Disease Prediction via a Weighted Ensemble Model: A Large-Scale Study with SHAP and Surrogate Decision Trees
von: Hasnat, Md Abrar, et al.
Veröffentlicht: (2025)
von: Hasnat, Md Abrar, et al.
Veröffentlicht: (2025)
RadBARTsum: Domain Specific Adaption of Denoising Sequence-to-Sequence Models for Abstractive Radiology Report Summarization
von: Wu, Jinge, et al.
Veröffentlicht: (2024)
von: Wu, Jinge, et al.
Veröffentlicht: (2024)
PC-SRGAN: Physically Consistent Super-Resolution Generative Adversarial Network for General Transient Simulations
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models on Historical Health Crisis Knowledge in Resource-Limited Settings: A Hybrid Multi-Metric Study
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
Are ASR foundation models generalized enough to capture features of regional dialects for low-resource languages?
von: Dipto, Tawsif Tashwar, et al.
Veröffentlicht: (2025)
von: Dipto, Tawsif Tashwar, et al.
Veröffentlicht: (2025)
Fine-Tuning Video Transformers for Word-Level Bangla Sign Language: A Comparative Analysis for Classification Tasks
von: Shawon, Jubayer Ahmed Bhuiyan, et al.
Veröffentlicht: (2025)
von: Shawon, Jubayer Ahmed Bhuiyan, et al.
Veröffentlicht: (2025)
The Carbon Cost of Conversation, Sustainability in the Age of Language Models
von: Amiri, Sayed Mahbub Hasan, et al.
Veröffentlicht: (2025)
von: Amiri, Sayed Mahbub Hasan, et al.
Veröffentlicht: (2025)
From Explanations to Architecture: Explainability-Driven CNN Refinement for Brain Tumor Classification in MRI
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
DE-KAN: A Kolmogorov Arnold Network with Dual Encoder for accurate 2D Teeth Segmentation
von: Mustakim, Md Mizanur Rahman, et al.
Veröffentlicht: (2025)
von: Mustakim, Md Mizanur Rahman, et al.
Veröffentlicht: (2025)
Empathy Detection from Text, Audiovisual, Audio or Physiological Signals: A Systematic Review of Task Formulations and Machine Learning Methods
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2023)
von: Hasan, Md Rakibul, et al.
Veröffentlicht: (2023)
BeHGAN: Bengali Handwritten Word Generation from Plain Text Using Generative Adversarial Networks
von: Islam, Md. Rakibul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Rakibul, et al.
Veröffentlicht: (2025)
SliceVision-F2I: A Synthetic Feature-to-Image Dataset for Visual Pattern Representation on Network Slices
von: Rafi, Md. Abid Hasan, et al.
Veröffentlicht: (2025)
von: Rafi, Md. Abid Hasan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CSTRL: Context-Driven Sequential Transfer Learning for Abstractive Radiology Report Summarization
von: Naznin, Mst. Fahmida Sultana, et al.
Veröffentlicht: (2025) -
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025) -
An Explainable Vision-Language Model Framework with Adaptive PID-Tversky Loss for Lumbar Spinal Stenosis Diagnosis
von: Sk., Md. Sajeebul Islam, et al.
Veröffentlicht: (2026) -
Context-Aware Asymmetric Ensembling for Interpretable Retinopathy of Prematurity Screening via Active Query and Vascular Attention
von: Hassan, Md. Mehedi, et al.
Veröffentlicht: (2026) -
Attentive Dilated Convolution for Automatic Sleep Staging using Force-directed Layout
von: Jobayer, Md, et al.
Veröffentlicht: (2024)