MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
Fuente:
arXiv
Guardado en:
| Autores principales: | Khan, Ufaq, Nawaz, Umair, Teja, L D M S S, Saeed, Numaan, Bilal, Muhammad, Xie, Yutong, Yaqub, Mohammad, Khan, Muhammad Haris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
por: Khan, Ufaq, et al.
Publicado: (2025)
por: Khan, Ufaq, et al.
Publicado: (2025)
AURORA: Adaptive Unified Representation for Robust Ultrasound Analysis
por: Khan, Ufaq, et al.
Publicado: (2026)
por: Khan, Ufaq, et al.
Publicado: (2026)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
por: Ding, Bonan, et al.
Publicado: (2026)
por: Ding, Bonan, et al.
Publicado: (2026)
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
por: Nawaz, Umair, et al.
Publicado: (2025)
por: Nawaz, Umair, et al.
Publicado: (2025)
WorldCache: Content-Aware Caching for Accelerated Video World Models
por: Nawaz, Umair, et al.
Publicado: (2026)
por: Nawaz, Umair, et al.
Publicado: (2026)
DAMASHA: Detecting AI in Mixed Adversarial Texts via Segmentation with Human-interpretable Attribution
por: Teja, L. D. M. S. Sai, et al.
Publicado: (2025)
por: Teja, L. D. M. S. Sai, et al.
Publicado: (2025)
Robust and Label-Efficient Deep Waste Detection
por: Abid, Hassan, et al.
Publicado: (2025)
por: Abid, Hassan, et al.
Publicado: (2025)
MedPromptX: Grounded Multimodal Prompting for Chest X-ray Diagnosis
por: Shaaban, Mai A., et al.
Publicado: (2024)
por: Shaaban, Mai A., et al.
Publicado: (2024)
AgriCLIP: Adapting CLIP for Agriculture and Livestock via Domain-Specialized Cross-Model Alignment
por: Nawaz, Umair, et al.
Publicado: (2024)
por: Nawaz, Umair, et al.
Publicado: (2024)
CountZES: Counting via Zero-Shot Exemplar Selection
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
por: Singha, Mainak, et al.
Publicado: (2026)
por: Singha, Mainak, et al.
Publicado: (2026)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
por: Ali, Eman, et al.
Publicado: (2023)
por: Ali, Eman, et al.
Publicado: (2023)
RAPTOR+: A Visually Grounded Vision-Language Framework to Improve Clinical Trust and Auditability in Automated Cancer Referral Processing
por: Abioye, Sofiat, et al.
Publicado: (2026)
por: Abioye, Sofiat, et al.
Publicado: (2026)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
por: Chokuwa, Sharon, et al.
Publicado: (2024)
por: Chokuwa, Sharon, et al.
Publicado: (2024)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
por: Khan, Adnan, et al.
Publicado: (2024)
por: Khan, Adnan, et al.
Publicado: (2024)
On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models
por: Malik, Hashmat Shadab, et al.
Publicado: (2024)
por: Malik, Hashmat Shadab, et al.
Publicado: (2024)
UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities
por: Khattak, Muhammad Uzair, et al.
Publicado: (2024)
por: Khattak, Muhammad Uzair, et al.
Publicado: (2024)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
por: Liaqat, Muhammad Irzam, et al.
Publicado: (2024)
por: Liaqat, Muhammad Irzam, et al.
Publicado: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
por: Saeed, Muhammad Saad, et al.
Publicado: (2024)
por: Saeed, Muhammad Saad, et al.
Publicado: (2024)
Efficient Parameter Adaptation for Multi-Modal Medical Image Segmentation and Prognosis
por: Saeed, Numan, et al.
Publicado: (2025)
por: Saeed, Numan, et al.
Publicado: (2025)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
por: Imam, Raza, et al.
Publicado: (2025)
por: Imam, Raza, et al.
Publicado: (2025)
Envisioning MedCLIP: A Deep Dive into Explainability for Medical Vision-Language Models
por: Hashmi, Anees Ur Rehman, et al.
Publicado: (2024)
por: Hashmi, Anees Ur Rehman, et al.
Publicado: (2024)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
por: Sharifdeen, Ashshak, et al.
Publicado: (2025)
por: Sharifdeen, Ashshak, et al.
Publicado: (2025)
HuLP: Human-in-the-Loop for Prognosis
por: Ridzuan, Muhammad, et al.
Publicado: (2024)
por: Ridzuan, Muhammad, et al.
Publicado: (2024)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
por: Basu, Abhishek, et al.
Publicado: (2025)
por: Basu, Abhishek, et al.
Publicado: (2025)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
por: Gunduboina, Hariseetharam, et al.
Publicado: (2025)
por: Gunduboina, Hariseetharam, et al.
Publicado: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
por: Ali, Eman, et al.
Publicado: (2024)
por: Ali, Eman, et al.
Publicado: (2024)
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2024)
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2024)
DuPLUS: Dual-Prompt Vision-Language Framework for Universal Medical Image Segmentation and Prognosis
por: Saeed, Numan, et al.
Publicado: (2025)
por: Saeed, Numan, et al.
Publicado: (2025)
MedROV: Towards Real-Time Open-Vocabulary Detection Across Diverse Medical Imaging Modalities
por: Sheikh, Tooba Tehreem, et al.
Publicado: (2025)
por: Sheikh, Tooba Tehreem, et al.
Publicado: (2025)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
por: Liu, Yu, et al.
Publicado: (2024)
por: Liu, Yu, et al.
Publicado: (2024)
Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan
por: Saeed, Muhammad Saad, et al.
Publicado: (2024)
por: Saeed, Muhammad Saad, et al.
Publicado: (2024)
Multi-modal Medical Image Fusion For Non-Small Cell Lung Cancer Classification
por: Hassan, Salma, et al.
Publicado: (2024)
por: Hassan, Salma, et al.
Publicado: (2024)
Depth Attention for Robust RGB Tracking
por: Liu, Yu, et al.
Publicado: (2024)
por: Liu, Yu, et al.
Publicado: (2024)
MedNNS: Supernet-based Medical Task-Adaptive Neural Network Search
por: Mecharbat, Lotfi Abdelkrim, et al.
Publicado: (2025)
por: Mecharbat, Lotfi Abdelkrim, et al.
Publicado: (2025)
MedMerge: Merging Models for Effective Transfer Learning to Medical Imaging Tasks
por: Almakky, Ibrahim, et al.
Publicado: (2024)
por: Almakky, Ibrahim, et al.
Publicado: (2024)
Agentic AI for Remote Sensing: Technical Challenges and Research Directions
por: Munir, Muhammad Akhtar, et al.
Publicado: (2026)
por: Munir, Muhammad Akhtar, et al.
Publicado: (2026)
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
ReConText3D: Replay-based Continual Text-to-3D Generation
por: Khan, Muhammad Ahmed Ullah, et al.
Publicado: (2026)
por: Khan, Muhammad Ahmed Ullah, et al.
Publicado: (2026)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
por: Danish, Muhammad Sohail, et al.
Publicado: (2025)
por: Danish, Muhammad Sohail, et al.
Publicado: (2025)
Ejemplares similares
-
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
por: Khan, Ufaq, et al.
Publicado: (2025) -
AURORA: Adaptive Unified Representation for Robust Ultrasound Analysis
por: Khan, Ufaq, et al.
Publicado: (2026) -
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
por: Ding, Bonan, et al.
Publicado: (2026) -
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
por: Nawaz, Umair, et al.
Publicado: (2025) -
WorldCache: Content-Aware Caching for Accelerated Video World Models
por: Nawaz, Umair, et al.
Publicado: (2026)