Cross-modal linkage risk in clinical vision-language models
Fuente:
arXiv
Salvato in:
| Autori principali: | Arasteh, Soroosh Tayebi, Lotfinia, Mahshad, Nebelung, Sven, Truhn, Daniel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
Differential privacy representation geometry for medical image analysis
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
Large Language Models Streamline Automated Machine Learning for Clinical Studies
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
The role of self-supervised pretraining in differentially private medical image analysis
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026)
Boosting multi-demographic federated learning for chest radiograph analysis using general-purpose self-supervised representations
di: Lotfinia, Mahshad, et al.
Pubblicazione: (2025)
di: Lotfinia, Mahshad, et al.
Pubblicazione: (2025)
Resolution scaling governs DINOv3 transfer performance in chest radiograph classification
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2025)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2025)
Safety and accuracy follow different scaling laws in clinical large language models
di: Wind, Sebastian, et al.
Pubblicazione: (2026)
di: Wind, Sebastian, et al.
Pubblicazione: (2026)
RadioRAG: Online Retrieval-augmented Generation for Radiology Question Answering
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2024)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2024)
Enhancing Network Initialization for Medical AI Models Using Large-Scale, Unlabeled Natural Images
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
Private, fair and accurate: Training large-scale, privacy-preserving AI models in medical imaging
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023)
Multi-step retrieval and reasoning improves radiology question answering with large language models
di: Wind, Sebastian, et al.
Pubblicazione: (2025)
di: Wind, Sebastian, et al.
Pubblicazione: (2025)
Differential privacy for medical deep learning: methods, tradeoffs, and deployment implications
di: Mohammadi, Marziyeh, et al.
Pubblicazione: (2025)
di: Mohammadi, Marziyeh, et al.
Pubblicazione: (2025)
PathAlign: A vision-language model for whole slide images in histopathology
di: Ahmed, Faruk, et al.
Pubblicazione: (2024)
di: Ahmed, Faruk, et al.
Pubblicazione: (2024)
Advancing vision-language models in front-end development via data synthesis
di: Ge, Tong, et al.
Pubblicazione: (2025)
di: Ge, Tong, et al.
Pubblicazione: (2025)
Bridging vision language model (VLM) evaluation gaps with a framework for scalable and cost-effective benchmark generation
di: Rädsch, Tim, et al.
Pubblicazione: (2025)
di: Rädsch, Tim, et al.
Pubblicazione: (2025)
Agentic retrieval-augmented reasoning reshapes collective reliability under model variability in radiology question answering
di: Farajiamiri, Mina, et al.
Pubblicazione: (2026)
di: Farajiamiri, Mina, et al.
Pubblicazione: (2026)
Owls are wise and foxes are unfaithful: Uncovering animal stereotypes in vision-language models
di: Aman, Tabinda, et al.
Pubblicazione: (2025)
di: Aman, Tabinda, et al.
Pubblicazione: (2025)
Compute-Efficient Medical Image Classification with Softmax-Free Transformers and Sequence Normalization
di: Khader, Firas, et al.
Pubblicazione: (2024)
di: Khader, Firas, et al.
Pubblicazione: (2024)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
di: Zhu, Mengdan, et al.
Pubblicazione: (2025)
di: Zhu, Mengdan, et al.
Pubblicazione: (2025)
When language and vision meet road safety: leveraging multimodal large language models for video-based traffic accident analysis
di: Zhang, Ruixuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruixuan, et al.
Pubblicazione: (2025)
The in-context inductive biases of vision-language models differ across modalities
di: Allen, Kelsey, et al.
Pubblicazione: (2025)
di: Allen, Kelsey, et al.
Pubblicazione: (2025)
e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings
di: Chen, Haonan, et al.
Pubblicazione: (2026)
di: Chen, Haonan, et al.
Pubblicazione: (2026)
MedicalPatchNet: A Patch-Based Self-Explainable AI Architecture for Chest X-ray Classification
di: Wienholt, Patrick, et al.
Pubblicazione: (2025)
di: Wienholt, Patrick, et al.
Pubblicazione: (2025)
HistGen: Histopathology Report Generation via Local-Global Feature Encoding and Cross-modal Context Interaction
di: Guo, Zhengrui, et al.
Pubblicazione: (2024)
di: Guo, Zhengrui, et al.
Pubblicazione: (2024)
Cross-modal Information Flow in Multimodal Large Language Models
di: Zhang, Zhi, et al.
Pubblicazione: (2024)
di: Zhang, Zhi, et al.
Pubblicazione: (2024)
VLA-Mark: A cross modal watermark for large vision-language alignment model
di: Liu, Shuliang, et al.
Pubblicazione: (2025)
di: Liu, Shuliang, et al.
Pubblicazione: (2025)
Visual Language Model based Cross-modal Semantic Communication Systems
di: Jiang, Feibo, et al.
Pubblicazione: (2024)
di: Jiang, Feibo, et al.
Pubblicazione: (2024)
BRAVE: Broadening the visual encoding of vision-language models
di: Kar, Oğuzhan Fatih, et al.
Pubblicazione: (2024)
di: Kar, Oğuzhan Fatih, et al.
Pubblicazione: (2024)
Medical Slice Transformer: Improved Diagnosis and Explainability on 3D Medical Images with DINOv2
di: Müller-Franzes, Gustav, et al.
Pubblicazione: (2024)
di: Müller-Franzes, Gustav, et al.
Pubblicazione: (2024)
SteuerLLM: Local specialized large language model for German tax law analysis
di: Wind, Sebastian, et al.
Pubblicazione: (2026)
di: Wind, Sebastian, et al.
Pubblicazione: (2026)
The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs
di: Li, Hong, et al.
Pubblicazione: (2024)
di: Li, Hong, et al.
Pubblicazione: (2024)
What explains the success of cross-modal fine-tuning with ORCA?
di: García-de-Herreros, Paloma, et al.
Pubblicazione: (2024)
di: García-de-Herreros, Paloma, et al.
Pubblicazione: (2024)
GET: Unlocking the Multi-modal Potential of CLIP for Generalized Category Discovery
di: Wang, Enguang, et al.
Pubblicazione: (2024)
di: Wang, Enguang, et al.
Pubblicazione: (2024)
Mask-aware Text-to-Image Retrieval: Referring Expression Segmentation Meets Cross-modal Retrieval
di: Shen, Li-Cheng, et al.
Pubblicazione: (2025)
di: Shen, Li-Cheng, et al.
Pubblicazione: (2025)
MMCTAgent: Multi-modal Critical Thinking Agent Framework for Complex Visual Reasoning
di: Kumar, Somnath, et al.
Pubblicazione: (2024)
di: Kumar, Somnath, et al.
Pubblicazione: (2024)
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind
di: Shi, Haojun, et al.
Pubblicazione: (2024)
di: Shi, Haojun, et al.
Pubblicazione: (2024)
Multi-modal Preference Alignment Remedies Degradation of Visual Instruction Tuning on Language Models
di: Li, Shengzhi, et al.
Pubblicazione: (2024)
di: Li, Shengzhi, et al.
Pubblicazione: (2024)
Preserving Pre-trained Representation Space: On Effectiveness of Prefix-tuning for Large Multi-modal Models
di: Kim, Donghoon, et al.
Pubblicazione: (2024)
di: Kim, Donghoon, et al.
Pubblicazione: (2024)
From Consistency to Complementarity: Aligned and Disentangled Multi-modal Learning for Time Series Understanding and Reasoning
di: Ni, Hang, et al.
Pubblicazione: (2026)
di: Ni, Hang, et al.
Pubblicazione: (2026)
Analyzing and Boosting the Power of Fine-Grained Visual Recognition for Multi-modal Large Language Models
di: He, Hulingxiao, et al.
Pubblicazione: (2025)
di: He, Hulingxiao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026) -
Differential privacy representation geometry for medical image analysis
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026) -
Large Language Models Streamline Automated Machine Learning for Clinical Studies
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2023) -
The role of self-supervised pretraining in differentially private medical image analysis
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2026) -
Boosting multi-demographic federated learning for chest radiograph analysis using general-purpose self-supervised representations
di: Lotfinia, Mahshad, et al.
Pubblicazione: (2025)