Your other Left! Vision-Language Models Fail to Identify Relative Positions in Medical Images
Fuente:
arXiv
Saved in:
| Main Authors: | Wolf, Daniel, Hillenhagen, Heiko, Taskin, Billurvan, Bäuerle, Alex, Beer, Meinrad, Götz, Michael, Ropinski, Timo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Vision Transformer with Prototypes for Interpretable Medical Image Classification
by: Gallée, Luisa, et al.
Published: (2025)
by: Gallée, Luisa, et al.
Published: (2025)
Less is More: Selective Reduction of CT Data for Self-Supervised Pre-Training of Deep Learning Models with Contrastive Learning Improves Downstream Classification Performance
by: Wolf, Daniel, et al.
Published: (2024)
by: Wolf, Daniel, et al.
Published: (2024)
FunnyNodules: A Customizable Medical Dataset Tailored for Evaluating Explainable AI
by: Gallée, Luisa, et al.
Published: (2025)
by: Gallée, Luisa, et al.
Published: (2025)
Interpretable Medical Image Classification using Prototype Learning and Privileged Information
by: Gallee, Luisa, et al.
Published: (2023)
by: Gallee, Luisa, et al.
Published: (2023)
A Survey on Quality Metrics for Text-to-Image Generation
by: Hartwig, Sebastian, et al.
Published: (2024)
by: Hartwig, Sebastian, et al.
Published: (2024)
Evaluating the Explainability of Attributes and Prototypes for a Medical Classification Model
by: Gallée, Luisa, et al.
Published: (2024)
by: Gallée, Luisa, et al.
Published: (2024)
Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design
by: Engel, Dominik, et al.
Published: (2023)
by: Engel, Dominik, et al.
Published: (2023)
Evaluating Graphical Perception Capabilities of Vision Transformers
by: Poonam, Poonam, et al.
Published: (2026)
by: Poonam, Poonam, et al.
Published: (2026)
Attention-Guided Masked Autoencoders For Learning Image Representations
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
by: Kniesel, Hannah, et al.
Published: (2025)
by: Kniesel, Hannah, et al.
Published: (2025)
Minimum Data, Maximum Impact: 20 annotated samples for explainable lung nodule classification
by: Gallée, Luisa, et al.
Published: (2025)
by: Gallée, Luisa, et al.
Published: (2025)
RelationField: Relate Anything in Radiance Fields
by: Koch, Sebastian, et al.
Published: (2024)
by: Koch, Sebastian, et al.
Published: (2024)
Das soziale Erbe
by: Ziegler, Meinrad
Published: (2015)
by: Ziegler, Meinrad
Published: (2015)
Unsupervised Semantic Segmentation Through Depth-Guided Feature Correlation and Sampling
by: Sick, Leon, et al.
Published: (2023)
by: Sick, Leon, et al.
Published: (2023)
Evaluating Foveated Frame Rate Reduction in Virtual Reality for Head-Mounted Displays
by: Flöter, Christopher, et al.
Published: (2025)
by: Flöter, Christopher, et al.
Published: (2025)
Humboldt: Metadata-Driven Extensible Data Discovery
by: Bäuerle, Alex, et al.
Published: (2024)
by: Bäuerle, Alex, et al.
Published: (2024)
Relative portfolio optimization via a value at risk based constraint
by: Bäuerle, Nicole, et al.
Published: (2025)
by: Bäuerle, Nicole, et al.
Published: (2025)
Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models
by: McLaughlin, Oliver, et al.
Published: (2026)
by: McLaughlin, Oliver, et al.
Published: (2026)
How Your Brain Filters Noise (And Why It Fails)
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Simulating the interplay of dipolar and quadrupolar interactions in NMR by spin dynamic mean-field theory
by: Gräßer, Timo, et al.
Published: (2025)
by: Gräßer, Timo, et al.
Published: (2025)
AutoTherm: A Dataset and Benchmark for Thermal Comfort Estimation Indoors and in Vehicles
by: Colley, Mark, et al.
Published: (2022)
by: Colley, Mark, et al.
Published: (2022)
Open3DSG: Open-Vocabulary 3D Scene Graphs from Point Clouds with Queryable Objects and Open-Set Relationships
by: Koch, Sebastian, et al.
Published: (2024)
by: Koch, Sebastian, et al.
Published: (2024)
S2D: Sparse-To-Dense Keymask Distillation for Unsupervised Video Instance Segmentation
by: Sick, Leon, et al.
Published: (2025)
by: Sick, Leon, et al.
Published: (2025)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
OpenHype: Hyperbolic Embeddings for Hierarchical Open-Vocabulary Radiance Fields
by: Weijler, Lisa, et al.
Published: (2025)
by: Weijler, Lisa, et al.
Published: (2025)
Design and Creation of Remote Temperature Measure System with An Esp32 to Evaluate Patient Health
by: Sylvain Meinrad Donkeng Voumo
Published: (2025)
by: Sylvain Meinrad Donkeng Voumo
Published: (2025)
Context-Aware Human Behavior Prediction Using Multimodal Large Language Models: Challenges and Insights
by: Liu, Yuchen, et al.
Published: (2025)
by: Liu, Yuchen, et al.
Published: (2025)
Ökonomie - Praxis - Subjektivierung
by: Bäuerle, Lukas
Published: (2022)
by: Bäuerle, Lukas
Published: (2022)
qbic-pipelines/vcftocounts: 2.0.0 - Rad Sepia
by: Famke Bäuerle
Published: (2025)
by: Famke Bäuerle
Published: (2025)
Mean Field Markov Decision Processes
by: Bäuerle, Nicole
Published: (2021)
by: Bäuerle, Nicole
Published: (2021)
Weakly Supervised Virus Capsid Detection with Image-Level Annotations in Electron Microscopy Images
by: Kniesel, Hannah, et al.
Published: (2025)
by: Kniesel, Hannah, et al.
Published: (2025)
HPSCAN: Human Perception-Based Scattered Data Clustering
by: Hartwig, Sebastian, et al.
Published: (2023)
by: Hartwig, Sebastian, et al.
Published: (2023)
Identifying and Mitigating Position Bias of Multi-image Vision-Language Models
by: Tian, Xinyu, et al.
Published: (2025)
by: Tian, Xinyu, et al.
Published: (2025)
Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization
by: Bharadwaj, Siddhant, et al.
Published: (2026)
by: Bharadwaj, Siddhant, et al.
Published: (2026)
How Do Medical MLLMs Fail? A Study on Visual Grounding in Medical Images
by: Liu, Guimeng, et al.
Published: (2026)
by: Liu, Guimeng, et al.
Published: (2026)
Unified Semantic Transformer for 3D Scene Understanding
by: Koch, Sebastian, et al.
Published: (2025)
by: Koch, Sebastian, et al.
Published: (2025)
Optimal investment under partial information and robust VaR-type constraint
by: Bäuerle, Nicole, et al.
Published: (2022)
by: Bäuerle, Nicole, et al.
Published: (2022)
Microscopic understanding of NMR signals by dynamic mean-field theory for spins
by: Gräßer, Timo, et al.
Published: (2024)
by: Gräßer, Timo, et al.
Published: (2024)
First-principles simulation of spin diffusion in static solids using dynamic mean-field theory
by: Gräßer, Timo, et al.
Published: (2025)
by: Gräßer, Timo, et al.
Published: (2025)
Dynamical mean-field theory for dense spin systems at finite temperature
by: Bieniek, Przemysław, et al.
Published: (2026)
by: Bieniek, Przemysław, et al.
Published: (2026)
Similar Items
-
Hierarchical Vision Transformer with Prototypes for Interpretable Medical Image Classification
by: Gallée, Luisa, et al.
Published: (2025) -
Less is More: Selective Reduction of CT Data for Self-Supervised Pre-Training of Deep Learning Models with Contrastive Learning Improves Downstream Classification Performance
by: Wolf, Daniel, et al.
Published: (2024) -
FunnyNodules: A Customizable Medical Dataset Tailored for Evaluating Explainable AI
by: Gallée, Luisa, et al.
Published: (2025) -
Interpretable Medical Image Classification using Prototype Learning and Privileged Information
by: Gallee, Luisa, et al.
Published: (2023) -
A Survey on Quality Metrics for Text-to-Image Generation
by: Hartwig, Sebastian, et al.
Published: (2024)