TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Varma, Maya, Delbrouck, Jean-Benoit, Ostmeier, Sophie, Chaudhari, Akshay, Langlotz, Curtis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
von: Varma, Maya, et al.
Veröffentlicht: (2024)
von: Varma, Maya, et al.
Veröffentlicht: (2024)
LieRE: Lie Rotational Positional Encodings
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-Supervision
von: Gao, Yunhe, et al.
Veröffentlicht: (2026)
von: Gao, Yunhe, et al.
Veröffentlicht: (2026)
From Detection to Mitigation: Addressing Bias in Deep Learning Models for Chest X-Ray Diagnosis
von: Mottez, Clemence, et al.
Veröffentlicht: (2025)
von: Mottez, Clemence, et al.
Veröffentlicht: (2025)
Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling
von: Wang, Chong, et al.
Veröffentlicht: (2026)
von: Wang, Chong, et al.
Veröffentlicht: (2026)
GREEN: Generative Radiology Report Evaluation and Error Notation
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders
von: Varma, Maya, et al.
Veröffentlicht: (2025)
von: Varma, Maya, et al.
Veröffentlicht: (2025)
CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography
von: Prakash, Eva, et al.
Veröffentlicht: (2026)
von: Prakash, Eva, et al.
Veröffentlicht: (2026)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
A Reasoning-Enabled Vision-Language Foundation Model for Chest X-ray Interpretation
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
Process Reward Models for Sentence-Level Verification of LVLM Radiology Reports
von: Thomas, Alois, et al.
Veröffentlicht: (2025)
von: Thomas, Alois, et al.
Veröffentlicht: (2025)
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
Improving Performance, Robustness, and Fairness of Radiographic AI Models with Finely-Controllable Synthetic Data
von: Moroianu, Stefania L., et al.
Veröffentlicht: (2025)
von: Moroianu, Stefania L., et al.
Veröffentlicht: (2025)
CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats
von: Chambon, Pierre, et al.
Veröffentlicht: (2024)
von: Chambon, Pierre, et al.
Veröffentlicht: (2024)
Attention Head Entropy of LLMs Predicts Answer Correctness
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2026)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2026)
A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
Merlin: A Computed Tomography Vision-Language Foundation Model and Dataset
von: Blankemeier, Louis, et al.
Veröffentlicht: (2024)
von: Blankemeier, Louis, et al.
Veröffentlicht: (2024)
Structuring Radiology Reports: Challenging LLMs with Lightweight Models
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging
von: Kumar, Ashwin, et al.
Veröffentlicht: (2026)
von: Kumar, Ashwin, et al.
Veröffentlicht: (2026)
Diffusion MRI Transformer with a Diffusion Space Rotary Positional Embedding (D-RoPE)
von: Kung, Gustavo Chau Loo, et al.
Veröffentlicht: (2026)
von: Kung, Gustavo Chau Loo, et al.
Veröffentlicht: (2026)
RadDiff: Describing Differences in Radiology Image Sets with Natural Language
von: Shen, Xiaoxian, et al.
Veröffentlicht: (2026)
von: Shen, Xiaoxian, et al.
Veröffentlicht: (2026)
CLoVe: Encoding Compositional Language in Contrastive Vision-Language Models
von: Castro, Santiago, et al.
Veröffentlicht: (2024)
von: Castro, Santiago, et al.
Veröffentlicht: (2024)
Prompt Triage: Structured Optimization Enhances Vision-Language Model Performance on Medical Imaging Benchmarks
von: Singhvi, Arnav, et al.
Veröffentlicht: (2025)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2025)
Time-to-Event Pretraining for 3D Medical Imaging
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
Evaluating Reasoning Faithfulness in Medical Vision-Language Models using Multimodal Perturbations
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
Vision Language Models are Biased
von: Vo, An, et al.
Veröffentlicht: (2025)
von: Vo, An, et al.
Veröffentlicht: (2025)
A Generative Foundation Model for Multimodal Histopathology
von: Xiang, Jinxi, et al.
Veröffentlicht: (2026)
von: Xiang, Jinxi, et al.
Veröffentlicht: (2026)
VeCAF: Vision-language Collaborative Active Finetuning with Training Objective Awareness
von: Zhang, Rongyu, et al.
Veröffentlicht: (2024)
von: Zhang, Rongyu, et al.
Veröffentlicht: (2024)
Discovering and Mitigating Visual Biases through Keyword Explanation
von: Kim, Younghyun, et al.
Veröffentlicht: (2023)
von: Kim, Younghyun, et al.
Veröffentlicht: (2023)
Discover and Mitigate Multiple Biased Subgroups in Image Classifiers
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
Automated Structured Radiology Report Generation
von: Delbrouck, Jean-Benoit, et al.
Veröffentlicht: (2025)
von: Delbrouck, Jean-Benoit, et al.
Veröffentlicht: (2025)
Evaluating and Improving the Effectiveness of Synthetic Chest X-Rays for Medical Image Analysis
von: Prakash, Eva, et al.
Veröffentlicht: (2024)
von: Prakash, Eva, et al.
Veröffentlicht: (2024)
Explaining 3D Computed Tomography Classifiers with Counterfactuals
von: Cohen, Joseph Paul, et al.
Veröffentlicht: (2025)
von: Cohen, Joseph Paul, et al.
Veröffentlicht: (2025)
VeGaS: Video Gaussian Splatting
von: Smolak-Dyżewska, Weronika, et al.
Veröffentlicht: (2024)
von: Smolak-Dyżewska, Weronika, et al.
Veröffentlicht: (2024)
BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users
von: Cheng, Wanyin, et al.
Veröffentlicht: (2025)
von: Cheng, Wanyin, et al.
Veröffentlicht: (2025)
How Reasoning Influences Intersectional Biases in Vision Language Models
von: Desai, Adit, et al.
Veröffentlicht: (2025)
von: Desai, Adit, et al.
Veröffentlicht: (2025)
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
von: Fang, Hongwei, et al.
Veröffentlicht: (2026)
von: Fang, Hongwei, et al.
Veröffentlicht: (2026)
Medical Vision Language Models as Policies for Robotic Surgery
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
von: Varma, Maya, et al.
Veröffentlicht: (2024) -
LieRE: Lie Rotational Positional Encodings
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024) -
Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-Supervision
von: Gao, Yunhe, et al.
Veröffentlicht: (2026) -
From Detection to Mitigation: Addressing Bias in Deep Learning Models for Chest X-Ray Diagnosis
von: Mottez, Clemence, et al.
Veröffentlicht: (2025) -
Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)