Detecting Performance Degradation under Data Shift in Pathology Vision-Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guan, Hao, Zhou, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhanced Diagnostic Performance via Large-Resolution Inference Optimization for Pathology Foundation Models
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
von: Ai, Wei, et al.
Veröffentlicht: (2026)
von: Ai, Wei, et al.
Veröffentlicht: (2026)
Efficient and Comprehensive Feature Extraction in Large Vision-Language Model for Pathology Analysis
von: Zhang, Shengxuming, et al.
Veröffentlicht: (2024)
von: Zhang, Shengxuming, et al.
Veröffentlicht: (2024)
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts
von: Li, Jiansheng, et al.
Veröffentlicht: (2025)
von: Li, Jiansheng, et al.
Veröffentlicht: (2025)
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models
von: Wang, Hongjun, et al.
Veröffentlicht: (2026)
von: Wang, Hongjun, et al.
Veröffentlicht: (2026)
Detecting Domain Shift in Multiple Instance Learning for Digital Pathology Using Fréchet Domain Distance
von: Pocevičiūtė, Milda, et al.
Veröffentlicht: (2024)
von: Pocevičiūtė, Milda, et al.
Veröffentlicht: (2024)
SocialFusion: Addressing Social Degradation in Pre-trained Vision-Language Models
von: Tahboub, Hamza, et al.
Veröffentlicht: (2025)
von: Tahboub, Hamza, et al.
Veröffentlicht: (2025)
Segment Anything in Pathology Images with Natural Language
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
Cost-effective Instruction Learning for Pathology Vision and Language Analysis
von: Chen, Kaitao, et al.
Veröffentlicht: (2024)
von: Chen, Kaitao, et al.
Veröffentlicht: (2024)
Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Generation
von: Yang, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Yang, Zhiyuan, et al.
Veröffentlicht: (2026)
Reasoning under Vision: Understanding Visual-Spatial Cognition in Vision-Language Models for CAPTCHA
von: Song, Python, et al.
Veröffentlicht: (2025)
von: Song, Python, et al.
Veröffentlicht: (2025)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
von: Chen, Ying, et al.
Veröffentlicht: (2024)
von: Chen, Ying, et al.
Veröffentlicht: (2024)
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models
von: Van, Minh-Hao, et al.
Veröffentlicht: (2025)
von: Van, Minh-Hao, et al.
Veröffentlicht: (2025)
PathGLS: Evaluating Pathology Vision-Language Models without Ground Truth through Multi-Dimensional Consistency
von: Chen, Minbing, et al.
Veröffentlicht: (2026)
von: Chen, Minbing, et al.
Veröffentlicht: (2026)
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
von: Zou, Zhengtao, et al.
Veröffentlicht: (2025)
von: Zou, Zhengtao, et al.
Veröffentlicht: (2025)
Flash-VL 2B: Optimizing Vision-Language Model Performance for Ultra-Low Latency and High Throughput
von: Zhang, Bo, et al.
Veröffentlicht: (2025)
von: Zhang, Bo, et al.
Veröffentlicht: (2025)
Dynamic Token Reduction during Generation for Vision Language Models
von: Liang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Liang, Xiaoyu, et al.
Veröffentlicht: (2025)
HookMIL: Revisiting Context Modeling in Multiple Instance Learning for Computational Pathology
von: Ling, Xitong, et al.
Veröffentlicht: (2025)
von: Ling, Xitong, et al.
Veröffentlicht: (2025)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
von: Cheng, Liwei, et al.
Veröffentlicht: (2026)
von: Cheng, Liwei, et al.
Veröffentlicht: (2026)
The Butterfly Effect in Pathology: Exploring Security in Pathology Foundation Models
von: Liu, Jiashuai, et al.
Veröffentlicht: (2025)
von: Liu, Jiashuai, et al.
Veröffentlicht: (2025)
TTP: Test-Time Padding for Adversarial Detection and Robust Adaptation on Vision-Language Models
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
A Multimodal Knowledge-enhanced Whole-slide Pathology Foundation Model
von: Xu, Yingxue, et al.
Veröffentlicht: (2024)
von: Xu, Yingxue, et al.
Veröffentlicht: (2024)
A Semantically Enhanced Generative Foundation Model Improves Pathological Image Synthesis
von: Guan, Xianchao, et al.
Veröffentlicht: (2025)
von: Guan, Xianchao, et al.
Veröffentlicht: (2025)
Uncovering Intrinsic Capabilities: A Paradigm for Data Curation in Vision-Language Models
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detection
von: Zhu, Haowen, et al.
Veröffentlicht: (2026)
von: Zhu, Haowen, et al.
Veröffentlicht: (2026)
Leveraging Vision-Language Models to Detect Attention in Educational Videos
von: Becquet, Gabriel, et al.
Veröffentlicht: (2026)
von: Becquet, Gabriel, et al.
Veröffentlicht: (2026)
Discovering Pathology Rationale and Token Allocation for Efficient Multimodal Pathology Reasoning
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
von: Sui, Elaine, et al.
Veröffentlicht: (2024)
von: Sui, Elaine, et al.
Veröffentlicht: (2024)
Enhancing Model Performance: Another Approach to Vision-Language Instruction Tuning
von: Vedanshu, et al.
Veröffentlicht: (2024)
von: Vedanshu, et al.
Veröffentlicht: (2024)
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models
von: Zhang, Jieyu, et al.
Veröffentlicht: (2024)
von: Zhang, Jieyu, et al.
Veröffentlicht: (2024)
Data Shift of Object Detection in Autonomous Driving
von: Xu, Lida
Veröffentlicht: (2025)
von: Xu, Lida
Veröffentlicht: (2025)
VLMine: Long-Tail Data Mining with Vision Language Models
von: Ye, Mao, et al.
Veröffentlicht: (2024)
von: Ye, Mao, et al.
Veröffentlicht: (2024)
On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
von: Zhang, Yunfei, et al.
Veröffentlicht: (2025)
von: Zhang, Yunfei, et al.
Veröffentlicht: (2025)
LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models
von: Jiang, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Jiang, Zhiyuan, et al.
Veröffentlicht: (2026)
EVLP:Learning Unified Embodied Vision-Language Planner with Reinforced Supervised Fine-Tuning
von: Cai, Xinyan, et al.
Veröffentlicht: (2025)
von: Cai, Xinyan, et al.
Veröffentlicht: (2025)
Self-Consistency as a Free Lunch: Reducing Hallucinations in Vision-Language Models via Self-Reflection
von: Han, Mingfei, et al.
Veröffentlicht: (2025)
von: Han, Mingfei, et al.
Veröffentlicht: (2025)
From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Enhanced Diagnostic Performance via Large-Resolution Inference Optimization for Pathology Foundation Models
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026) -
The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
von: Ai, Wei, et al.
Veröffentlicht: (2026) -
Efficient and Comprehensive Feature Extraction in Large Vision-Language Model for Pathology Analysis
von: Zhang, Shengxuming, et al.
Veröffentlicht: (2024) -
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts
von: Li, Jiansheng, et al.
Veröffentlicht: (2025) -
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models
von: Wang, Hongjun, et al.
Veröffentlicht: (2026)