Detecting Performance Degradation under Data Shift in Pathology Vision-Language Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Guan, Hao, Zhou, Li |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhanced Diagnostic Performance via Large-Resolution Inference Optimization for Pathology Foundation Models
di: Hu, Mengxuan, et al.
Pubblicazione: (2026)
di: Hu, Mengxuan, et al.
Pubblicazione: (2026)
The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
di: Ai, Wei, et al.
Pubblicazione: (2026)
di: Ai, Wei, et al.
Pubblicazione: (2026)
Efficient and Comprehensive Feature Extraction in Large Vision-Language Model for Pathology Analysis
di: Zhang, Shengxuming, et al.
Pubblicazione: (2024)
di: Zhang, Shengxuming, et al.
Pubblicazione: (2024)
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts
di: Li, Jiansheng, et al.
Pubblicazione: (2025)
di: Li, Jiansheng, et al.
Pubblicazione: (2025)
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models
di: Wang, Hongjun, et al.
Pubblicazione: (2026)
di: Wang, Hongjun, et al.
Pubblicazione: (2026)
Detecting Domain Shift in Multiple Instance Learning for Digital Pathology Using Fréchet Domain Distance
di: Pocevičiūtė, Milda, et al.
Pubblicazione: (2024)
di: Pocevičiūtė, Milda, et al.
Pubblicazione: (2024)
SocialFusion: Addressing Social Degradation in Pre-trained Vision-Language Models
di: Tahboub, Hamza, et al.
Pubblicazione: (2025)
di: Tahboub, Hamza, et al.
Pubblicazione: (2025)
Segment Anything in Pathology Images with Natural Language
di: Chen, Zhixuan, et al.
Pubblicazione: (2025)
di: Chen, Zhixuan, et al.
Pubblicazione: (2025)
Cost-effective Instruction Learning for Pathology Vision and Language Analysis
di: Chen, Kaitao, et al.
Pubblicazione: (2024)
di: Chen, Kaitao, et al.
Pubblicazione: (2024)
Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Generation
di: Yang, Zhiyuan, et al.
Pubblicazione: (2026)
di: Yang, Zhiyuan, et al.
Pubblicazione: (2026)
Reasoning under Vision: Understanding Visual-Spatial Cognition in Vision-Language Models for CAPTCHA
di: Song, Python, et al.
Pubblicazione: (2025)
di: Song, Python, et al.
Pubblicazione: (2025)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
di: Chen, Ying, et al.
Pubblicazione: (2024)
di: Chen, Ying, et al.
Pubblicazione: (2024)
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models
di: Van, Minh-Hao, et al.
Pubblicazione: (2025)
di: Van, Minh-Hao, et al.
Pubblicazione: (2025)
PathGLS: Evaluating Pathology Vision-Language Models without Ground Truth through Multi-Dimensional Consistency
di: Chen, Minbing, et al.
Pubblicazione: (2026)
di: Chen, Minbing, et al.
Pubblicazione: (2026)
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
di: Liu, Xiao, et al.
Pubblicazione: (2024)
di: Liu, Xiao, et al.
Pubblicazione: (2024)
Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
di: Zou, Zhengtao, et al.
Pubblicazione: (2025)
di: Zou, Zhengtao, et al.
Pubblicazione: (2025)
Flash-VL 2B: Optimizing Vision-Language Model Performance for Ultra-Low Latency and High Throughput
di: Zhang, Bo, et al.
Pubblicazione: (2025)
di: Zhang, Bo, et al.
Pubblicazione: (2025)
Dynamic Token Reduction during Generation for Vision Language Models
di: Liang, Xiaoyu, et al.
Pubblicazione: (2025)
di: Liang, Xiaoyu, et al.
Pubblicazione: (2025)
HookMIL: Revisiting Context Modeling in Multiple Instance Learning for Computational Pathology
di: Ling, Xitong, et al.
Pubblicazione: (2025)
di: Ling, Xitong, et al.
Pubblicazione: (2025)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
The Butterfly Effect in Pathology: Exploring Security in Pathology Foundation Models
di: Liu, Jiashuai, et al.
Pubblicazione: (2025)
di: Liu, Jiashuai, et al.
Pubblicazione: (2025)
TTP: Test-Time Padding for Adversarial Detection and Robust Adaptation on Vision-Language Models
di: Li, Zhiwei, et al.
Pubblicazione: (2025)
di: Li, Zhiwei, et al.
Pubblicazione: (2025)
A Multimodal Knowledge-enhanced Whole-slide Pathology Foundation Model
di: Xu, Yingxue, et al.
Pubblicazione: (2024)
di: Xu, Yingxue, et al.
Pubblicazione: (2024)
A Semantically Enhanced Generative Foundation Model Improves Pathological Image Synthesis
di: Guan, Xianchao, et al.
Pubblicazione: (2025)
di: Guan, Xianchao, et al.
Pubblicazione: (2025)
Uncovering Intrinsic Capabilities: A Paradigm for Data Curation in Vision-Language Models
di: Li, Junjie, et al.
Pubblicazione: (2025)
di: Li, Junjie, et al.
Pubblicazione: (2025)
3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detection
di: Zhu, Haowen, et al.
Pubblicazione: (2026)
di: Zhu, Haowen, et al.
Pubblicazione: (2026)
Leveraging Vision-Language Models to Detect Attention in Educational Videos
di: Becquet, Gabriel, et al.
Pubblicazione: (2026)
di: Becquet, Gabriel, et al.
Pubblicazione: (2026)
Discovering Pathology Rationale and Token Allocation for Efficient Multimodal Pathology Reasoning
di: Xu, Zhe, et al.
Pubblicazione: (2025)
di: Xu, Zhe, et al.
Pubblicazione: (2025)
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
di: Sui, Elaine, et al.
Pubblicazione: (2024)
di: Sui, Elaine, et al.
Pubblicazione: (2024)
Enhancing Model Performance: Another Approach to Vision-Language Instruction Tuning
di: Vedanshu, et al.
Pubblicazione: (2024)
di: Vedanshu, et al.
Pubblicazione: (2024)
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
di: Zhang, Jingyuan, et al.
Pubblicazione: (2025)
di: Zhang, Jingyuan, et al.
Pubblicazione: (2025)
ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models
di: Zhang, Jieyu, et al.
Pubblicazione: (2024)
di: Zhang, Jieyu, et al.
Pubblicazione: (2024)
Data Shift of Object Detection in Autonomous Driving
di: Xu, Lida
Pubblicazione: (2025)
di: Xu, Lida
Pubblicazione: (2025)
VLMine: Long-Tail Data Mining with Vision Language Models
di: Ye, Mao, et al.
Pubblicazione: (2024)
di: Ye, Mao, et al.
Pubblicazione: (2024)
On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression
di: Zhang, Xinwei, et al.
Pubblicazione: (2026)
di: Zhang, Xinwei, et al.
Pubblicazione: (2026)
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
di: Zhang, Yunfei, et al.
Pubblicazione: (2025)
di: Zhang, Yunfei, et al.
Pubblicazione: (2025)
LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models
di: Jiang, Zhiyuan, et al.
Pubblicazione: (2026)
di: Jiang, Zhiyuan, et al.
Pubblicazione: (2026)
EVLP:Learning Unified Embodied Vision-Language Planner with Reinforced Supervised Fine-Tuning
di: Cai, Xinyan, et al.
Pubblicazione: (2025)
di: Cai, Xinyan, et al.
Pubblicazione: (2025)
Self-Consistency as a Free Lunch: Reducing Hallucinations in Vision-Language Models via Self-Reflection
di: Han, Mingfei, et al.
Pubblicazione: (2025)
di: Han, Mingfei, et al.
Pubblicazione: (2025)
From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration
di: Song, Mingyang, et al.
Pubblicazione: (2025)
di: Song, Mingyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Enhanced Diagnostic Performance via Large-Resolution Inference Optimization for Pathology Foundation Models
di: Hu, Mengxuan, et al.
Pubblicazione: (2026) -
The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
di: Ai, Wei, et al.
Pubblicazione: (2026) -
Efficient and Comprehensive Feature Extraction in Large Vision-Language Model for Pathology Analysis
di: Zhang, Shengxuming, et al.
Pubblicazione: (2024) -
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts
di: Li, Jiansheng, et al.
Pubblicazione: (2025) -
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models
di: Wang, Hongjun, et al.
Pubblicazione: (2026)