OODBench: Out-of-Distribution Benchmark for Large Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lin, Ling, Bai, Yang, Su, Heng, Zhu, Congcong, Wang, Yaoxing, Zhou, Yang, Fu, Huazhu, Chen, Jingrun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment
di: Li, Zhuangzi, et al.
Pubblicazione: (2026)
di: Li, Zhuangzi, et al.
Pubblicazione: (2026)
MarineEval: Assessing the Marine Intelligence of Vision-Language Models
di: Wong, YuK-Kwan, et al.
Pubblicazione: (2025)
di: Wong, YuK-Kwan, et al.
Pubblicazione: (2025)
Vision meets algae: A novel way for microalgae recognization and health monitor
di: Zhou, Shizheng, et al.
Pubblicazione: (2022)
di: Zhou, Shizheng, et al.
Pubblicazione: (2022)
Harnessing Large Language and Vision-Language Models for Robust Out-of-Distribution Detection
di: Lee, Pei-Kang, et al.
Pubblicazione: (2025)
di: Lee, Pei-Kang, et al.
Pubblicazione: (2025)
OS-W2S: An Automatic Labeling Engine for Language-Guided Open-Set Aerial Object Detection
di: Wei, Guoting, et al.
Pubblicazione: (2025)
di: Wei, Guoting, et al.
Pubblicazione: (2025)
FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting
di: Ji, Fengxian, et al.
Pubblicazione: (2026)
di: Ji, Fengxian, et al.
Pubblicazione: (2026)
Delving into Out-of-Distribution Detection with Medical Vision-Language Models
di: Ju, Lie, et al.
Pubblicazione: (2025)
di: Ju, Lie, et al.
Pubblicazione: (2025)
Vision-Language Model IP Protection via Prompt-based Learning
di: Wang, Lianyu, et al.
Pubblicazione: (2025)
di: Wang, Lianyu, et al.
Pubblicazione: (2025)
Backdooring Vision-Language Models with Out-Of-Distribution Data
di: Lyu, Weimin, et al.
Pubblicazione: (2024)
di: Lyu, Weimin, et al.
Pubblicazione: (2024)
YUV20K: A Complexity-Driven Benchmark and Trajectory-Aware Alignment Model for Video Camouflaged Object Detection
di: Liu, Yiyu, et al.
Pubblicazione: (2026)
di: Liu, Yiyu, et al.
Pubblicazione: (2026)
DPCD: A Quality Assessment Database for Dynamic Point Clouds
di: Liu, Yating, et al.
Pubblicazione: (2025)
di: Liu, Yating, et al.
Pubblicazione: (2025)
Out-Of-Distribution Detection with Diversification (Provably)
di: Yao, Haiyun, et al.
Pubblicazione: (2024)
di: Yao, Haiyun, et al.
Pubblicazione: (2024)
STSA: Spatial-Temporal Semantic Alignment for Visual Dubbing
di: Ding, Zijun, et al.
Pubblicazione: (2025)
di: Ding, Zijun, et al.
Pubblicazione: (2025)
Phase Matching for Out-of-Distribution Generalization
di: Hu, Chengming, et al.
Pubblicazione: (2023)
di: Hu, Chengming, et al.
Pubblicazione: (2023)
Vision Also You Need: Navigating Out-of-Distribution Detection with Multimodal Large Language Model
di: Xu, Haoran, et al.
Pubblicazione: (2026)
di: Xu, Haoran, et al.
Pubblicazione: (2026)
Vivim: a Video Vision Mamba for Medical Video Segmentation
di: Yang, Yijun, et al.
Pubblicazione: (2024)
di: Yang, Yijun, et al.
Pubblicazione: (2024)
Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning
di: Zhu, Yingjie, et al.
Pubblicazione: (2024)
di: Zhu, Yingjie, et al.
Pubblicazione: (2024)
ORIC: Benchmarking Object Recognition under Contextual Incongruity in Large Vision-Language Models
di: Li, Zhaoyang, et al.
Pubblicazione: (2025)
di: Li, Zhaoyang, et al.
Pubblicazione: (2025)
Few-Shot Learning from Gigapixel Images via Hierarchical Vision-Language Alignment and Modeling
di: Wong, Bryan, et al.
Pubblicazione: (2025)
di: Wong, Bryan, et al.
Pubblicazione: (2025)
A Benchmark and Evaluation for Real-World Out-of-Distribution Detection Using Vision-Language Models
di: Noda, Shiho, et al.
Pubblicazione: (2025)
di: Noda, Shiho, et al.
Pubblicazione: (2025)
K-FACE: A Large-Scale KIST Face Database in Consideration with Unconstrained Environments
di: Choi, Yeji, et al.
Pubblicazione: (2021)
di: Choi, Yeji, et al.
Pubblicazione: (2021)
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
Out-of-Distribution Detection Using Peer-Class Generated by Large Language Model
di: Huang, K, et al.
Pubblicazione: (2024)
di: Huang, K, et al.
Pubblicazione: (2024)
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift
di: Xu, Qinwu
Pubblicazione: (2026)
di: Xu, Qinwu
Pubblicazione: (2026)
TokBench: Evaluating Your Visual Tokenizer before Visual Generation
di: Wu, Junfeng, et al.
Pubblicazione: (2025)
di: Wu, Junfeng, et al.
Pubblicazione: (2025)
LEAML: Label-Efficient Adaptation to Out-of-Distribution Visual Tasks for Multimodal Large Language Models
di: Lin, Ci-Siang, et al.
Pubblicazione: (2025)
di: Lin, Ci-Siang, et al.
Pubblicazione: (2025)
AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving
di: Xing, Shuo, et al.
Pubblicazione: (2024)
di: Xing, Shuo, et al.
Pubblicazione: (2024)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
di: Cui, Mingxuan, et al.
Pubblicazione: (2026)
di: Cui, Mingxuan, et al.
Pubblicazione: (2026)
STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering
di: Sun, Guohao, et al.
Pubblicazione: (2024)
di: Sun, Guohao, et al.
Pubblicazione: (2024)
DarkDriving: A Real-World Day and Night Aligned Dataset for Autonomous Driving in the Dark Environment
di: Wang, Wuqi, et al.
Pubblicazione: (2026)
di: Wang, Wuqi, et al.
Pubblicazione: (2026)
MLLM4TS: Leveraging Vision and Multimodal Language Models for General Time-Series Analysis
di: Liu, Qinghua, et al.
Pubblicazione: (2025)
di: Liu, Qinghua, et al.
Pubblicazione: (2025)
TWIX: Automatically Reconstructing Structured Data from Templatized Documents
di: Lin, Yiming, et al.
Pubblicazione: (2025)
di: Lin, Yiming, et al.
Pubblicazione: (2025)
Is Dataset Quality Still a Concern in Diagnosis Using Large Foundation Model?
di: Lin, Ziqin, et al.
Pubblicazione: (2024)
di: Lin, Ziqin, et al.
Pubblicazione: (2024)
Knowledge Regularized Negative Feature Tuning of Vision-Language Models for Out-of-Distribution Detection
di: Zhu, Wenjie, et al.
Pubblicazione: (2025)
di: Zhu, Wenjie, et al.
Pubblicazione: (2025)
On the Out-Of-Distribution Generalization of Multimodal Large Language Models
di: Zhang, Xingxuan, et al.
Pubblicazione: (2024)
di: Zhang, Xingxuan, et al.
Pubblicazione: (2024)
LOVO: Efficient Complex Object Query in Large-Scale Video Datasets
di: Liu, Yuxin, et al.
Pubblicazione: (2025)
di: Liu, Yuxin, et al.
Pubblicazione: (2025)
BACON: Improving Clarity of Image Captions via Bag-of-Concept Graphs
di: Yang, Zhantao, et al.
Pubblicazione: (2024)
di: Yang, Zhantao, et al.
Pubblicazione: (2024)
MMVR: Millimeter-wave Multi-View Radar Dataset and Benchmark for Indoor Perception
di: Rahman, M. Mahbubur, et al.
Pubblicazione: (2024)
di: Rahman, M. Mahbubur, et al.
Pubblicazione: (2024)
ViLa-MIL: Dual-scale Vision-Language Multiple Instance Learning for Whole Slide Image Classification
di: Shi, Jiangbo, et al.
Pubblicazione: (2025)
di: Shi, Jiangbo, et al.
Pubblicazione: (2025)
UniVRSE: Unified Vision-conditioned Response Semantic Entropy for Hallucination Detection in Medical Vision-Language Models
di: Liao, Zehui, et al.
Pubblicazione: (2025)
di: Liao, Zehui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment
di: Li, Zhuangzi, et al.
Pubblicazione: (2026) -
MarineEval: Assessing the Marine Intelligence of Vision-Language Models
di: Wong, YuK-Kwan, et al.
Pubblicazione: (2025) -
Vision meets algae: A novel way for microalgae recognization and health monitor
di: Zhou, Shizheng, et al.
Pubblicazione: (2022) -
Harnessing Large Language and Vision-Language Models for Robust Out-of-Distribution Detection
di: Lee, Pei-Kang, et al.
Pubblicazione: (2025) -
OS-W2S: An Automatic Labeling Engine for Language-Guided Open-Set Aerial Object Detection
di: Wei, Guoting, et al.
Pubblicazione: (2025)