MEBench: A Novel Benchmark for Understanding Mutual Exclusivity Bias in Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Thai, Anh, Stojanov, Stefan, Huang, Zixuan, Boote, Bikram, Rehg, James M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ZeroShape: Regression-based Zero-shot Shape Reconstruction
di: Huang, Zixuan, et al.
Pubblicazione: (2023)
di: Huang, Zixuan, et al.
Pubblicazione: (2023)
Leveraging Object Priors for Point Tracking
di: Boote, Bikram, et al.
Pubblicazione: (2024)
di: Boote, Bikram, et al.
Pubblicazione: (2024)
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
di: Thai, Anh, et al.
Pubblicazione: (2024)
di: Thai, Anh, et al.
Pubblicazione: (2024)
Symmetry Strikes Back: From Single-Image Symmetry Detection to 3D Generation
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions
di: Kim, Junho, et al.
Pubblicazione: (2026)
di: Kim, Junho, et al.
Pubblicazione: (2026)
Vinedresser3D: Agentic Text-guided 3D Editing
di: Chi, Yankuan, et al.
Pubblicazione: (2026)
di: Chi, Yankuan, et al.
Pubblicazione: (2026)
Cue3D: Quantifying the Role of Image Cues in Single-Image 3D Generation
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
How Much 3D Do Video Foundation Models Encode?
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
PointInfinity: Resolution-Invariant Point Diffusion Models
di: Huang, Zixuan, et al.
Pubblicazione: (2024)
di: Huang, Zixuan, et al.
Pubblicazione: (2024)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
GlyphPattern: An Abstract Pattern Recognition Benchmark for Vision-Language Models
di: Wu, Zixuan, et al.
Pubblicazione: (2024)
di: Wu, Zixuan, et al.
Pubblicazione: (2024)
Benchmarking and Mitigating MCQA Selection Bias of Large Vision-Language Models
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
Tendency-driven Mutual Exclusivity for Weakly Supervised Incremental Semantic Segmentation
di: Si, Chongjie, et al.
Pubblicazione: (2024)
di: Si, Chongjie, et al.
Pubblicazione: (2024)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models
di: Guo, Xianda, et al.
Pubblicazione: (2024)
di: Guo, Xianda, et al.
Pubblicazione: (2024)
Towards Online Multi-Modal Social Interaction Understanding
di: Li, Xinpeng, et al.
Pubblicazione: (2025)
di: Li, Xinpeng, et al.
Pubblicazione: (2025)
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
di: Hong, Wenyi, et al.
Pubblicazione: (2025)
di: Hong, Wenyi, et al.
Pubblicazione: (2025)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
di: Nguyen, Thong, et al.
Pubblicazione: (2025)
di: Nguyen, Thong, et al.
Pubblicazione: (2025)
Benchmarking Vision Language Models for Cultural Understanding
di: Nayak, Shravan, et al.
Pubblicazione: (2024)
di: Nayak, Shravan, et al.
Pubblicazione: (2024)
DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural Images
di: Kara, Ozgur, et al.
Pubblicazione: (2025)
di: Kara, Ozgur, et al.
Pubblicazione: (2025)
Pathological Truth Bias in Vision-Language Models
di: Thube, Yash
Pubblicazione: (2025)
di: Thube, Yash
Pubblicazione: (2025)
HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning
di: Wei, Yanbin, et al.
Pubblicazione: (2026)
di: Wei, Yanbin, et al.
Pubblicazione: (2026)
UWBench: A Comprehensive Vision-Language Benchmark for Underwater Understanding
di: Zhang, Da, et al.
Pubblicazione: (2025)
di: Zhang, Da, et al.
Pubblicazione: (2025)
GS-Bias: Global-Spatial Bias Learner for Single-Image Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
SocialGesture: Delving into Multi-person Gesture Understanding
di: Cao, Xu, et al.
Pubblicazione: (2025)
di: Cao, Xu, et al.
Pubblicazione: (2025)
Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models
di: Yang, Shijun, et al.
Pubblicazione: (2025)
di: Yang, Shijun, et al.
Pubblicazione: (2025)
Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking
di: Chung, Sangyun, et al.
Pubblicazione: (2024)
di: Chung, Sangyun, et al.
Pubblicazione: (2024)
SARLANG-1M: A Benchmark for Vision-Language Modeling in SAR Image Understanding
di: Wei, Yimin, et al.
Pubblicazione: (2025)
di: Wei, Yimin, et al.
Pubblicazione: (2025)
DetectiumFire: A Comprehensive Multi-modal Dataset Bridging Vision and Language for Fire Understanding
di: Liu, Zixuan, et al.
Pubblicazione: (2025)
di: Liu, Zixuan, et al.
Pubblicazione: (2025)
GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
The Bias of Harmful Label Associations in Vision-Language Models
di: Hazirbas, Caner, et al.
Pubblicazione: (2024)
di: Hazirbas, Caner, et al.
Pubblicazione: (2024)
Beyond Memorization: A Multi-Modal Ordinal Regression Benchmark to Expose Popularity Bias in Vision-Language Models
di: Szu-Tu, Li-Zhong, et al.
Pubblicazione: (2025)
di: Szu-Tu, Li-Zhong, et al.
Pubblicazione: (2025)
MIRL: Mutual Information-Guided Reinforcement Learning for Vision-Language Models
di: Zhang, Yin, et al.
Pubblicazione: (2026)
di: Zhang, Yin, et al.
Pubblicazione: (2026)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
μ-Bench: A Vision-Language Benchmark for Microscopy Understanding
di: Lozano, Alejandro, et al.
Pubblicazione: (2024)
di: Lozano, Alejandro, et al.
Pubblicazione: (2024)
CVLUE: A New Benchmark Dataset for Chinese Vision-Language Understanding Evaluation
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
Understanding Degradation with Vision Language Model
di: Lan, Guanzhou, et al.
Pubblicazione: (2026)
di: Lan, Guanzhou, et al.
Pubblicazione: (2026)
Uncovering Bias in Large Vision-Language Models at Scale with Counterfactuals
di: Howard, Phillip, et al.
Pubblicazione: (2024)
di: Howard, Phillip, et al.
Pubblicazione: (2024)
VLBiasBench: A Comprehensive Benchmark for Evaluating Bias in Large Vision-Language Model
di: Wang, Sibo, et al.
Pubblicazione: (2024)
di: Wang, Sibo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ZeroShape: Regression-based Zero-shot Shape Reconstruction
di: Huang, Zixuan, et al.
Pubblicazione: (2023) -
Leveraging Object Priors for Point Tracking
di: Boote, Bikram, et al.
Pubblicazione: (2024) -
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
di: Thai, Anh, et al.
Pubblicazione: (2024) -
Symmetry Strikes Back: From Single-Image Symmetry Detection to 3D Generation
di: Li, Xiang, et al.
Pubblicazione: (2024) -
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
di: Lee, Sangmin, et al.
Pubblicazione: (2024)