Blind to Position, Biased in Language: Probing Mid-Layer Representational Bias in Vision-Language Encoders for Zero-Shot Language-Grounded Spatial Understanding

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: An, Na Min, Kang, Inha, Lee, Minhyun, Shim, Hyunjung
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!