Can Vision-Language Models Infer Speaker's Ignorance? The Role of Visual and Linguistic Cues
Fuente:
arXiv
Guardado en:
| Autores principales: | Cho, Ye-eun, Maeng, Yunho |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Pragmatic Reasoning in Large Language Models: Evidence from Scalar Diversity
por: Cho, Ye-eun
Publicado: (2026)
por: Cho, Ye-eun
Publicado: (2026)
Continuous Interpretive Steering for Scalar Diversity
por: Cho, Ye-eun
Publicado: (2026)
por: Cho, Ye-eun
Publicado: (2026)
Cross-Linguistic Transcription and Phonological Representation in the Huìtóngguǎnxì Huáyíyìyǔ
por: Kim, Ji-eun
Publicado: (2026)
por: Kim, Ji-eun
Publicado: (2026)
Gender Bias in LLM-generated Interview Responses
por: Kong, Haein, et al.
Publicado: (2024)
por: Kong, Haein, et al.
Publicado: (2024)
Pragmatic inference of scalar implicature by LLMs
por: Cho, Ye-eun, et al.
Publicado: (2024)
por: Cho, Ye-eun, et al.
Publicado: (2024)
It's Not the Capability: Harness Sensitivity Is Non-Monotone Across LLM Agent Tiers
por: Cho, Yong-eun
Publicado: (2026)
por: Cho, Yong-eun
Publicado: (2026)
Typed-RAG: Type-Aware Decomposition of Non-Factoid Questions for Retrieval-Augmented Generation
por: Lee, DongGeon, et al.
Publicado: (2025)
por: Lee, DongGeon, et al.
Publicado: (2025)
SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?
por: Lee, Jonggeun, et al.
Publicado: (2026)
por: Lee, Jonggeun, et al.
Publicado: (2026)
Can Structural Cues Save LLMs? Evaluating Language Models in Massive Document Streams
por: Lee, Yukyung, et al.
Publicado: (2026)
por: Lee, Yukyung, et al.
Publicado: (2026)
Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts
por: Nooralahzadeh, Farhad, et al.
Publicado: (2026)
por: Nooralahzadeh, Farhad, et al.
Publicado: (2026)
Can Vision Language Models Understand Mimed Actions?
por: Cho, Hyundong, et al.
Publicado: (2025)
por: Cho, Hyundong, et al.
Publicado: (2025)
Can Large Language Models Infer Causation from Correlation?
por: Jin, Zhijing, et al.
Publicado: (2023)
por: Jin, Zhijing, et al.
Publicado: (2023)
Can Vision-Language Models Solve Visual Math Equations?
por: Choudhury, Monjoy Narayan, et al.
Publicado: (2025)
por: Choudhury, Monjoy Narayan, et al.
Publicado: (2025)
Pretraining Vision-Language Model for Difference Visual Question Answering in Longitudinal Chest X-rays
por: Cho, Yeongjae, et al.
Publicado: (2024)
por: Cho, Yeongjae, et al.
Publicado: (2024)
Pixology: Probing the Linguistic and Visual Capabilities of Pixel-based Language Models
por: Tatariya, Kushal, et al.
Publicado: (2024)
por: Tatariya, Kushal, et al.
Publicado: (2024)
Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues
por: Zhang, Zory, et al.
Publicado: (2025)
por: Zhang, Zory, et al.
Publicado: (2025)
Vision-Grounded Machine Interpreting: Improving the Translation Process through Visual Cues
por: Fantinuoli, Claudio
Publicado: (2025)
por: Fantinuoli, Claudio
Publicado: (2025)
Can Vision Language Models Learn from Visual Demonstrations of Ambiguous Spatial Reasoning?
por: Zhao, Bowen, et al.
Publicado: (2024)
por: Zhao, Bowen, et al.
Publicado: (2024)
Predicting States of Understanding in Explanatory Interactions Using Cognitive Load-Related Linguistic Cues
por: Wang, Yu, et al.
Publicado: (2026)
por: Wang, Yu, et al.
Publicado: (2026)
Beyond Training for Cultural Awareness: The Role of Dataset Linguistic Structure in Large Language Models
por: Masoud, Reem I., et al.
Publicado: (2026)
por: Masoud, Reem I., et al.
Publicado: (2026)
Vision Language Models Cannot Plan, but Can They Formalize?
por: He, Muyu, et al.
Publicado: (2025)
por: He, Muyu, et al.
Publicado: (2025)
How Language Models Prioritize Contextual Grammatical Cues?
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning
por: Wei, Yana, et al.
Publicado: (2025)
por: Wei, Yana, et al.
Publicado: (2025)
Linguistic Minimal Pairs Elicit Linguistic Similarity in Large Language Models
por: Zhou, Xinyu, et al.
Publicado: (2024)
por: Zhou, Xinyu, et al.
Publicado: (2024)
Imperfect Language, Artificial Intelligence, and the Human Mind: An Interdisciplinary Approach to Linguistic Errors in Native Spanish Speakers
por: López, Francisco Portillo
Publicado: (2025)
por: López, Francisco Portillo
Publicado: (2025)
Can Large Language Models Infer Causal Relationships from Real-World Text?
por: Saklad, Ryan, et al.
Publicado: (2025)
por: Saklad, Ryan, et al.
Publicado: (2025)
Large Language Models Can Infer Personality from Free-Form User Interactions
por: Peters, Heinrich, et al.
Publicado: (2024)
por: Peters, Heinrich, et al.
Publicado: (2024)
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
por: Hinojosa, Carlos, et al.
Publicado: (2026)
por: Hinojosa, Carlos, et al.
Publicado: (2026)
Cross-Linguistic Transfer in Multilingual NLP: The Role of Language Families and Morphology
por: Bankula, Ajitesh, et al.
Publicado: (2025)
por: Bankula, Ajitesh, et al.
Publicado: (2025)
SimLM: Can Language Models Infer Parameters of Physical Systems?
por: Memery, Sean, et al.
Publicado: (2023)
por: Memery, Sean, et al.
Publicado: (2023)
Distinguishing Ignorance from Error in LLM Hallucinations
por: Simhi, Adi, et al.
Publicado: (2024)
por: Simhi, Adi, et al.
Publicado: (2024)
Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It)
por: Miaschi, Alessio, et al.
Publicado: (2024)
por: Miaschi, Alessio, et al.
Publicado: (2024)
Enhancing Spoken Discourse Modeling in Language Models Using Gestural Cues
por: Suresh, Varsha, et al.
Publicado: (2025)
por: Suresh, Varsha, et al.
Publicado: (2025)
CALM: Joint Contextual Acoustic-Linguistic Modeling for Personalization of Multi-Speaker ASR
por: Shakeel, Muhammad, et al.
Publicado: (2026)
por: Shakeel, Muhammad, et al.
Publicado: (2026)
Can Authorship Attribution Models Distinguish Speakers in Speech Transcripts?
por: Aggazzotti, Cristina, et al.
Publicado: (2023)
por: Aggazzotti, Cristina, et al.
Publicado: (2023)
Can LLMs Infer Personality from Real World Conversations?
por: Zhu, Jianfeng, et al.
Publicado: (2025)
por: Zhu, Jianfeng, et al.
Publicado: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
por: Seo, Hoigi, et al.
Publicado: (2025)
por: Seo, Hoigi, et al.
Publicado: (2025)
Vision-Language Modeling in PET/CT for Visual Grounding of Positive Findings
por: Huemann, Zachary, et al.
Publicado: (2025)
por: Huemann, Zachary, et al.
Publicado: (2025)
Large Language Models Can Infer Psychological Dispositions of Social Media Users
por: Peters, Heinrich, et al.
Publicado: (2023)
por: Peters, Heinrich, et al.
Publicado: (2023)
Analogical Structure, Minimal Contextual Cues and Contrastive Distractors: Input Design for Sample-Efficient Linguistic Rule Induction
por: Jiang, Chunyang, et al.
Publicado: (2025)
por: Jiang, Chunyang, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating Pragmatic Reasoning in Large Language Models: Evidence from Scalar Diversity
por: Cho, Ye-eun
Publicado: (2026) -
Continuous Interpretive Steering for Scalar Diversity
por: Cho, Ye-eun
Publicado: (2026) -
Cross-Linguistic Transcription and Phonological Representation in the Huìtóngguǎnxì Huáyíyìyǔ
por: Kim, Ji-eun
Publicado: (2026) -
Gender Bias in LLM-generated Interview Responses
por: Kong, Haein, et al.
Publicado: (2024) -
Pragmatic inference of scalar implicature by LLMs
por: Cho, Ye-eun, et al.
Publicado: (2024)