LVLMs are Bad at Overhearing Human Referential Communication
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zhengxiang, Li, Weiling, Kaliosis, Panagiotis, Rambow, Owen, Brennan, Susan E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LVLMs and Humans Ground Differently in Referential Communication
by: Zeng, Peter, et al.
Published: (2026)
by: Zeng, Peter, et al.
Published: (2026)
Clustering Document Parts: Detecting and Characterizing Influence Campaigns from Documents
by: Wang, Zhengxiang, et al.
Published: (2024)
by: Wang, Zhengxiang, et al.
Published: (2024)
Evaluating LLMs with Multiple Problems at once
by: Wang, Zhengxiang, et al.
Published: (2024)
by: Wang, Zhengxiang, et al.
Published: (2024)
LLMs can Perform Multi-Dimensional Analytic Writing Assessments: A Case Study of L2 Graduate-Level Academic English Writing
by: Wang, Zhengxiang, et al.
Published: (2025)
by: Wang, Zhengxiang, et al.
Published: (2025)
OmniVox: Zero-Shot Emotion Recognition with Omni-LLMs
by: Murzaku, John, et al.
Published: (2025)
by: Murzaku, John, et al.
Published: (2025)
Zero-Shot Belief: A Hard Problem for LLMs
by: Murzaku, John, et al.
Published: (2025)
by: Murzaku, John, et al.
Published: (2025)
Intention and Face in Dialog
by: Soubki, Adil, et al.
Published: (2024)
by: Soubki, Adil, et al.
Published: (2024)
Training LLMs to Recognize Hedges in Spontaneous Narratives
by: Paige, Amie J., et al.
Published: (2024)
by: Paige, Amie J., et al.
Published: (2024)
It Couldn't Help But Overhear: On the Limits of Modelling Meta-Communicative Grounding Acts with Supervised Learning
by: Madureira, Brielen, et al.
Published: (2024)
by: Madureira, Brielen, et al.
Published: (2024)
Examining Gender and Power on Wikipedia Through Face and Politeness
by: Soubki, Adil, et al.
Published: (2024)
by: Soubki, Adil, et al.
Published: (2024)
Overhearing LLM Agents: A Survey, Taxonomy, and Roadmap
by: Zhu, Andrew, et al.
Published: (2025)
by: Zhu, Andrew, et al.
Published: (2025)
Learning Transductions and Alignments with RNN Seq2seq Models
by: Wang, Zhengxiang
Published: (2023)
by: Wang, Zhengxiang
Published: (2023)
A Data-Driven Guided Decoding Mechanism for Diagnostic Captioning
by: Kaliosis, Panagiotis, et al.
Published: (2024)
by: Kaliosis, Panagiotis, et al.
Published: (2024)
Multimodal Belief Prediction
by: Murzaku, John, et al.
Published: (2024)
by: Murzaku, John, et al.
Published: (2024)
Greek2MathTex: A Greek Speech-to-Text Framework for LaTeX Equations Generation
by: Gkritzali, Evangelia, et al.
Published: (2024)
by: Gkritzali, Evangelia, et al.
Published: (2024)
Synthetic Audio Helps for Cognitive State Tasks
by: Soubki, Adil, et al.
Published: (2025)
by: Soubki, Adil, et al.
Published: (2025)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
by: Kaliosis, Panagiotis, et al.
Published: (2025)
by: Kaliosis, Panagiotis, et al.
Published: (2025)
First Steps Towards Overhearing LLM Agents: A Case Study With Dungeons & Dragons Gameplay
by: Zhu, Andrew, et al.
Published: (2025)
by: Zhu, Andrew, et al.
Published: (2025)
Grounding Language in Multi-Perspective Referential Communication
by: Tang, Zineng, et al.
Published: (2024)
by: Tang, Zineng, et al.
Published: (2024)
XAM: Interactive Explainability for Authorship Attribution Models
by: Alshomary, Milad, et al.
Published: (2025)
by: Alshomary, Milad, et al.
Published: (2025)
Measuring Iterative Temporal Reasoning with Time Puzzles
by: Wang, Zhengxiang, et al.
Published: (2026)
by: Wang, Zhengxiang, et al.
Published: (2026)
Gram2Vec: An Interpretable Document Vectorizer
by: Zeng, Peter, et al.
Published: (2024)
by: Zeng, Peter, et al.
Published: (2024)
NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly
by: Fung, Yi R., et al.
Published: (2022)
by: Fung, Yi R., et al.
Published: (2022)
Do LVLMs Know What They Know? A Systematic Study of Knowledge Boundary Perception in LVLMs
by: Ding, Zhikai, et al.
Published: (2025)
by: Ding, Zhikai, et al.
Published: (2025)
Automated Safety Benchmarking: A Multi-agent Pipeline for LVLMs
by: Zhu, Xiangyang, et al.
Published: (2026)
by: Zhu, Xiangyang, et al.
Published: (2026)
Residualized Similarity for Faithfully Explainable Authorship Verification
by: Zeng, Peter, et al.
Published: (2025)
by: Zeng, Peter, et al.
Published: (2025)
Views Are My Own, but Also Yours: Benchmarking Theory of Mind Using Common Ground
by: Soubki, Adil, et al.
Published: (2024)
by: Soubki, Adil, et al.
Published: (2024)
Referencing Where to Focus: Improving VisualGrounding with Referential Query
by: Wang, Yabing, et al.
Published: (2024)
by: Wang, Yabing, et al.
Published: (2024)
Benchmarking and Improving LVLMs on Event Extraction from Multimedia Documents
by: Xing, Fuyu, et al.
Published: (2025)
by: Xing, Fuyu, et al.
Published: (2025)
It Depends: Resolving Referential Ambiguity in Minimal Contexts with Commonsense Knowledge
by: Ellinger, Lukas, et al.
Published: (2025)
by: Ellinger, Lukas, et al.
Published: (2025)
SRPO: Self-Referential Policy Optimization for Vision-Language-Action Models
by: Fei, Senyu, et al.
Published: (2025)
by: Fei, Senyu, et al.
Published: (2025)
Common Objects Out of Context (COOCo): Investigating Multimodal Context and Semantic Scene Violations in Referential Communication
by: Merlo, Filippo, et al.
Published: (2025)
by: Merlo, Filippo, et al.
Published: (2025)
SHIELD: Classifier-Guided Prompting for Robust and Safer LVLMs
by: Ren, Juan, et al.
Published: (2025)
by: Ren, Juan, et al.
Published: (2025)
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
by: Shi, Zhengxiang, et al.
Published: (2023)
by: Shi, Zhengxiang, et al.
Published: (2023)
URL: Universal Referential Knowledge Linking via Task-instructed Representation Compression
by: Li, Zhuoqun, et al.
Published: (2024)
by: Li, Zhuoqun, et al.
Published: (2024)
Personality Editing for Language Models through Adjusting Self-Referential Queries
by: Hwang, Seojin, et al.
Published: (2025)
by: Hwang, Seojin, et al.
Published: (2025)
RAcQUEt: Unveiling the Dangers of Overlooked Referential Ambiguity in Visual LLMs
by: Testoni, Alberto, et al.
Published: (2024)
by: Testoni, Alberto, et al.
Published: (2024)
Active Few-Shot Learning for Text Classification
by: Ahmadnia, Saeed, et al.
Published: (2025)
by: Ahmadnia, Saeed, et al.
Published: (2025)
Improving Alignment in LVLMs with Debiased Self-Judgment
by: Yang, Sihan, et al.
Published: (2025)
by: Yang, Sihan, et al.
Published: (2025)
FKA-Owl: Advancing Multimodal Fake News Detection through Knowledge-Augmented LVLMs
by: Liu, Xuannan, et al.
Published: (2024)
by: Liu, Xuannan, et al.
Published: (2024)
Similar Items
-
LVLMs and Humans Ground Differently in Referential Communication
by: Zeng, Peter, et al.
Published: (2026) -
Clustering Document Parts: Detecting and Characterizing Influence Campaigns from Documents
by: Wang, Zhengxiang, et al.
Published: (2024) -
Evaluating LLMs with Multiple Problems at once
by: Wang, Zhengxiang, et al.
Published: (2024) -
LLMs can Perform Multi-Dimensional Analytic Writing Assessments: A Case Study of L2 Graduate-Level Academic English Writing
by: Wang, Zhengxiang, et al.
Published: (2025) -
OmniVox: Zero-Shot Emotion Recognition with Omni-LLMs
by: Murzaku, John, et al.
Published: (2025)