Resilience through Scene Context in Visual Referring Expression Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Junker, Simeon, Zarrieß, Sina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SceneGram: Conceptualizing and Describing Tangrams in Scene Context
por: Junker, Simeon, et al.
Publicado: (2025)
por: Junker, Simeon, et al.
Publicado: (2025)
Are Multimodal Large Language Models Pragmatically Competent Listeners in Simple Reference Resolution Tasks?
por: Junker, Simeon, et al.
Publicado: (2025)
por: Junker, Simeon, et al.
Publicado: (2025)
The Illusion of Competence: Evaluating the Effect of Explanations on Users' Mental Models of Visual Question Answering Systems
por: Sieker, Judith, et al.
Publicado: (2024)
por: Sieker, Judith, et al.
Publicado: (2024)
Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests
por: Ali, Manar, et al.
Publicado: (2026)
por: Ali, Manar, et al.
Publicado: (2026)
Subword models struggle with word learning, but surprisal hides it
por: Bunzeck, Bastian, et al.
Publicado: (2025)
por: Bunzeck, Bastian, et al.
Publicado: (2025)
Child-directed speech facilitates production, not comprehension, in BabyLMs
por: Bunzeck, Bastian, et al.
Publicado: (2026)
por: Bunzeck, Bastian, et al.
Publicado: (2026)
The Frequency Confound in Language-Model Surprisal and Metaphor Novelty
por: Momen, Omar, et al.
Publicado: (2026)
por: Momen, Omar, et al.
Publicado: (2026)
How Hypocritical Is Your LLM judge? Listener-Speaker Asymmetries in the Pragmatic Competence of Large Language Models
por: Sieker, Judith, et al.
Publicado: (2026)
por: Sieker, Judith, et al.
Publicado: (2026)
SemCSE: Semantic Contrastive Sentence Embeddings Using LLM-Generated Summaries For Scientific Abstracts
por: Brinner, Marc, et al.
Publicado: (2025)
por: Brinner, Marc, et al.
Publicado: (2025)
Rationalizing Transformer Predictions via End-To-End Differentiable Self-Training
por: Brinner, Marc, et al.
Publicado: (2025)
por: Brinner, Marc, et al.
Publicado: (2025)
Efficient Scientific Full Text Classification: The Case of EICAT Impact Assessments
por: Brinner, Marc Felix, et al.
Publicado: (2025)
por: Brinner, Marc Felix, et al.
Publicado: (2025)
Model Interpretability and Rationale Extraction by Input Mask Optimization
por: Brinner, Marc, et al.
Publicado: (2025)
por: Brinner, Marc, et al.
Publicado: (2025)
Evaluating Diversity in Automatic Poetry Generation
por: Chen, Yanran, et al.
Publicado: (2024)
por: Chen, Yanran, et al.
Publicado: (2024)
LLMs Struggle to Reject False Presuppositions when Misinformation Stakes are High
por: Sieker, Judith, et al.
Publicado: (2025)
por: Sieker, Judith, et al.
Publicado: (2025)
Do Construction Distributions Shape Formal Language Learning In German BabyLMs?
por: Bunzeck, Bastian, et al.
Publicado: (2025)
por: Bunzeck, Bastian, et al.
Publicado: (2025)
SemCSE-Multi: Multifaceted and Decodable Embeddings for Aspect-Specific and Interpretable Scientific Domain Mapping
por: Brinner, Marc, et al.
Publicado: (2025)
por: Brinner, Marc, et al.
Publicado: (2025)
Enhancing Domain-Specific Encoder Models with LLM-Generated Data: How to Leverage Ontologies, and How to Do Without Them
por: Brinner, Marc, et al.
Publicado: (2025)
por: Brinner, Marc, et al.
Publicado: (2025)
Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs
por: Lachenmaier, Clara, et al.
Publicado: (2026)
por: Lachenmaier, Clara, et al.
Publicado: (2026)
Can LLMs Ground when they (Don't) Know: A Study on Direct and Loaded Political Questions
por: Lachenmaier, Clara, et al.
Publicado: (2025)
por: Lachenmaier, Clara, et al.
Publicado: (2025)
Small Language Models Also Work With Small Vocabularies: Probing the Linguistic Abilities of Grapheme- and Phoneme-Based Baby Llamas
por: Bunzeck, Bastian, et al.
Publicado: (2024)
por: Bunzeck, Bastian, et al.
Publicado: (2024)
Implicit Causality-biases in humans and LLMs as a tool for benchmarking LLM discourse capabilities
por: Kankowski, Florian, et al.
Publicado: (2025)
por: Kankowski, Florian, et al.
Publicado: (2025)
Are BabyLMs Deaf to Gricean Maxims? A Pragmatic Evaluation of Sample-efficient Language Models
por: Askari, Raha, et al.
Publicado: (2025)
por: Askari, Raha, et al.
Publicado: (2025)
The InviTE Corpus: Annotating Invectives in Tudor English Texts for Computational Modeling
por: Spliethoff, Sophie, et al.
Publicado: (2025)
por: Spliethoff, Sophie, et al.
Publicado: (2025)
Surprisal and Metaphor Novelty Judgments: Moderate Correlations and Divergent Scaling Effects Revealed by Corpus-Based and Synthetic Datasets
por: Momen, Omar, et al.
Publicado: (2026)
por: Momen, Omar, et al.
Publicado: (2026)
VAGUE: Visual Contexts Clarify Ambiguous Expressions
por: Nam, Heejeong, et al.
Publicado: (2024)
por: Nam, Heejeong, et al.
Publicado: (2024)
Referring Expression Generation in Visually Grounded Dialogue with Discourse-aware Comprehension Guiding
por: Willemsen, Bram, et al.
Publicado: (2024)
por: Willemsen, Bram, et al.
Publicado: (2024)
Intrinsic Task-based Evaluation for Referring Expression Generation
por: Chen, Guanyi, et al.
Publicado: (2024)
por: Chen, Guanyi, et al.
Publicado: (2024)
Vision-Language Models Are Not Pragmatically Competent in Referring Expression Generation
por: Ma, Ziqiao, et al.
Publicado: (2025)
por: Ma, Ziqiao, et al.
Publicado: (2025)
Dialogue Is Not Enough to Make a Communicative BabyLM (But Neither Is Developmentally Inspired Reinforcement Learning)
por: Padovani, Francesca, et al.
Publicado: (2025)
por: Padovani, Francesca, et al.
Publicado: (2025)
Detecting Referring Expressions in Visually Grounded Dialogue with Autoregressive Language Models
por: Willemsen, Bram, et al.
Publicado: (2025)
por: Willemsen, Bram, et al.
Publicado: (2025)
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks
por: Dong, Qihua, et al.
Publicado: (2026)
por: Dong, Qihua, et al.
Publicado: (2026)
GerPS-Compare: Comparing NER methods for legal norm analysis
por: Bachinger, Sarah T., et al.
Publicado: (2024)
por: Bachinger, Sarah T., et al.
Publicado: (2024)
CoT Referring: Improving Referring Expression Tasks with Grounded Reasoning
por: Dong, Qihua, et al.
Publicado: (2025)
por: Dong, Qihua, et al.
Publicado: (2025)
Transcrib3D: 3D Referring Expression Resolution through Large Language Models
por: Fang, Jiading, et al.
Publicado: (2024)
por: Fang, Jiading, et al.
Publicado: (2024)
Mining for Species, Locations, Habitats, and Ecosystems from Scientific Papers in Invasion Biology: A Large-Scale Exploratory Study with Large Language Models
por: D'Souza, Jennifer, et al.
Publicado: (2025)
por: D'Souza, Jennifer, et al.
Publicado: (2025)
Do they mean 'us'? Interpreting Referring Expressions in Intergroup Bias
por: Govindarajan, Venkata S, et al.
Publicado: (2024)
por: Govindarajan, Venkata S, et al.
Publicado: (2024)
Harlequin: Color-driven Generation of Synthetic Data for Referring Expression Comprehension
por: Parolari, Luca, et al.
Publicado: (2024)
por: Parolari, Luca, et al.
Publicado: (2024)
Generative Visual Commonsense Answering and Explaining with Generative Scene Graph Constructing
por: Yuan, Fan, et al.
Publicado: (2025)
por: Yuan, Fan, et al.
Publicado: (2025)
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language
por: Sennrich, Kilian, et al.
Publicado: (2025)
por: Sennrich, Kilian, et al.
Publicado: (2025)
Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures
por: Inadumi, Shun, et al.
Publicado: (2025)
por: Inadumi, Shun, et al.
Publicado: (2025)
Ejemplares similares
-
SceneGram: Conceptualizing and Describing Tangrams in Scene Context
por: Junker, Simeon, et al.
Publicado: (2025) -
Are Multimodal Large Language Models Pragmatically Competent Listeners in Simple Reference Resolution Tasks?
por: Junker, Simeon, et al.
Publicado: (2025) -
The Illusion of Competence: Evaluating the Effect of Explanations on Users' Mental Models of Visual Question Answering Systems
por: Sieker, Judith, et al.
Publicado: (2024) -
Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests
por: Ali, Manar, et al.
Publicado: (2026) -
Subword models struggle with word learning, but surprisal hides it
por: Bunzeck, Bastian, et al.
Publicado: (2025)