Don't Learn, Ground: A Case for Natural Language Inference with Visual Grounding
Fuente:
arXiv
Salvato in:
| Autori principali: | Ignatev, Daniil, Santeer, Ayman, Gatt, Albert, Paperno, Denis |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning
di: Song, Yingjin, et al.
Pubblicazione: (2024)
di: Song, Yingjin, et al.
Pubblicazione: (2024)
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
di: Song, Yingjin, et al.
Pubblicazione: (2025)
di: Song, Yingjin, et al.
Pubblicazione: (2025)
VAQUUM: Are Vague Quantifiers Grounded in Visual Data?
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025)
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025)
Hypernetworks for Perspectivist Adaptation
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
A Systematic Analysis of Large Language Models as Soft Reasoners: The Case of Syllogistic Inferences
di: Bertolazzi, Leonardo, et al.
Pubblicazione: (2024)
di: Bertolazzi, Leonardo, et al.
Pubblicazione: (2024)
Disentangling the Roles of Representation and Selection in Data Pruning
di: Du, Yupei, et al.
Pubblicazione: (2025)
di: Du, Yupei, et al.
Pubblicazione: (2025)
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
di: Li, Nan, et al.
Pubblicazione: (2025)
di: Li, Nan, et al.
Pubblicazione: (2025)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
di: Evelo, Bart, et al.
Pubblicazione: (2026)
di: Evelo, Bart, et al.
Pubblicazione: (2026)
Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth
di: Gur-Arieh, Yoav, et al.
Pubblicazione: (2026)
di: Gur-Arieh, Yoav, et al.
Pubblicazione: (2026)
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization
di: Zhou, Jin Peng, et al.
Pubblicazione: (2024)
di: Zhou, Jin Peng, et al.
Pubblicazione: (2024)
Can LLMs Ground when they (Don't) Know: A Study on Direct and Loaded Political Questions
di: Lachenmaier, Clara, et al.
Pubblicazione: (2025)
di: Lachenmaier, Clara, et al.
Pubblicazione: (2025)
From Image Captioning to Visual Storytelling
di: Passadakis, Admitos, et al.
Pubblicazione: (2025)
di: Passadakis, Admitos, et al.
Pubblicazione: (2025)
Morphological Analysis for the Maltese Language: The Challenges of a Hybrid System
di: Borg, Claudia, et al.
Pubblicazione: (2017)
di: Borg, Claudia, et al.
Pubblicazione: (2017)
Language Models Don't Learn the Physical Manifestation of Language
di: Lee, Bruce W., et al.
Pubblicazione: (2024)
di: Lee, Bruce W., et al.
Pubblicazione: (2024)
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
di: Schaffelder, Max, et al.
Pubblicazione: (2025)
di: Schaffelder, Max, et al.
Pubblicazione: (2025)
Hatevolution: What Static Benchmarks Don't Tell Us
di: Di Bonaventura, Chiara, et al.
Pubblicazione: (2025)
di: Di Bonaventura, Chiara, et al.
Pubblicazione: (2025)
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
di: Goloburda, Maiya, et al.
Pubblicazione: (2026)
di: Goloburda, Maiya, et al.
Pubblicazione: (2026)
Toward Culturally Grounded Natural Language Processing
di: Nezhad, Sina Bagheri
Pubblicazione: (2026)
di: Nezhad, Sina Bagheri
Pubblicazione: (2026)
Think, But Don't Overthink: Reproducing Recursive Language Models
di: Wang, Daren
Pubblicazione: (2026)
di: Wang, Daren
Pubblicazione: (2026)
Don't Touch My Diacritics
di: Gorman, Kyle, et al.
Pubblicazione: (2024)
di: Gorman, Kyle, et al.
Pubblicazione: (2024)
Language-Conditioned Visual Grounding with CLIP Multilingual
di: de Curtò, J., et al.
Pubblicazione: (2026)
di: de Curtò, J., et al.
Pubblicazione: (2026)
Human Label Variation in Implicit Discourse Relation Recognition
di: Yung, Frances, et al.
Pubblicazione: (2026)
di: Yung, Frances, et al.
Pubblicazione: (2026)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
di: Tyukin, Georgy, et al.
Pubblicazione: (2024)
di: Tyukin, Georgy, et al.
Pubblicazione: (2024)
Don't Pay Attention
di: Hammoud, Mohammad, et al.
Pubblicazione: (2025)
di: Hammoud, Mohammad, et al.
Pubblicazione: (2025)
Towards Knowledge-Grounded Natural Language Understanding and Generation
di: Whitehouse, Chenxi
Pubblicazione: (2024)
di: Whitehouse, Chenxi
Pubblicazione: (2024)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
di: Parmar, Jupinder, et al.
Pubblicazione: (2024)
di: Parmar, Jupinder, et al.
Pubblicazione: (2024)
CV-Probes: Studying the interplay of lexical and world knowledge in visually grounded verb understanding
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
Probing Omissions and Distortions in Transformer-based RDF-to-Text Models
di: Faille, Juliette, et al.
Pubblicazione: (2024)
di: Faille, Juliette, et al.
Pubblicazione: (2024)
DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
Reasoning-Grounded Natural Language Explanations for Language Models
di: Cahlik, Vojtech, et al.
Pubblicazione: (2025)
di: Cahlik, Vojtech, et al.
Pubblicazione: (2025)
When Models Decide and When They Bind: A Two-Stage Computation for Multiple-Choice Question-Answering
di: Wong, Hugh Mee, et al.
Pubblicazione: (2026)
di: Wong, Hugh Mee, et al.
Pubblicazione: (2026)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
di: Yang, Shuo, et al.
Pubblicazione: (2025)
di: Yang, Shuo, et al.
Pubblicazione: (2025)
Grounding Language Models for Visual Entity Recognition
di: Xiao, Zilin, et al.
Pubblicazione: (2024)
di: Xiao, Zilin, et al.
Pubblicazione: (2024)
Automatic Metrics in Natural Language Generation: A Survey of Current Evaluation Practices
di: Schmidtová, Patrícia, et al.
Pubblicazione: (2024)
di: Schmidtová, Patrícia, et al.
Pubblicazione: (2024)
If You Don't Understand It, Don't Use It: Eliminating Trojans with Filters Between Layers
di: Hernandez, Adriano
Pubblicazione: (2024)
di: Hernandez, Adriano
Pubblicazione: (2024)
ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs
di: Wang, Zhipin, et al.
Pubblicazione: (2026)
di: Wang, Zhipin, et al.
Pubblicazione: (2026)
Naturally Supervised 3D Visual Grounding with Language-Regularized Concept Learners
di: Feng, Chun, et al.
Pubblicazione: (2024)
di: Feng, Chun, et al.
Pubblicazione: (2024)
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
di: Tian, Yuan, et al.
Pubblicazione: (2024)
di: Tian, Yuan, et al.
Pubblicazione: (2024)
Don't Throw Away Your Pretrained Model
di: Feng, Shangbin, et al.
Pubblicazione: (2025)
di: Feng, Shangbin, et al.
Pubblicazione: (2025)
Don't Say No: Jailbreaking LLM by Suppressing Refusal
di: Zhou, Yukai, et al.
Pubblicazione: (2024)
di: Zhou, Yukai, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning
di: Song, Yingjin, et al.
Pubblicazione: (2024) -
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
di: Song, Yingjin, et al.
Pubblicazione: (2025) -
VAQUUM: Are Vague Quantifiers Grounded in Visual Data?
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025) -
Hypernetworks for Perspectivist Adaptation
di: Ignatev, Daniil, et al.
Pubblicazione: (2025) -
A Systematic Analysis of Large Language Models as Soft Reasoners: The Case of Syllogistic Inferences
di: Bertolazzi, Leonardo, et al.
Pubblicazione: (2024)