NLKI: A lightweight Natural Language Knowledge Integration Framework for Improving Small VLMs in Commonsense VQA Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Dutta, Aritra, Mukherjee, Swapnanil, Ghosal, Deepanway, Aditya, Somak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
di: Dutta, Aritra, et al.
Pubblicazione: (2026)
di: Dutta, Aritra, et al.
Pubblicazione: (2026)
LOGICPO: Efficient Translation of NL-based Logical Problems to FOL using LLMs and Preference Optimization
di: Viswanadha, Koushik, et al.
Pubblicazione: (2025)
di: Viswanadha, Koushik, et al.
Pubblicazione: (2025)
Filling the Gap: Is Commonsense Knowledge Generation useful for Natural Language Inference?
di: Jayaweera, Chathuri, et al.
Pubblicazione: (2025)
di: Jayaweera, Chathuri, et al.
Pubblicazione: (2025)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
di: Li, Jiachun, et al.
Pubblicazione: (2024)
di: Li, Jiachun, et al.
Pubblicazione: (2024)
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
di: Hong, Pengfei, et al.
Pubblicazione: (2024)
di: Hong, Pengfei, et al.
Pubblicazione: (2024)
What Really is Commonsense Knowledge?
di: Do, Quyet V., et al.
Pubblicazione: (2024)
di: Do, Quyet V., et al.
Pubblicazione: (2024)
The Jumping Reasoning Curve? Tracking the Evolution of Reasoning Performance in GPT-[n] and o-[n] Models on Multimodal Puzzles
di: Toh, Vernon Y. H., et al.
Pubblicazione: (2025)
di: Toh, Vernon Y. H., et al.
Pubblicazione: (2025)
ConstraintChecker: A Plugin for Large Language Models to Reason on Commonsense Knowledge Bases
di: Do, Quyet V., et al.
Pubblicazione: (2024)
di: Do, Quyet V., et al.
Pubblicazione: (2024)
Acquiring and Modelling Abstract Commonsense Knowledge via Conceptualization
di: He, Mutian, et al.
Pubblicazione: (2022)
di: He, Mutian, et al.
Pubblicazione: (2022)
Commonsense Knowledge Editing Based on Free-Text in LLMs
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
Multimodal Commonsense Knowledge Distillation for Visual Question Answering
di: Yang, Shuo, et al.
Pubblicazione: (2024)
di: Yang, Shuo, et al.
Pubblicazione: (2024)
PragWorld: A Benchmark Evaluating LLMs' Local World Model under Minimal Linguistic Alterations and Conversational Dynamics
di: Vashistha, Sachin, et al.
Pubblicazione: (2025)
di: Vashistha, Sachin, et al.
Pubblicazione: (2025)
Complex Reasoning over Logical Queries on Commonsense Knowledge Graphs
di: Fang, Tianqing, et al.
Pubblicazione: (2024)
di: Fang, Tianqing, et al.
Pubblicazione: (2024)
Commonsense for Zero-Shot Natural Language Video Localization
di: Holla, Meghana, et al.
Pubblicazione: (2023)
di: Holla, Meghana, et al.
Pubblicazione: (2023)
Reporting and Analysing the Environmental Impact of Language Models on the Example of Commonsense Question Answering with External Knowledge
di: Usmanova, Aida, et al.
Pubblicazione: (2024)
di: Usmanova, Aida, et al.
Pubblicazione: (2024)
Knowledge Graph Structure as Prompt: Improving Small Language Models Capabilities for Knowledge-based Causal Discovery
di: Susanti, Yuni, et al.
Pubblicazione: (2024)
di: Susanti, Yuni, et al.
Pubblicazione: (2024)
Zero-Shot Commonsense Validation and Reasoning with Large Language Models: An Evaluation on SemEval-2020 Task 4 Dataset
di: Alfugaha, Rawand, et al.
Pubblicazione: (2025)
di: Alfugaha, Rawand, et al.
Pubblicazione: (2025)
Not All Votes Count! Programs as Verifiers Improve Self-Consistency of Language Models for Math Reasoning
di: Toh, Vernon Y. H., et al.
Pubblicazione: (2024)
di: Toh, Vernon Y. H., et al.
Pubblicazione: (2024)
Towards LogiGLUE: A Brief Survey and A Benchmark for Analyzing Logical Reasoning Capabilities of Language Models
di: Luo, Man, et al.
Pubblicazione: (2023)
di: Luo, Man, et al.
Pubblicazione: (2023)
Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling
di: Zou, Hongjian, et al.
Pubblicazione: (2026)
di: Zou, Hongjian, et al.
Pubblicazione: (2026)
Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
di: Majumder, Navonil, et al.
Pubblicazione: (2024)
di: Majumder, Navonil, et al.
Pubblicazione: (2024)
Knowledge Generation for Zero-shot Knowledge-based VQA
di: Cao, Rui, et al.
Pubblicazione: (2024)
di: Cao, Rui, et al.
Pubblicazione: (2024)
Towards Faithful Knowledge Graph Explanation Through Deep Alignment in Commonsense Question Answering
di: Zhai, Weihe, et al.
Pubblicazione: (2023)
di: Zhai, Weihe, et al.
Pubblicazione: (2023)
A Commonsense-Infused Language-Agnostic Learning Framework for Enhancing Prediction of Political Polarity in Multilingual News Headlines
di: Swati, Swati, et al.
Pubblicazione: (2022)
di: Swati, Swati, et al.
Pubblicazione: (2022)
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2024)
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2024)
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
di: Karakaş, Sercan
Pubblicazione: (2026)
di: Karakaş, Sercan
Pubblicazione: (2026)
BrainBench: Exposing the Commonsense Reasoning Gap in Large Language Models
di: Tang, Yuzhe
Pubblicazione: (2026)
di: Tang, Yuzhe
Pubblicazione: (2026)
Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning
di: Sun, Qi, et al.
Pubblicazione: (2024)
di: Sun, Qi, et al.
Pubblicazione: (2024)
Measuring What VLMs Don't Say: Validation Metrics Hide Clinical Terminology Erasure in Radiology Report Generation
di: Parikh, Aditya, et al.
Pubblicazione: (2026)
di: Parikh, Aditya, et al.
Pubblicazione: (2026)
G-SAP: Graph-based Structure-Aware Prompt Learning over Heterogeneous Knowledge for Commonsense Reasoning
di: Dai, Ruiting, et al.
Pubblicazione: (2024)
di: Dai, Ruiting, et al.
Pubblicazione: (2024)
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning
di: Juneja, Gurusha, et al.
Pubblicazione: (2023)
di: Juneja, Gurusha, et al.
Pubblicazione: (2023)
Improving Automatic VQA Evaluation Using Large Language Models
di: Mañas, Oscar, et al.
Pubblicazione: (2023)
di: Mañas, Oscar, et al.
Pubblicazione: (2023)
mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA
di: Zhang, Tao, et al.
Pubblicazione: (2024)
di: Zhang, Tao, et al.
Pubblicazione: (2024)
Can Language Models Take A Hint? Prompting for Controllable Contextualized Commonsense Inference
di: Colon-Hernandez, Pedro, et al.
Pubblicazione: (2024)
di: Colon-Hernandez, Pedro, et al.
Pubblicazione: (2024)
PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions
di: Jin, Sicheng, et al.
Pubblicazione: (2026)
di: Jin, Sicheng, et al.
Pubblicazione: (2026)
Explainability in Neural Networks for Natural Language Processing Tasks
di: Mersha, Melkamu, et al.
Pubblicazione: (2024)
di: Mersha, Melkamu, et al.
Pubblicazione: (2024)
ViCLSR: A Supervised Contrastive Learning Framework with Natural Language Inference for Natural Language Understanding Tasks
di: Van Huynh, Tin, et al.
Pubblicazione: (2026)
di: Van Huynh, Tin, et al.
Pubblicazione: (2026)
Detecting Emotional Incongruity of Sarcasm by Commonsense Reasoning
di: Qiu, Ziqi, et al.
Pubblicazione: (2024)
di: Qiu, Ziqi, et al.
Pubblicazione: (2024)
Estimating Commonsense Plausibility through Semantic Shifts
di: Cui, Wanqing, et al.
Pubblicazione: (2025)
di: Cui, Wanqing, et al.
Pubblicazione: (2025)
LLMBind: A Unified Modality-Task Integration Framework
di: Zhu, Bin, et al.
Pubblicazione: (2024)
di: Zhu, Bin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
di: Dutta, Aritra, et al.
Pubblicazione: (2026) -
LOGICPO: Efficient Translation of NL-based Logical Problems to FOL using LLMs and Preference Optimization
di: Viswanadha, Koushik, et al.
Pubblicazione: (2025) -
Filling the Gap: Is Commonsense Knowledge Generation useful for Natural Language Inference?
di: Jayaweera, Chathuri, et al.
Pubblicazione: (2025) -
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
di: Li, Jiachun, et al.
Pubblicazione: (2024) -
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
di: Hong, Pengfei, et al.
Pubblicazione: (2024)