MMCOMET: A Large-Scale Multimodal Commonsense Knowledge Graph for Contextual Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Eileen, Arnaout, Hiba, Pratama, Dhita, Yang, Shuo, Liu, Dangyang, Yang, Jie, Poon, Josiah, Pan, Jeff, Han, Caren |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling
by: Wang, Eileen, et al.
Published: (2024)
by: Wang, Eileen, et al.
Published: (2024)
GEM-VPC: A dual Graph-Enhanced Multimodal integration for Video Paragraph Captioning
by: Wang, Eileen, et al.
Published: (2024)
by: Wang, Eileen, et al.
Published: (2024)
Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs
by: Pratama, Dhita Putri, et al.
Published: (2026)
by: Pratama, Dhita Putri, et al.
Published: (2026)
Multimodal Commonsense Knowledge Distillation for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Multimodal Large Language Models and Tunings: Vision, Language, Sensors, Audio, and Beyond
by: Han, Soyeon Caren, et al.
Published: (2024)
by: Han, Soyeon Caren, et al.
Published: (2024)
3M-Health: Multimodal Multi-Teacher Knowledge Distillation for Mental Health Detection
by: Cabral, Rina Carines, et al.
Published: (2024)
by: Cabral, Rina Carines, et al.
Published: (2024)
Beyond Size and Class Balance: Alpha as a New Dataset Quality Metric for Deep Learning
by: Couch, Josiah, et al.
Published: (2024)
by: Couch, Josiah, et al.
Published: (2024)
Local Interpretations for Explainable Natural Language Processing: A Survey
by: Luo, Siwen, et al.
Published: (2021)
by: Luo, Siwen, et al.
Published: (2021)
Game-MUG: Multimodal Oriented Game Situation Understanding and Commentary Generation Dataset
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
VRD-IU: Lessons from Visually Rich Document Intelligence and Understanding
by: Ding, Yihao, et al.
Published: (2025)
by: Ding, Yihao, et al.
Published: (2025)
When More Is Less: A Systematic Analysis of Spatial and Commonsense Information for Visual Spatial Reasoning
by: Akasaka, Muku, et al.
Published: (2026)
by: Akasaka, Muku, et al.
Published: (2026)
X-Factor: Quality Is a Dataset-Intrinsic Property
by: Couch, Josiah, et al.
Published: (2025)
by: Couch, Josiah, et al.
Published: (2025)
3MVRD: Multimodal Multi-task Multi-teacher Visually-Rich Form Document Understanding
by: Ding, Yihao, et al.
Published: (2024)
by: Ding, Yihao, et al.
Published: (2024)
DiffuCOMET: Contextual Commonsense Knowledge Diffusion
by: Gao, Silin, et al.
Published: (2024)
by: Gao, Silin, et al.
Published: (2024)
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning
by: Zheng, Danna, et al.
Published: (2024)
by: Zheng, Danna, et al.
Published: (2024)
Complex Reasoning over Logical Queries on Commonsense Knowledge Graphs
by: Fang, Tianqing, et al.
Published: (2024)
by: Fang, Tianqing, et al.
Published: (2024)
Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning
by: Zhao, Yu, et al.
Published: (2025)
by: Zhao, Yu, et al.
Published: (2025)
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models
by: Wang, Yuqing, et al.
Published: (2023)
by: Wang, Yuqing, et al.
Published: (2023)
MSG-Chart: Multimodal Scene Graph for ChartQA
by: Dai, Yue, et al.
Published: (2024)
by: Dai, Yue, et al.
Published: (2024)
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
by: Toroghi, Armin, et al.
Published: (2024)
by: Toroghi, Armin, et al.
Published: (2024)
In-depth Research Impact Summarization through Fine-Grained Temporal Citation Analysis
by: Arnaout, Hiba, et al.
Published: (2025)
by: Arnaout, Hiba, et al.
Published: (2025)
Utilizing Sequential Information of General Lab-test Results and Diagnoses History for Differential Diagnosis of Dementia
by: Xing, Yizong, et al.
Published: (2025)
by: Xing, Yizong, et al.
Published: (2025)
Graph-Based Multimodal Contrastive Learning for Chart Question Answering
by: Dai, Yue, et al.
Published: (2025)
by: Dai, Yue, et al.
Published: (2025)
Joint Multi-Facts Reasoning Network For Complex Temporal Question Answering Over Knowledge Graph
by: Huang, Rikui, et al.
Published: (2024)
by: Huang, Rikui, et al.
Published: (2024)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
by: Li, Jiachun, et al.
Published: (2024)
by: Li, Jiachun, et al.
Published: (2024)
TriG-NER: Triplet-Grid Framework for Discontinuous Named Entity Recognition
by: Cabral, Rina Carines, et al.
Published: (2024)
by: Cabral, Rina Carines, et al.
Published: (2024)
Isoperimetric Inequality for degenerate elliptic operators of Grushin type
by: He, Dangyang
Published: (2026)
by: He, Dangyang
Published: (2026)
Reverse Riesz Inequality on Manifolds with Ends
by: He, Dangyang
Published: (2024)
by: He, Dangyang
Published: (2024)
On the Riesz transform and its reverse inequality on manifolds with quadratically decaying curvature
by: He, Dangyang
Published: (2025)
by: He, Dangyang
Published: (2025)
On the Reverse Inequality of Riesz transform on metric cone with potential
by: He, Dangyang
Published: (2025)
by: He, Dangyang
Published: (2025)
Some Remarks on the Riesz and reverse Riesz transforms on Broken Line
by: He, Dangyang
Published: (2025)
by: He, Dangyang
Published: (2025)
Isoperimetric Inequality on Manifolds with Quadratically Decaying Curvature
by: He, Dangyang
Published: (2025)
by: He, Dangyang
Published: (2025)
BUCA: A Binary Classification Approach to Unsupervised Commonsense Question Answering
by: He, Jie, et al.
Published: (2023)
by: He, Jie, et al.
Published: (2023)
Meta-RTL: Reinforcement-Based Meta-Transfer Learning for Low-Resource Commonsense Reasoning
by: Fu, Yu, et al.
Published: (2024)
by: Fu, Yu, et al.
Published: (2024)
Using Large Language Models to Create Personalized Networks From Therapy Sessions
by: Ong, Clarissa W., et al.
Published: (2025)
by: Ong, Clarissa W., et al.
Published: (2025)
Abductive Reasoning with Probabilistic Commonsense
by: Cotnareanu, Joseph, et al.
Published: (2026)
by: Cotnareanu, Joseph, et al.
Published: (2026)
PEACH: Pretrained-embedding Explanation Across Contextual and Hierarchical Structure
by: Cao, Feiqi, et al.
Published: (2024)
by: Cao, Feiqi, et al.
Published: (2024)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
by: Tonga, Junior Cedric, et al.
Published: (2026)
by: Tonga, Junior Cedric, et al.
Published: (2026)
Which Similarity-Sensitive Entropy (Sentropy)?
by: Nguyen, Phuc, et al.
Published: (2025)
by: Nguyen, Phuc, et al.
Published: (2025)
Similar Items
-
SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling
by: Wang, Eileen, et al.
Published: (2024) -
GEM-VPC: A dual Graph-Enhanced Multimodal integration for Video Paragraph Captioning
by: Wang, Eileen, et al.
Published: (2024) -
Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs
by: Pratama, Dhita Putri, et al.
Published: (2026) -
Multimodal Commonsense Knowledge Distillation for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2024) -
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2025)