Data Distribution Bottlenecks in Grounding Language Models to Knowledge Bases
Fuente:
arXiv
Saved in:
| Main Authors: | Shu, Yiheng, Yu, Zhiwei |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2024)
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2024)
Language Bottleneck Models for Qualitative Knowledge State Modeling
by: Berthon, Antonin, et al.
Published: (2025)
by: Berthon, Antonin, et al.
Published: (2025)
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2025)
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2025)
Improving Scientific Hypothesis Generation with Knowledge Grounded Large Language Models
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Diagnosing Bottlenecks in Data Visualization Understanding by Vision-Language Models
by: Tartaglini, Alexa R., et al.
Published: (2025)
by: Tartaglini, Alexa R., et al.
Published: (2025)
Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2026)
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2026)
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
by: He, Bowei, et al.
Published: (2026)
by: He, Bowei, et al.
Published: (2026)
AGENTCL: Toward Rigorous Evaluation of Continual Learning in Language Agents
by: Shu, Yiheng, et al.
Published: (2026)
by: Shu, Yiheng, et al.
Published: (2026)
Grounding Synthetic Data Evaluations of Language Models in Unsupervised Document Corpora
by: Majurski, Michael, et al.
Published: (2025)
by: Majurski, Michael, et al.
Published: (2025)
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
by: Li, Yinghui, et al.
Published: (2025)
by: Li, Yinghui, et al.
Published: (2025)
Contrastive Learning for Knowledge-Based Question Generation in Large Language Models
by: Zhang, Zhenhong, et al.
Published: (2024)
by: Zhang, Zhenhong, et al.
Published: (2024)
Knowledge-Grounded Agentic Large Language Models for Multi-Hazard Understanding from Reconnaissance Reports
by: Kuai, Chenchen, et al.
Published: (2025)
by: Kuai, Chenchen, et al.
Published: (2025)
AdaCultureSafe: Adaptive Cultural Safety Grounded by Cultural Knowledge in Large Language Models
by: Kang, Hankun, et al.
Published: (2026)
by: Kang, Hankun, et al.
Published: (2026)
Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models
by: He-Yueya, Joy, et al.
Published: (2024)
by: He-Yueya, Joy, et al.
Published: (2024)
The Alignment Bottleneck in Decomposition-Based Claim Verification
by: Akhter, Mahmud Elahi, et al.
Published: (2026)
by: Akhter, Mahmud Elahi, et al.
Published: (2026)
Exploring Information Processing in Large Language Models: Insights from Information Bottleneck Theory
by: Yang, Zhou, et al.
Published: (2025)
by: Yang, Zhou, et al.
Published: (2025)
Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents
by: Gou, Boyu, et al.
Published: (2024)
by: Gou, Boyu, et al.
Published: (2024)
KBLaM: Knowledge Base augmented Language Model
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
Reasoning Factual Knowledge in Structured Data with Large Language Models
by: Huang, Sirui, et al.
Published: (2024)
by: Huang, Sirui, et al.
Published: (2024)
KICGPT: Large Language Model with Knowledge in Context for Knowledge Graph Completion
by: Wei, Yanbin, et al.
Published: (2024)
by: Wei, Yanbin, et al.
Published: (2024)
The Tokenization Bottleneck: How Vocabulary Extension Improves Chemistry Representation Learning in Pretrained Language Models
by: Kalamkar, Prathamesh, et al.
Published: (2025)
by: Kalamkar, Prathamesh, et al.
Published: (2025)
Symbolic Grounding Reveals Representational Bottlenecks in Abstract Visual Reasoning
by: Vaishnav, Mohit, et al.
Published: (2026)
by: Vaishnav, Mohit, et al.
Published: (2026)
Language Models Benefit from Preparation with Elicited Knowledge
by: Yu, Jiacan, et al.
Published: (2024)
by: Yu, Jiacan, et al.
Published: (2024)
Language Models as Knowledge Bases for Visual Word Sense Disambiguation
by: Kritharoula, Anastasia, et al.
Published: (2023)
by: Kritharoula, Anastasia, et al.
Published: (2023)
DeepWriter: A Fact-Grounded Multimodal Writing Assistant Based On Offline Knowledge Base
by: Mao, Song, et al.
Published: (2025)
by: Mao, Song, et al.
Published: (2025)
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders
by: Zhu, Xiaofeng, et al.
Published: (2024)
by: Zhu, Xiaofeng, et al.
Published: (2024)
Culturally-Grounded Governance for Multilingual Language Models: Rights, Data Boundaries, and Accountable AI Design
by: Shi, Hanjing, et al.
Published: (2026)
by: Shi, Hanjing, et al.
Published: (2026)
Knowledge Reasoning Language Model: Unifying Knowledge and Language for Inductive Knowledge Graph Reasoning
by: Zhuo, Xingrui, et al.
Published: (2025)
by: Zhuo, Xingrui, et al.
Published: (2025)
Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
by: Choi, Minje, et al.
Published: (2023)
by: Choi, Minje, et al.
Published: (2023)
Policy Learning with a Language Bottleneck
by: Srivastava, Megha, et al.
Published: (2024)
by: Srivastava, Megha, et al.
Published: (2024)
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
by: Kim, Minsung, et al.
Published: (2025)
by: Kim, Minsung, et al.
Published: (2025)
Cross-Lingual Consistency: A Novel Inference Framework for Advancing Reasoning in Large Language Models
by: Yu, Zhiwei, et al.
Published: (2025)
by: Yu, Zhiwei, et al.
Published: (2025)
Knowledge Bases in Support of Large Language Models for Processing Web News
by: Zhang, Yihe, et al.
Published: (2024)
by: Zhang, Yihe, et al.
Published: (2024)
Knowledge-tuning Large Language Models with Structured Medical Knowledge Bases for Reliable Response Generation in Chinese
by: Wang, Haochun, et al.
Published: (2023)
by: Wang, Haochun, et al.
Published: (2023)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Improving Neural Topic Modeling with Semantically-Grounded Soft Label Distributions
by: Li, Raymond, et al.
Published: (2026)
by: Li, Raymond, et al.
Published: (2026)
Identifying Knowledge Editing Types in Large Language Models
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
"Understanding AI": Semantic Grounding in Large Language Models
by: Lyre, Holger
Published: (2024)
by: Lyre, Holger
Published: (2024)
Leveraging Online Data to Enhance Medical Knowledge in a Small Persian Language Model
by: Ghassabi, Mehrdad, et al.
Published: (2025)
by: Ghassabi, Mehrdad, et al.
Published: (2025)
Integrating Large Language Models and Knowledge Graphs for Extraction and Validation of Textual Test Data
by: De Santis, Antonio, et al.
Published: (2024)
by: De Santis, Antonio, et al.
Published: (2024)
Similar Items
-
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2024) -
Language Bottleneck Models for Qualitative Knowledge State Modeling
by: Berthon, Antonin, et al.
Published: (2025) -
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2025) -
Improving Scientific Hypothesis Generation with Knowledge Grounded Large Language Models
by: Xiong, Guangzhi, et al.
Published: (2024) -
Diagnosing Bottlenecks in Data Visualization Understanding by Vision-Language Models
by: Tartaglini, Alexa R., et al.
Published: (2025)