Pre-training Limited Memory Language Models with Internal and External Knowledge
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Linxi, Zalouk, Sofian, Belardi, Christian K., Lovelace, Justin, Zhou, Jin Peng, Noonan, Ryan Thomas, Go, Dongyoung, Weinberger, Kilian Q., Artzi, Yoav, Sun, Jennifer J. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023)
by: Kishore, Varsha, et al.
Published: (2023)
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
by: Belardi, Christian, et al.
Published: (2026)
by: Belardi, Christian, et al.
Published: (2026)
Prescriptive Scaling Laws for Data Constrained Training
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Diffusion Guided Language Modeling
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
Sample-Efficient Diffusion for Text-To-Speech Synthesis
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
SpeechOp: Inference-Time Task Composition for Generative Speech Processing
by: Lovelace, Justin, et al.
Published: (2025)
by: Lovelace, Justin, et al.
Published: (2025)
Improving Multislice Electron Ptychography with a Generative Prior
by: Belardi, Christian K., et al.
Published: (2025)
by: Belardi, Christian K., et al.
Published: (2025)
Post-training for Efficient Communication via Convention Formation
by: Hua, Yilun, et al.
Published: (2025)
by: Hua, Yilun, et al.
Published: (2025)
Talk Less, Interact Better: Evaluating In-context Conversational Adaptation in Multimodal LLMs
by: Hua, Yilun, et al.
Published: (2024)
by: Hua, Yilun, et al.
Published: (2024)
No Mean Feat: Simple, Strong Baselines for Context Compression
by: Feldman, Yair, et al.
Published: (2025)
by: Feldman, Yair, et al.
Published: (2025)
Lost in Backpropagation: The LM Head is a Gradient Bottleneck
by: Godey, Nathan, et al.
Published: (2026)
by: Godey, Nathan, et al.
Published: (2026)
Knot So Simple: A Minimalistic Environment for Spatial Reasoning
by: Chen, Zizhao, et al.
Published: (2025)
by: Chen, Zizhao, et al.
Published: (2025)
Music Transcription with (Almost) No Supervision
by: Shin, Saebyeol, et al.
Published: (2026)
by: Shin, Saebyeol, et al.
Published: (2026)
On Speeding Up Language Model Evaluation
by: Zhou, Jin Peng, et al.
Published: (2024)
by: Zhou, Jin Peng, et al.
Published: (2024)
CoGen: Learning from Feedback with Coupled Comprehension and Generation
by: Gul, Mustafa Omer, et al.
Published: (2024)
by: Gul, Mustafa Omer, et al.
Published: (2024)
Breadcrumbs Reasoning: Memory-Efficient Reasoning with Compression Beacons
by: Monea, Giovanni, et al.
Published: (2025)
by: Monea, Giovanni, et al.
Published: (2025)
Learning from Synthetic Data Improves Multi-hop Reasoning
by: Kabra, Anmol, et al.
Published: (2026)
by: Kabra, Anmol, et al.
Published: (2026)
A Surprising Failure? Multimodal LLMs and the NLVR Challenge
by: Wu, Anne, et al.
Published: (2024)
by: Wu, Anne, et al.
Published: (2024)
INPROVF: Leveraging Large Language Models to Repair High-level Robot Controllers from Assumption Violations
by: Meng, Qian, et al.
Published: (2025)
by: Meng, Qian, et al.
Published: (2025)
SteP: Stacked LLM Policies for Web Actions
by: Sodhi, Paloma, et al.
Published: (2023)
by: Sodhi, Paloma, et al.
Published: (2023)
A Joint Study of Phrase Grounding and Task Performance in Vision and Language Models
by: Kojima, Noriyuki, et al.
Published: (2023)
by: Kojima, Noriyuki, et al.
Published: (2023)
LLMs Are In-Context Bandit Reinforcement Learners
by: Monea, Giovanni, et al.
Published: (2024)
by: Monea, Giovanni, et al.
Published: (2024)
Success and Cost Elicit Convention Formation for Efficient Communication
by: Vaduguru, Saujas, et al.
Published: (2025)
by: Vaduguru, Saujas, et al.
Published: (2025)
Internal sums for synthetic fibered $(\infty,1)$-categories
by: Weinberger, Jonathan
Published: (2022)
by: Weinberger, Jonathan
Published: (2022)
Correction with Backtracking Reduces Hallucination in Summarization
by: Liu, Zhenzhen, et al.
Published: (2023)
by: Liu, Zhenzhen, et al.
Published: (2023)
Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond
by: Wang, Qizhou, et al.
Published: (2025)
by: Wang, Qizhou, et al.
Published: (2025)
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization
by: Zhou, Jin Peng, et al.
Published: (2024)
by: Zhou, Jin Peng, et al.
Published: (2024)
Learning to decode logical circuits
by: Zhou, Yiqing, et al.
Published: (2025)
by: Zhou, Yiqing, et al.
Published: (2025)
Orchestrating LLMs with Different Personalizations
by: Zhou, Jin Peng, et al.
Published: (2024)
by: Zhou, Jin Peng, et al.
Published: (2024)
Imitation Learning from a Single Temporally Misaligned Video
by: Huey, William, et al.
Published: (2025)
by: Huey, William, et al.
Published: (2025)
Sparsity-Based Interpolation of External, Internal and Swap Regret
by: Lu, Zhou, et al.
Published: (2025)
by: Lu, Zhou, et al.
Published: (2025)
Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory
by: Li, Ruiyin, et al.
Published: (2026)
by: Li, Ruiyin, et al.
Published: (2026)
The Utilization of Internal and External Memory Strategies in Evidence-Based Practice. EBP Briefs. Volume 14, Issue 1
by: Davis, Emily, et al.
Published: (2019)
by: Davis, Emily, et al.
Published: (2019)
Transparentize the Internal and External Knowledge Utilization in LLMs with Trustworthy Citation
by: Shen, Jiajun, et al.
Published: (2025)
by: Shen, Jiajun, et al.
Published: (2025)
Internal and External Knowledge Interactive Refinement Framework for Knowledge-Intensive Question Answering
by: Du, Haowei, et al.
Published: (2024)
by: Du, Haowei, et al.
Published: (2024)
Graders should cheat: privileged information enables expert-level automated evaluations
by: Zhou, Jin Peng, et al.
Published: (2025)
by: Zhou, Jin Peng, et al.
Published: (2025)
Chapter L’invenzione dei percorsi pedonali meccanizzati. Dalla città delle automobili alla città dei pedoni
by: Belardi, Paolo
Published: (2024)
by: Belardi, Paolo
Published: (2024)
Chapter Da Perugia a Genova e poi ancora a Perugia: sui “disegni regolatori” di Galeazzo Alessi
by: Belardi, Paolo
Published: (2022)
by: Belardi, Paolo
Published: (2022)
Dr. Jorge Adolfo Albertal (1934-2007)
by: Jorge Belardi
Published: (2007)
by: Jorge Belardi
Published: (2007)
Similar Items
-
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026) -
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023) -
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
by: Belardi, Christian, et al.
Published: (2026) -
Prescriptive Scaling Laws for Data Constrained Training
by: Lovelace, Justin, et al.
Published: (2026) -
Diffusion Guided Language Modeling
by: Lovelace, Justin, et al.
Published: (2024)