Language Diffusion Models are Associative Memories Capable of Retrieving Unseen Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Pham, Bao, Zaki, Mohammed J., Ambrogioni, Luca, Krotov, Dmitry, Negri, Matteo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Memorization to Generalization: Emergence of Diffusion Models from Associative Memory
di: Pham, Bao, et al.
Pubblicazione: (2025)
di: Pham, Bao, et al.
Pubblicazione: (2025)
The Information Dynamics of Generative Diffusion
di: Stancevic, Dejan, et al.
Pubblicazione: (2025)
di: Stancevic, Dejan, et al.
Pubblicazione: (2025)
Memory in Plain Sight: Surveying the Uncanny Resemblances of Associative Memories and Diffusion Models
di: Hoover, Benjamin, et al.
Pubblicazione: (2023)
di: Hoover, Benjamin, et al.
Pubblicazione: (2023)
Entropic Time Schedulers for Generative Diffusion Models
di: Stancevic, Dejan, et al.
Pubblicazione: (2025)
di: Stancevic, Dejan, et al.
Pubblicazione: (2025)
KERL: Knowledge-Enhanced Personalized Recipe Recommendation using Large Language Models
di: Mohbat, Fnu, et al.
Pubblicazione: (2025)
di: Mohbat, Fnu, et al.
Pubblicazione: (2025)
Deep Clustering with Associative Memories
di: Saha, Bishwajit, et al.
Pubblicazione: (2026)
di: Saha, Bishwajit, et al.
Pubblicazione: (2026)
Modern Methods in Associative Memory
di: Krotov, Dmitry, et al.
Pubblicazione: (2025)
di: Krotov, Dmitry, et al.
Pubblicazione: (2025)
Benchmarking the Medical Understanding and Reasoning of Large Language Models in Arabic Healthcare Tasks
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
Can Interpretation Predict Behavior on Unseen Data?
di: Li, Victoria R., et al.
Pubblicazione: (2025)
di: Li, Victoria R., et al.
Pubblicazione: (2025)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
di: Yang, Zonglin, et al.
Pubblicazione: (2024)
di: Yang, Zonglin, et al.
Pubblicazione: (2024)
Language Model Memory and Memory Models for Language
di: Badger, Benjamin L.
Pubblicazione: (2026)
di: Badger, Benjamin L.
Pubblicazione: (2026)
A Benchmark for Procedural Memory Retrieval in Language Agents
di: Kohar, Ishant, et al.
Pubblicazione: (2025)
di: Kohar, Ishant, et al.
Pubblicazione: (2025)
Generalization v.s. Memorization: Tracing Language Models' Capabilities Back to Pretraining Data
di: Wang, Xinyi, et al.
Pubblicazione: (2024)
di: Wang, Xinyi, et al.
Pubblicazione: (2024)
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?
di: Sakai, Yusuke, et al.
Pubblicazione: (2023)
di: Sakai, Yusuke, et al.
Pubblicazione: (2023)
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data
di: Xiong, Zheyang, et al.
Pubblicazione: (2024)
di: Xiong, Zheyang, et al.
Pubblicazione: (2024)
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
di: Alizadeh, Keivan, et al.
Pubblicazione: (2023)
di: Alizadeh, Keivan, et al.
Pubblicazione: (2023)
Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space
di: Vishwakarma, Gowrav, et al.
Pubblicazione: (2026)
di: Vishwakarma, Gowrav, et al.
Pubblicazione: (2026)
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models
di: Lin, Nianyi, et al.
Pubblicazione: (2025)
di: Lin, Nianyi, et al.
Pubblicazione: (2025)
MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models
di: Xia, Kejing, et al.
Pubblicazione: (2026)
di: Xia, Kejing, et al.
Pubblicazione: (2026)
Diffusion Language Models Can Perform Many Tasks with Scaling and Instruction-Finetuning
di: Ye, Jiasheng, et al.
Pubblicazione: (2023)
di: Ye, Jiasheng, et al.
Pubblicazione: (2023)
On Calibration of Large Language Models: From Response To Capability
di: Yang, Sin-Han, et al.
Pubblicazione: (2026)
di: Yang, Sin-Han, et al.
Pubblicazione: (2026)
Prescriptive Scaling Reveals the Evolution of Language Model Capabilities
di: Zhang, Hanlin, et al.
Pubblicazione: (2026)
di: Zhang, Hanlin, et al.
Pubblicazione: (2026)
Exploring and Benchmarking the Planning Capabilities of Large Language Models
di: Bohnet, Bernd, et al.
Pubblicazione: (2024)
di: Bohnet, Bernd, et al.
Pubblicazione: (2024)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
di: Nguyen, Dang, et al.
Pubblicazione: (2024)
di: Nguyen, Dang, et al.
Pubblicazione: (2024)
Adaptive Guidance for Retrieval-Augmented Masked Diffusion Models
di: Kim, Jaemin, et al.
Pubblicazione: (2026)
di: Kim, Jaemin, et al.
Pubblicazione: (2026)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
di: Wang, Siwei, et al.
Pubblicazione: (2024)
di: Wang, Siwei, et al.
Pubblicazione: (2024)
Instruction Diversity Drives Generalization To Unseen Tasks
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
di: Zhang, Dylan, et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Suicide Detection on Social Media with Limited Labels
di: Nguyen, Vy, et al.
Pubblicazione: (2024)
di: Nguyen, Vy, et al.
Pubblicazione: (2024)
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence
di: Zhang, Xingxuan, et al.
Pubblicazione: (2025)
di: Zhang, Xingxuan, et al.
Pubblicazione: (2025)
Quantifying and Improving the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data
di: Yang, Shiping, et al.
Pubblicazione: (2025)
di: Yang, Shiping, et al.
Pubblicazione: (2025)
CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks
di: Feng, Jie, et al.
Pubblicazione: (2024)
di: Feng, Jie, et al.
Pubblicazione: (2024)
Disentangling Logic: The Role of Context in Large Language Model Reasoning Capabilities
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
EffGen: Enabling Small Language Models as Capable Autonomous Agents
di: Srivastava, Gaurav, et al.
Pubblicazione: (2026)
di: Srivastava, Gaurav, et al.
Pubblicazione: (2026)
Non-instructional Fine-tuning: Enabling Instruction-Following Capabilities in Pre-trained Language Models without Instruction-Following Data
di: Xie, Juncheng, et al.
Pubblicazione: (2024)
di: Xie, Juncheng, et al.
Pubblicazione: (2024)
DiffLM: Controllable Synthetic Data Generation via Diffusion Language Models
di: Zhou, Ying, et al.
Pubblicazione: (2024)
di: Zhou, Ying, et al.
Pubblicazione: (2024)
Reliable, Adaptable, and Attributable Language Models with Retrieval
di: Asai, Akari, et al.
Pubblicazione: (2024)
di: Asai, Akari, et al.
Pubblicazione: (2024)
Evaluating Interventional Reasoning Capabilities of Large Language Models
di: Kasetty, Tejas, et al.
Pubblicazione: (2024)
di: Kasetty, Tejas, et al.
Pubblicazione: (2024)
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series
di: Zhang, Ge, et al.
Pubblicazione: (2024)
di: Zhang, Ge, et al.
Pubblicazione: (2024)
Fundamental Safety-Capability Trade-offs in Fine-tuning Large Language Models
di: Chen, Pin-Yu, et al.
Pubblicazione: (2025)
di: Chen, Pin-Yu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Memorization to Generalization: Emergence of Diffusion Models from Associative Memory
di: Pham, Bao, et al.
Pubblicazione: (2025) -
The Information Dynamics of Generative Diffusion
di: Stancevic, Dejan, et al.
Pubblicazione: (2025) -
Memory in Plain Sight: Surveying the Uncanny Resemblances of Associative Memories and Diffusion Models
di: Hoover, Benjamin, et al.
Pubblicazione: (2023) -
Entropic Time Schedulers for Generative Diffusion Models
di: Stancevic, Dejan, et al.
Pubblicazione: (2025) -
KERL: Knowledge-Enhanced Personalized Recipe Recommendation using Large Language Models
di: Mohbat, Fnu, et al.
Pubblicazione: (2025)