If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
Fuente:
arXiv
Salvato in:
| Autori principali: | Yoshida, Ryo, Isono, Shinnosuke, Kajikawa, Kohei, Someya, Taiga, Sugimoto, Yushi, Oseki, Yohei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
di: Yoshida, Ryo, et al.
Pubblicazione: (2026)
di: Yoshida, Ryo, et al.
Pubblicazione: (2026)
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
di: Yoshida, Ryo, et al.
Pubblicazione: (2024)
di: Yoshida, Ryo, et al.
Pubblicazione: (2024)
Language Acquisition Device in Large Language Models
di: Mita, Masato, et al.
Pubblicazione: (2026)
di: Mita, Masato, et al.
Pubblicazione: (2026)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
di: Someya, Taiga, et al.
Pubblicazione: (2025)
di: Someya, Taiga, et al.
Pubblicazione: (2025)
Syntactically-guided Information Maintenance in Sentence Comprehension
di: Isono, Shinnosuke, et al.
Pubblicazione: (2026)
di: Isono, Shinnosuke, et al.
Pubblicazione: (2026)
Rethinking the Relationship between the Power Law and Hierarchical Structures
di: Nakaishi, Kai, et al.
Pubblicazione: (2025)
di: Nakaishi, Kai, et al.
Pubblicazione: (2025)
Composition, Attention, or Both?
di: Yoshida, Ryo, et al.
Pubblicazione: (2022)
di: Yoshida, Ryo, et al.
Pubblicazione: (2022)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
di: Mita, Masato, et al.
Pubblicazione: (2025)
di: Mita, Masato, et al.
Pubblicazione: (2025)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
di: Kajikawa, Kohei, et al.
Pubblicazione: (2024)
di: Kajikawa, Kohei, et al.
Pubblicazione: (2024)
Information-Theoretic Storage Cost in Sentence Comprehension
di: Kajikawa, Kohei, et al.
Pubblicazione: (2026)
di: Kajikawa, Kohei, et al.
Pubblicazione: (2026)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
di: Yoshida, Ryo, et al.
Pubblicazione: (2021)
di: Yoshida, Ryo, et al.
Pubblicazione: (2021)
Emergent Word Order Universals from Cognitively-Motivated Language Models
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2024)
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2026)
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2026)
Psychometric Predictive Power of Large Language Models
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2023)
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2023)
Large Language Models Are Human-Like Internally
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2025)
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
di: Haga, Akari, et al.
Pubblicazione: (2024)
di: Haga, Akari, et al.
Pubblicazione: (2024)
Can Language Models Learn Typologically Implausible Languages?
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
Leveraging Human Production-Interpretation Asymmetries to Test LLM Cognitive Plausibility
di: Lam, Suet-Ying, et al.
Pubblicazione: (2025)
di: Lam, Suet-Ying, et al.
Pubblicazione: (2025)
Exclusive Unlearning
di: Sasaki, Mutsumi, et al.
Pubblicazione: (2026)
di: Sasaki, Mutsumi, et al.
Pubblicazione: (2026)
MLP Memory: A Retriever-Pretrained Memory for Large Language Models
di: Wei, Rubin, et al.
Pubblicazione: (2025)
di: Wei, Rubin, et al.
Pubblicazione: (2025)
How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders
di: Inaba, Tatsuro, et al.
Pubblicazione: (2025)
di: Inaba, Tatsuro, et al.
Pubblicazione: (2025)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
di: Lepori, Michael A., et al.
Pubblicazione: (2025)
di: Lepori, Michael A., et al.
Pubblicazione: (2025)
Information Locality as an Inductive Bias for Neural Language Models
di: Someya, Taiga, et al.
Pubblicazione: (2025)
di: Someya, Taiga, et al.
Pubblicazione: (2025)
Cognitive Memory in Large Language Models
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
Learning What to Remember: Adaptive Probabilistic Memory Retention for Memory-Efficient Language Models
di: Rafiuddin, S M, et al.
Pubblicazione: (2025)
di: Rafiuddin, S M, et al.
Pubblicazione: (2025)
S$^3$-Attention:Attention-Aligned Endogenous Retrieval for Memory-Bounded Long-Context Inference
di: Ma, Qingsen, et al.
Pubblicazione: (2026)
di: Ma, Qingsen, et al.
Pubblicazione: (2026)
Can Language Models Induce Grammatical Knowledge from Indirect Evidence?
di: Oba, Miyu, et al.
Pubblicazione: (2024)
di: Oba, Miyu, et al.
Pubblicazione: (2024)
Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory
di: Xu, Derong, et al.
Pubblicazione: (2026)
di: Xu, Derong, et al.
Pubblicazione: (2026)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
di: Harada, Yuto, et al.
Pubblicazione: (2025)
di: Harada, Yuto, et al.
Pubblicazione: (2025)
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
di: Li, Jinze, et al.
Pubblicazione: (2026)
di: Li, Jinze, et al.
Pubblicazione: (2026)
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
di: Chao, Hanxiang, et al.
Pubblicazione: (2026)
di: Chao, Hanxiang, et al.
Pubblicazione: (2026)
Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility
di: Wang, Sheng-Fu, et al.
Pubblicazione: (2025)
di: Wang, Sheng-Fu, et al.
Pubblicazione: (2025)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
di: Li, Zihan, et al.
Pubblicazione: (2026)
di: Li, Zihan, et al.
Pubblicazione: (2026)
Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibility
di: Palta, Shramay, et al.
Pubblicazione: (2025)
di: Palta, Shramay, et al.
Pubblicazione: (2025)
Regularization, Semi-supervision, and Supervision for a Plausible Attention-Based Explanation
di: Nguyen, Duc Hau, et al.
Pubblicazione: (2025)
di: Nguyen, Duc Hau, et al.
Pubblicazione: (2025)
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
di: Sun, Yushi, et al.
Pubblicazione: (2026)
di: Sun, Yushi, et al.
Pubblicazione: (2026)
Structured Memory Mechanisms for Stable Context Representation in Large Language Models
di: Xing, Yue, et al.
Pubblicazione: (2025)
di: Xing, Yue, et al.
Pubblicazione: (2025)
Deciphering the Interplay of Parametric and Non-parametric Memory in Retrieval-augmented Language Models
di: Farahani, Mehrdad, et al.
Pubblicazione: (2024)
di: Farahani, Mehrdad, et al.
Pubblicazione: (2024)
MemLong: Memory-Augmented Retrieval for Long Text Modeling
di: Liu, Weijie, et al.
Pubblicazione: (2024)
di: Liu, Weijie, et al.
Pubblicazione: (2024)
Trellis: Learning to Compress Key-Value Memory in Attention Models
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
di: Yoshida, Ryo, et al.
Pubblicazione: (2026) -
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
di: Yoshida, Ryo, et al.
Pubblicazione: (2024) -
Language Acquisition Device in Large Language Models
di: Mita, Masato, et al.
Pubblicazione: (2026) -
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
di: Someya, Taiga, et al.
Pubblicazione: (2025) -
Syntactically-guided Information Maintenance in Sentence Comprehension
di: Isono, Shinnosuke, et al.
Pubblicazione: (2026)