If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
Fuente:
arXiv
Saved in:
| Main Authors: | Yoshida, Ryo, Isono, Shinnosuke, Kajikawa, Kohei, Someya, Taiga, Sugimoto, Yushi, Oseki, Yohei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026)
by: Yoshida, Ryo, et al.
Published: (2026)
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
by: Yoshida, Ryo, et al.
Published: (2024)
by: Yoshida, Ryo, et al.
Published: (2024)
Language Acquisition Device in Large Language Models
by: Mita, Masato, et al.
Published: (2026)
by: Mita, Masato, et al.
Published: (2026)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
by: Someya, Taiga, et al.
Published: (2025)
by: Someya, Taiga, et al.
Published: (2025)
Syntactically-guided Information Maintenance in Sentence Comprehension
by: Isono, Shinnosuke, et al.
Published: (2026)
by: Isono, Shinnosuke, et al.
Published: (2026)
Rethinking the Relationship between the Power Law and Hierarchical Structures
by: Nakaishi, Kai, et al.
Published: (2025)
by: Nakaishi, Kai, et al.
Published: (2025)
Composition, Attention, or Both?
by: Yoshida, Ryo, et al.
Published: (2022)
by: Yoshida, Ryo, et al.
Published: (2022)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
by: Mita, Masato, et al.
Published: (2025)
by: Mita, Masato, et al.
Published: (2025)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
by: Kajikawa, Kohei, et al.
Published: (2024)
by: Kajikawa, Kohei, et al.
Published: (2024)
Information-Theoretic Storage Cost in Sentence Comprehension
by: Kajikawa, Kohei, et al.
Published: (2026)
by: Kajikawa, Kohei, et al.
Published: (2026)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
by: Yoshida, Ryo, et al.
Published: (2021)
by: Yoshida, Ryo, et al.
Published: (2021)
Emergent Word Order Universals from Cognitively-Motivated Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2024)
by: Kuribayashi, Tatsuki, et al.
Published: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
Psychometric Predictive Power of Large Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
by: Haga, Akari, et al.
Published: (2024)
by: Haga, Akari, et al.
Published: (2024)
Can Language Models Learn Typologically Implausible Languages?
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
Leveraging Human Production-Interpretation Asymmetries to Test LLM Cognitive Plausibility
by: Lam, Suet-Ying, et al.
Published: (2025)
by: Lam, Suet-Ying, et al.
Published: (2025)
Exclusive Unlearning
by: Sasaki, Mutsumi, et al.
Published: (2026)
by: Sasaki, Mutsumi, et al.
Published: (2026)
MLP Memory: A Retriever-Pretrained Memory for Large Language Models
by: Wei, Rubin, et al.
Published: (2025)
by: Wei, Rubin, et al.
Published: (2025)
How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders
by: Inaba, Tatsuro, et al.
Published: (2025)
by: Inaba, Tatsuro, et al.
Published: (2025)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
by: Lepori, Michael A., et al.
Published: (2025)
by: Lepori, Michael A., et al.
Published: (2025)
Information Locality as an Inductive Bias for Neural Language Models
by: Someya, Taiga, et al.
Published: (2025)
by: Someya, Taiga, et al.
Published: (2025)
Cognitive Memory in Large Language Models
by: Shan, Lianlei, et al.
Published: (2025)
by: Shan, Lianlei, et al.
Published: (2025)
Learning What to Remember: Adaptive Probabilistic Memory Retention for Memory-Efficient Language Models
by: Rafiuddin, S M, et al.
Published: (2025)
by: Rafiuddin, S M, et al.
Published: (2025)
S$^3$-Attention:Attention-Aligned Endogenous Retrieval for Memory-Bounded Long-Context Inference
by: Ma, Qingsen, et al.
Published: (2026)
by: Ma, Qingsen, et al.
Published: (2026)
Can Language Models Induce Grammatical Knowledge from Indirect Evidence?
by: Oba, Miyu, et al.
Published: (2024)
by: Oba, Miyu, et al.
Published: (2024)
Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory
by: Xu, Derong, et al.
Published: (2026)
by: Xu, Derong, et al.
Published: (2026)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
by: Harada, Yuto, et al.
Published: (2025)
by: Harada, Yuto, et al.
Published: (2025)
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
by: Li, Jinze, et al.
Published: (2026)
by: Li, Jinze, et al.
Published: (2026)
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
by: Chao, Hanxiang, et al.
Published: (2026)
by: Chao, Hanxiang, et al.
Published: (2026)
Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility
by: Wang, Sheng-Fu, et al.
Published: (2025)
by: Wang, Sheng-Fu, et al.
Published: (2025)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
by: Li, Zihan, et al.
Published: (2026)
by: Li, Zihan, et al.
Published: (2026)
Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibility
by: Palta, Shramay, et al.
Published: (2025)
by: Palta, Shramay, et al.
Published: (2025)
Regularization, Semi-supervision, and Supervision for a Plausible Attention-Based Explanation
by: Nguyen, Duc Hau, et al.
Published: (2025)
by: Nguyen, Duc Hau, et al.
Published: (2025)
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
by: Sun, Yushi, et al.
Published: (2026)
by: Sun, Yushi, et al.
Published: (2026)
Structured Memory Mechanisms for Stable Context Representation in Large Language Models
by: Xing, Yue, et al.
Published: (2025)
by: Xing, Yue, et al.
Published: (2025)
Deciphering the Interplay of Parametric and Non-parametric Memory in Retrieval-augmented Language Models
by: Farahani, Mehrdad, et al.
Published: (2024)
by: Farahani, Mehrdad, et al.
Published: (2024)
MemLong: Memory-Augmented Retrieval for Long Text Modeling
by: Liu, Weijie, et al.
Published: (2024)
by: Liu, Weijie, et al.
Published: (2024)
Trellis: Learning to Compress Key-Value Memory in Attention Models
by: Karami, Mahdi, et al.
Published: (2025)
by: Karami, Mahdi, et al.
Published: (2025)
Similar Items
-
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026) -
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
by: Yoshida, Ryo, et al.
Published: (2024) -
Language Acquisition Device in Large Language Models
by: Mita, Masato, et al.
Published: (2026) -
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
by: Someya, Taiga, et al.
Published: (2025) -
Syntactically-guided Information Maintenance in Sentence Comprehension
by: Isono, Shinnosuke, et al.
Published: (2026)