If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yoshida, Ryo, Isono, Shinnosuke, Kajikawa, Kohei, Someya, Taiga, Sugimoto, Yushi, Oseki, Yohei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
von: Yoshida, Ryo, et al.
Veröffentlicht: (2026)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2026)
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
von: Yoshida, Ryo, et al.
Veröffentlicht: (2024)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2024)
Language Acquisition Device in Large Language Models
von: Mita, Masato, et al.
Veröffentlicht: (2026)
von: Mita, Masato, et al.
Veröffentlicht: (2026)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
Syntactically-guided Information Maintenance in Sentence Comprehension
von: Isono, Shinnosuke, et al.
Veröffentlicht: (2026)
von: Isono, Shinnosuke, et al.
Veröffentlicht: (2026)
Rethinking the Relationship between the Power Law and Hierarchical Structures
von: Nakaishi, Kai, et al.
Veröffentlicht: (2025)
von: Nakaishi, Kai, et al.
Veröffentlicht: (2025)
Composition, Attention, or Both?
von: Yoshida, Ryo, et al.
Veröffentlicht: (2022)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2022)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
von: Mita, Masato, et al.
Veröffentlicht: (2025)
von: Mita, Masato, et al.
Veröffentlicht: (2025)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2024)
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2024)
Information-Theoretic Storage Cost in Sentence Comprehension
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2026)
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2026)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
von: Yoshida, Ryo, et al.
Veröffentlicht: (2021)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2021)
Emergent Word Order Universals from Cognitively-Motivated Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2024)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
Psychometric Predictive Power of Large Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
Large Language Models Are Human-Like Internally
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
von: Haga, Akari, et al.
Veröffentlicht: (2024)
von: Haga, Akari, et al.
Veröffentlicht: (2024)
Can Language Models Learn Typologically Implausible Languages?
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
Leveraging Human Production-Interpretation Asymmetries to Test LLM Cognitive Plausibility
von: Lam, Suet-Ying, et al.
Veröffentlicht: (2025)
von: Lam, Suet-Ying, et al.
Veröffentlicht: (2025)
Exclusive Unlearning
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2026)
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2026)
MLP Memory: A Retriever-Pretrained Memory for Large Language Models
von: Wei, Rubin, et al.
Veröffentlicht: (2025)
von: Wei, Rubin, et al.
Veröffentlicht: (2025)
How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders
von: Inaba, Tatsuro, et al.
Veröffentlicht: (2025)
von: Inaba, Tatsuro, et al.
Veröffentlicht: (2025)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
Information Locality as an Inductive Bias for Neural Language Models
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
Cognitive Memory in Large Language Models
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
Learning What to Remember: Adaptive Probabilistic Memory Retention for Memory-Efficient Language Models
von: Rafiuddin, S M, et al.
Veröffentlicht: (2025)
von: Rafiuddin, S M, et al.
Veröffentlicht: (2025)
S$^3$-Attention:Attention-Aligned Endogenous Retrieval for Memory-Bounded Long-Context Inference
von: Ma, Qingsen, et al.
Veröffentlicht: (2026)
von: Ma, Qingsen, et al.
Veröffentlicht: (2026)
Can Language Models Induce Grammatical Knowledge from Indirect Evidence?
von: Oba, Miyu, et al.
Veröffentlicht: (2024)
von: Oba, Miyu, et al.
Veröffentlicht: (2024)
Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory
von: Xu, Derong, et al.
Veröffentlicht: (2026)
von: Xu, Derong, et al.
Veröffentlicht: (2026)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
von: Harada, Yuto, et al.
Veröffentlicht: (2025)
von: Harada, Yuto, et al.
Veröffentlicht: (2025)
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
von: Li, Jinze, et al.
Veröffentlicht: (2026)
von: Li, Jinze, et al.
Veröffentlicht: (2026)
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
von: Chao, Hanxiang, et al.
Veröffentlicht: (2026)
von: Chao, Hanxiang, et al.
Veröffentlicht: (2026)
Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility
von: Wang, Sheng-Fu, et al.
Veröffentlicht: (2025)
von: Wang, Sheng-Fu, et al.
Veröffentlicht: (2025)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
von: Li, Zihan, et al.
Veröffentlicht: (2026)
von: Li, Zihan, et al.
Veröffentlicht: (2026)
Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibility
von: Palta, Shramay, et al.
Veröffentlicht: (2025)
von: Palta, Shramay, et al.
Veröffentlicht: (2025)
Regularization, Semi-supervision, and Supervision for a Plausible Attention-Based Explanation
von: Nguyen, Duc Hau, et al.
Veröffentlicht: (2025)
von: Nguyen, Duc Hau, et al.
Veröffentlicht: (2025)
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
von: Sun, Yushi, et al.
Veröffentlicht: (2026)
von: Sun, Yushi, et al.
Veröffentlicht: (2026)
Structured Memory Mechanisms for Stable Context Representation in Large Language Models
von: Xing, Yue, et al.
Veröffentlicht: (2025)
von: Xing, Yue, et al.
Veröffentlicht: (2025)
Deciphering the Interplay of Parametric and Non-parametric Memory in Retrieval-augmented Language Models
von: Farahani, Mehrdad, et al.
Veröffentlicht: (2024)
von: Farahani, Mehrdad, et al.
Veröffentlicht: (2024)
MemLong: Memory-Augmented Retrieval for Long Text Modeling
von: Liu, Weijie, et al.
Veröffentlicht: (2024)
von: Liu, Weijie, et al.
Veröffentlicht: (2024)
Trellis: Learning to Compress Key-Value Memory in Attention Models
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
von: Yoshida, Ryo, et al.
Veröffentlicht: (2026) -
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
von: Yoshida, Ryo, et al.
Veröffentlicht: (2024) -
Language Acquisition Device in Large Language Models
von: Mita, Masato, et al.
Veröffentlicht: (2026) -
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
von: Someya, Taiga, et al.
Veröffentlicht: (2025) -
Syntactically-guided Information Maintenance in Sentence Comprehension
von: Isono, Shinnosuke, et al.
Veröffentlicht: (2026)