Linking In-context Learning in Transformers to Human Episodic Memory
Fuente:
arXiv
Salvato in:
| Autori principali: | Ji-An, Li, Zhou, Corey Y., Benna, Marcus K., Mattar, Marcelo G. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Deep Learning without Weight Symmetry
di: Ji-An, Li, et al.
Pubblicazione: (2024)
di: Ji-An, Li, et al.
Pubblicazione: (2024)
Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations
di: Ji-An, Li, et al.
Pubblicazione: (2025)
di: Ji-An, Li, et al.
Pubblicazione: (2025)
Echo: A Large Language Model with Temporal Episodic Memory
di: Liu, WenTao, et al.
Pubblicazione: (2025)
di: Liu, WenTao, et al.
Pubblicazione: (2025)
Human-inspired Episodic Memory for Infinite Context LLMs
di: Fountas, Zafeirios, et al.
Pubblicazione: (2024)
di: Fountas, Zafeirios, et al.
Pubblicazione: (2024)
The Position Curse: LLMs Struggle to Locate the Last Few Items in a List
di: Zhang, Zhanqi, et al.
Pubblicazione: (2026)
di: Zhang, Zhanqi, et al.
Pubblicazione: (2026)
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training
di: Mistry, Deven Mahesh, et al.
Pubblicazione: (2025)
di: Mistry, Deven Mahesh, et al.
Pubblicazione: (2025)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
di: Pink, Mathis, et al.
Pubblicazione: (2024)
di: Pink, Mathis, et al.
Pubblicazione: (2024)
Episodic Memories Generation and Evaluation Benchmark for Large Language Models
di: Huet, Alexis, et al.
Pubblicazione: (2025)
di: Huet, Alexis, et al.
Pubblicazione: (2025)
Transformers are Universal In-context Learners
di: Furuya, Takashi, et al.
Pubblicazione: (2024)
di: Furuya, Takashi, et al.
Pubblicazione: (2024)
The broader spectrum of in-context learning
di: Lampinen, Andrew Kyle, et al.
Pubblicazione: (2024)
di: Lampinen, Andrew Kyle, et al.
Pubblicazione: (2024)
Curse of High Dimensionality Issue in Transformer for Long-context Modeling
di: Zhang, Shuhai, et al.
Pubblicazione: (2025)
di: Zhang, Shuhai, et al.
Pubblicazione: (2025)
RNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
di: Wen, Kaiyue, et al.
Pubblicazione: (2024)
di: Wen, Kaiyue, et al.
Pubblicazione: (2024)
Cognitively-Inspired Episodic Memory Architectures for Accurate and Efficient Character AI
di: Gonzalez, Rafael Arias, et al.
Pubblicazione: (2025)
di: Gonzalez, Rafael Arias, et al.
Pubblicazione: (2025)
Learning from Supervision with Semantic and Episodic Memory: A Reflective Approach to Agent Adaptation
di: Hassell, Jackson, et al.
Pubblicazione: (2025)
di: Hassell, Jackson, et al.
Pubblicazione: (2025)
In-context Learning in Presence of Spurious Correlations
di: Harutyunyan, Hrayr, et al.
Pubblicazione: (2024)
di: Harutyunyan, Hrayr, et al.
Pubblicazione: (2024)
Guideline Learning for In-context Information Extraction
di: Pang, Chaoxu, et al.
Pubblicazione: (2023)
di: Pang, Chaoxu, et al.
Pubblicazione: (2023)
In-context Learning and Gradient Descent Revisited
di: Deutch, Gilad, et al.
Pubblicazione: (2023)
di: Deutch, Gilad, et al.
Pubblicazione: (2023)
Implicit In-context Learning
di: Li, Zhuowei, et al.
Pubblicazione: (2024)
di: Li, Zhuowei, et al.
Pubblicazione: (2024)
On the Relation between Sensitivity and Accuracy in In-context Learning
di: Chen, Yanda, et al.
Pubblicazione: (2022)
di: Chen, Yanda, et al.
Pubblicazione: (2022)
Learning without training: The implicit dynamics of in-context learning
di: Dherin, Benoit, et al.
Pubblicazione: (2025)
di: Dherin, Benoit, et al.
Pubblicazione: (2025)
Understanding the Thinking Process of Reasoning Models: A Perspective from Schoenfeld's Episode Theory
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
Memory-Efficient Fine-Tuning of Transformers via Token Selection
di: Simoulin, Antoine, et al.
Pubblicazione: (2025)
di: Simoulin, Antoine, et al.
Pubblicazione: (2025)
How Many Human Judgments Are Enough? Feasibility Limits of Human Preference Evaluation
di: Lee, Wilson Y.
Pubblicazione: (2026)
di: Lee, Wilson Y.
Pubblicazione: (2026)
Towards Understanding the Relationship between In-context Learning and Compositional Generalization
di: Han, Sungjun, et al.
Pubblicazione: (2024)
di: Han, Sungjun, et al.
Pubblicazione: (2024)
Graph-based Molecular In-context Learning Grounded on Morgan Fingerprints
di: Al-Lawati, Ali, et al.
Pubblicazione: (2025)
di: Al-Lawati, Ali, et al.
Pubblicazione: (2025)
Scaling In-Context Online Learning Capability of LLMs via Cross-Episode Meta-RL
di: Lin, Xiaofeng, et al.
Pubblicazione: (2026)
di: Lin, Xiaofeng, et al.
Pubblicazione: (2026)
HMT: Hierarchical Memory Transformer for Efficient Long Context Language Processing
di: He, Zifan, et al.
Pubblicazione: (2024)
di: He, Zifan, et al.
Pubblicazione: (2024)
Diagonal Batching Unlocks Parallelism in Recurrent Memory Transformers for Long Contexts
di: Sivtsov, Danil, et al.
Pubblicazione: (2025)
di: Sivtsov, Danil, et al.
Pubblicazione: (2025)
Enhancing In-context Learning via Linear Probe Calibration
di: Abbas, Momin, et al.
Pubblicazione: (2024)
di: Abbas, Momin, et al.
Pubblicazione: (2024)
An Evolved Universal Transformer Memory
di: Cetin, Edoardo, et al.
Pubblicazione: (2024)
di: Cetin, Edoardo, et al.
Pubblicazione: (2024)
Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity
di: Madhyastha, Pranava, et al.
Pubblicazione: (2026)
di: Madhyastha, Pranava, et al.
Pubblicazione: (2026)
Learnable Permutation for Structured Sparsity on Transformer Models
di: Li, Zekai, et al.
Pubblicazione: (2026)
di: Li, Zekai, et al.
Pubblicazione: (2026)
Breaking through the learning plateaus of in-context learning in Transformer
di: Fu, Jingwen, et al.
Pubblicazione: (2023)
di: Fu, Jingwen, et al.
Pubblicazione: (2023)
DETAIL: Task DEmonsTration Attribution for Interpretable In-context Learning
di: Zhou, Zijian, et al.
Pubblicazione: (2024)
di: Zhou, Zijian, et al.
Pubblicazione: (2024)
Polynomial Regression as a Task for Understanding In-context Learning Through Finetuning and Alignment
di: Wilcoxson, Max, et al.
Pubblicazione: (2024)
di: Wilcoxson, Max, et al.
Pubblicazione: (2024)
Sub-SA: Strengthen In-context Learning via Submodular Selective Annotation
di: Qian, Jian, et al.
Pubblicazione: (2024)
di: Qian, Jian, et al.
Pubblicazione: (2024)
Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective
di: Ghosh, Bishwamittra, et al.
Pubblicazione: (2026)
di: Ghosh, Bishwamittra, et al.
Pubblicazione: (2026)
A Study on the Calibration of In-context Learning
di: Zhang, Hanlin, et al.
Pubblicazione: (2023)
di: Zhang, Hanlin, et al.
Pubblicazione: (2023)
Mechanistic Fine-tuning for In-context Learning
di: Cho, Hakaze, et al.
Pubblicazione: (2025)
di: Cho, Hakaze, et al.
Pubblicazione: (2025)
Large language models reorganize representational geometry during in-context learning
di: Xiong, Hua-Dong, et al.
Pubblicazione: (2026)
di: Xiong, Hua-Dong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Deep Learning without Weight Symmetry
di: Ji-An, Li, et al.
Pubblicazione: (2024) -
Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations
di: Ji-An, Li, et al.
Pubblicazione: (2025) -
Echo: A Large Language Model with Temporal Episodic Memory
di: Liu, WenTao, et al.
Pubblicazione: (2025) -
Human-inspired Episodic Memory for Infinite Context LLMs
di: Fountas, Zafeirios, et al.
Pubblicazione: (2024) -
The Position Curse: LLMs Struggle to Locate the Last Few Items in a List
di: Zhang, Zhanqi, et al.
Pubblicazione: (2026)