MemLLM: Finetuning LLMs to Use An Explicit Read-Write Memory
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Modarressi, Ali, Köksal, Abdullatif, Imani, Ayyoob, Fayyaz, Mohsen, Schütze, Hinrich |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
RET-LLM: Towards a General Read-Write Memory for Large Language Models
par: Modarressi, Ali, et autres
Publié: (2023)
par: Modarressi, Ali, et autres
Publié: (2023)
Consistent Document-Level Relation Extraction via Counterfactuals
par: Modarressi, Ali, et autres
Publié: (2024)
par: Modarressi, Ali, et autres
Publié: (2024)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
par: Zhao, Raoyuan, et autres
Publié: (2025)
par: Zhao, Raoyuan, et autres
Publié: (2025)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
par: Köksal, Abdullatif, et autres
Publié: (2024)
par: Köksal, Abdullatif, et autres
Publié: (2024)
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena
par: Weissweiler, Leonie, et autres
Publié: (2024)
par: Weissweiler, Leonie, et autres
Publié: (2024)
Through the LLM Looking Glass: A Socratic Probing of Donkeys, Elephants, and Markets
par: Kennedy, Molly, et autres
Publié: (2025)
par: Kennedy, Molly, et autres
Publié: (2025)
Collapse of Dense Retrievers: Short, Early, and Literal Biases Outranking Factual Evidence
par: Fayyaz, Mohsen, et autres
Publié: (2025)
par: Fayyaz, Mohsen, et autres
Publié: (2025)
GlotLID: Language Identification for Low-Resource Languages
par: Kargaran, Amir Hossein, et autres
Publié: (2023)
par: Kargaran, Amir Hossein, et autres
Publié: (2023)
Steering MoE LLMs via Expert (De)Activation
par: Fayyaz, Mohsen, et autres
Publié: (2025)
par: Fayyaz, Mohsen, et autres
Publié: (2025)
CRAFT Your Dataset: Task-Specific Synthetic Dataset Generation Through Corpus Retrieval and Augmentation
par: Ziegler, Ingo, et autres
Publié: (2024)
par: Ziegler, Ingo, et autres
Publié: (2024)
LongForm: Effective Instruction Tuning with Reverse Instructions
par: Köksal, Abdullatif, et autres
Publié: (2023)
par: Köksal, Abdullatif, et autres
Publié: (2023)
How far can bias go? Tracing bias from pretraining data to alignment
par: Thaler, Marion, et autres
Publié: (2024)
par: Thaler, Marion, et autres
Publié: (2024)
TurkishMMLU: Measuring Massive Multitask Language Understanding in Turkish
par: Yüksel, Arda, et autres
Publié: (2024)
par: Yüksel, Arda, et autres
Publié: (2024)
SYNTHEVAL: Hybrid Behavioral Testing of NLP Models with Synthetic CheckLists
par: Zhao, Raoyuan, et autres
Publié: (2024)
par: Zhao, Raoyuan, et autres
Publié: (2024)
Time Course MechInterp: Analyzing the Evolution of Components and Knowledge in Large Language Models
par: Hakimi, Ahmad Dawar, et autres
Publié: (2025)
par: Hakimi, Ahmad Dawar, et autres
Publié: (2025)
G-MemLLM: Gated Latent Memory Augmentation for Long-Context Reasoning in Large Language Models
par: Xu, Xun
Publié: (2026)
par: Xu, Xun
Publié: (2026)
Taxi1500: A Multilingual Dataset for Text Classification in 1500 Languages
par: Ma, Chunlan, et autres
Publié: (2023)
par: Ma, Chunlan, et autres
Publié: (2023)
ImpliRet: Benchmarking the Implicit Fact Retrieval Challenge
par: Taghavi, Zeinab Sadat, et autres
Publié: (2025)
par: Taghavi, Zeinab Sadat, et autres
Publié: (2025)
MEXA: Multilingual Evaluation of English-Centric LLMs via Cross-Lingual Alignment
par: Kargaran, Amir Hossein, et autres
Publié: (2024)
par: Kargaran, Amir Hossein, et autres
Publié: (2024)
DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights
par: Mikaelyan, Liana, et autres
Publié: (2025)
par: Mikaelyan, Liana, et autres
Publié: (2025)
How Transliterations Improve Crosslingual Alignment
par: Liu, Yihong, et autres
Publié: (2024)
par: Liu, Yihong, et autres
Publié: (2024)
NoLiMa: Long-Context Evaluation Beyond Literal Matching
par: Modarressi, Ali, et autres
Publié: (2025)
par: Modarressi, Ali, et autres
Publié: (2025)
GlotScript: A Resource and Tool for Low Resource Writing System Identification
par: Kargaran, Amir Hossein, et autres
Publié: (2023)
par: Kargaran, Amir Hossein, et autres
Publié: (2023)
Left, Right, or Center? Evaluating LLM Framing in News Classification and Generation
par: Kennedy, Molly, et autres
Publié: (2026)
par: Kennedy, Molly, et autres
Publié: (2026)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
par: Veloso, Leonor, et autres
Publié: (2026)
par: Veloso, Leonor, et autres
Publié: (2026)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
par: Gerstner, Sebastian, et autres
Publié: (2026)
par: Gerstner, Sebastian, et autres
Publié: (2026)
Understanding Gated Neurons in Transformers from Their Input-Output Functionality
par: Gerstner, Sebastian, et autres
Publié: (2025)
par: Gerstner, Sebastian, et autres
Publié: (2025)
PsyMem: Fine-grained psychological alignment and Explicit Memory Control for Advanced Role-Playing LLMs
par: Cheng, Xilong, et autres
Publié: (2025)
par: Cheng, Xilong, et autres
Publié: (2025)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
par: Özeren, Enes, et autres
Publié: (2025)
par: Özeren, Enes, et autres
Publié: (2025)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
par: Xhelili, Orgest, et autres
Publié: (2024)
par: Xhelili, Orgest, et autres
Publié: (2024)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
par: Nie, Ercong, et autres
Publié: (2025)
par: Nie, Ercong, et autres
Publié: (2025)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
par: Liu, Yihong, et autres
Publié: (2026)
par: Liu, Yihong, et autres
Publié: (2026)
Why Better Cross-Lingual Alignment Fails for Better Cross-Lingual Transfer: Case of Encoders
par: Veitsman, Yana, et autres
Publié: (2026)
par: Veitsman, Yana, et autres
Publié: (2026)
The Anatomy of an Edit: Mechanism-Guided Activation Steering for Knowledge Editing
par: Cao, Yuan, et autres
Publié: (2026)
par: Cao, Yuan, et autres
Publié: (2026)
MemInsight: Autonomous Memory Augmentation for LLM Agents
par: Salama, Rana, et autres
Publié: (2025)
par: Salama, Rana, et autres
Publié: (2025)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
par: Fayyaz, Mohsen, et autres
Publié: (2024)
par: Fayyaz, Mohsen, et autres
Publié: (2024)
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs
par: Mortensen, David R., et autres
Publié: (2024)
par: Mortensen, David R., et autres
Publié: (2024)
Persistent Personas? Role-Playing, Instruction Following, and Safety in Extended Interactions
par: de Araujo, Pedro Henrique Luz, et autres
Publié: (2025)
par: de Araujo, Pedro Henrique Luz, et autres
Publié: (2025)
MaskLID: Code-Switching Language Identification through Iterative Masking
par: Kargaran, Amir Hossein, et autres
Publié: (2024)
par: Kargaran, Amir Hossein, et autres
Publié: (2024)
XAMPLER: Learning to Retrieve Cross-Lingual In-Context Examples
par: Lin, Peiqin, et autres
Publié: (2024)
par: Lin, Peiqin, et autres
Publié: (2024)
Documents similaires
-
RET-LLM: Towards a General Read-Write Memory for Large Language Models
par: Modarressi, Ali, et autres
Publié: (2023) -
Consistent Document-Level Relation Extraction via Counterfactuals
par: Modarressi, Ali, et autres
Publié: (2024) -
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
par: Zhao, Raoyuan, et autres
Publié: (2025) -
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
par: Köksal, Abdullatif, et autres
Publié: (2024) -
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena
par: Weissweiler, Leonie, et autres
Publié: (2024)