Studying Large Language Model Behaviors Under Context-Memory Conflicts With Real Documents
Fuente:
arXiv
Saved in:
| Main Authors: | Kortukov, Evgenii, Rubinstein, Alexander, Nguyen, Elisa, Oh, Seong Joon |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards User-Focused Research in Training Data Attribution for Human-Centered Explainable AI
by: Nguyen, Elisa, et al.
Published: (2024)
by: Nguyen, Elisa, et al.
Published: (2024)
MEME: Multi-entity & Evolving Memory Evaluation
by: Jung, Seokwon, et al.
Published: (2026)
by: Jung, Seokwon, et al.
Published: (2026)
DISCO: Diversifying Sample Condensation for Efficient Model Evaluation
by: Rubinstein, Alexander, et al.
Published: (2025)
by: Rubinstein, Alexander, et al.
Published: (2025)
Do Deep Neural Network Solutions Form a Star Domain?
by: Sonthalia, Ankit, et al.
Published: (2024)
by: Sonthalia, Ankit, et al.
Published: (2024)
Scalable Ensemble Diversification for OOD Generalization and Detection
by: Rubinstein, Alexander, et al.
Published: (2024)
by: Rubinstein, Alexander, et al.
Published: (2024)
Are We Done with Object-Centric Learning?
by: Rubinstein, Alexander, et al.
Published: (2025)
by: Rubinstein, Alexander, et al.
Published: (2025)
ASIDE: Architectural Separation of Instructions and Data in Language Models
by: Zverev, Egor, et al.
Published: (2025)
by: Zverev, Egor, et al.
Published: (2025)
Mitigating Shortcut Learning with Diffusion Counterfactuals and Diverse Ensembles
by: Scimeca, Luca, et al.
Published: (2023)
by: Scimeca, Luca, et al.
Published: (2023)
Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models
by: Puerto, Haritz, et al.
Published: (2024)
by: Puerto, Haritz, et al.
Published: (2024)
MASEval: Extending Multi-Agent Evaluation from Models to Systems
by: Emde, Cornelius, et al.
Published: (2026)
by: Emde, Cornelius, et al.
Published: (2026)
Intermediate Layer Classifiers for OOD generalization
by: Uselis, Arnas, et al.
Published: (2025)
by: Uselis, Arnas, et al.
Published: (2025)
Calibrating Large Language Models Using Their Generations Only
by: Ulmer, Dennis, et al.
Published: (2024)
by: Ulmer, Dennis, et al.
Published: (2024)
LLM generation novelty through the lens of semantic similarity
by: Davydov, Philipp, et al.
Published: (2025)
by: Davydov, Philipp, et al.
Published: (2025)
First Hallucination Tokens Are Different from Conditional Ones
by: Snel, Jakob, et al.
Published: (2025)
by: Snel, Jakob, et al.
Published: (2025)
Reasoning Under 1 Billion: Memory-Augmented Reinforcement Learning for Large Language Models
by: Le, Hung, et al.
Published: (2025)
by: Le, Hung, et al.
Published: (2025)
Benchmarking Uncertainty Disentanglement: Specialized Uncertainties for Specialized Tasks
by: Mucsányi, Bálint, et al.
Published: (2024)
by: Mucsányi, Bálint, et al.
Published: (2024)
Does Data Scaling Lead to Visual Compositional Generalization?
by: Uselis, Arnas, et al.
Published: (2025)
by: Uselis, Arnas, et al.
Published: (2025)
Compressed Context Memory For Online Language Model Interaction
by: Kim, Jang-Hyun, et al.
Published: (2023)
by: Kim, Jang-Hyun, et al.
Published: (2023)
Compositional Generalization Requires Linear, Orthogonal Representations in Vision Embedding Models
by: Uselis, Arnas, et al.
Published: (2026)
by: Uselis, Arnas, et al.
Published: (2026)
CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally
by: Koishigarina, Darina, et al.
Published: (2025)
by: Koishigarina, Darina, et al.
Published: (2025)
Mixture of Scales: Memory-Efficient Token-Adaptive Binarization for Large Language Models
by: Jo, Dongwon, et al.
Published: (2024)
by: Jo, Dongwon, et al.
Published: (2024)
Optimizing Humor Generation in Large Language Models: Temperature Configurations and Architectural Trade-offs
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents
by: Arghal, Raghu, et al.
Published: (2026)
by: Arghal, Raghu, et al.
Published: (2026)
How can embedding models bind concepts?
by: Uselis, Arnas, et al.
Published: (2026)
by: Uselis, Arnas, et al.
Published: (2026)
Memory-Augmented Architecture for Long-Term Context Handling in Large Language Models
by: Shinwari, Haseeb Ullah Khan, et al.
Published: (2025)
by: Shinwari, Haseeb Ullah Khan, et al.
Published: (2025)
Challenges in Understanding Modality Conflict in Vision-Language Models
by: Nguyen, Trang, et al.
Published: (2025)
by: Nguyen, Trang, et al.
Published: (2025)
Dynamics Reveals Structure: Challenging the Linear Propagation Assumption
by: Chang, Hoyeon, et al.
Published: (2026)
by: Chang, Hoyeon, et al.
Published: (2026)
Strategic Dishonesty Can Undermine AI Safety Evaluations of Frontier LLMs
by: Panfilov, Alexander, et al.
Published: (2025)
by: Panfilov, Alexander, et al.
Published: (2025)
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
by: Shim, Jung-Woo, et al.
Published: (2025)
by: Shim, Jung-Woo, et al.
Published: (2025)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
by: Shim, Jung-Woo, et al.
Published: (2025)
by: Shim, Jung-Woo, et al.
Published: (2025)
Evaluating Large Language Models in Vulnerability Detection Under Variable Context Windows
by: Lin, Jie, et al.
Published: (2025)
by: Lin, Jie, et al.
Published: (2025)
Pretrained Visual Uncertainties
by: Kirchhof, Michael, et al.
Published: (2024)
by: Kirchhof, Michael, et al.
Published: (2024)
Universal Algorithm-Implicit Learning
by: Woerner, Stefano, et al.
Published: (2026)
by: Woerner, Stefano, et al.
Published: (2026)
In-Context Clustering with Large Language Models
by: Wang, Ying, et al.
Published: (2025)
by: Wang, Ying, et al.
Published: (2025)
LICO: Large Language Models for In-Context Molecular Optimization
by: Nguyen, Tung, et al.
Published: (2024)
by: Nguyen, Tung, et al.
Published: (2024)
In-Context Function Learning in Large Language Models
by: Akata, Elif, et al.
Published: (2026)
by: Akata, Elif, et al.
Published: (2026)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
by: Hengle, Amey, et al.
Published: (2024)
by: Hengle, Amey, et al.
Published: (2024)
Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models
by: Qiao, Boyu, et al.
Published: (2026)
by: Qiao, Boyu, et al.
Published: (2026)
The Mosaic Memory of Large Language Models
by: Shilov, Igor, et al.
Published: (2024)
by: Shilov, Igor, et al.
Published: (2024)
From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model
by: Ju, Yeong-Joon, et al.
Published: (2025)
by: Ju, Yeong-Joon, et al.
Published: (2025)
Similar Items
-
Towards User-Focused Research in Training Data Attribution for Human-Centered Explainable AI
by: Nguyen, Elisa, et al.
Published: (2024) -
MEME: Multi-entity & Evolving Memory Evaluation
by: Jung, Seokwon, et al.
Published: (2026) -
DISCO: Diversifying Sample Condensation for Efficient Model Evaluation
by: Rubinstein, Alexander, et al.
Published: (2025) -
Do Deep Neural Network Solutions Form a Star Domain?
by: Sonthalia, Ankit, et al.
Published: (2024) -
Scalable Ensemble Diversification for OOD Generalization and Detection
by: Rubinstein, Alexander, et al.
Published: (2024)