Needle in the Haystack for Memory Based Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nelson, Elliot, Kollias, Georgios, Das, Payel, Chaudhury, Subhajit, Dan, Soham |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generation Constraint Scaling Can Mitigate Hallucination
von: Kollias, Georgios, et al.
Veröffentlicht: (2024)
von: Kollias, Georgios, et al.
Veröffentlicht: (2024)
Can Memory-Augmented Language Models Generalize on Reasoning-in-a-Haystack Tasks?
von: Das, Payel, et al.
Veröffentlicht: (2025)
von: Das, Payel, et al.
Veröffentlicht: (2025)
Large Language Models can be Strong Self-Detoxifiers
von: Ko, Ching-Yun, et al.
Veröffentlicht: (2024)
von: Ko, Ching-Yun, et al.
Veröffentlicht: (2024)
Larimar: Large Language Models with Episodic Memory Control
von: Das, Payel, et al.
Veröffentlicht: (2024)
von: Das, Payel, et al.
Veröffentlicht: (2024)
EpMAN: Episodic Memory AttentioN for Generalizing to Longer Contexts
von: Chaudhury, Subhajit, et al.
Veröffentlicht: (2025)
von: Chaudhury, Subhajit, et al.
Veröffentlicht: (2025)
NeuroPrune: A Neuro-inspired Topological Sparse Training Algorithm for Large Language Models
von: Dhurandhar, Amit, et al.
Veröffentlicht: (2024)
von: Dhurandhar, Amit, et al.
Veröffentlicht: (2024)
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities
von: Dai, Hui, et al.
Veröffentlicht: (2024)
von: Dai, Hui, et al.
Veröffentlicht: (2024)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
In Search of Needles in a 11M Haystack: Recurrent Memory Finds What LLMs Miss
von: Kuratov, Yuri, et al.
Veröffentlicht: (2024)
von: Kuratov, Yuri, et al.
Veröffentlicht: (2024)
Hidden in the Haystack: Smaller Needles are More Difficult for LLMs to Find
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
From Haystack to Needle: Label Space Reduction for Zero-shot Classification
von: Vandemoortele, Nathan, et al.
Veröffentlicht: (2025)
von: Vandemoortele, Nathan, et al.
Veröffentlicht: (2025)
Fundamental Safety-Capability Trade-offs in Fine-tuning Large Language Models
von: Chen, Pin-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Pin-Yu, et al.
Veröffentlicht: (2025)
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
Large Language Model Confidence Estimation via Black-Box Access
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval
von: Yang, David H., et al.
Veröffentlicht: (2026)
von: Yang, David H., et al.
Veröffentlicht: (2026)
Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack
von: Xu, Xiaoyue, et al.
Veröffentlicht: (2024)
von: Xu, Xiaoyue, et al.
Veröffentlicht: (2024)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
von: Aksoy, Sinan G., et al.
Veröffentlicht: (2026)
von: Aksoy, Sinan G., et al.
Veröffentlicht: (2026)
Jailbreaking in the Haystack
von: Shah, Rishi Rajesh, et al.
Veröffentlicht: (2025)
von: Shah, Rishi Rajesh, et al.
Veröffentlicht: (2025)
Reasoning on Multiple Needles In A Haystack
von: Wang, Yidong
Veröffentlicht: (2025)
von: Wang, Yidong
Veröffentlicht: (2025)
Correlated Errors in Large Language Models
von: Kim, Elliot, et al.
Veröffentlicht: (2025)
von: Kim, Elliot, et al.
Veröffentlicht: (2025)
SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
von: Shen, Han, et al.
Veröffentlicht: (2024)
von: Shen, Han, et al.
Veröffentlicht: (2024)
Language Model Memory and Memory Models for Language
von: Badger, Benjamin L.
Veröffentlicht: (2026)
von: Badger, Benjamin L.
Veröffentlicht: (2026)
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design
von: Neehal, Nafis, et al.
Veröffentlicht: (2024)
von: Neehal, Nafis, et al.
Veröffentlicht: (2024)
CoLa: Learning to Interactively Collaborate with Large Language Models
von: Sharma, Abhishek, et al.
Veröffentlicht: (2025)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2025)
Clustering-driven Memory Compression for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
Episodic Memories Generation and Evaluation Benchmark for Large Language Models
von: Huet, Alexis, et al.
Veröffentlicht: (2025)
von: Huet, Alexis, et al.
Veröffentlicht: (2025)
Echo: A Large Language Model with Temporal Episodic Memory
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2026)
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2026)
Beyond Next Word Prediction: Developing Comprehensive Evaluation Frameworks for measuring LLM performance on real world applications
von: Agrawal, Vishakha, et al.
Veröffentlicht: (2025)
von: Agrawal, Vishakha, et al.
Veröffentlicht: (2025)
Learning Mathematical Rules with Large Language Models
von: Gorceix, Antoine, et al.
Veröffentlicht: (2024)
von: Gorceix, Antoine, et al.
Veröffentlicht: (2024)
Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
von: Deng, Xinle, et al.
Veröffentlicht: (2026)
von: Deng, Xinle, et al.
Veröffentlicht: (2026)
Memory Tokens: Large Language Models Can Generate Reversible Sentence Embeddings
von: Sastre, Ignacio, et al.
Veröffentlicht: (2025)
von: Sastre, Ignacio, et al.
Veröffentlicht: (2025)
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
von: Alizadeh, Keivan, et al.
Veröffentlicht: (2023)
von: Alizadeh, Keivan, et al.
Veröffentlicht: (2023)
ME-Switch: A Memory-Efficient Expert Switching Framework for Large Language Models
von: Liu, Jing, et al.
Veröffentlicht: (2024)
von: Liu, Jing, et al.
Veröffentlicht: (2024)
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models
von: Lin, Nianyi, et al.
Veröffentlicht: (2025)
von: Lin, Nianyi, et al.
Veröffentlicht: (2025)
Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
Revisiting the Scaling Properties of Downstream Metrics in Large Language Model Training
von: Krajewski, Jakub, et al.
Veröffentlicht: (2025)
von: Krajewski, Jakub, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Generation Constraint Scaling Can Mitigate Hallucination
von: Kollias, Georgios, et al.
Veröffentlicht: (2024) -
Can Memory-Augmented Language Models Generalize on Reasoning-in-a-Haystack Tasks?
von: Das, Payel, et al.
Veröffentlicht: (2025) -
Large Language Models can be Strong Self-Detoxifiers
von: Ko, Ching-Yun, et al.
Veröffentlicht: (2024) -
Larimar: Large Language Models with Episodic Memory Control
von: Das, Payel, et al.
Veröffentlicht: (2024) -
EpMAN: Episodic Memory AttentioN for Generalizing to Longer Contexts
von: Chaudhury, Subhajit, et al.
Veröffentlicht: (2025)