EpMAN: Episodic Memory AttentioN for Generalizing to Longer Contexts
Fuente:
arXiv
Saved in:
| Main Authors: | Chaudhury, Subhajit, Das, Payel, Swaminathan, Sarathkrishna, Kollias, Georgios, Nelson, Elliot, Pahwa, Khushbu, Pedapati, Tejaswini, Melnyk, Igor, Riemer, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Needle in the Haystack for Memory Based Large Language Models
by: Nelson, Elliot, et al.
Published: (2024)
by: Nelson, Elliot, et al.
Published: (2024)
Larimar: Large Language Models with Episodic Memory Control
by: Das, Payel, et al.
Published: (2024)
by: Das, Payel, et al.
Published: (2024)
Generation Constraint Scaling Can Mitigate Hallucination
by: Kollias, Georgios, et al.
Published: (2024)
by: Kollias, Georgios, et al.
Published: (2024)
Large Language Models can be Strong Self-Detoxifiers
by: Ko, Ching-Yun, et al.
Published: (2024)
by: Ko, Ching-Yun, et al.
Published: (2024)
Can Memory-Augmented Language Models Generalize on Reasoning-in-a-Haystack Tasks?
by: Das, Payel, et al.
Published: (2025)
by: Das, Payel, et al.
Published: (2025)
NeuroPrune: A Neuro-inspired Topological Sparse Training Algorithm for Large Language Models
by: Dhurandhar, Amit, et al.
Published: (2024)
by: Dhurandhar, Amit, et al.
Published: (2024)
Sparse Gradient Compression for Fine-Tuning Large Language Models
by: Yang, David H., et al.
Published: (2025)
by: Yang, David H., et al.
Published: (2025)
ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval
by: Yang, David H., et al.
Published: (2026)
by: Yang, David H., et al.
Published: (2026)
LongFuncEval: Measuring the effectiveness of long context models for function calling
by: Kate, Kiran, et al.
Published: (2025)
by: Kate, Kiran, et al.
Published: (2025)
TabSketchFM: Sketch-based Tabular Representation Learning for Data Discovery over Data Lakes
by: Khatiwada, Aamod, et al.
Published: (2024)
by: Khatiwada, Aamod, et al.
Published: (2024)
From PEFT to DEFT: Parameter Efficient Finetuning for Reducing Activation Density in Transformers
by: Runwal, Bharat, et al.
Published: (2024)
by: Runwal, Bharat, et al.
Published: (2024)
Modular Prompt Learning Improves Vision-Language Models
by: Huang, Zhenhan, et al.
Published: (2025)
by: Huang, Zhenhan, et al.
Published: (2025)
Differentiable Prompt Learning for Vision Language Models
by: Huang, Zhenhan, et al.
Published: (2024)
by: Huang, Zhenhan, et al.
Published: (2024)
Intermediate Representations are Strong AI-Generated Image Detectors
by: Huang, Zhenhan, et al.
Published: (2026)
by: Huang, Zhenhan, et al.
Published: (2026)
Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training
by: Barakat, Anas, et al.
Published: (2026)
by: Barakat, Anas, et al.
Published: (2026)
Large Language Model Confidence Estimation via Black-Box Access
by: Pedapati, Tejaswini, et al.
Published: (2024)
by: Pedapati, Tejaswini, et al.
Published: (2024)
Koopman Learning with Episodic Memory
by: Redman, William T., et al.
Published: (2023)
by: Redman, William T., et al.
Published: (2023)
Design Multiband Monopole and Microstrip Patch Antennas using High Frequency Structure Simulator
by: Giannakopoulos, Georgios, et al.
Published: (2024)
by: Giannakopoulos, Georgios, et al.
Published: (2024)
Securing the Future of IVR: AI-Driven Innovation with Agile Security, Data Regulation, and Ethical AI Integration
by: Shaikh, Khushbu Mehboob, et al.
Published: (2025)
by: Shaikh, Khushbu Mehboob, et al.
Published: (2025)
Evolution of IVR building techniques: from code writing to AI-powered automation
by: Shaikh, Khushbu Mehboob, et al.
Published: (2024)
by: Shaikh, Khushbu Mehboob, et al.
Published: (2024)
Graph is all you need? Lightweight data-agnostic neural architecture search without training
by: Huang, Zhenhan, et al.
Published: (2024)
by: Huang, Zhenhan, et al.
Published: (2024)
On the Effects of Fine-tuning Language Models for Text-Based Reinforcement Learning
by: Gruppi, Mauricio, et al.
Published: (2024)
by: Gruppi, Mauricio, et al.
Published: (2024)
CoFrNets: Interpretable Neural Architecture Inspired by Continued Fractions
by: Puri, Isha, et al.
Published: (2025)
by: Puri, Isha, et al.
Published: (2025)
Multi-modal brain encoding models for multi-modal stimuli
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
A Neuro-Symbolic Approach to Multi-Agent RL for Interpretability and Probabilistic Decision Making
by: Subramanian, Chitra, et al.
Published: (2024)
by: Subramanian, Chitra, et al.
Published: (2024)
Position: Theory of Mind Benchmarks are Broken for Large Language Models
by: Riemer, Matthew, et al.
Published: (2024)
by: Riemer, Matthew, et al.
Published: (2024)
Combining Domain and Alignment Vectors to Achieve Better Knowledge-Safety Trade-offs in LLMs
by: Thakkar, Megh, et al.
Published: (2024)
by: Thakkar, Megh, et al.
Published: (2024)
REMem: Reasoning with Episodic Memory in Language Agent
by: Shu, Yiheng, et al.
Published: (2026)
by: Shu, Yiheng, et al.
Published: (2026)
TOPJoin: A Context-Aware Multi-Criteria Approach for Joinable Column Search
by: Kokel, Harsha, et al.
Published: (2025)
by: Kokel, Harsha, et al.
Published: (2025)
Evaluating Joinable Column Discovery Approaches for Context-Aware Search
by: Kokel, Harsha, et al.
Published: (2025)
by: Kokel, Harsha, et al.
Published: (2025)
InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows
by: Ataallah, Kirolos, et al.
Published: (2024)
by: Ataallah, Kirolos, et al.
Published: (2024)
GNNX-BENCH: Unravelling the Utility of Perturbation-based GNN Explainers through In-depth Benchmarking
by: Kosan, Mert, et al.
Published: (2023)
by: Kosan, Mert, et al.
Published: (2023)
Theaters of Citizenship
by: Pahwa, Sonali
Published: (2021)
by: Pahwa, Sonali
Published: (2021)
Decentralized Decision Making in Two Sided Manufacturing-as-a-Service Marketplaces
by: Pahwa, Deepak
Published: (2025)
by: Pahwa, Deepak
Published: (2025)
Design of a Multidimensional Model Using Object Oriented Features in UML
by: Payal Pahwa
Published: (2011)
by: Payal Pahwa
Published: (2011)
A Deep Dive into the Trade-Offs of Parameter-Efficient Preference Alignment Techniques
by: Thakkar, Megh, et al.
Published: (2024)
by: Thakkar, Megh, et al.
Published: (2024)
OjaKV: Context-Aware Online Low-Rank KV Cache Compression
by: Zhu, Yuxuan, et al.
Published: (2025)
by: Zhu, Yuxuan, et al.
Published: (2025)
Intergenerational Transmission of Between‐Group Occupational Disparity: Some Indian Evidence
by: Dipankar Das, et al.
Published: (2025)
by: Dipankar Das, et al.
Published: (2025)
Structured Episodic Event Memory
by: Lu, Zhengxuan, et al.
Published: (2026)
by: Lu, Zhengxuan, et al.
Published: (2026)
Constructing Memories, Episodic and Semantic
by: Hunter Gentry
Published: (2025)
by: Hunter Gentry
Published: (2025)
Similar Items
-
Needle in the Haystack for Memory Based Large Language Models
by: Nelson, Elliot, et al.
Published: (2024) -
Larimar: Large Language Models with Episodic Memory Control
by: Das, Payel, et al.
Published: (2024) -
Generation Constraint Scaling Can Mitigate Hallucination
by: Kollias, Georgios, et al.
Published: (2024) -
Large Language Models can be Strong Self-Detoxifiers
by: Ko, Ching-Yun, et al.
Published: (2024) -
Can Memory-Augmented Language Models Generalize on Reasoning-in-a-Haystack Tasks?
by: Das, Payel, et al.
Published: (2025)