Saved in:
| Main Authors: | Ganguli, Anish, Deb, Prabal, Banerjee, Debleena |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.05177 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale
by: Ganguli, Anurup
Published: (2026)
by: Ganguli, Anurup
Published: (2026)
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
by: Go, Eun, et al.
Published: (2026)
by: Go, Eun, et al.
Published: (2026)
Conservative Contextual Bandits: Beyond Linear Representations
by: Deb, Rohan, et al.
Published: (2024)
by: Deb, Rohan, et al.
Published: (2024)
Beyond Johnson-Lindenstrauss: Uniform Bounds for Sketched Bilinear Forms
by: Deb, Rohan, et al.
Published: (2025)
by: Deb, Rohan, et al.
Published: (2025)
MedXAI: A Retrieval-Augmented and Self-Verifying Framework for Knowledge-Guided Medical Image Analysis
by: Urooj, Midhat, et al.
Published: (2025)
by: Urooj, Midhat, et al.
Published: (2025)
Agentic Learner with Grow-and-Refine Multimodal Semantic Memory
by: Bo, Weihao, et al.
Published: (2025)
by: Bo, Weihao, et al.
Published: (2025)
AdaptBot: Combining LLM with Knowledge Graphs and Human Input for Generic-to-Specific Task Decomposition and Knowledge Refinement
by: Singh, Shivam, et al.
Published: (2025)
by: Singh, Shivam, et al.
Published: (2025)
AtMan: Understanding Transformer Predictions Through Memory Efficient Attention Manipulation
by: Deiseroth, Björn, et al.
Published: (2023)
by: Deiseroth, Björn, et al.
Published: (2023)
AdaKD: Dynamic Knowledge Distillation of ASR models using Adaptive Loss Weighting
by: Ganguly, Shreyan, et al.
Published: (2024)
by: Ganguly, Shreyan, et al.
Published: (2024)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
by: Chen, Feng, et al.
Published: (2023)
by: Chen, Feng, et al.
Published: (2023)
Deriving Neural Scaling Laws from the statistics of natural language
by: Cagnetta, Francesco, et al.
Published: (2026)
by: Cagnetta, Francesco, et al.
Published: (2026)
MemoryKT: An Integrative Memory-and-Forgetting Method for Knowledge Tracing
by: Lin, Mingrong, et al.
Published: (2025)
by: Lin, Mingrong, et al.
Published: (2025)
State Contamination in Memory-Augmented LLM Agents
by: Wang, Yian, et al.
Published: (2026)
by: Wang, Yian, et al.
Published: (2026)
TARDiS : Text Augmentation for Refining Diversity and Separability
by: Kim, Kyungmin, et al.
Published: (2025)
by: Kim, Kyungmin, et al.
Published: (2025)
Rethinking Fine-Tuning when Scaling Test-Time Compute: Limiting Confidence Improves Mathematical Reasoning
by: Chen, Feng, et al.
Published: (2025)
by: Chen, Feng, et al.
Published: (2025)
Learning Hierarchical Procedural Memory for LLM Agents through Bayesian Selection and Contrastive Refinement
by: Forouzandeh, Saman, et al.
Published: (2025)
by: Forouzandeh, Saman, et al.
Published: (2025)
Back to the Future: Look-ahead Augmentation and Parallel Self-Refinement for Time Series Forecasting
by: Kim, Sunho, et al.
Published: (2026)
by: Kim, Sunho, et al.
Published: (2026)
Heterogeneous Knowledge for Augmented Modular Reinforcement Learning
by: Wolf, Lorenz, et al.
Published: (2023)
by: Wolf, Lorenz, et al.
Published: (2023)
Graph-level Protein Representation Learning by Structure Knowledge Refinement
by: Wang, Ge, et al.
Published: (2024)
by: Wang, Ge, et al.
Published: (2024)
Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs
by: Feng, Guangyu, et al.
Published: (2026)
by: Feng, Guangyu, et al.
Published: (2026)
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
by: Malviya, Pranshu, et al.
Published: (2023)
by: Malviya, Pranshu, et al.
Published: (2023)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Bypassing the Rationale: Causal Auditing of Implicit Reasoning in Language Models
by: Sathyanarayanan, Anish, et al.
Published: (2026)
by: Sathyanarayanan, Anish, et al.
Published: (2026)
An Imperfect Verifier is Good Enough: Learning with Noisy Rewards
by: Plesner, Andreas, et al.
Published: (2026)
by: Plesner, Andreas, et al.
Published: (2026)
KV Cache Quantization for Self-Forcing Video Generation: A 33-Method Empirical Study
by: Ranganath, Suraj, et al.
Published: (2026)
by: Ranganath, Suraj, et al.
Published: (2026)
Geometric Median (GM) Matching for Robust Data Pruning
by: Acharya, Anish, et al.
Published: (2024)
by: Acharya, Anish, et al.
Published: (2024)
From Kepler to Newton: Inductive Biases Guide Learned World Models in Transformers
by: Liu, Ziming, et al.
Published: (2026)
by: Liu, Ziming, et al.
Published: (2026)
SIGMA: Refining Large Language Model Reasoning via Sibling-Guided Monte Carlo Augmentation
by: Ren, Yanwei, et al.
Published: (2025)
by: Ren, Yanwei, et al.
Published: (2025)
Chain-of-Specificity: An Iteratively Refining Method for Eliciting Knowledge from Large Language Models
by: Wei, Kaiwen, et al.
Published: (2024)
by: Wei, Kaiwen, et al.
Published: (2024)
Temporal Embeddings: Scalable Self-Supervised Temporal Representation Learning from Spatiotemporal Data for Multimodal Computer Vision
by: Cao, Yi, et al.
Published: (2023)
by: Cao, Yi, et al.
Published: (2023)
Agentic Feature Augmentation: Unifying Selection and Generation with Teaming, Planning, and Memories
by: Gong, Nanxu, et al.
Published: (2025)
by: Gong, Nanxu, et al.
Published: (2025)
Efficient On-Policy Reinforcement Learning via Exploration of Sparse Parameter Space
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
by: Bao, Kaixi, et al.
Published: (2025)
by: Bao, Kaixi, et al.
Published: (2025)
Memory Self-Regeneration: Uncovering Hidden Knowledge in Unlearned Models
by: Polowczyk, Agnieszka, et al.
Published: (2025)
by: Polowczyk, Agnieszka, et al.
Published: (2025)
AdaTKG: Adaptive Memory for Temporal Knowledge Graph Reasoning
by: Lee, Seunghan, et al.
Published: (2026)
by: Lee, Seunghan, et al.
Published: (2026)
Temporal Knowledge-Graph Memory in a Partially Observable Environment
by: Kim, Taewoon, et al.
Published: (2024)
by: Kim, Taewoon, et al.
Published: (2024)
Knowledge Base Construction for Knowledge-Augmented Text-to-SQL
by: Baek, Jinheon, et al.
Published: (2025)
by: Baek, Jinheon, et al.
Published: (2025)
Knowledge-Augmented Explainable and Interpretable Learning for Anomaly Detection and Diagnosis
by: Atzmueller, Martin, et al.
Published: (2024)
by: Atzmueller, Martin, et al.
Published: (2024)
Joint Hypergraph Rewiring and Memory-Augmented Forecasting Techniques in Digital Twin Technology
by: Sakhinana, Sagar Srinivas, et al.
Published: (2024)
by: Sakhinana, Sagar Srinivas, et al.
Published: (2024)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
by: Liu, Zeyuan, et al.
Published: (2026)
by: Liu, Zeyuan, et al.
Published: (2026)
Similar Items
-
TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale
by: Ganguli, Anurup
Published: (2026) -
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
by: Go, Eun, et al.
Published: (2026) -
Conservative Contextual Bandits: Beyond Linear Representations
by: Deb, Rohan, et al.
Published: (2024) -
Beyond Johnson-Lindenstrauss: Uniform Bounds for Sketched Bilinear Forms
by: Deb, Rohan, et al.
Published: (2025) -
MedXAI: A Retrieval-Augmented and Self-Verifying Framework for Knowledge-Guided Medical Image Analysis
by: Urooj, Midhat, et al.
Published: (2025)