SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Sijia, Brahma, Dhanajit, Henao, Ricardo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SAGE-32B: Agentic Reasoning via Iterative Distillation
von: Jha, Basab, et al.
Veröffentlicht: (2026)
von: Jha, Basab, et al.
Veröffentlicht: (2026)
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
von: Lee, Chanuk, et al.
Veröffentlicht: (2026)
von: Lee, Chanuk, et al.
Veröffentlicht: (2026)
Calibrating Model-Based Evaluation Metrics for Summarization
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
TxGemma: Efficient and Agentic LLMs for Therapeutics
von: Wang, Eric, et al.
Veröffentlicht: (2025)
von: Wang, Eric, et al.
Veröffentlicht: (2025)
MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization
von: Wang, Ziqing, et al.
Veröffentlicht: (2026)
von: Wang, Ziqing, et al.
Veröffentlicht: (2026)
GEM: A Gym for Agentic LLMs
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
SciMON: Scientific Inspiration Machines Optimized for Novelty
von: Wang, Qingyun, et al.
Veröffentlicht: (2023)
von: Wang, Qingyun, et al.
Veröffentlicht: (2023)
Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models
von: Arteaga, Gabriel Y., et al.
Veröffentlicht: (2024)
von: Arteaga, Gabriel Y., et al.
Veröffentlicht: (2024)
Thinking About Thinking: SAGE-nano's Inverse Reasoning for Self-Aware Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2025)
von: Jha, Basab, et al.
Veröffentlicht: (2025)
AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
Agentic Adversarial QA for Improving Domain-Specific LLMs
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
Agentic Policy Optimization via Instruction-Policy Co-Evolution
von: Zhou, Han, et al.
Veröffentlicht: (2025)
von: Zhou, Han, et al.
Veröffentlicht: (2025)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
von: Deiseroth, Björn, et al.
Veröffentlicht: (2024)
von: Deiseroth, Björn, et al.
Veröffentlicht: (2024)
Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework
von: Guan, Zihan, et al.
Veröffentlicht: (2026)
von: Guan, Zihan, et al.
Veröffentlicht: (2026)
What's New in My Data? Novelty Exploration via Contrastive Generation
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
Tool Preferences in Agentic LLMs are Unreliable
von: Faghih, Kazem, et al.
Veröffentlicht: (2025)
von: Faghih, Kazem, et al.
Veröffentlicht: (2025)
Hypertokens: Holographic Associative Memory in Tokenized LLMs
von: Augeri, Christopher James
Veröffentlicht: (2025)
von: Augeri, Christopher James
Veröffentlicht: (2025)
Sparse-RL: Breaking the Memory Wall in LLM Reinforcement Learning via Stable Sparse Rollouts
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
von: Duan, Jinhao, et al.
Veröffentlicht: (2025)
von: Duan, Jinhao, et al.
Veröffentlicht: (2025)
Toward Adaptive Reasoning in Large Language Models with Thought Rollback
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
FlashSampling: Fast and Memory-Efficient Exact Sampling
von: Ruiz, Tomas, et al.
Veröffentlicht: (2026)
von: Ruiz, Tomas, et al.
Veröffentlicht: (2026)
LLMem: Estimating GPU Memory Usage for Fine-Tuning Pre-Trained LLMs
von: Kim, Taeho, et al.
Veröffentlicht: (2024)
von: Kim, Taeho, et al.
Veröffentlicht: (2024)
How to Train Data-Efficient LLMs
von: Sachdeva, Noveen, et al.
Veröffentlicht: (2024)
von: Sachdeva, Noveen, et al.
Veröffentlicht: (2024)
General Agentic Memory Via Deep Research
von: Yan, B. Y., et al.
Veröffentlicht: (2025)
von: Yan, B. Y., et al.
Veröffentlicht: (2025)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
von: Pink, Mathis, et al.
Veröffentlicht: (2024)
von: Pink, Mathis, et al.
Veröffentlicht: (2024)
Sample-Efficient Alignment for LLMs
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Wikipedia in the Era of LLMs: Evolution and Risks
von: Huang, Siming, et al.
Veröffentlicht: (2025)
von: Huang, Siming, et al.
Veröffentlicht: (2025)
Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models
von: Vendrell, Victor Conchello, et al.
Veröffentlicht: (2026)
von: Vendrell, Victor Conchello, et al.
Veröffentlicht: (2026)
Agentic Critical Training
von: Liu, Weize, et al.
Veröffentlicht: (2026)
von: Liu, Weize, et al.
Veröffentlicht: (2026)
HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation
von: Wu, Peilin, et al.
Veröffentlicht: (2025)
von: Wu, Peilin, et al.
Veröffentlicht: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
von: Wang, Ganghua, et al.
Veröffentlicht: (2025)
von: Wang, Ganghua, et al.
Veröffentlicht: (2025)
Efficient multi-prompt evaluation of LLMs
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
Efficiently Distilling LLMs for Edge Applications
von: Kundu, Achintya, et al.
Veröffentlicht: (2024)
von: Kundu, Achintya, et al.
Veröffentlicht: (2024)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning
von: Wang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyi, et al.
Veröffentlicht: (2025)
Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SAGE-32B: Agentic Reasoning via Iterative Distillation
von: Jha, Basab, et al.
Veröffentlicht: (2026) -
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
von: Lee, Chanuk, et al.
Veröffentlicht: (2026) -
Calibrating Model-Based Evaluation Metrics for Summarization
von: Liu, Hongye, et al.
Veröffentlicht: (2026) -
TxGemma: Efficient and Agentic LLMs for Therapeutics
von: Wang, Eric, et al.
Veröffentlicht: (2025) -
MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization
von: Wang, Ziqing, et al.
Veröffentlicht: (2026)