Toward Efficient Agents: Memory, Tool learning, and Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Xiaofang, Li, Lijun, Zhou, Heng, Zhu, Tong, Qu, Xiaoye, Fan, Yuchen, Wei, Qianshan, Ye, Rui, Kang, Li, Qin, Yiran, Kou, Zhiqiang, Liu, Daizong, Li, Qi, Ding, Ning, Chen, Siheng, Shao, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
von: Wei, Qianshan, et al.
Veröffentlicht: (2025)
von: Wei, Qianshan, et al.
Veröffentlicht: (2025)
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs
von: Huang, Siyuan, et al.
Veröffentlicht: (2026)
von: Huang, Siyuan, et al.
Veröffentlicht: (2026)
SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents
von: Feng, Xinshun, et al.
Veröffentlicht: (2026)
von: Feng, Xinshun, et al.
Veröffentlicht: (2026)
GEMS: Agent-Native Multimodal Generation with Memory and Skills
von: He, Zefeng, et al.
Veröffentlicht: (2026)
von: He, Zefeng, et al.
Veröffentlicht: (2026)
VideoSSR: Video Self-Supervised Reinforcement Learning
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
Rethinking Video-Language Model from the Language Input Perspective
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
Spotlight on Token Perception for Multimodal Reinforcement Learning
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
von: Li, Sijia, et al.
Veröffentlicht: (2025)
von: Li, Sijia, et al.
Veröffentlicht: (2025)
A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
von: Liu, Daizong, et al.
Veröffentlicht: (2024)
von: Liu, Daizong, et al.
Veröffentlicht: (2024)
Tree-based Credit Assignment for Multi-Agent Memory System
von: Mao, Marina, et al.
Veröffentlicht: (2026)
von: Mao, Marina, et al.
Veröffentlicht: (2026)
LatentMem: Customizing Latent Memory for Multi-Agent Systems
von: Fu, Muxin, et al.
Veröffentlicht: (2026)
von: Fu, Muxin, et al.
Veröffentlicht: (2026)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
von: Li, Zihan, et al.
Veröffentlicht: (2026)
von: Li, Zihan, et al.
Veröffentlicht: (2026)
Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models
von: Lei, Zixing, et al.
Veröffentlicht: (2026)
von: Lei, Zixing, et al.
Veröffentlicht: (2026)
Reading $\neq$ Seeing: Diagnosing and Closing the Typography Gap in Vision-Language Models
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
von: Yin, Sheng, et al.
Veröffentlicht: (2024)
von: Yin, Sheng, et al.
Veröffentlicht: (2024)
LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid
von: Sun, Weigao, et al.
Veröffentlicht: (2025)
von: Sun, Weigao, et al.
Veröffentlicht: (2025)
BrowseMaster: Towards Scalable Web Browsing via Tool-Augmented Programmatic Agent Pair
von: Pang, Xianghe, et al.
Veröffentlicht: (2025)
von: Pang, Xianghe, et al.
Veröffentlicht: (2025)
ReSo: A Reward-driven Self-organizing LLM-based Multi-Agent System for Reasoning Tasks
von: Zhou, Heng, et al.
Veröffentlicht: (2025)
von: Zhou, Heng, et al.
Veröffentlicht: (2025)
Audio Does Matter: Importance-Aware Multi-Granularity Fusion for Video Moment Retrieval
von: Lin, Junan, et al.
Veröffentlicht: (2025)
von: Lin, Junan, et al.
Veröffentlicht: (2025)
ToolMem: Enhancing Multimodal Agents with Learnable Tool Capability Memory
von: Xiao, Yunzhong, et al.
Veröffentlicht: (2025)
von: Xiao, Yunzhong, et al.
Veröffentlicht: (2025)
Agent-Oriented Planning in Multi-Agent Systems
von: Li, Ao, et al.
Veröffentlicht: (2024)
von: Li, Ao, et al.
Veröffentlicht: (2024)
MemReader: From Passive to Active Extraction for Long-Term Agent Memory
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration
von: Ding, Keyan, et al.
Veröffentlicht: (2025)
von: Ding, Keyan, et al.
Veröffentlicht: (2025)
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
von: Kang, Li, et al.
Veröffentlicht: (2025)
von: Kang, Li, et al.
Veröffentlicht: (2025)
To transfer or not transfer: Unified transferability metric and analysis
von: Zhan, Qianshan, et al.
Veröffentlicht: (2023)
von: Zhan, Qianshan, et al.
Veröffentlicht: (2023)
Beyond Tools and Persons: Who Are They? Classifying Robots and AI Agents for Proportional Governance
von: Ning, Huansheng, et al.
Veröffentlicht: (2026)
von: Ning, Huansheng, et al.
Veröffentlicht: (2026)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
DiffThinker: Towards Generative Multimodal Reasoning with Diffusion Models
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
SEE: Continual Fine-tuning with Sequential Ensemble of Experts
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems
von: Ye, Rui, et al.
Veröffentlicht: (2025)
von: Ye, Rui, et al.
Veröffentlicht: (2025)
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
von: Li, Jinze, et al.
Veröffentlicht: (2026)
von: Li, Jinze, et al.
Veröffentlicht: (2026)
Visual Contextual Attack: Jailbreaking MLLMs with Image-Driven Context Injection
von: Miao, Ziqi, et al.
Veröffentlicht: (2025)
von: Miao, Ziqi, et al.
Veröffentlicht: (2025)
Rethinking Bottlenecks in Safety Fine-Tuning of Vision Language Models
von: Ding, Yi, et al.
Veröffentlicht: (2025)
von: Ding, Yi, et al.
Veröffentlicht: (2025)
Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
von: Li, Yafu, et al.
Veröffentlicht: (2025)
von: Li, Yafu, et al.
Veröffentlicht: (2025)
MemoryFormer: Minimize Transformer Computation by Removing Fully-Connected Layers
von: Ding, Ning, et al.
Veröffentlicht: (2024)
von: Ding, Ning, et al.
Veröffentlicht: (2024)
EvoMem: Improving Multi-Agent Planning with Dual-Evolving Memory
von: Fan, Wenzhe, et al.
Veröffentlicht: (2025)
von: Fan, Wenzhe, et al.
Veröffentlicht: (2025)
When Does Memory Help Multi-Trajectory Inference for Tool-Use LLM Agents?
von: Li, Xinzhe, et al.
Veröffentlicht: (2026)
von: Li, Xinzhe, et al.
Veröffentlicht: (2026)
Improving Pseudo Labels with Global-Local Denoising Framework for Cross-lingual Named Entity Recognition
von: Ding, Zhuojun, et al.
Veröffentlicht: (2024)
von: Ding, Zhuojun, et al.
Veröffentlicht: (2024)
Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval Using Language
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
von: Wei, Qianshan, et al.
Veröffentlicht: (2025) -
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs
von: Huang, Siyuan, et al.
Veröffentlicht: (2026) -
SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents
von: Feng, Xinshun, et al.
Veröffentlicht: (2026) -
GEMS: Agent-Native Multimodal Generation with Memory and Skills
von: He, Zefeng, et al.
Veröffentlicht: (2026) -
VideoSSR: Video Self-Supervised Reinforcement Learning
von: He, Zefeng, et al.
Veröffentlicht: (2025)