AgentCPM-Explore: Realizing Long-Horizon Deep Exploration for Edge-Scale Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Haotian, Cong, Xin, Fan, Shengda, Fu, Yuyang, Gong, Ziqin, Lu, Yaxi, Li, Yishan, Niu, Boye, Pan, Chengjun, Song, Zijun, Wang, Huadong, Wu, Yesai, Wu, Yueying, Xie, Zihao, Yan, Yukun, Zhang, Zhong, Lin, Yankai, Liu, Zhiyuan, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
von: Li, Yishan, et al.
Veröffentlicht: (2026)
von: Li, Yishan, et al.
Veröffentlicht: (2026)
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
von: Lu, Yaxi, et al.
Veröffentlicht: (2024)
von: Lu, Yaxi, et al.
Veröffentlicht: (2024)
Learning to Generate Structured Output with Schema Reinforcement Learning
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
Test-Time Deep Thinking to Explore Implicit Rules
von: Chen, Wentong, et al.
Veröffentlicht: (2026)
von: Chen, Wentong, et al.
Veröffentlicht: (2026)
ToLeaP: Rethinking Development of Tool Learning with Large Language Models
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
von: Luo, Qinyu, et al.
Veröffentlicht: (2024)
von: Luo, Qinyu, et al.
Veröffentlicht: (2024)
WorkflowLLM: Enhancing Workflow Orchestration Capability of Large Language Models
von: Fan, Shengda, et al.
Veröffentlicht: (2024)
von: Fan, Shengda, et al.
Veröffentlicht: (2024)
AgentRM: Enhancing Agent Generalization with Reward Modeling
von: Xia, Yu, et al.
Veröffentlicht: (2025)
von: Xia, Yu, et al.
Veröffentlicht: (2025)
DebugBench: Evaluating Debugging Capability of Large Language Models
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
MiniCPM4: Ultra-Efficient LLMs on End Devices
von: MiniCPM Team, et al.
Veröffentlicht: (2025)
von: MiniCPM Team, et al.
Veröffentlicht: (2025)
Representation Learning for Natural Language Processing
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2020)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2020)
Rational Decision-Making Agent with Internalized Utility Judgment
von: Ye, Yining, et al.
Veröffentlicht: (2023)
von: Ye, Yining, et al.
Veröffentlicht: (2023)
AtomMem : Learnable Dynamic Agentic Memory with Atomic Memory Operation
von: Huo, Yupeng, et al.
Veröffentlicht: (2026)
von: Huo, Yupeng, et al.
Veröffentlicht: (2026)
AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents
von: Fan, Shengda, et al.
Veröffentlicht: (2026)
von: Fan, Shengda, et al.
Veröffentlicht: (2026)
DARC: Decoupled Asymmetric Reasoning Curriculum for LLM Evolution
von: Fan, Shengda, et al.
Veröffentlicht: (2026)
von: Fan, Shengda, et al.
Veröffentlicht: (2026)
KG-Infused RAG: Augmenting Corpus-Based RAG with External Knowledge Graphs
von: Wu, Dingjun, et al.
Veröffentlicht: (2025)
von: Wu, Dingjun, et al.
Veröffentlicht: (2025)
Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
Autono: A ReAct-Based Highly Robust Autonomous Agent Framework
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
Reliable Conversational Agents under ASP Control that Understand Natural Language
von: Zeng, Yankai
Veröffentlicht: (2025)
von: Zeng, Yankai
Veröffentlicht: (2025)
The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
Tell Me More! Towards Implicit User Intention Understanding of Language Model Driven Agents
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
Complete characterization of perfect quantum strategies in quantum magic rectangle games
von: Wu, Yueying
Veröffentlicht: (2026)
von: Wu, Yueying
Veröffentlicht: (2026)
Gradient Manipulation in Distributed Stochastic Gradient Descent with Strategic Agents: Truthful Incentives with Convergence Guarantees
von: Chen, Ziqin, et al.
Veröffentlicht: (2026)
von: Chen, Ziqin, et al.
Veröffentlicht: (2026)
MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction
von: Cui, Junbo, et al.
Veröffentlicht: (2026)
von: Cui, Junbo, et al.
Veröffentlicht: (2026)
Learning Refined Document Representations for Dense Retrieval via Deliberate Thinking
von: Ji, Yifan, et al.
Veröffentlicht: (2025)
von: Ji, Yifan, et al.
Veröffentlicht: (2025)
Learning Evolving Tools for Large Language Models
von: Chen, Guoxin, et al.
Veröffentlicht: (2024)
von: Chen, Guoxin, et al.
Veröffentlicht: (2024)
AgentMixer: Multi-Agent Correlated Policy Factorization
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
Iterative Experience Refinement of Software-Developing Agents
von: Qian, Chen, et al.
Veröffentlicht: (2024)
von: Qian, Chen, et al.
Veröffentlicht: (2024)
SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
von: MiniCPM Team, et al.
Veröffentlicht: (2026)
von: MiniCPM Team, et al.
Veröffentlicht: (2026)
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
von: Chen, Weize, et al.
Veröffentlicht: (2024)
von: Chen, Weize, et al.
Veröffentlicht: (2024)
Distance between Relevant Information Pieces Causes Bias in Long-Context LLMs
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
MiCo: End-to-End Mixed Precision Neural Network Co-Exploration Framework for Edge AI
von: Jiang, Zijun, et al.
Veröffentlicht: (2025)
von: Jiang, Zijun, et al.
Veröffentlicht: (2025)
Experiential Co-Learning of Software-Developing Agents
von: Qian, Chen, et al.
Veröffentlicht: (2023)
von: Qian, Chen, et al.
Veröffentlicht: (2023)
Learning Agent-Compatible Context Management for Long-Horizon Tasks
von: Yi, Lu, et al.
Veröffentlicht: (2026)
von: Yi, Lu, et al.
Veröffentlicht: (2026)
UltraHorizon: Benchmarking Agent Capabilities in Ultra Long-Horizon Scenarios
von: Luo, Haotian, et al.
Veröffentlicht: (2025)
von: Luo, Haotian, et al.
Veröffentlicht: (2025)
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
von: Hu, Shengding, et al.
Veröffentlicht: (2024)
von: Hu, Shengding, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
von: Li, Yishan, et al.
Veröffentlicht: (2026) -
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025) -
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
von: Lu, Yaxi, et al.
Veröffentlicht: (2024) -
Learning to Generate Structured Output with Schema Reinforcement Learning
von: Lu, Yaxi, et al.
Veröffentlicht: (2025) -
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution
von: Qian, Cheng, et al.
Veröffentlicht: (2024)