Guardado en:
| Autores principales: | Fang, Wei, Zhang, Yang, Qian, Kaizhi, Glass, James, Zhu, Yada |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2503.14432 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Single-Shot: Multi-step Tool Retrieval via Query Planning
por: Fang, Wei, et al.
Publicado: (2026)
por: Fang, Wei, et al.
Publicado: (2026)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
por: Dong, Guanting, et al.
Publicado: (2025)
por: Dong, Guanting, et al.
Publicado: (2025)
LLM Agents Making Agent Tools
por: Wölflein, Georg, et al.
Publicado: (2025)
por: Wölflein, Georg, et al.
Publicado: (2025)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
por: Song, Xiaoshuai, et al.
Publicado: (2026)
por: Song, Xiaoshuai, et al.
Publicado: (2026)
M$^2$PT: Multimodal Prompt Tuning for Zero-shot Instruction Learning
por: Wang, Taowen, et al.
Publicado: (2024)
por: Wang, Taowen, et al.
Publicado: (2024)
SMART: Self-Aware Agent for Tool Overuse Mitigation
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation
por: Zhu, Dongsheng, et al.
Publicado: (2025)
por: Zhu, Dongsheng, et al.
Publicado: (2025)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
por: Parekh, Tanmay, et al.
Publicado: (2025)
por: Parekh, Tanmay, et al.
Publicado: (2025)
ToolPlanner: A Tool Augmented LLM for Multi Granularity Instructions with Path Planning and Feedback
por: Wu, Qinzhuo, et al.
Publicado: (2024)
por: Wu, Qinzhuo, et al.
Publicado: (2024)
ToolRL: Reward is All Tool Learning Needs
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
LoopTool: Closing the Data-Training Loop for Robust LLM Tool Calls
por: Zhang, Kangning, et al.
Publicado: (2025)
por: Zhang, Kangning, et al.
Publicado: (2025)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
por: Xu, Ran, et al.
Publicado: (2025)
por: Xu, Ran, et al.
Publicado: (2025)
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
por: Qian, Cheng, et al.
Publicado: (2026)
por: Qian, Cheng, et al.
Publicado: (2026)
ChemAmp: Amplified Chemistry Tools via Composable Agents
por: Li, Zhucong, et al.
Publicado: (2025)
por: Li, Zhucong, et al.
Publicado: (2025)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
por: Luo, Haipeng, et al.
Publicado: (2025)
por: Luo, Haipeng, et al.
Publicado: (2025)
Current Agents Fail to Leverage World Model as Tool for Foresight
por: Qian, Cheng, et al.
Publicado: (2026)
por: Qian, Cheng, et al.
Publicado: (2026)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
por: Labrak, Yanis, et al.
Publicado: (2023)
por: Labrak, Yanis, et al.
Publicado: (2023)
ToolSandbox: A Stateful, Conversational, Interactive Evaluation Benchmark for LLM Tool Use Capabilities
por: Lu, Jiarui, et al.
Publicado: (2024)
por: Lu, Jiarui, et al.
Publicado: (2024)
MCP-AgentBench: Evaluating Real-World Language Agent Performance with MCP-Mediated Tools
por: Guo, Zikang, et al.
Publicado: (2025)
por: Guo, Zikang, et al.
Publicado: (2025)
Tool Unlearning for Tool-Augmented LLMs
por: Cheng, Jiali, et al.
Publicado: (2025)
por: Cheng, Jiali, et al.
Publicado: (2025)
Zero-Shot Dense Retrieval with Embeddings from Relevance Feedback
por: Jedidi, Nour, et al.
Publicado: (2024)
por: Jedidi, Nour, et al.
Publicado: (2024)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
por: Baidya, Avinash, et al.
Publicado: (2025)
por: Baidya, Avinash, et al.
Publicado: (2025)
DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition
por: Kim, Siun, et al.
Publicado: (2026)
por: Kim, Siun, et al.
Publicado: (2026)
STAC: When Innocent Tools Form Dangerous Chains to Jailbreak LLM Agents
por: Li, Jing-Jing, et al.
Publicado: (2025)
por: Li, Jing-Jing, et al.
Publicado: (2025)
SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation
por: Liu, Hao, et al.
Publicado: (2026)
por: Liu, Hao, et al.
Publicado: (2026)
An LLM-Tool Compiler for Fused Parallel Function Calling
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
ToolACE: Winning the Points of LLM Function Calling
por: Liu, Weiwen, et al.
Publicado: (2024)
por: Liu, Weiwen, et al.
Publicado: (2024)
To Code or not to Code? Adaptive Tool Integration for Math Language Models via Expectation-Maximization
por: Wang, Haozhe, et al.
Publicado: (2025)
por: Wang, Haozhe, et al.
Publicado: (2025)
Evaluating Tool-Augmented Agents in Remote Sensing Platforms
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning
por: Wu, Feijie, et al.
Publicado: (2025)
por: Wu, Feijie, et al.
Publicado: (2025)
Tool Learning with Foundation Models
por: Qin, Yujia, et al.
Publicado: (2023)
por: Qin, Yujia, et al.
Publicado: (2023)
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model
por: Jiang, Haitao, et al.
Publicado: (2023)
por: Jiang, Haitao, et al.
Publicado: (2023)
PathCoT: Chain-of-Thought Prompting for Zero-shot Pathology Visual Reasoning
por: Zhou, Junjie, et al.
Publicado: (2025)
por: Zhou, Junjie, et al.
Publicado: (2025)
MeNTi: Bridging Medical Calculator and LLM Agent with Nested Tool Calling
por: Zhu, Yakun, et al.
Publicado: (2024)
por: Zhu, Yakun, et al.
Publicado: (2024)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
por: Ni, Xinyi, et al.
Publicado: (2025)
por: Ni, Xinyi, et al.
Publicado: (2025)
Tools Fail: Detecting Silent Errors in Faulty Tools
por: Sun, Jimin, et al.
Publicado: (2024)
por: Sun, Jimin, et al.
Publicado: (2024)
Better LLM Reasoning via Dual-Play
por: Zhang, Zhengxin, et al.
Publicado: (2025)
por: Zhang, Zhengxin, et al.
Publicado: (2025)
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM
por: Zhang, Ruohong, et al.
Publicado: (2023)
por: Zhang, Ruohong, et al.
Publicado: (2023)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
por: Kim, Zae Myung, et al.
Publicado: (2026)
por: Kim, Zae Myung, et al.
Publicado: (2026)
SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
por: Liu, Bo, et al.
Publicado: (2025)
por: Liu, Bo, et al.
Publicado: (2025)
Ejemplares similares
-
Beyond Single-Shot: Multi-step Tool Retrieval via Query Planning
por: Fang, Wei, et al.
Publicado: (2026) -
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
por: Dong, Guanting, et al.
Publicado: (2025) -
LLM Agents Making Agent Tools
por: Wölflein, Georg, et al.
Publicado: (2025) -
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
por: Song, Xiaoshuai, et al.
Publicado: (2026) -
M$^2$PT: Multimodal Prompt Tuning for Zero-shot Instruction Learning
por: Wang, Taowen, et al.
Publicado: (2024)