ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Feng, Jiazhan, Huang, Shijue, Qu, Xingwei, Zhang, Ge, Qin, Yujia, Zhong, Baoquan, Jiang, Chengquan, Chi, Jinxin, Zhong, Wanjun |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
par: Huang, Shijue, et autres
Publié: (2025)
par: Huang, Shijue, et autres
Publié: (2025)
ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding
par: Liu, Xiao, et autres
Publié: (2026)
par: Liu, Xiao, et autres
Publié: (2026)
Planning, Creation, Usage: Benchmarking LLMs for Comprehensive Tool Utilization in Real-World Complex Scenarios
par: Huang, Shijue, et autres
Publié: (2024)
par: Huang, Shijue, et autres
Publié: (2024)
CLongEval: A Chinese Benchmark for Evaluating Long-Context Large Language Models
par: Qiu, Zexuan, et autres
Publié: (2024)
par: Qiu, Zexuan, et autres
Publié: (2024)
Process-Supervised Reinforcement Learning for Interactive Multimodal Tool-Use Agents
par: Tan, Weiting, et autres
Publié: (2025)
par: Tan, Weiting, et autres
Publié: (2025)
VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use
par: Jiang, Dongfu, et autres
Publié: (2025)
par: Jiang, Dongfu, et autres
Publié: (2025)
Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst
par: Wang, Hongru, et autres
Publié: (2025)
par: Wang, Hongru, et autres
Publié: (2025)
Agentic Tool Use in Large Language Models
par: Hu, Jinchao, et autres
Publié: (2026)
par: Hu, Jinchao, et autres
Publié: (2026)
CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents
par: Liu, Jiayu, et autres
Publié: (2025)
par: Liu, Jiayu, et autres
Publié: (2025)
iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use
par: Zeng, Yirong, et autres
Publié: (2025)
par: Zeng, Yirong, et autres
Publié: (2025)
ToolExpander: Extending the Frontiers of Tool-Using Reinforcement Learning to Weak LLMs
par: Chen, Fu, et autres
Publié: (2025)
par: Chen, Fu, et autres
Publié: (2025)
GAP: Graph-Based Agent Planning with Parallel Tool Use and Reinforcement Learning
par: Wu, Jiaqi, et autres
Publié: (2025)
par: Wu, Jiaqi, et autres
Publié: (2025)
Advancing SLM Tool-Use Capability using Reinforcement Learning
par: Paprunia, Dhruvi, et autres
Publié: (2025)
par: Paprunia, Dhruvi, et autres
Publié: (2025)
CREATOR: Tool Creation for Disentangling Abstract and Concrete Reasoning of Large Language Models
par: Qian, Cheng, et autres
Publié: (2023)
par: Qian, Cheng, et autres
Publié: (2023)
StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning
par: Yu, Yuanqing, et autres
Publié: (2024)
par: Yu, Yuanqing, et autres
Publié: (2024)
Concise and Precise Context Compression for Tool-Using Language Models
par: Xu, Yang, et autres
Publié: (2024)
par: Xu, Yang, et autres
Publié: (2024)
StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models
par: Guo, Zhicheng, et autres
Publié: (2024)
par: Guo, Zhicheng, et autres
Publié: (2024)
Adaptive Tool Generation with Models as Tools and Reinforcement Learning
par: Wang, Chenpeng, et autres
Publié: (2025)
par: Wang, Chenpeng, et autres
Publié: (2025)
Reverse-Engineered Reasoning for Open-Ended Generation
par: Wang, Haozhe, et autres
Publié: (2025)
par: Wang, Haozhe, et autres
Publié: (2025)
PEARL: Plan Exploration and Adaptive Reinforcement Learning for Multihop Tool Use
par: Wang, Qihao, et autres
Publié: (2026)
par: Wang, Qihao, et autres
Publié: (2026)
Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning
par: Xu, Ningning, et autres
Publié: (2025)
par: Xu, Ningning, et autres
Publié: (2025)
Learning Evolving Tools for Large Language Models
par: Chen, Guoxin, et autres
Publié: (2024)
par: Chen, Guoxin, et autres
Publié: (2024)
Learning to Use Tools via Cooperative and Interactive Agents
par: Shi, Zhengliang, et autres
Publié: (2024)
par: Shi, Zhengliang, et autres
Publié: (2024)
COIG-Writer: A High-Quality Dataset for Chinese Creative Writing with Thought Processes
par: Li, Yunwen, et autres
Publié: (2025)
par: Li, Yunwen, et autres
Publié: (2025)
Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations
par: Liu, Yilong, et autres
Publié: (2026)
par: Liu, Yilong, et autres
Publié: (2026)
The Evolution of Tool Use in LLM Agents: From Single-Tool Call to Multi-Tool Orchestration
par: Xu, Haoyuan, et autres
Publié: (2026)
par: Xu, Haoyuan, et autres
Publié: (2026)
FinToolSyn: A forward synthesis Framework for Financial Tool-Use Dialogue Data with Dynamic Tool Retrieval
par: Huang, Caishuang, et autres
Publié: (2026)
par: Huang, Caishuang, et autres
Publié: (2026)
The Tool Illusion: Rethinking Tool Use in Web Agents
par: Lou, Renze, et autres
Publié: (2026)
par: Lou, Renze, et autres
Publié: (2026)
ToolRM: Towards Agentic Tool-Use Reward Modeling
par: Li, Renhao, et autres
Publié: (2025)
par: Li, Renhao, et autres
Publié: (2025)
Re-Invoke: Tool Invocation Rewriting for Zero-Shot Tool Retrieval
par: Chen, Yanfei, et autres
Publié: (2024)
par: Chen, Yanfei, et autres
Publié: (2024)
Tool Unlearning for Tool-Augmented LLMs
par: Cheng, Jiali, et autres
Publié: (2025)
par: Cheng, Jiali, et autres
Publié: (2025)
ToolGate: Contract-Grounded and Verified Tool Execution for LLMs
par: Liu, Yanming, et autres
Publié: (2026)
par: Liu, Yanming, et autres
Publié: (2026)
Acting Less is Reasoning More! Teaching Model to Act Efficiently
par: Wang, Hongru, et autres
Publié: (2025)
par: Wang, Hongru, et autres
Publié: (2025)
Learning to Edit: Aligning LLMs with Knowledge Editing
par: Jiang, Yuxin, et autres
Publié: (2024)
par: Jiang, Yuxin, et autres
Publié: (2024)
FamilyTool: A Multi-hop Personalized Tool Use Benchmark
par: Wang, Yuxin, et autres
Publié: (2025)
par: Wang, Yuxin, et autres
Publié: (2025)
Benchmarking LLM Tool-Use in the Wild
par: Yu, Peijie, et autres
Publié: (2026)
par: Yu, Peijie, et autres
Publié: (2026)
Towards Practical Tool Usage for Continually Learning LLMs
par: Huang, Jerry, et autres
Publié: (2024)
par: Huang, Jerry, et autres
Publié: (2024)
Try, Check and Retry: A Divide-and-Conquer Framework for Boosting Long-context Tool-Calling Performance of LLMs
par: Chen, Kunfeng, et autres
Publié: (2026)
par: Chen, Kunfeng, et autres
Publié: (2026)
MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
par: Huang, Yue, et autres
Publié: (2023)
par: Huang, Yue, et autres
Publié: (2023)
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
par: Dong, Guanting, et autres
Publié: (2026)
par: Dong, Guanting, et autres
Publié: (2026)
Documents similaires
-
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
par: Huang, Shijue, et autres
Publié: (2025) -
ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding
par: Liu, Xiao, et autres
Publié: (2026) -
Planning, Creation, Usage: Benchmarking LLMs for Comprehensive Tool Utilization in Real-World Complex Scenarios
par: Huang, Shijue, et autres
Publié: (2024) -
CLongEval: A Chinese Benchmark for Evaluating Long-Context Large Language Models
par: Qiu, Zexuan, et autres
Publié: (2024) -
Process-Supervised Reinforcement Learning for Interactive Multimodal Tool-Use Agents
par: Tan, Weiting, et autres
Publié: (2025)