Guardado en:
| Autores principales: | Chen, Fu, Wang, Peng, Li, Xiyin, Li, Wen, Lei, Shichi, Xiang, Dongdong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2510.07737 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
por: Dong, Guanting, et al.
Publicado: (2025)
por: Dong, Guanting, et al.
Publicado: (2025)
Tool Unlearning for Tool-Augmented LLMs
por: Cheng, Jiali, et al.
Publicado: (2025)
por: Cheng, Jiali, et al.
Publicado: (2025)
RLFR: Extending Reinforcement Learning for LLMs with Flow Environment
por: Zhang, Jinghao, et al.
Publicado: (2025)
por: Zhang, Jinghao, et al.
Publicado: (2025)
Reinforcement Learning for Tool-Integrated Interleaved Thinking towards Cross-Domain Generalization
por: Chen, Zhengyu, et al.
Publicado: (2025)
por: Chen, Zhengyu, et al.
Publicado: (2025)
Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe
por: Wu, Xixi, et al.
Publicado: (2026)
por: Wu, Xixi, et al.
Publicado: (2026)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
por: Wang, Xingyao, et al.
Publicado: (2023)
por: Wang, Xingyao, et al.
Publicado: (2023)
ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs
por: Ding, Peng, et al.
Publicado: (2025)
por: Ding, Peng, et al.
Publicado: (2025)
ToolRL: Reward is All Tool Learning Needs
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision
por: Lu, Yifei, et al.
Publicado: (2025)
por: Lu, Yifei, et al.
Publicado: (2025)
Training LLMs for Multi-Step Tool Orchestration with Constrained Data Synthesis and Graduated Rewards
por: Jiayang, Cheng, et al.
Publicado: (2026)
por: Jiayang, Cheng, et al.
Publicado: (2026)
ToolSample: Dual Dynamic Sampling Methods with Curriculum Learning for RL-based Tool Learning
por: Feng, Zihao, et al.
Publicado: (2025)
por: Feng, Zihao, et al.
Publicado: (2025)
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error
por: Wang, Boshi, et al.
Publicado: (2024)
por: Wang, Boshi, et al.
Publicado: (2024)
Quality Matters: Evaluating Synthetic Data for Tool-Using LLMs
por: Iskander, Shadi, et al.
Publicado: (2024)
por: Iskander, Shadi, et al.
Publicado: (2024)
Towards Practical Tool Usage for Continually Learning LLMs
por: Huang, Jerry, et al.
Publicado: (2024)
por: Huang, Jerry, et al.
Publicado: (2024)
ToolACE-R: Model-aware Iterative Training and Adaptive Refinement for Tool Learning
por: Zeng, Xingshan, et al.
Publicado: (2025)
por: Zeng, Xingshan, et al.
Publicado: (2025)
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning
por: Zou, Jiaru, et al.
Publicado: (2025)
por: Zou, Jiaru, et al.
Publicado: (2025)
Unified Tool Integration for LLMs: A Protocol-Agnostic Approach to Function Calling
por: Ding, Peng, et al.
Publicado: (2025)
por: Ding, Peng, et al.
Publicado: (2025)
iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use
por: Zeng, Yirong, et al.
Publicado: (2025)
por: Zeng, Yirong, et al.
Publicado: (2025)
Deep Learning and Machine Learning, Advancing Big Data Analytics and Management: Unveiling AI's Potential Through Tools, Techniques, and Applications
por: Feng, Pohsun, et al.
Publicado: (2024)
por: Feng, Pohsun, et al.
Publicado: (2024)
MetaTool: Facilitating Large Language Models to Master Tools with Meta-task Augmentation
por: Wang, Xiaohan, et al.
Publicado: (2024)
por: Wang, Xiaohan, et al.
Publicado: (2024)
Tool Learning with Foundation Models
por: Qin, Yujia, et al.
Publicado: (2023)
por: Qin, Yujia, et al.
Publicado: (2023)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
por: Kim, Zae Myung, et al.
Publicado: (2026)
por: Kim, Zae Myung, et al.
Publicado: (2026)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
por: Xu, Ran, et al.
Publicado: (2025)
por: Xu, Ran, et al.
Publicado: (2025)
More Vulnerable than You Think: On the Stability of Tool-Integrated LLM Agents
por: Xiong, Weimin, et al.
Publicado: (2025)
por: Xiong, Weimin, et al.
Publicado: (2025)
MATATA: Weakly Supervised End-to-End MAthematical Tool-Augmented Reasoning for Tabular Applications
por: Vinayagame, Vishnou, et al.
Publicado: (2024)
por: Vinayagame, Vishnou, et al.
Publicado: (2024)
Tool Preferences in Agentic LLMs are Unreliable
por: Faghih, Kazem, et al.
Publicado: (2025)
por: Faghih, Kazem, et al.
Publicado: (2025)
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation
por: Zhu, Dongsheng, et al.
Publicado: (2025)
por: Zhu, Dongsheng, et al.
Publicado: (2025)
Reasoning and Tools for Human-Level Forecasting
por: Hsieh, Elvis, et al.
Publicado: (2024)
por: Hsieh, Elvis, et al.
Publicado: (2024)
SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning
por: Liang, Xiao, et al.
Publicado: (2025)
por: Liang, Xiao, et al.
Publicado: (2025)
ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings
por: Hao, Shibo, et al.
Publicado: (2023)
por: Hao, Shibo, et al.
Publicado: (2023)
AutoTriton: Automatic Triton Programming with Reinforcement Learning in LLMs
por: Li, Shangzhan, et al.
Publicado: (2025)
por: Li, Shangzhan, et al.
Publicado: (2025)
ConfClip: Confidence-Weighted and Clipped Reward for Reinforcement Learning in LLMs
por: Zhang, Bonan, et al.
Publicado: (2025)
por: Zhang, Bonan, et al.
Publicado: (2025)
The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs
por: Nawrot, Piotr, et al.
Publicado: (2025)
por: Nawrot, Piotr, et al.
Publicado: (2025)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
por: Zhao, Xuandong, et al.
Publicado: (2024)
por: Zhao, Xuandong, et al.
Publicado: (2024)
CheMatAgent: Enhancing LLMs for Chemistry and Materials Science through Tree-Search Based Tool Learning
por: Wu, Mengsong, et al.
Publicado: (2025)
por: Wu, Mengsong, et al.
Publicado: (2025)
Reinforcement Learning Enhanced LLMs: A Survey
por: Wang, Shuhe, et al.
Publicado: (2024)
por: Wang, Shuhe, et al.
Publicado: (2024)
Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2025)
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2025)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
por: Dong, Wenhan, et al.
Publicado: (2025)
por: Dong, Wenhan, et al.
Publicado: (2025)
Conveyor: Efficient Tool-aware LLM Serving with Tool Partial Execution
por: Xu, Yechen, et al.
Publicado: (2024)
por: Xu, Yechen, et al.
Publicado: (2024)
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
por: Song, Kefan, et al.
Publicado: (2025)
por: Song, Kefan, et al.
Publicado: (2025)
Ejemplares similares
-
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
por: Dong, Guanting, et al.
Publicado: (2025) -
Tool Unlearning for Tool-Augmented LLMs
por: Cheng, Jiali, et al.
Publicado: (2025) -
RLFR: Extending Reinforcement Learning for LLMs with Flow Environment
por: Zhang, Jinghao, et al.
Publicado: (2025) -
Reinforcement Learning for Tool-Integrated Interleaved Thinking towards Cross-Domain Generalization
por: Chen, Zhengyu, et al.
Publicado: (2025) -
Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe
por: Wu, Xixi, et al.
Publicado: (2026)