From Failure to Mastery: Generating Hard Samples for Tool-use Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hao, Bingguang, Xu, Zengzhuang, Wen, Yuntao, Xu, Xinyi, Liu, Yang, Zhao, Tong, Wang, Maolin, Chen, Long, Wang, Dong, Chen, Yicheng, Peng, Cunyin, Zhao, Xiangyu, Zhuang, Chenyi, Zhang, Ji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning through Exploration: A Reinforcement Learning Framework for Robust Function Calling
von: Hao, Bingguang, et al.
Veröffentlicht: (2025)
von: Hao, Bingguang, et al.
Veröffentlicht: (2025)
FunReason-MT Technical Report: Advanced Data Synthesis Solution for Real-world Multi-Turn Tool-use
von: Xu, Zengzhuang, et al.
Veröffentlicht: (2025)
von: Xu, Zengzhuang, et al.
Veröffentlicht: (2025)
FunReason: Enhancing Large Language Models' Function Calling via Self-Refinement Multiscale Loss and Automated Data Refinement
von: Hao, Bingguang, et al.
Veröffentlicht: (2025)
von: Hao, Bingguang, et al.
Veröffentlicht: (2025)
StressWeb: A Diagnostic Benchmark for Web Agent Robustness under Realistic Interaction Variability
von: Bai, Haoyue, et al.
Veröffentlicht: (2026)
von: Bai, Haoyue, et al.
Veröffentlicht: (2026)
Tensorized Hypergraph Neural Networks
von: Wang, Maolin, et al.
Veröffentlicht: (2023)
von: Wang, Maolin, et al.
Veröffentlicht: (2023)
Large Multimodal Model Compression via Efficient Pruning and Distillation at AntGroup
von: Wang, Maolin, et al.
Veröffentlicht: (2023)
von: Wang, Maolin, et al.
Veröffentlicht: (2023)
MetaLoRA: Tensor-Enhanced Adaptive Low-Rank Fine-tuning
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
V2P: Visual Attention Calibration for GUI Grounding via Background Suppression and Center Peaking
von: Chen, Jikai, et al.
Veröffentlicht: (2026)
von: Chen, Jikai, et al.
Veröffentlicht: (2026)
V2P: Visual Attention Calibration for GUI Grounding via Background Suppression and Center Peaking
von: Chen, Jikai, et al.
Veröffentlicht: (2025)
von: Chen, Jikai, et al.
Veröffentlicht: (2025)
From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
From Exploration to Mastery: Enabling LLMs to Master Tools via Self-Driven Interactions
von: Qu, Changle, et al.
Veröffentlicht: (2024)
von: Qu, Changle, et al.
Veröffentlicht: (2024)
DNS-Rec: Data-aware Neural Architecture Search for Recommender Systems
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
Transformer-Enhanced Motion Planner: Attention-Guided Sampling for State-Specific Decision Making
von: Zhuang, Lei, et al.
Veröffentlicht: (2024)
von: Zhuang, Lei, et al.
Veröffentlicht: (2024)
Recon-Act: A Self-Evolving Multi-Agent Browser-Use System via Web Reconnaissance, Tool Generation, and Task Execution
von: He, Kaiwen, et al.
Veröffentlicht: (2025)
von: He, Kaiwen, et al.
Veröffentlicht: (2025)
SQLCritic: Correcting Text-to-SQL Generation via Clause-wise Critic
von: Chen, Jikai, et al.
Veröffentlicht: (2025)
von: Chen, Jikai, et al.
Veröffentlicht: (2025)
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents
von: Yang, Min, et al.
Veröffentlicht: (2026)
von: Yang, Min, et al.
Veröffentlicht: (2026)
Multi-Agent Tool-Integrated Policy Optimization
von: Mo, Zhanfeng, et al.
Veröffentlicht: (2025)
von: Mo, Zhanfeng, et al.
Veröffentlicht: (2025)
Solving Jigsaw Puzzles using Iterative Random Sampling: Parallels with Development of Skill Mastery
von: Zhao, Neil, et al.
Veröffentlicht: (2024)
von: Zhao, Neil, et al.
Veröffentlicht: (2024)
To Search or Not to Search: Aligning the Decision Boundary of Deep Search Agents via Causal Intervention
von: Zhang, Wenlin, et al.
Veröffentlicht: (2026)
von: Zhang, Wenlin, et al.
Veröffentlicht: (2026)
Empowering Denoising Sequential Recommendation with Large Language Model Embeddings
von: Wu, Tongzhou, et al.
Veröffentlicht: (2025)
von: Wu, Tongzhou, et al.
Veröffentlicht: (2025)
Large π‐Conjugated Structure Strategy for Optimizing Polyimide Film‐Forming Properties and Developing Gas Fluorescence Sensors
von: Xuecheng Wang, et al.
Veröffentlicht: (2024)
von: Xuecheng Wang, et al.
Veröffentlicht: (2024)
A Simulation Analysis of the Slumping Failure of Unstable Rocks With a Weak Base
von: Yuntao Zhou, et al.
Veröffentlicht: (2024)
von: Yuntao Zhou, et al.
Veröffentlicht: (2024)
GLINT-RU: Gated Lightweight Intelligent Recurrent Units for Sequential Recommender Systems
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
Improving LLM-based Recommendation with Self-Hard Negatives from Intermediate Layers
von: Li, Bingqian, et al.
Veröffentlicht: (2026)
von: Li, Bingqian, et al.
Veröffentlicht: (2026)
An Open and Comprehensive Pipeline for Unified Object Grounding and Detection
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
A Visual Reinforcement Learning-Based Separate Primitive Policy for Peg-in-Hole Tasks
von: Xu, Zichun, et al.
Veröffentlicht: (2025)
von: Xu, Zichun, et al.
Veröffentlicht: (2025)
Align-GRAG: Anchor and Rationale Guided Dual Alignment for Graph Retrieval-Augmented Generation
von: Xu, Derong, et al.
Veröffentlicht: (2025)
von: Xu, Derong, et al.
Veröffentlicht: (2025)
From Single to Multi-Granularity: Toward Long-Term Memory Association and Selection of Conversational Agents
von: Xu, Derong, et al.
Veröffentlicht: (2025)
von: Xu, Derong, et al.
Veröffentlicht: (2025)
GTA: A Benchmark for General Tool Agents
von: Wang, Jize, et al.
Veröffentlicht: (2024)
von: Wang, Jize, et al.
Veröffentlicht: (2024)
Measure Domain's Gap: A Similar Domain Selection Principle for Multi-Domain Recommendation
von: Wen, Yi, et al.
Veröffentlicht: (2025)
von: Wen, Yi, et al.
Veröffentlicht: (2025)
Charting a Path to Tumor Therapy: Mastery of Luminescent Metal–Organic Frameworks for Precise Spatiotemporal Control
von: Renchi Gao, et al.
Veröffentlicht: (2024)
von: Renchi Gao, et al.
Veröffentlicht: (2024)
MOSAIC: Composable Safety Alignment with Modular Control Tokens
von: Peng, Jingyu, et al.
Veröffentlicht: (2026)
von: Peng, Jingyu, et al.
Veröffentlicht: (2026)
Open-Source Reinforcement Learning Environments Implemented in MuJoCo with Franka Manipulator
von: Xu, Zichun, et al.
Veröffentlicht: (2023)
von: Xu, Zichun, et al.
Veröffentlicht: (2023)
Sample Complexity of Neural Policy Mirror Descent for Policy Optimization on Low-Dimensional Manifolds
von: Xu, Zhenghao, et al.
Veröffentlicht: (2023)
von: Xu, Zhenghao, et al.
Veröffentlicht: (2023)
MKGL: Mastery of a Three-Word Language
von: Guo, Lingbing, et al.
Veröffentlicht: (2024)
von: Guo, Lingbing, et al.
Veröffentlicht: (2024)
Hard Magnetic Graphene Nanocomposite for Multimodal, Reconfigurable Soft Electronics
von: Zehua Xiang, et al.
Veröffentlicht: (2024)
von: Zehua Xiang, et al.
Veröffentlicht: (2024)
Data Efficient Adaptation in Large Language Models via Continuous Low-Rank Fine-Tuning
von: Han, Xiao, et al.
Veröffentlicht: (2025)
von: Han, Xiao, et al.
Veröffentlicht: (2025)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
von: Zhao, Yao, et al.
Veröffentlicht: (2023)
von: Zhao, Yao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Reasoning through Exploration: A Reinforcement Learning Framework for Robust Function Calling
von: Hao, Bingguang, et al.
Veröffentlicht: (2025) -
FunReason-MT Technical Report: Advanced Data Synthesis Solution for Real-world Multi-Turn Tool-use
von: Xu, Zengzhuang, et al.
Veröffentlicht: (2025) -
FunReason: Enhancing Large Language Models' Function Calling via Self-Refinement Multiscale Loss and Automated Data Refinement
von: Hao, Bingguang, et al.
Veröffentlicht: (2025) -
StressWeb: A Diagnostic Benchmark for Web Agent Robustness under Realistic Interaction Variability
von: Bai, Haoyue, et al.
Veröffentlicht: (2026) -
Tensorized Hypergraph Neural Networks
von: Wang, Maolin, et al.
Veröffentlicht: (2023)