From Exploration to Mastery: Enabling LLMs to Master Tools via Self-Driven Interactions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qu, Changle, Dai, Sunhao, Wei, Xiaochi, Cai, Hengyi, Wang, Shuaiqiang, Yin, Dawei, Xu, Jun, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tool Learning with Large Language Models: A Survey
von: Qu, Changle, et al.
Veröffentlicht: (2024)
von: Qu, Changle, et al.
Veröffentlicht: (2024)
Towards Completeness-Oriented Tool Retrieval for Large Language Models
von: Qu, Changle, et al.
Veröffentlicht: (2024)
von: Qu, Changle, et al.
Veröffentlicht: (2024)
MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching
von: Qu, Changle, et al.
Veröffentlicht: (2026)
von: Qu, Changle, et al.
Veröffentlicht: (2026)
Efficient Thought Space Exploration Through Strategic Intervention
von: Li, Ziheng, et al.
Veröffentlicht: (2025)
von: Li, Ziheng, et al.
Veröffentlicht: (2025)
XL$^2$Bench: A Benchmark for Extremely Long Context Understanding with Long-range Dependencies
von: Ni, Xuanfan, et al.
Veröffentlicht: (2024)
von: Ni, Xuanfan, et al.
Veröffentlicht: (2024)
Towards Verifiable Text Generation with Evolving Memory and Self-Reflection
von: Sun, Hao, et al.
Veröffentlicht: (2023)
von: Sun, Hao, et al.
Veröffentlicht: (2023)
Learning to Retrieve from Agent Trajectories
von: Zhou, Yuqi, et al.
Veröffentlicht: (2026)
von: Zhou, Yuqi, et al.
Veröffentlicht: (2026)
ReCODE: Modeling Repeat Consumption with Neural ODE
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
CurES: From Gradient Analysis to Efficient Curriculum Learning for Reasoning LLMs
von: Zeng, Yongcheng, et al.
Veröffentlicht: (2025)
von: Zeng, Yongcheng, et al.
Veröffentlicht: (2025)
AdaSwitch: Balancing Exploration and Guidance in Knowledge Distillation via Adaptive Switching
von: Peng, Jingyu, et al.
Veröffentlicht: (2025)
von: Peng, Jingyu, et al.
Veröffentlicht: (2025)
LLMs + Persona-Plug = Personalized LLMs
von: Liu, Jiongnan, et al.
Veröffentlicht: (2024)
von: Liu, Jiongnan, et al.
Veröffentlicht: (2024)
Exploring the Potential of Large Language Models (LLMs) in Learning on Graphs
von: Chen, Zhikai, et al.
Veröffentlicht: (2023)
von: Chen, Zhikai, et al.
Veröffentlicht: (2023)
Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM Agents
von: Gao, Heyang, et al.
Veröffentlicht: (2025)
von: Gao, Heyang, et al.
Veröffentlicht: (2025)
KuaiLive: A Real-time Interactive Dataset for Live Streaming Recommendation
von: Qu, Changle, et al.
Veröffentlicht: (2025)
von: Qu, Changle, et al.
Veröffentlicht: (2025)
Cross-model Control: Improving Multiple Large Language Models in One-time Training
von: Wu, Jiayi, et al.
Veröffentlicht: (2024)
von: Wu, Jiayi, et al.
Veröffentlicht: (2024)
AdaFuse: Accelerating Dynamic Adapter Inference via Token-Level Pre-Gating and Fused Kernel Optimization
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization
von: Wu, Jiayi, et al.
Veröffentlicht: (2024)
von: Wu, Jiayi, et al.
Veröffentlicht: (2024)
From Failure to Mastery: Generating Hard Samples for Tool-use Agents
von: Hao, Bingguang, et al.
Veröffentlicht: (2026)
von: Hao, Bingguang, et al.
Veröffentlicht: (2026)
From Prompting to Alignment: A Generative Framework for Query Recommendation
von: Min, Erxue, et al.
Veröffentlicht: (2025)
von: Min, Erxue, et al.
Veröffentlicht: (2025)
Not All Preferences Are Created Equal: Stability-Aware and Gradient-Efficient Alignment for Reasoning Models
von: Wu, Hui, et al.
Veröffentlicht: (2026)
von: Wu, Hui, et al.
Veröffentlicht: (2026)
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation
von: Zhu, Dongsheng, et al.
Veröffentlicht: (2025)
von: Zhu, Dongsheng, et al.
Veröffentlicht: (2025)
Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
Exploring the Escalation of Source Bias in User, Data, and Recommender System Feedback Loop
von: Zhou, Yuqi, et al.
Veröffentlicht: (2024)
von: Zhou, Yuqi, et al.
Veröffentlicht: (2024)
AgentSkiller: Scaling Generalist Agent Intelligence through Semantically Integrated Cross-Domain Data Synthesis
von: Sun, Zexu, et al.
Veröffentlicht: (2026)
von: Sun, Zexu, et al.
Veröffentlicht: (2026)
UOEP: User-Oriented Exploration Policy for Enhancing Long-Term User Experiences in Recommender Systems
von: Zhang, Changshuo, et al.
Veröffentlicht: (2024)
von: Zhang, Changshuo, et al.
Veröffentlicht: (2024)
Mirage of Mastery: Memorization Tricks LLMs into Artificially Inflated Self-Knowledge
von: Kale, Sahil
Veröffentlicht: (2025)
von: Kale, Sahil
Veröffentlicht: (2025)
Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method
von: Zhao, Yukun, et al.
Veröffentlicht: (2023)
von: Zhao, Yukun, et al.
Veröffentlicht: (2023)
CTR-Guided Generative Query Suggestion in Conversational Search
von: Min, Erxue, et al.
Veröffentlicht: (2025)
von: Min, Erxue, et al.
Veröffentlicht: (2025)
NExT-Search: Rebuilding User Feedback Ecosystem for Generative AI Search
von: Dai, Sunhao, et al.
Veröffentlicht: (2025)
von: Dai, Sunhao, et al.
Veröffentlicht: (2025)
MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering
von: Wu, Hui, et al.
Veröffentlicht: (2026)
von: Wu, Hui, et al.
Veröffentlicht: (2026)
CitaLaw: Enhancing LLM with Citations in Legal Domain
von: Zhang, Kepu, et al.
Veröffentlicht: (2024)
von: Zhang, Kepu, et al.
Veröffentlicht: (2024)
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents
von: Yang, Min, et al.
Veröffentlicht: (2026)
von: Yang, Min, et al.
Veröffentlicht: (2026)
Replication and Exploration of Generative Retrieval over Dynamic Corpora
von: Zhang, Zhen, et al.
Veröffentlicht: (2025)
von: Zhang, Zhen, et al.
Veröffentlicht: (2025)
Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding
von: Li, Ziheng, et al.
Veröffentlicht: (2025)
von: Li, Ziheng, et al.
Veröffentlicht: (2025)
Leveraging Generative Models for Real-Time Query-Driven Text Summarization in Large-Scale Web Search
von: Xiong, Zeyu, et al.
Veröffentlicht: (2025)
von: Xiong, Zeyu, et al.
Veröffentlicht: (2025)
Towards Next-Generation Recommender Systems: A Benchmark for Personalized Recommendation Assistant with LLMs
von: Huang, Jiani, et al.
Veröffentlicht: (2025)
von: Huang, Jiani, et al.
Veröffentlicht: (2025)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity Documents
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking
von: Zhu, Junda, et al.
Veröffentlicht: (2025)
von: Zhu, Junda, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tool Learning with Large Language Models: A Survey
von: Qu, Changle, et al.
Veröffentlicht: (2024) -
Towards Completeness-Oriented Tool Retrieval for Large Language Models
von: Qu, Changle, et al.
Veröffentlicht: (2024) -
MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching
von: Qu, Changle, et al.
Veröffentlicht: (2026) -
Efficient Thought Space Exploration Through Strategic Intervention
von: Li, Ziheng, et al.
Veröffentlicht: (2025) -
XL$^2$Bench: A Benchmark for Extremely Long Context Understanding with Long-range Dependencies
von: Ni, Xuanfan, et al.
Veröffentlicht: (2024)