SkillCraft: Can LLM Agents Learn to Use Tools Skillfully?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Shiqi, Gai, Jingze, Zhou, Ruochen, Zhang, Jinghan, Zhu, Tongyao, Li, Junlong, Wang, Kangrui, Wang, Zihan, Chen, Zhengyu, Kaleb, Klara, Miao, Ning, Gao, Siyang, Lu, Cong, Li, Manling, He, Junxian, Teh, Yee Whye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Internalizing World Models via Self-Play Finetuning for Agentic RL
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
Why Is Spatial Reasoning Hard for VLMs? An Attention Mechanism Perspective on Focus Areas
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
von: Li, Qinyu, et al.
Veröffentlicht: (2025)
von: Li, Qinyu, et al.
Veröffentlicht: (2025)
StochasTok: Improving Fine-Grained Subword Understanding in LLMs
von: Sims, Anya, et al.
Veröffentlicht: (2025)
von: Sims, Anya, et al.
Veröffentlicht: (2025)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024)
von: Sims, Anya, et al.
Veröffentlicht: (2024)
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
Does Learning Mathematical Problem-Solving Generalize to Broader Reasoning?
von: Zhou, Ruochen, et al.
Veröffentlicht: (2025)
von: Zhou, Ruochen, et al.
Veröffentlicht: (2025)
LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories
von: Kang, Liwei, et al.
Veröffentlicht: (2026)
von: Kang, Liwei, et al.
Veröffentlicht: (2026)
Semi-supervised Multiscale Matching for SAR-Optical Image
von: Gai, Jingze, et al.
Veröffentlicht: (2025)
von: Gai, Jingze, et al.
Veröffentlicht: (2025)
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
L3Ms -- Lagrange Large Language Models
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
Incorporating Unlabelled Data into Bayesian Neural Networks
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
SymDiff: Equivariant Diffusion via Stochastic Symmetrisation
von: Zhang, Leo, et al.
Veröffentlicht: (2024)
von: Zhang, Leo, et al.
Veröffentlicht: (2024)
HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness
von: Wang, Jianing, et al.
Veröffentlicht: (2026)
von: Wang, Jianing, et al.
Veröffentlicht: (2026)
Crafting India's Skill Ecology
Veröffentlicht: (2025)
Veröffentlicht: (2025)
When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems
von: Wang, Su, et al.
Veröffentlicht: (2026)
von: Wang, Su, et al.
Veröffentlicht: (2026)
Enhancing Large Language Model Reasoning with Reward Models: An Analytical Survey
von: Liu, Qiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Qiyuan, et al.
Veröffentlicht: (2025)
Manifold Aware Denoising Score Matching (MAD)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
Library Media Skills and Holiday Crafts.
Veröffentlicht: (1987)
Veröffentlicht: (1987)
Rao-Blackwellised Reparameterisation Gradients
von: Lam, Kevin H., et al.
Veröffentlicht: (2025)
von: Lam, Kevin H., et al.
Veröffentlicht: (2025)
EmbodiSkill: Skill-Aware Reflection for Self-Evolving Embodied Agents
von: Ju, Ruofei, et al.
Veröffentlicht: (2026)
von: Ju, Ruofei, et al.
Veröffentlicht: (2026)
Sub-Scaling Laws: On the Role of Data Density and Training Strategies in LLMs
von: Chen, Zhengyu, et al.
Veröffentlicht: (2025)
von: Chen, Zhengyu, et al.
Veröffentlicht: (2025)
Unleashing the Power of Meta-tuning for Few-shot Generalization Through Sparse Interpolated Experts
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2024)
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2024)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
von: Bai, Andrew, et al.
Veröffentlicht: (2025)
von: Bai, Andrew, et al.
Veröffentlicht: (2025)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
Selective Safety Steering via Value-Filtered Decoding
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
Meta-Learning Objectives for Preference Optimization
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
EvIL: Evolution Strategies for Generalisable Imitation Learning
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
Unlock Reliable Skill Inference for Quadruped Adaptive Behavior by Skill Graph
von: Zhang, Hongyin, et al.
Veröffentlicht: (2023)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2023)
Variational Flow Maps: Make Some Noise for One-Step Conditional Generation
von: Mammadov, Abbas, et al.
Veröffentlicht: (2026)
von: Mammadov, Abbas, et al.
Veröffentlicht: (2026)
Amortized Probabilistic Detection of Communities in Graphs
von: Wang, Yueqi, et al.
Veröffentlicht: (2020)
von: Wang, Yueqi, et al.
Veröffentlicht: (2020)
SigmaDock: Untwisting Molecular Docking With Fragment-Based SE(3) Diffusion
von: Prat, Alvaro, et al.
Veröffentlicht: (2025)
von: Prat, Alvaro, et al.
Veröffentlicht: (2025)
Meta Flow Maps enable scalable reward alignment
von: Potaptchik, Peter, et al.
Veröffentlicht: (2026)
von: Potaptchik, Peter, et al.
Veröffentlicht: (2026)
Prompting Strategies for Enabling Large Language Models to Infer Causation from Correlation
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
RAGEN-2: Reasoning Collapse in Agentic RL
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
SkillProbe: Security Auditing for Emerging Agent Skill Marketplaces via Multi-Agent Collaboration
von: Guo, Zihan, et al.
Veröffentlicht: (2026)
von: Guo, Zihan, et al.
Veröffentlicht: (2026)
SkillGraph: Skill-Augmented Reinforcement Learning for Agents via Evolving Skill Graphs
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2026)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Internalizing World Models via Self-Play Finetuning for Agentic RL
von: Chen, Shiqi, et al.
Veröffentlicht: (2025) -
Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging
von: Chen, Shiqi, et al.
Veröffentlicht: (2025) -
Why Is Spatial Reasoning Hard for VLMs? An Attention Mechanism Perspective on Focus Areas
von: Chen, Shiqi, et al.
Veröffentlicht: (2025) -
NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
von: Li, Qinyu, et al.
Veröffentlicht: (2025) -
StochasTok: Improving Fine-Grained Subword Understanding in LLMs
von: Sims, Anya, et al.
Veröffentlicht: (2025)