Recoverability Has a Law: The ERR Measure for Tool-Augmented Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Vuddanti, Sri Vatsa, Chittiprolu, Satwik Kumar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PALADIN: Self-Correcting Language Model Agents to Cure Tool-Failure Cases
di: Vuddanti, Sri Vatsa, et al.
Pubblicazione: (2025)
di: Vuddanti, Sri Vatsa, et al.
Pubblicazione: (2025)
Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use
di: Thaman, Kunvar
Pubblicazione: (2026)
di: Thaman, Kunvar
Pubblicazione: (2026)
Evaluating Tool-Augmented Agents in Remote Sensing Platforms
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
Has the Deep Neural Network learned the Stochastic Process? An Evaluation Viewpoint
di: Kumar, Harshit, et al.
Pubblicazione: (2024)
di: Kumar, Harshit, et al.
Pubblicazione: (2024)
ML-Tool-Bench: Tool-Augmented Planning for ML Tasks
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
Exploring Multi-Modal Data with Tool-Augmented LLM Agents for Precise Causal Discovery
di: Shen, ChengAo, et al.
Pubblicazione: (2024)
di: Shen, ChengAo, et al.
Pubblicazione: (2024)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
di: Luo, Haipeng, et al.
Pubblicazione: (2025)
di: Luo, Haipeng, et al.
Pubblicazione: (2025)
FCOS: A Two-Stage Recoverable Model Pruning Framework for Automatic Modulation Recognition
di: Lu, Yao, et al.
Pubblicazione: (2025)
di: Lu, Yao, et al.
Pubblicazione: (2025)
Tool Unlearning for Tool-Augmented LLMs
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
Augmented Reinforcement Learning Framework For Enhancing Decision-Making In Machine Learning Models Using External Agents
di: Singh, Sandesh Kumar
Pubblicazione: (2025)
di: Singh, Sandesh Kumar
Pubblicazione: (2025)
Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
di: Zeng, Yirong, et al.
Pubblicazione: (2025)
di: Zeng, Yirong, et al.
Pubblicazione: (2025)
Weightless Neural Networks for Continuously Trainable Personalized Recommendation Systems
di: Latif, Rafayel, et al.
Pubblicazione: (2025)
di: Latif, Rafayel, et al.
Pubblicazione: (2025)
Hidden-State Privacy Has an Empty Middle
di: Bell, Alexander Okezue
Pubblicazione: (2026)
di: Bell, Alexander Okezue
Pubblicazione: (2026)
TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools
di: Gao, Shanghua, et al.
Pubblicazione: (2025)
di: Gao, Shanghua, et al.
Pubblicazione: (2025)
Scaling Laws for Pre-training Agents and World Models
di: Pearce, Tim, et al.
Pubblicazione: (2024)
di: Pearce, Tim, et al.
Pubblicazione: (2024)
Scaling Laws for Imitation Learning in Single-Agent Games
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
Identifiability Matters: Revealing the Hidden Recoverable Condition in Unbiased Learning to Rank
di: Chen, Mouxiang, et al.
Pubblicazione: (2023)
di: Chen, Mouxiang, et al.
Pubblicazione: (2023)
Optimal Scaling Laws for Efficiency Gains in a Theoretical Transformer-Augmented Sectional MoE Framework
di: Sane, Soham
Pubblicazione: (2025)
di: Sane, Soham
Pubblicazione: (2025)
State Contamination in Memory-Augmented LLM Agents
di: Wang, Yian, et al.
Pubblicazione: (2026)
di: Wang, Yian, et al.
Pubblicazione: (2026)
DART: Semantic Recoverability for Structured Tool Agents
di: Yang, Ke, et al.
Pubblicazione: (2026)
di: Yang, Ke, et al.
Pubblicazione: (2026)
Understanding Tool-Augmented Agents for Lean Formalization: A Factorial Analysis
di: Zhang, Ke, et al.
Pubblicazione: (2026)
di: Zhang, Ke, et al.
Pubblicazione: (2026)
LawPal : A Retrieval Augmented Generation Based System for Enhanced Legal Accessibility in India
di: Panchal, Dnyanesh, et al.
Pubblicazione: (2025)
di: Panchal, Dnyanesh, et al.
Pubblicazione: (2025)
OR-Toolformer: Modeling and Solving Operations Research Problems with Tool Augmented Large Language Models
di: Zhang, Jianzhang, et al.
Pubblicazione: (2025)
di: Zhang, Jianzhang, et al.
Pubblicazione: (2025)
Machine Learning as a Tool (MLAT): A Framework for Integrating Statistical ML Models as Callable Tools within LLM Agent Workflows
di: Chen, Edwin, et al.
Pubblicazione: (2026)
di: Chen, Edwin, et al.
Pubblicazione: (2026)
What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
di: Vafa, Keyon, et al.
Pubblicazione: (2025)
di: Vafa, Keyon, et al.
Pubblicazione: (2025)
On the Structural Limitations of Weight-Based Neural Adaptation and the Role of Reversible Behavioral Learning
di: Konduru, Pardhu Sri Rushi Varma
Pubblicazione: (2026)
di: Konduru, Pardhu Sri Rushi Varma
Pubblicazione: (2026)
SparK: Query-Aware Unstructured Sparsity with Recoverable KV Cache Channel Pruning
di: Liao, Huanxuan, et al.
Pubblicazione: (2025)
di: Liao, Huanxuan, et al.
Pubblicazione: (2025)
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
di: Huang, Xinting, et al.
Pubblicazione: (2026)
di: Huang, Xinting, et al.
Pubblicazione: (2026)
MetaTool: Facilitating Large Language Models to Master Tools with Meta-task Augmentation
di: Wang, Xiaohan, et al.
Pubblicazione: (2024)
di: Wang, Xiaohan, et al.
Pubblicazione: (2024)
Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models
di: Kumar, Sachin
Pubblicazione: (2026)
di: Kumar, Sachin
Pubblicazione: (2026)
Has LLM Reached the Scaling Ceiling Yet? Unified Insights into LLM Regularities and Constraints
di: Luo, Charles
Pubblicazione: (2024)
di: Luo, Charles
Pubblicazione: (2024)
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
di: Goldie, Anna, et al.
Pubblicazione: (2024)
di: Goldie, Anna, et al.
Pubblicazione: (2024)
Towards Automated Patent Workflows: AI-Orchestrated Multi-Agent Framework for Intellectual Property Management and Analysis
di: Srinivas, Sakhinana Sagar, et al.
Pubblicazione: (2024)
di: Srinivas, Sakhinana Sagar, et al.
Pubblicazione: (2024)
SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents
di: Feng, Xinshun, et al.
Pubblicazione: (2026)
di: Feng, Xinshun, et al.
Pubblicazione: (2026)
Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR)
di: Knorr, Marius S., et al.
Pubblicazione: (2026)
di: Knorr, Marius S., et al.
Pubblicazione: (2026)
Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents
di: Zhen, Shuai, et al.
Pubblicazione: (2026)
di: Zhen, Shuai, et al.
Pubblicazione: (2026)
SynthTools: A Framework for Scaling Synthetic Tools for Agent Development
di: Castellani, Tommaso, et al.
Pubblicazione: (2025)
di: Castellani, Tommaso, et al.
Pubblicazione: (2025)
UrbanMind: Towards Urban General Intelligence via Tool-Enhanced Retrieval-Augmented Generation and Multilevel Optimization
di: Yang, Kai, et al.
Pubblicazione: (2025)
di: Yang, Kai, et al.
Pubblicazione: (2025)
Reinforced Agent: Inference-Time Feedback for Tool-Calling Agents
di: Ta, Anh, et al.
Pubblicazione: (2026)
di: Ta, Anh, et al.
Pubblicazione: (2026)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
di: Li, Sijia, et al.
Pubblicazione: (2025)
di: Li, Sijia, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PALADIN: Self-Correcting Language Model Agents to Cure Tool-Failure Cases
di: Vuddanti, Sri Vatsa, et al.
Pubblicazione: (2025) -
Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use
di: Thaman, Kunvar
Pubblicazione: (2026) -
Evaluating Tool-Augmented Agents in Remote Sensing Platforms
di: Singh, Simranjit, et al.
Pubblicazione: (2024) -
Has the Deep Neural Network learned the Stochastic Process? An Evaluation Viewpoint
di: Kumar, Harshit, et al.
Pubblicazione: (2024) -
ML-Tool-Bench: Tool-Augmented Planning for ML Tasks
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)