Gespeichert in:
| Hauptverfasser: | Delassi, Khaled Bachir, Zeggane, Lakhdar, Cherroun, Hadda, Haouhat, Abdelhamid, Bouzouad, Kaoutar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2508.03488 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Arabic Multimodal Machine Learning: Datasets, Applications, Approaches, and Challenges
von: Haouhat, Abdelhamid, et al.
Veröffentlicht: (2025)
von: Haouhat, Abdelhamid, et al.
Veröffentlicht: (2025)
Semantic Tool Discovery for Large Language Models: A Vector-Based Approach to MCP Tool Selection
von: Mudunuri, Sarat, et al.
Veröffentlicht: (2026)
von: Mudunuri, Sarat, et al.
Veröffentlicht: (2026)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
von: Sherifi, Betim, et al.
Veröffentlicht: (2024)
von: Sherifi, Betim, et al.
Veröffentlicht: (2024)
Tool-integrated Reinforcement Learning for Repo Deep Search
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
ToolFuzz -- Automated Agent Tool Testing
von: Milev, Ivan, et al.
Veröffentlicht: (2025)
von: Milev, Ivan, et al.
Veröffentlicht: (2025)
JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents
von: Ghoshal, Sandip, et al.
Veröffentlicht: (2026)
von: Ghoshal, Sandip, et al.
Veröffentlicht: (2026)
Investigating Tool-Memory Conflicts in Tool-Augmented LLMs
von: Cheng, Jiali, et al.
Veröffentlicht: (2026)
von: Cheng, Jiali, et al.
Veröffentlicht: (2026)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
Rethinking the Role of Entropy in Optimizing Tool-Use Behaviors for Large Language Model Agents
von: Li, Zeping, et al.
Veröffentlicht: (2026)
von: Li, Zeping, et al.
Veröffentlicht: (2026)
ParaTool: Shifting Tool Representations from Context to Parameters
von: Yu, Zekai, et al.
Veröffentlicht: (2026)
von: Yu, Zekai, et al.
Veröffentlicht: (2026)
LLMs Integration in Software Engineering Team Projects: Roles, Impact, and a Pedagogical Design Space for AI Tools in Computing Education
von: Kharrufa, Ahmed, et al.
Veröffentlicht: (2024)
von: Kharrufa, Ahmed, et al.
Veröffentlicht: (2024)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
von: Li, Dawei, et al.
Veröffentlicht: (2026)
von: Li, Dawei, et al.
Veröffentlicht: (2026)
Comparison of Static Application Security Testing Tools and Large Language Models for Repo-level Vulnerability Detection
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
EvolveTool-Bench: Evaluating the Quality of LLM-Generated Tool Libraries as Software Artifacts
von: Kaliyev, Alibek T., et al.
Veröffentlicht: (2026)
von: Kaliyev, Alibek T., et al.
Veröffentlicht: (2026)
ToolMisuseBench: An Offline Deterministic Benchmark for Tool Misuse and Recovery in Agentic Systems
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment
von: Yu, Shasha, et al.
Veröffentlicht: (2026)
von: Yu, Shasha, et al.
Veröffentlicht: (2026)
Class-Level Code Generation from Natural Language Using Iterative, Tool-Enhanced Reasoning over Repository
von: Deshpande, Ajinkya, et al.
Veröffentlicht: (2024)
von: Deshpande, Ajinkya, et al.
Veröffentlicht: (2024)
The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
Breaking the Illusion of Identity in LLM Tooling
von: Miller, Marek
Veröffentlicht: (2026)
von: Miller, Marek
Veröffentlicht: (2026)
Schema First Tool APIs for LLM Agents: A Controlled Study of Tool Misuse, Recovery, and Budgeted Performance
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
User Centric Evaluation of Code Generation Tools
von: Miah, Tanha, et al.
Veröffentlicht: (2024)
von: Miah, Tanha, et al.
Veröffentlicht: (2024)
Towards a Digital Twin Modeling Method for Container Terminal Port
von: Hakimi, Faouzi, et al.
Veröffentlicht: (2025)
von: Hakimi, Faouzi, et al.
Veröffentlicht: (2025)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
ReXCL: A Tool for Requirement Document Extraction and Classification
von: Bhattacharya, Paheli, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Paheli, et al.
Veröffentlicht: (2025)
AISysRev -- LLM-based Tool for Title-abstract Screening
von: Huotala, Aleksi, et al.
Veröffentlicht: (2025)
von: Huotala, Aleksi, et al.
Veröffentlicht: (2025)
MCP-Zero: Active Tool Discovery for Autonomous LLM Agents
von: Fei, Xiang, et al.
Veröffentlicht: (2025)
von: Fei, Xiang, et al.
Veröffentlicht: (2025)
ASA: Training-Free Representation Engineering for Tool-Calling Agents
von: Wang, Youjin, et al.
Veröffentlicht: (2026)
von: Wang, Youjin, et al.
Veröffentlicht: (2026)
Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents
von: Kovács, Ádám
Veröffentlicht: (2026)
von: Kovács, Ádám
Veröffentlicht: (2026)
Applying an Agentic Coding Tool for Improving Published Algorithm Implementations
von: Suwannik, Worasait
Veröffentlicht: (2026)
von: Suwannik, Worasait
Veröffentlicht: (2026)
Revisiting Software Engineering Education in the Era of Large Language Models: A Curriculum Adaptation and Academic Integrity Framework
von: Degerli, Mustafa
Veröffentlicht: (2026)
von: Degerli, Mustafa
Veröffentlicht: (2026)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
von: Tsay, Jason, et al.
Veröffentlicht: (2025)
von: Tsay, Jason, et al.
Veröffentlicht: (2025)
Solver-Aided Verification of Policy Compliance in Tool-Augmented LLM Agents
von: Winston, Cailin, et al.
Veröffentlicht: (2026)
von: Winston, Cailin, et al.
Veröffentlicht: (2026)
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
von: Chen, Aili, et al.
Veröffentlicht: (2026)
von: Chen, Aili, et al.
Veröffentlicht: (2026)
Trajectory Supervision for Continual Tool-Use Learning in LLMs
von: Reddy, Vishnu Vardhan, et al.
Veröffentlicht: (2026)
von: Reddy, Vishnu Vardhan, et al.
Veröffentlicht: (2026)
Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs: A Systematic Evaluation
von: Zhang, Weixing, et al.
Veröffentlicht: (2026)
von: Zhang, Weixing, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Arabic Multimodal Machine Learning: Datasets, Applications, Approaches, and Challenges
von: Haouhat, Abdelhamid, et al.
Veröffentlicht: (2025) -
Semantic Tool Discovery for Large Language Models: A Vector-Based Approach to MCP Tool Selection
von: Mudunuri, Sarat, et al.
Veröffentlicht: (2026) -
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
von: Sherifi, Betim, et al.
Veröffentlicht: (2024) -
Tool-integrated Reinforcement Learning for Repo Deep Search
von: Ma, Zexiong, et al.
Veröffentlicht: (2025) -
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)