A Framework for Testing and Adapting REST APIs as LLM Tools
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bandlamudi, Jayachandu, Chaudhuri, Ritwik, Gantayat, Neelamadhav, Ghosh, Sambit, Mukherjee, Kushal, Agarwal, Prerna, Sindhgatta, Renuka, Mehta, Sameep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
Framework to enable and test conversational assistant for APIs and RPAs
von: Jayachandu Bandlamudi, et al.
Veröffentlicht: (2024)
von: Jayachandu Bandlamudi, et al.
Veröffentlicht: (2024)
Mind the GAP: Text Safety Does Not Transfer to Tool-Call Safety in LLM Agents
von: Cartagena, Arnold, et al.
Veröffentlicht: (2026)
von: Cartagena, Arnold, et al.
Veröffentlicht: (2026)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026)
von: Souza, Débora, et al.
Veröffentlicht: (2026)
AdaDec: A Uncertainty-Guided Lookahead Decoding Framework for LLM-Based Code Generation
von: He, Kaifeng, et al.
Veröffentlicht: (2025)
von: He, Kaifeng, et al.
Veröffentlicht: (2025)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Comparative Analysis of LLM Abliteration Methods: A Cross-Architecture Evaluation
von: Young, Richard J.
Veröffentlicht: (2025)
von: Young, Richard J.
Veröffentlicht: (2025)
ContractBench: Can LLM Agents Preserve Observation Contracts?
von: Wang, Jicheng, et al.
Veröffentlicht: (2026)
von: Wang, Jicheng, et al.
Veröffentlicht: (2026)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
von: Kohl, Jens, et al.
Veröffentlicht: (2024)
von: Kohl, Jens, et al.
Veröffentlicht: (2024)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
von: More, Riddhi, et al.
Veröffentlicht: (2025)
von: More, Riddhi, et al.
Veröffentlicht: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
SATORI: Static Test Oracle Generation for REST APIs
von: Alonso, Juan C., et al.
Veröffentlicht: (2025)
von: Alonso, Juan C., et al.
Veröffentlicht: (2025)
Test Amplification for REST APIs via Single and Multi-Agent LLM Systems
von: Nooyens, Robbe, et al.
Veröffentlicht: (2025)
von: Nooyens, Robbe, et al.
Veröffentlicht: (2025)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
RepoLaunch: Automating Build&Test Pipeline of Code Repositories on ANY Language and ANY Platform
von: Li, Kenan, et al.
Veröffentlicht: (2026)
von: Li, Kenan, et al.
Veröffentlicht: (2026)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
von: Kessel, Marcus
Veröffentlicht: (2024)
von: Kessel, Marcus
Veröffentlicht: (2024)
DeepREST: Automated Test Case Generation for REST APIs Exploiting Deep Reinforcement Learning
von: Corradini, Davide, et al.
Veröffentlicht: (2024)
von: Corradini, Davide, et al.
Veröffentlicht: (2024)
Benchmarking Energy Efficiency of Large Language Models Using vLLM
von: Pronk, K., et al.
Veröffentlicht: (2025)
von: Pronk, K., et al.
Veröffentlicht: (2025)
Can AI Assist in Olympiad Coding
von: Ren, Samuel
Veröffentlicht: (2025)
von: Ren, Samuel
Veröffentlicht: (2025)
ReDel: A Toolkit for LLM-Powered Recursive Multi-Agent Systems
von: Zhu, Andrew, et al.
Veröffentlicht: (2024)
von: Zhu, Andrew, et al.
Veröffentlicht: (2024)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
Learning Software Bug Reports: A Systematic Literature Review
von: Long, Guoming, et al.
Veröffentlicht: (2025)
von: Long, Guoming, et al.
Veröffentlicht: (2025)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
von: Bradbury, Jeremy S., et al.
Veröffentlicht: (2024)
von: Bradbury, Jeremy S., et al.
Veröffentlicht: (2024)
Enhancing LLM Code Generation Capabilities through Test-Driven Development and Code Interpreter
von: Jalil, Sajed, et al.
Veröffentlicht: (2025)
von: Jalil, Sajed, et al.
Veröffentlicht: (2025)
Achieving Tool Calling Functionality in LLMs Using Only Prompt Engineering Without Fine-Tuning
von: He, Shengtao
Veröffentlicht: (2024)
von: He, Shengtao
Veröffentlicht: (2024)
Engineering A Large Language Model From Scratch
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
von: Terragni, Valerio
Veröffentlicht: (2026)
von: Terragni, Valerio
Veröffentlicht: (2026)
The Appeal and Reality of Recycling LoRAs with Adaptive Merging
von: Liu, Haokun, et al.
Veröffentlicht: (2026)
von: Liu, Haokun, et al.
Veröffentlicht: (2026)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research
von: Savenkov, Vladislav
Veröffentlicht: (2026)
von: Savenkov, Vladislav
Veröffentlicht: (2026)
FREYR: A Framework for Recognizing and Executing Your Requests
von: Gallotta, Roberto, et al.
Veröffentlicht: (2025)
von: Gallotta, Roberto, et al.
Veröffentlicht: (2025)
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
von: More, Riddhi, et al.
Veröffentlicht: (2025)
von: More, Riddhi, et al.
Veröffentlicht: (2025)
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2025)
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2025)
Predicting 3D Rigid Body Dynamics with Deep Residual Network
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
Ähnliche Einträge
-
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025) -
Framework to enable and test conversational assistant for APIs and RPAs
von: Jayachandu Bandlamudi, et al.
Veröffentlicht: (2024) -
Mind the GAP: Text Safety Does Not Transfer to Tool-Call Safety in LLM Agents
von: Cartagena, Arnold, et al.
Veröffentlicht: (2026) -
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026) -
AdaDec: A Uncertainty-Guided Lookahead Decoding Framework for LLM-Based Code Generation
von: He, Kaifeng, et al.
Veröffentlicht: (2025)