Live API-Bench: 2500+ Live APIs for Testing Multi-Step Tool Calling
Fuente:
arXiv
Saved in:
| Main Authors: | Elder, Benjamin, Murthi, Anupama, Kang, Jungkoo, Naik, Ankita Rajaram, Kate, Kiran, Basu, Kinjal, Contractor, Danish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025)
by: Jain, Kush, et al.
Published: (2025)
HarnessAPI: A Skill-First Framework for Unified Streaming APIs and MCP Tools
by: Jose, Edwin
Published: (2026)
by: Jose, Edwin
Published: (2026)
Code2API: A Tool for Generating Reusable APIs from Stack Overflow Code Snippets
by: Mai, Yubo, et al.
Published: (2025)
by: Mai, Yubo, et al.
Published: (2025)
LongFuncEval: Measuring the effectiveness of long context models for function calling
by: Kate, Kiran, et al.
Published: (2025)
by: Kate, Kiran, et al.
Published: (2025)
Pricing4APIs: A Rigorous Model for RESTful API Pricings
by: Fresno-Aranda, Rafael, et al.
Published: (2023)
by: Fresno-Aranda, Rafael, et al.
Published: (2023)
Semantic API Alignment: Linking High-level User Goals to APIs
by: Feldt, Robert, et al.
Published: (2024)
by: Feldt, Robert, et al.
Published: (2024)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
by: Tsay, Jason, et al.
Published: (2025)
by: Tsay, Jason, et al.
Published: (2025)
Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments
by: Crouse, Maxwell, et al.
Published: (2026)
by: Crouse, Maxwell, et al.
Published: (2026)
PandasBench: A Benchmark for the Pandas API
by: Broihier, Alex, et al.
Published: (2025)
by: Broihier, Alex, et al.
Published: (2025)
Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs
by: Lu, Yuxuan, et al.
Published: (2026)
by: Lu, Yuxuan, et al.
Published: (2026)
Do RESTful API Design Rules Have an Impact on the Understandability of Web APIs? A Web-Based Experiment with API Descriptions
by: Bogner, Justus, et al.
Published: (2023)
by: Bogner, Justus, et al.
Published: (2023)
A Framework for Testing and Adapting REST APIs as LLM Tools
by: Bandlamudi, Jayachandu, et al.
Published: (2025)
by: Bandlamudi, Jayachandu, et al.
Published: (2025)
Test Plan Generation for Live Testing of Cloud Services
by: Jebbar, Oussama, et al.
Published: (2025)
by: Jebbar, Oussama, et al.
Published: (2025)
Evaluating LLMs on Sequential API Call Through Automated Test Generation
by: Huang, Yuheng, et al.
Published: (2025)
by: Huang, Yuheng, et al.
Published: (2025)
SATORI: Static Test Oracle Generation for REST APIs
by: Alonso, Juan C., et al.
Published: (2025)
by: Alonso, Juan C., et al.
Published: (2025)
Security Testing of RESTful APIs With Test Case Mutation
by: Salva, Sebastien, et al.
Published: (2024)
by: Salva, Sebastien, et al.
Published: (2024)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025)
by: Stennett, Tyler, et al.
Published: (2025)
OASBuilder: Generating OpenAPI Specifications from Online API Documentation with Large Language Models
by: Lazar, Koren, et al.
Published: (2025)
by: Lazar, Koren, et al.
Published: (2025)
Utilizing API Response for Test Refinement
by: Sondhi, Devika, et al.
Published: (2025)
by: Sondhi, Devika, et al.
Published: (2025)
APITestGenie: Generating Web API Tests from Requirements and API Specifications with LLMs
by: Pereira, André, et al.
Published: (2026)
by: Pereira, André, et al.
Published: (2026)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
by: Agarwal, Prerna, et al.
Published: (2025)
by: Agarwal, Prerna, et al.
Published: (2025)
AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
Test Amplification for REST APIs Using "Out-of-the-box" Large Language Models
by: Bardakci, Tolgahan, et al.
Published: (2025)
by: Bardakci, Tolgahan, et al.
Published: (2025)
Test Amplification for REST APIs via Single and Multi-Agent LLM Systems
by: Nooyens, Robbe, et al.
Published: (2025)
by: Nooyens, Robbe, et al.
Published: (2025)
Evaluating the Effectiveness of Coverage-Guided Fuzzing for Testing Deep Learning Library APIs
by: Qin, Feiran, et al.
Published: (2025)
by: Qin, Feiran, et al.
Published: (2025)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
by: Huynh, Hieu, et al.
Published: (2025)
by: Huynh, Hieu, et al.
Published: (2025)
Generating REST API Tests With Descriptive Names
by: Garrett, Philip, et al.
Published: (2025)
by: Garrett, Philip, et al.
Published: (2025)
PrediQL: Automated Testing of GraphQL APIs with LLMs
by: Liu, Shaolun, et al.
Published: (2025)
by: Liu, Shaolun, et al.
Published: (2025)
FastFI: Enhancing API Call-Site Robustness in Microservice-Based Systems with Fault Injection
by: Tan, Yuzhen, et al.
Published: (2026)
by: Tan, Yuzhen, et al.
Published: (2026)
R+R: Reassessing Java Security API Misuse in Current LLMs: A Replication on JCA and JSSE APIs with External Security Knowledge
by: Lu, Tianhe, et al.
Published: (2026)
by: Lu, Tianhe, et al.
Published: (2026)
They Call Her 'Miss' and Him 'Professor': Lived Experiences of Women Teaching Support Staff in IT/SE Education
by: Malhotra, Vasudha, et al.
Published: (2026)
by: Malhotra, Vasudha, et al.
Published: (2026)
Developer Perspectives on REST API Usability: A Study of REST API Guidelines
by: Peldszus, Sven, et al.
Published: (2026)
by: Peldszus, Sven, et al.
Published: (2026)
Client--Library Compatibility Testing with API Interaction Snapshots
by: Monce, Gustave, et al.
Published: (2025)
by: Monce, Gustave, et al.
Published: (2025)
Log-based, Business-aware REST API Testing
by: Yang, Ding, et al.
Published: (2026)
by: Yang, Ding, et al.
Published: (2026)
Automated Testing of Broken Authentication Vulnerabilities in Web APIs with AuthREST
by: Corradini, Davide, et al.
Published: (2025)
by: Corradini, Davide, et al.
Published: (2025)
Schema First Tool APIs for LLM Agents: A Controlled Study of Tool Misuse, Recovery, and Budgeted Performance
by: Sigdel, Akshey, et al.
Published: (2026)
by: Sigdel, Akshey, et al.
Published: (2026)
RESTestBench: A Benchmark for Evaluating the Effectiveness of LLM-Generated REST API Test Cases from NL Requirements
by: Kogler, Leon, et al.
Published: (2026)
by: Kogler, Leon, et al.
Published: (2026)
SC2Tools: StarCraft II Toolset and Dataset API
by: Białecki, Andrzej, et al.
Published: (2025)
by: Białecki, Andrzej, et al.
Published: (2025)
DeepREST: Automated Test Case Generation for REST APIs Exploiting Deep Reinforcement Learning
by: Corradini, Davide, et al.
Published: (2024)
by: Corradini, Davide, et al.
Published: (2024)
What's in a Package? Getting Visibility Into Dependencies Using Security-Sensitive API Calls
by: Rahman, Imranur, et al.
Published: (2024)
by: Rahman, Imranur, et al.
Published: (2024)
Similar Items
-
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025) -
HarnessAPI: A Skill-First Framework for Unified Streaming APIs and MCP Tools
by: Jose, Edwin
Published: (2026) -
Code2API: A Tool for Generating Reusable APIs from Stack Overflow Code Snippets
by: Mai, Yubo, et al.
Published: (2025) -
LongFuncEval: Measuring the effectiveness of long context models for function calling
by: Kate, Kiran, et al.
Published: (2025) -
Pricing4APIs: A Rigorous Model for RESTful API Pricings
by: Fresno-Aranda, Rafael, et al.
Published: (2023)