API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Basu, Kinjal, Abdelaziz, Ibrahim, Chaudhury, Subhajit, Dan, Soham, Crouse, Maxwell, Munawar, Asim, Kumaravel, Sadhana, Muthusamy, Vinod, Kapanipathi, Pavan, Lastras, Luis A. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Formally Specifying the High-Level Behavior of LLM-Based Agents
by: Crouse, Maxwell, et al.
Published: (2023)
by: Crouse, Maxwell, et al.
Published: (2023)
Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments
by: Crouse, Maxwell, et al.
Published: (2026)
by: Crouse, Maxwell, et al.
Published: (2026)
ToolRM: Outcome Reward Models for Tool-Calling Large Language Models
by: Agarwal, Mayank, et al.
Published: (2025)
by: Agarwal, Mayank, et al.
Published: (2025)
R2D2: Remembering, Replaying and Dynamic Decision Making with a Reflective Agentic Memory
by: Huang, Tenghao, et al.
Published: (2025)
by: Huang, Tenghao, et al.
Published: (2025)
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks
by: Abdelaziz, Ibrahim, et al.
Published: (2024)
by: Abdelaziz, Ibrahim, et al.
Published: (2024)
LongFuncEval: Measuring the effectiveness of long context models for function calling
by: Kate, Kiran, et al.
Published: (2025)
by: Kate, Kiran, et al.
Published: (2025)
Putting It All into Context: Simplifying Agents with LCLMs
by: Jiang, Mingjian, et al.
Published: (2025)
by: Jiang, Mingjian, et al.
Published: (2025)
On the Effects of Fine-tuning Language Models for Text-Based Reinforcement Learning
by: Gruppi, Mauricio, et al.
Published: (2024)
by: Gruppi, Mauricio, et al.
Published: (2024)
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Live API-Bench: 2500+ Live APIs for Testing Multi-Step Tool Calling
by: Elder, Benjamin, et al.
Published: (2025)
by: Elder, Benjamin, et al.
Published: (2025)
Needle in the Haystack for Memory Based Large Language Models
by: Nelson, Elliot, et al.
Published: (2024)
by: Nelson, Elliot, et al.
Published: (2024)
OpenAI for OpenAPI: Automated generation of REST API specification via LLMs
by: Chen, Hao, et al.
Published: (2026)
by: Chen, Hao, et al.
Published: (2026)
APITestGenie: Generating Web API Tests from Requirements and API Specifications with LLMs
by: Pereira, André, et al.
Published: (2026)
by: Pereira, André, et al.
Published: (2026)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
by: Huynh, Hieu, et al.
Published: (2025)
by: Huynh, Hieu, et al.
Published: (2025)
API Security Based on Automatic OpenAPI Mapping
by: Levi, Yarin, et al.
Published: (2026)
by: Levi, Yarin, et al.
Published: (2026)
SEAL: Suite for Evaluating API-use of LLMs
by: Kim, Woojeong, et al.
Published: (2024)
by: Kim, Woojeong, et al.
Published: (2024)
PandasBench: A Benchmark for the Pandas API
by: Broihier, Alex, et al.
Published: (2025)
by: Broihier, Alex, et al.
Published: (2025)
Transparent API
by: Luis Garreta
Published: (2008)
by: Luis Garreta
Published: (2008)
IOLBENCH: Benchmarking LLMs on Linguistic Reasoning
by: Goyal, Satyam, et al.
Published: (2025)
by: Goyal, Satyam, et al.
Published: (2025)
Compositional Program Generation for Few-Shot Systematic Generalization
by: Klinger, Tim, et al.
Published: (2023)
by: Klinger, Tim, et al.
Published: (2023)
$μ$Ed API: Towards a Shared API for Education Microservices
by: Sölch, Maximillan, et al.
Published: (2026)
by: Sölch, Maximillan, et al.
Published: (2026)
Generating API Parameter Security Rules with LLM for API Misuse Detection
by: Liu, Jinghua, et al.
Published: (2024)
by: Liu, Jinghua, et al.
Published: (2024)
An Android API Recommendation Approach Based on API Dependency Paths Learning
by: Jia Deng, et al.
Published: (2025)
by: Jia Deng, et al.
Published: (2025)
Logits of API-Protected LLMs Leak Proprietary Information
by: Finlayson, Matthew, et al.
Published: (2024)
by: Finlayson, Matthew, et al.
Published: (2024)
Automating API Documentation with LLMs: A BERTopic Approach
by: Naghshzan, AmirHossein
Published: (2025)
by: Naghshzan, AmirHossein
Published: (2025)
Contextual API Completion for Unseen Repositories Using LLMs
by: Nashid, Noor, et al.
Published: (2024)
by: Nashid, Noor, et al.
Published: (2024)
oneAPI.jl
by: Besard, Tim
Published: (2025)
by: Besard, Tim
Published: (2025)
Developer Perspectives on REST API Usability: A Study of REST API Guidelines
by: Peldszus, Sven, et al.
Published: (2026)
by: Peldszus, Sven, et al.
Published: (2026)
Speculate: Generating REST API Specifications Using LLMs — Artifact
by: Singh, Krishanu
Published: (2026)
by: Singh, Krishanu
Published: (2026)
Lightweight Model Editing for LLMs to Correct Deprecated API Recommendations
by: Lin, Guancheng, et al.
Published: (2025)
by: Lin, Guancheng, et al.
Published: (2025)
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs
by: Dutta, Sujan, et al.
Published: (2024)
by: Dutta, Sujan, et al.
Published: (2024)
CodeUpdateArena: Benchmarking Knowledge Editing on API Updates
by: Liu, Zeyu Leo, et al.
Published: (2024)
by: Liu, Zeyu Leo, et al.
Published: (2024)
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation
by: Guo, Zhen, et al.
Published: (2024)
by: Guo, Zhen, et al.
Published: (2024)
ASAP-Repair: API-Specific Automated Program Repair Based on API Usage Graphs
by: Nielebock, Sebastian, et al.
Published: (2024)
by: Nielebock, Sebastian, et al.
Published: (2024)
DLT Service Federation API
by: Universidad Carlos III de Madrid (UC3)
Published: (2024)
by: Universidad Carlos III de Madrid (UC3)
Published: (2024)
Consistency Evaluators for the GAME API
by: Luthra, Ishika
Published: (2026)
by: Luthra, Ishika
Published: (2026)
GigaAPI for GPU Parallelization
by: Suvarna, M., et al.
Published: (2025)
by: Suvarna, M., et al.
Published: (2025)
OASBuilder: Generating OpenAPI Specifications from Online API Documentation with Large Language Models
by: Lazar, Koren, et al.
Published: (2025)
by: Lazar, Koren, et al.
Published: (2025)
Validating API Design Requirements for Interoperability: A Static Analysis Approach Using OpenAPI
by: Sundberg, Edwin, et al.
Published: (2025)
by: Sundberg, Edwin, et al.
Published: (2025)
Similar Items
-
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
by: Basu, Kinjal, et al.
Published: (2024) -
Formally Specifying the High-Level Behavior of LLM-Based Agents
by: Crouse, Maxwell, et al.
Published: (2023) -
Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments
by: Crouse, Maxwell, et al.
Published: (2026) -
ToolRM: Outcome Reward Models for Tool-Calling Large Language Models
by: Agarwal, Mayank, et al.
Published: (2025) -
R2D2: Remembering, Replaying and Dynamic Decision Making with a Reflective Agentic Memory
by: Huang, Tenghao, et al.
Published: (2025)