In-N-Out: A Parameter-Level API Graph Dataset for Tool Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Seungkyu, Kim, Nalim, Jo, Yohan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SimuHome: A Temporal- and Environment-Aware Benchmark for Smart Home LLM Agents
di: Seo, Gyuhyeon, et al.
Pubblicazione: (2025)
di: Seo, Gyuhyeon, et al.
Pubblicazione: (2025)
KMI: A Dataset of Korean Motivational Interviewing Dialogues for Psychotherapy
di: Kim, Hyunjong, et al.
Pubblicazione: (2025)
di: Kim, Hyunjong, et al.
Pubblicazione: (2025)
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
di: Lee, Yooseop, et al.
Pubblicazione: (2025)
di: Lee, Yooseop, et al.
Pubblicazione: (2025)
Thinking Like a Doctor: Conversational Diagnosis through the Exploration of Diagnostic Knowledge Graphs
di: Won, Jeongmoon, et al.
Pubblicazione: (2026)
di: Won, Jeongmoon, et al.
Pubblicazione: (2026)
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
di: Kim, Junseo, et al.
Pubblicazione: (2025)
di: Kim, Junseo, et al.
Pubblicazione: (2025)
Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacement
di: Kong, Injin, et al.
Pubblicazione: (2026)
di: Kong, Injin, et al.
Pubblicazione: (2026)
Infant Agent: A Tool-Integrated, Logic-Driven Agent with Cost-Effective API Usage
di: Lei, Bin, et al.
Pubblicazione: (2024)
di: Lei, Bin, et al.
Pubblicazione: (2024)
PKG API: A Tool for Personal Knowledge Graph Management
di: Bernard, Nolwenn, et al.
Pubblicazione: (2024)
di: Bernard, Nolwenn, et al.
Pubblicazione: (2024)
Context-Robust Knowledge Editing for Language Models
di: Park, Haewon, et al.
Pubblicazione: (2025)
di: Park, Haewon, et al.
Pubblicazione: (2025)
Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
di: Chae, Kyubyung, et al.
Pubblicazione: (2024)
di: Chae, Kyubyung, et al.
Pubblicazione: (2024)
Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models
di: Kong, Injin, et al.
Pubblicazione: (2026)
di: Kong, Injin, et al.
Pubblicazione: (2026)
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
di: Lim, Sungjib, et al.
Pubblicazione: (2025)
di: Lim, Sungjib, et al.
Pubblicazione: (2025)
Model-based Preference Optimization in Abstractive Summarization without Human Feedback
di: Choi, Jaepill, et al.
Pubblicazione: (2024)
di: Choi, Jaepill, et al.
Pubblicazione: (2024)
ThinkBrake: Efficient Reasoning via Log-Probability Margin Guided Decoding
di: Song, Sangjun, et al.
Pubblicazione: (2025)
di: Song, Sangjun, et al.
Pubblicazione: (2025)
SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration
di: Ding, Keyan, et al.
Pubblicazione: (2025)
di: Ding, Keyan, et al.
Pubblicazione: (2025)
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation
di: Guo, Zhen, et al.
Pubblicazione: (2024)
di: Guo, Zhen, et al.
Pubblicazione: (2024)
DialSim: A Dialogue Simulator for Evaluating Long-Term Multi-Party Dialogue Understanding of Conversational Agents
di: Kim, Jiho, et al.
Pubblicazione: (2024)
di: Kim, Jiho, et al.
Pubblicazione: (2024)
Ever-Evolving Memory by Blending and Refining the Past
di: Kim, Seo Hyun, et al.
Pubblicazione: (2024)
di: Kim, Seo Hyun, et al.
Pubblicazione: (2024)
R2-KG: General-Purpose Dual-Agent Framework for Reliable Reasoning on Knowledge Graphs
di: Jo, Sumin, et al.
Pubblicazione: (2025)
di: Jo, Sumin, et al.
Pubblicazione: (2025)
Learning to Retrieve User History and Generate User Profiles for Personalized Persuasiveness Prediction
di: Park, Sejun, et al.
Pubblicazione: (2026)
di: Park, Sejun, et al.
Pubblicazione: (2026)
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
di: Kim, Wonjoong, et al.
Pubblicazione: (2025)
di: Kim, Wonjoong, et al.
Pubblicazione: (2025)
Pre-Storage Reasoning for Episodic Memory: Shifting Inference Burden to Memory for Personalized Dialogue
di: Kim, Sangyeop, et al.
Pubblicazione: (2025)
di: Kim, Sangyeop, et al.
Pubblicazione: (2025)
StableToolBench-MirrorAPI: Modeling Tool Environments as Mirrors of 7,000+ Real-World APIs
di: Guo, Zhicheng, et al.
Pubblicazione: (2025)
di: Guo, Zhicheng, et al.
Pubblicazione: (2025)
Improving Dialogue State Tracking through Combinatorial Search for In-Context Examples
di: Pyun, Haesung, et al.
Pubblicazione: (2025)
di: Pyun, Haesung, et al.
Pubblicazione: (2025)
API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs
di: Basu, Kinjal, et al.
Pubblicazione: (2024)
di: Basu, Kinjal, et al.
Pubblicazione: (2024)
Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API Complexity
di: Kim, Doyoung, et al.
Pubblicazione: (2026)
di: Kim, Doyoung, et al.
Pubblicazione: (2026)
Human Psychometric Questionnaires Mischaracterize LLM Behavior
di: Song, Woojung, et al.
Pubblicazione: (2025)
di: Song, Woojung, et al.
Pubblicazione: (2025)
Deterministic Legal Agents: A Canonical Primitive API for Auditable Reasoning over Temporal Knowledge Graphs
di: de Martim, Hudson
Pubblicazione: (2025)
di: de Martim, Hudson
Pubblicazione: (2025)
Dialogue Systems for Emotional Support via Value Reinforcement
di: Kim, Juhee, et al.
Pubblicazione: (2025)
di: Kim, Juhee, et al.
Pubblicazione: (2025)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
di: Kim, Zae Myung, et al.
Pubblicazione: (2026)
di: Kim, Zae Myung, et al.
Pubblicazione: (2026)
DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
GAP: Graph-Based Agent Planning with Parallel Tool Use and Reinforcement Learning
di: Wu, Jiaqi, et al.
Pubblicazione: (2025)
di: Wu, Jiaqi, et al.
Pubblicazione: (2025)
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
di: Qiao, Yuxuan, et al.
Pubblicazione: (2025)
di: Qiao, Yuxuan, et al.
Pubblicazione: (2025)
Ontology-Free General-Domain Knowledge Graph-to-Text Generation Dataset Synthesis using Large Language Model
di: Kim, Daehee, et al.
Pubblicazione: (2024)
di: Kim, Daehee, et al.
Pubblicazione: (2024)
Distilling LLM Agent into Small Models with Retrieval and Code Tools
di: Kang, Minki, et al.
Pubblicazione: (2025)
di: Kang, Minki, et al.
Pubblicazione: (2025)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
di: Ni, Xinyi, et al.
Pubblicazione: (2025)
di: Ni, Xinyi, et al.
Pubblicazione: (2025)
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
di: Han, Jongwook, et al.
Pubblicazione: (2025)
di: Han, Jongwook, et al.
Pubblicazione: (2025)
ToolBridge: An Open-Source Dataset to Equip LLMs with External Tool Capabilities
di: Jin, Zhenchao, et al.
Pubblicazione: (2024)
di: Jin, Zhenchao, et al.
Pubblicazione: (2024)
GTA: A Benchmark for General Tool Agents
di: Wang, Jize, et al.
Pubblicazione: (2024)
di: Wang, Jize, et al.
Pubblicazione: (2024)
LLM-C3MOD: A Human-LLM Collaborative System for Cross-Cultural Hate Speech Moderation
di: Park, Junyeong, et al.
Pubblicazione: (2025)
di: Park, Junyeong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SimuHome: A Temporal- and Environment-Aware Benchmark for Smart Home LLM Agents
di: Seo, Gyuhyeon, et al.
Pubblicazione: (2025) -
KMI: A Dataset of Korean Motivational Interviewing Dialogues for Psychotherapy
di: Kim, Hyunjong, et al.
Pubblicazione: (2025) -
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
di: Lee, Yooseop, et al.
Pubblicazione: (2025) -
Thinking Like a Doctor: Conversational Diagnosis through the Exploration of Diagnostic Knowledge Graphs
di: Won, Jeongmoon, et al.
Pubblicazione: (2026) -
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
di: Kim, Junseo, et al.
Pubblicazione: (2025)