Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Jonggeun, Song, Woojung, Han, Jongwook, Pyun, Haesung, Jo, Yohan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Quantifying Data Contamination in Psychometric Evaluations of LLMs
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
Bridging the Knowledge-Prediction Gap in LLMs on Multiple-Choice Questions
von: Park, Yoonah, et al.
Veröffentlicht: (2025)
von: Park, Yoonah, et al.
Veröffentlicht: (2025)
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
Improving Dialogue State Tracking through Combinatorial Search for In-Context Examples
von: Pyun, Haesung, et al.
Veröffentlicht: (2025)
von: Pyun, Haesung, et al.
Veröffentlicht: (2025)
SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?
von: Lee, Jonggeun, et al.
Veröffentlicht: (2026)
von: Lee, Jonggeun, et al.
Veröffentlicht: (2026)
Non-Collaborative User Simulators for Tool Agents
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
Human Psychometric Questionnaires Mischaracterize LLM Behavior
von: Song, Woojung, et al.
Veröffentlicht: (2025)
von: Song, Woojung, et al.
Veröffentlicht: (2025)
SpokenUS: A Spoken User Simulator for Task-Oriented Dialogue
von: Lee, Jonggeun, et al.
Veröffentlicht: (2026)
von: Lee, Jonggeun, et al.
Veröffentlicht: (2026)
ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
von: Lim, Sungjib, et al.
Veröffentlicht: (2025)
von: Lim, Sungjib, et al.
Veröffentlicht: (2025)
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
In-N-Out: A Parameter-Level API Graph Dataset for Tool Agents
von: Lee, Seungkyu, et al.
Veröffentlicht: (2025)
von: Lee, Seungkyu, et al.
Veröffentlicht: (2025)
SimuHome: A Temporal- and Environment-Aware Benchmark for Smart Home LLM Agents
von: Seo, Gyuhyeon, et al.
Veröffentlicht: (2025)
von: Seo, Gyuhyeon, et al.
Veröffentlicht: (2025)
Language Models Don't Learn the Physical Manifestation of Language
von: Lee, Bruce W., et al.
Veröffentlicht: (2024)
von: Lee, Bruce W., et al.
Veröffentlicht: (2024)
To Adapt or not to Adapt, Rethinking the Value of Medical Knowledge-Aware Large Language Models
von: Domingo-Aldama, Ane G., et al.
Veröffentlicht: (2026)
von: Domingo-Aldama, Ane G., et al.
Veröffentlicht: (2026)
Domain-Adapted Small Language Models for Reliable Clinical Triage
von: Aljohani, Manar, et al.
Veröffentlicht: (2026)
von: Aljohani, Manar, et al.
Veröffentlicht: (2026)
ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding
von: Xia, Heming, et al.
Veröffentlicht: (2026)
von: Xia, Heming, et al.
Veröffentlicht: (2026)
AdaptMI: Adaptive Skill-based In-context Math Instruction for Small Language Models
von: He, Yinghui, et al.
Veröffentlicht: (2025)
von: He, Yinghui, et al.
Veröffentlicht: (2025)
Think, But Don't Overthink: Reproducing Recursive Language Models
von: Wang, Daren
Veröffentlicht: (2026)
von: Wang, Daren
Veröffentlicht: (2026)
Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacement
von: Kong, Injin, et al.
Veröffentlicht: (2026)
von: Kong, Injin, et al.
Veröffentlicht: (2026)
PARM: Pipeline-Adapted Reward Model
von: Fan, Xingyu, et al.
Veröffentlicht: (2026)
von: Fan, Xingyu, et al.
Veröffentlicht: (2026)
Adapt-Pruner: Adaptive Structural Pruning for Efficient Small Language Model Training
von: Pan, Rui, et al.
Veröffentlicht: (2025)
von: Pan, Rui, et al.
Veröffentlicht: (2025)
Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models
von: Kong, Injin, et al.
Veröffentlicht: (2026)
von: Kong, Injin, et al.
Veröffentlicht: (2026)
Adapting Small Language Models to Low-Resource Domains: A Case Study in Hindi Tourism QA
von: Majhi, Sandipan, et al.
Veröffentlicht: (2025)
von: Majhi, Sandipan, et al.
Veröffentlicht: (2025)
Family Matters: Language Transfer and Merging for Adapting Small LLMs to Faroese
von: Kunz, Jenny, et al.
Veröffentlicht: (2025)
von: Kunz, Jenny, et al.
Veröffentlicht: (2025)
Don't Fine-Tune, Decode: Syntax Error-Free Tool Use via Constrained Decoding
von: Zhang, Kexun, et al.
Veröffentlicht: (2023)
von: Zhang, Kexun, et al.
Veröffentlicht: (2023)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
von: Parmar, Jupinder, et al.
Veröffentlicht: (2024)
von: Parmar, Jupinder, et al.
Veröffentlicht: (2024)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
Don't Throw Away Your Pretrained Model
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models
von: Kumar, Sachin
Veröffentlicht: (2026)
von: Kumar, Sachin
Veröffentlicht: (2026)
LMFusion: Adapting Pretrained Language Models for Multimodal Generation
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
Adapting Large Language Models to Domains via Reading Comprehension
von: Cheng, Daixuan, et al.
Veröffentlicht: (2023)
von: Cheng, Daixuan, et al.
Veröffentlicht: (2023)
Adapting Large Language Models for Document-Level Machine Translation
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
von: Wu, Minghao, et al.
Veröffentlicht: (2024)
RoQLlama: A Lightweight Romanian Adapted Language Model
von: Dima, George-Andrei, et al.
Veröffentlicht: (2024)
von: Dima, George-Andrei, et al.
Veröffentlicht: (2024)
NesTools: A Dataset for Evaluating Nested Tool Learning Abilities of Large Language Models
von: Han, Han, et al.
Veröffentlicht: (2024)
von: Han, Han, et al.
Veröffentlicht: (2024)
Tell, Don't Show: Leveraging Language Models' Abstractive Retellings to Model Literary Themes
von: Lucy, Li, et al.
Veröffentlicht: (2025)
von: Lucy, Li, et al.
Veröffentlicht: (2025)
RAFT: Adapting Language Model to Domain Specific RAG
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
Schema as Parameterized Tools for Universal Information Extraction
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
FastAdaSP: Multitask-Adapted Efficient Inference for Large Speech Language Model
von: Lu, Yichen, et al.
Veröffentlicht: (2024)
von: Lu, Yichen, et al.
Veröffentlicht: (2024)
Adapting Language Models via Token Translation
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Quantifying Data Contamination in Psychometric Evaluations of LLMs
von: Han, Jongwook, et al.
Veröffentlicht: (2025) -
Bridging the Knowledge-Prediction Gap in LLMs on Multiple-Choice Questions
von: Park, Yoonah, et al.
Veröffentlicht: (2025) -
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
von: Han, Jongwook, et al.
Veröffentlicht: (2025) -
Improving Dialogue State Tracking through Combinatorial Search for In-Context Examples
von: Pyun, Haesung, et al.
Veröffentlicht: (2025) -
SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?
von: Lee, Jonggeun, et al.
Veröffentlicht: (2026)