Linguistic and Argument Diversity in Synthetic Data for Function-Calling Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Greenstein, Dan, Karnin, Zohar, Amiraz, Chen, Somekh, Oren |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Cross-Lingual Cost: Retrieval Biases in RAG over Arabic-English Corpora
von: Amiraz, Chen, et al.
Veröffentlicht: (2025)
von: Amiraz, Chen, et al.
Veröffentlicht: (2025)
The Distracting Effect: Understanding Irrelevant Passages in RAG
von: Amiraz, Chen, et al.
Veröffentlicht: (2025)
von: Amiraz, Chen, et al.
Veröffentlicht: (2025)
Visual Editing with LLM-based Tool Chaining: An Efficient Distillation Approach for Real-Time Applications
von: Sultan, Oren, et al.
Veröffentlicht: (2024)
von: Sultan, Oren, et al.
Veröffentlicht: (2024)
GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling
von: Xu, Hao-Xiang, et al.
Veröffentlicht: (2026)
von: Xu, Hao-Xiang, et al.
Veröffentlicht: (2026)
An Empirical Analysis of Diversity in Argument Summarization
von: van der Meer, Michiel, et al.
Veröffentlicht: (2024)
von: van der Meer, Michiel, et al.
Veröffentlicht: (2024)
Debate-to-Write: A Persona-Driven Multi-Agent Framework for Diverse Argument Generation
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
PILOT: Steering Synthetic Data Generation with Psychological & Linguistic Output Targeting
von: Cisar, Caitlin, et al.
Veröffentlicht: (2025)
von: Cisar, Caitlin, et al.
Veröffentlicht: (2025)
CLEAR: A Comprehensive Linguistic Evaluation of Argument Rewriting by Large Language Models
von: Huber, Thomas, et al.
Veröffentlicht: (2025)
von: Huber, Thomas, et al.
Veröffentlicht: (2025)
EigenData: A Self-Evolving Multi-Agent Platform for Function-Calling Data Synthesis, Auditing, and Repair
von: Chen, Jiaao, et al.
Veröffentlicht: (2026)
von: Chen, Jiaao, et al.
Veröffentlicht: (2026)
Quality Matters: Evaluating Synthetic Data for Tool-Using LLMs
von: Iskander, Shadi, et al.
Veröffentlicht: (2024)
von: Iskander, Shadi, et al.
Veröffentlicht: (2024)
Measuring Diversity in Synthetic Datasets
von: Zhu, Yuchang, et al.
Veröffentlicht: (2025)
von: Zhu, Yuchang, et al.
Veröffentlicht: (2025)
Asynchronous LLM Function Calling
von: Gim, In, et al.
Veröffentlicht: (2024)
von: Gim, In, et al.
Veröffentlicht: (2024)
Repurposing Synthetic Data for Fine-grained Search Agent Supervision
von: Zhao, Yida, et al.
Veröffentlicht: (2025)
von: Zhao, Yida, et al.
Veröffentlicht: (2025)
Multi-Agent Dialectical Refinement for Enhanced Argument Classification
von: Bąba, Jakub, et al.
Veröffentlicht: (2026)
von: Bąba, Jakub, et al.
Veröffentlicht: (2026)
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
Splits! Flexible Sociocultural Linguistic Investigation at Scale
von: Caplan, Eylon, et al.
Veröffentlicht: (2025)
von: Caplan, Eylon, et al.
Veröffentlicht: (2025)
Limited Linguistic Diversity in Embodied AI Datasets
von: Wanna, Selma, et al.
Veröffentlicht: (2026)
von: Wanna, Selma, et al.
Veröffentlicht: (2026)
Understanding Enthymemes in Argument Maps: Bridging Argument Mining and Logic-based Argumentation
von: Ben-Naim, Jonathan, et al.
Veröffentlicht: (2024)
von: Ben-Naim, Jonathan, et al.
Veröffentlicht: (2024)
Agent Data Protocol: Unifying Datasets for Diverse, Effective Fine-tuning of LLM Agents
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
Palm: A Culturally Inclusive and Linguistically Diverse Dataset for Arabic LLMs
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2025)
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2025)
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana
von: Filice, Simone, et al.
Veröffentlicht: (2025)
von: Filice, Simone, et al.
Veröffentlicht: (2025)
Improving Linguistic Diversity of Large Language Models with Possibility Exploration Fine-Tuning
von: Mai, Long, et al.
Veröffentlicht: (2024)
von: Mai, Long, et al.
Veröffentlicht: (2024)
Improving Bilingual Capabilities of Language Models to Support Diverse Linguistic Practices in Education
von: Syamkumar, Anand, et al.
Veröffentlicht: (2024)
von: Syamkumar, Anand, et al.
Veröffentlicht: (2024)
LLM for Barcodes: Generating Diverse Synthetic Data for Identity Documents
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2024)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2024)
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
Diverse and Fine-Grained Instruction-Following Ability Exploration with Synthetic Data
von: Gu, Zihui, et al.
Veröffentlicht: (2024)
von: Gu, Zihui, et al.
Veröffentlicht: (2024)
End-to-End Argument Mining through Autoregressive Argumentative Structure Prediction
von: Das, Nilmadhab, et al.
Veröffentlicht: (2025)
von: Das, Nilmadhab, et al.
Veröffentlicht: (2025)
A Multi-Task Role-Playing Agent Capable of Imitating Character Linguistic Styles
von: Chen, Siyuan, et al.
Veröffentlicht: (2024)
von: Chen, Siyuan, et al.
Veröffentlicht: (2024)
Exploring and Controlling Diversity in LLM-Agent Conversation
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
Call for Rigor in Reporting Quality of Instruction Tuning Data
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
Personality Expression Across Contexts: Linguistic and Behavioral Variation in LLM Agents
von: Han, Bin, et al.
Veröffentlicht: (2026)
von: Han, Bin, et al.
Veröffentlicht: (2026)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
Democratizing LLMs for Low-Resource Languages by Leveraging their English Dominant Abilities with Linguistically-Diverse Prompts
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
Large Language Models as Zero-shot Dialogue State Tracker through Function Calling
von: Li, Zekun, et al.
Veröffentlicht: (2024)
von: Li, Zekun, et al.
Veröffentlicht: (2024)
Alopex: A Computational Framework for Enabling On-Device Function Calls with LLMs
von: Ran, Yide, et al.
Veröffentlicht: (2024)
von: Ran, Yide, et al.
Veröffentlicht: (2024)
LinguistAgent: A Reflective Multi-Model Platform for Automated Linguistic Annotation
von: Li, Bingru
Veröffentlicht: (2026)
von: Li, Bingru
Veröffentlicht: (2026)
Experiences Build Characters: The Linguistic Origins and Functional Impact of LLM Personality
von: Wang, Xi, et al.
Veröffentlicht: (2026)
von: Wang, Xi, et al.
Veröffentlicht: (2026)
From Self-Evolving Synthetic Data to Verifiable-Reward RL: Post-Training Multi-turn Interactive Tool-Using Agents
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
MeNTi: Bridging Medical Calculator and LLM Agent with Nested Tool Calling
von: Zhu, Yakun, et al.
Veröffentlicht: (2024)
von: Zhu, Yakun, et al.
Veröffentlicht: (2024)
Improving Synthetic Data Training for Contextual Biasing Models with a Keyword-Aware Cost Function
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Cross-Lingual Cost: Retrieval Biases in RAG over Arabic-English Corpora
von: Amiraz, Chen, et al.
Veröffentlicht: (2025) -
The Distracting Effect: Understanding Irrelevant Passages in RAG
von: Amiraz, Chen, et al.
Veröffentlicht: (2025) -
Visual Editing with LLM-based Tool Chaining: An Efficient Distillation Approach for Real-Time Applications
von: Sultan, Oren, et al.
Veröffentlicht: (2024) -
GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling
von: Xu, Hao-Xiang, et al.
Veröffentlicht: (2026) -
An Empirical Analysis of Diversity in Argument Summarization
von: van der Meer, Michiel, et al.
Veröffentlicht: (2024)