On the Robustness of Agentic Function Calling
Fuente:
arXiv
Salvato in:
| Autori principali: | Rabinovich, Ella, Anaby-Tavor, Ateret |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Near-Miss: Latent Policy Failure Detection in Agentic Workflows
di: Rabinovich, Ella, et al.
Pubblicazione: (2026)
di: Rabinovich, Ella, et al.
Pubblicazione: (2026)
A Novel Metric for Measuring the Robustness of Large Language Models in Non-adversarial Scenarios
di: Ackerman, Samuel, et al.
Pubblicazione: (2024)
di: Ackerman, Samuel, et al.
Pubblicazione: (2024)
Towards Enforcing Company Policy Adherence in Agentic Workflows
di: Zwerdling, Naama, et al.
Pubblicazione: (2025)
di: Zwerdling, Naama, et al.
Pubblicazione: (2025)
What's the Plan? Evaluating and Developing Planning-Aware Techniques for Language Models
di: Hirsch, Eran, et al.
Pubblicazione: (2024)
di: Hirsch, Eran, et al.
Pubblicazione: (2024)
Think Again! The Effect of Test-Time Compute on Preferences, Opinions, and Beliefs of Large Language Models
di: Kour, George, et al.
Pubblicazione: (2025)
di: Kour, George, et al.
Pubblicazione: (2025)
From Zero to Hero: Cold-Start Anomaly Detection
di: Reiss, Tal, et al.
Pubblicazione: (2024)
di: Reiss, Tal, et al.
Pubblicazione: (2024)
CRISP: Complex Reasoning with Interpretable Step-based Plans
di: Vetzler, Matan, et al.
Pubblicazione: (2025)
di: Vetzler, Matan, et al.
Pubblicazione: (2025)
Exploring Straightforward Conversational Red-Teaming
di: Kour, George, et al.
Pubblicazione: (2024)
di: Kour, George, et al.
Pubblicazione: (2024)
SpeCrawler: Generating OpenAPI Specifications from API Documentation Using Large Language Models
di: Lazar, Koren, et al.
Pubblicazione: (2024)
di: Lazar, Koren, et al.
Pubblicazione: (2024)
Effective Red-Teaming of Policy-Adherent Agents
di: Nakash, Itay, et al.
Pubblicazione: (2025)
di: Nakash, Itay, et al.
Pubblicazione: (2025)
Efficient Agent Evaluation via Diversity-Guided User Simulation
di: Nakash, Itay, et al.
Pubblicazione: (2026)
di: Nakash, Itay, et al.
Pubblicazione: (2026)
That's Optional: A Contemporary Exploration of "that" Omission in English Subordinate Clauses
di: Rabinovich, Ella
Pubblicazione: (2024)
di: Rabinovich, Ella
Pubblicazione: (2024)
Breaking ReAct Agents: Foot-in-the-Door Attack Will Get You In
di: Nakash, Itay, et al.
Pubblicazione: (2024)
di: Nakash, Itay, et al.
Pubblicazione: (2024)
Who are you, ChatGPT? Personality and Demographic Style in LLM-Generated Content
di: Porat, Dana Sotto, et al.
Pubblicazione: (2025)
di: Porat, Dana Sotto, et al.
Pubblicazione: (2025)
On the Interplay between Musical Preferences and Personality through the Lens of Language
di: Shem-Tov, Eliran, et al.
Pubblicazione: (2025)
di: Shem-Tov, Eliran, et al.
Pubblicazione: (2025)
Unveiling Affective Polarization Trends in Parliamentary Proceedings
di: Goldin, Gili, et al.
Pubblicazione: (2025)
di: Goldin, Gili, et al.
Pubblicazione: (2025)
An Annotation Scheme for Factuality and its Application to Parliamentary Proceedings
di: Goldin, Gili, et al.
Pubblicazione: (2025)
di: Goldin, Gili, et al.
Pubblicazione: (2025)
Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models
di: Belkhiter, Yannis, et al.
Pubblicazione: (2026)
di: Belkhiter, Yannis, et al.
Pubblicazione: (2026)
Rethinking Selective Knowledge Distillation
di: Tavor, Almog, et al.
Pubblicazione: (2026)
di: Tavor, Almog, et al.
Pubblicazione: (2026)
Automatic Extraction of Disease Risk Factors from Medical Publications
di: Rubchinsky, Maxim, et al.
Pubblicazione: (2024)
di: Rubchinsky, Maxim, et al.
Pubblicazione: (2024)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
di: Goldin, Gili, et al.
Pubblicazione: (2024)
di: Goldin, Gili, et al.
Pubblicazione: (2024)
An LLM Compiler for Parallel Function Calling
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
Uncertainty Quantification for LLM Function-Calling
di: Ye, Zihuiwen, et al.
Pubblicazione: (2026)
di: Ye, Zihuiwen, et al.
Pubblicazione: (2026)
Asynchronous LLM Function Calling
di: Gim, In, et al.
Pubblicazione: (2024)
di: Gim, In, et al.
Pubblicazione: (2024)
Reasoning through Exploration: A Reinforcement Learning Framework for Robust Function Calling
di: Hao, Bingguang, et al.
Pubblicazione: (2025)
di: Hao, Bingguang, et al.
Pubblicazione: (2025)
CallNavi, A Challenge and Empirical Study on LLM Function Calling and Routing
di: Song, Yewei, et al.
Pubblicazione: (2025)
di: Song, Yewei, et al.
Pubblicazione: (2025)
TinyAgent: Function Calling at the Edge
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
CONFETTI: Conversational Function-Calling Evaluation Through Turn-Level Interactions
di: Alkhouli, Tamer, et al.
Pubblicazione: (2025)
di: Alkhouli, Tamer, et al.
Pubblicazione: (2025)
SimpleTool: Parallel Decoding for Real-Time LLM Function Calling
di: Shi, Xiaoxin, et al.
Pubblicazione: (2026)
di: Shi, Xiaoxin, et al.
Pubblicazione: (2026)
Robust Native Language Identification through Agentic Decomposition
di: Uluslu, Ahmet Yavuz, et al.
Pubblicazione: (2025)
di: Uluslu, Ahmet Yavuz, et al.
Pubblicazione: (2025)
RASTeR: Robust, Agentic, and Structured Temporal Reasoning
di: Schumacher, Dan, et al.
Pubblicazione: (2024)
di: Schumacher, Dan, et al.
Pubblicazione: (2024)
Utilizing Multimodal Data for Edge Case Robust Call-sign Recognition and Understanding
di: Blatt, Alexander, et al.
Pubblicazione: (2024)
di: Blatt, Alexander, et al.
Pubblicazione: (2024)
When2Call: When (not) to Call Tools
di: Ross, Hayley, et al.
Pubblicazione: (2025)
di: Ross, Hayley, et al.
Pubblicazione: (2025)
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks
di: Abdelaziz, Ibrahim, et al.
Pubblicazione: (2024)
di: Abdelaziz, Ibrahim, et al.
Pubblicazione: (2024)
Facilitating Multi-turn Function Calling for LLMs via Compositional Instruction Tuning
di: Chen, Mingyang, et al.
Pubblicazione: (2024)
di: Chen, Mingyang, et al.
Pubblicazione: (2024)
Self-Guided Function Calling in Large Language Models via Stepwise Experience Recall
di: Cui, Sijia, et al.
Pubblicazione: (2025)
di: Cui, Sijia, et al.
Pubblicazione: (2025)
Linguistic and Argument Diversity in Synthetic Data for Function-Calling Agents
di: Greenstein, Dan, et al.
Pubblicazione: (2026)
di: Greenstein, Dan, et al.
Pubblicazione: (2026)
Towards Reliable Benchmarking: A Contamination Free, Controllable Evaluation Framework for Multi-step LLM Function Calling
di: Maekawa, Seiji, et al.
Pubblicazione: (2025)
di: Maekawa, Seiji, et al.
Pubblicazione: (2025)
Team of Thoughts: Efficient Test-time Scaling of Agentic Systems through Orchestrated Tool Calling
di: Wong, Jeffrey T. H., et al.
Pubblicazione: (2026)
di: Wong, Jeffrey T. H., et al.
Pubblicazione: (2026)
Brief Is Better: Non-Monotonic Chain-of-Thought Budget Effects in Function-Calling Language Agents
di: Qi, Xuan
Pubblicazione: (2026)
di: Qi, Xuan
Pubblicazione: (2026)
Documenti analoghi
-
Near-Miss: Latent Policy Failure Detection in Agentic Workflows
di: Rabinovich, Ella, et al.
Pubblicazione: (2026) -
A Novel Metric for Measuring the Robustness of Large Language Models in Non-adversarial Scenarios
di: Ackerman, Samuel, et al.
Pubblicazione: (2024) -
Towards Enforcing Company Policy Adherence in Agentic Workflows
di: Zwerdling, Naama, et al.
Pubblicazione: (2025) -
What's the Plan? Evaluating and Developing Planning-Aware Techniques for Language Models
di: Hirsch, Eran, et al.
Pubblicazione: (2024) -
Think Again! The Effect of Test-Time Compute on Preferences, Opinions, and Beliefs of Large Language Models
di: Kour, George, et al.
Pubblicazione: (2025)