Mind the GAP: Text Safety Does Not Transfer to Tool-Call Safety in LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cartagena, Arnold, Teixeira, Ariane |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
A Framework for Testing and Adapting REST APIs as LLM Tools
von: Bandlamudi, Jayachandu, et al.
Veröffentlicht: (2025)
von: Bandlamudi, Jayachandu, et al.
Veröffentlicht: (2025)
ContractBench: Can LLM Agents Preserve Observation Contracts?
von: Wang, Jicheng, et al.
Veröffentlicht: (2026)
von: Wang, Jicheng, et al.
Veröffentlicht: (2026)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
Achieving Tool Calling Functionality in LLMs Using Only Prompt Engineering Without Fine-Tuning
von: He, Shengtao
Veröffentlicht: (2024)
von: He, Shengtao
Veröffentlicht: (2024)
Learning Software Bug Reports: A Systematic Literature Review
von: Long, Guoming, et al.
Veröffentlicht: (2025)
von: Long, Guoming, et al.
Veröffentlicht: (2025)
Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety
von: Gringras, David
Veröffentlicht: (2026)
von: Gringras, David
Veröffentlicht: (2026)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
von: Kohl, Jens, et al.
Veröffentlicht: (2024)
von: Kohl, Jens, et al.
Veröffentlicht: (2024)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026)
von: Souza, Débora, et al.
Veröffentlicht: (2026)
Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
Benchmarking Energy Efficiency of Large Language Models Using vLLM
von: Pronk, K., et al.
Veröffentlicht: (2025)
von: Pronk, K., et al.
Veröffentlicht: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research
von: Savenkov, Vladislav
Veröffentlicht: (2026)
von: Savenkov, Vladislav
Veröffentlicht: (2026)
FREYR: A Framework for Recognizing and Executing Your Requests
von: Gallotta, Roberto, et al.
Veröffentlicht: (2025)
von: Gallotta, Roberto, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
XPath Agent: An Efficient XPath Programming Agent Based on LLM for Web Crawler
von: Li, Yu, et al.
Veröffentlicht: (2024)
von: Li, Yu, et al.
Veröffentlicht: (2024)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
von: More, Riddhi, et al.
Veröffentlicht: (2025)
von: More, Riddhi, et al.
Veröffentlicht: (2025)
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
Beyond Greenfield: The D3 Framework for AI-Driven Productivity in Brownfield Engineering
von: Sharma, Krishna Kumaar
Veröffentlicht: (2025)
von: Sharma, Krishna Kumaar
Veröffentlicht: (2025)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
SiliconMind-V1: Multi-Agent Distillation and Debug-Reasoning Workflows for Verilog Code Generation
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2026)
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2026)
The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot
von: Yeverechyahu, Doron, et al.
Veröffentlicht: (2024)
von: Yeverechyahu, Doron, et al.
Veröffentlicht: (2024)
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents
von: Mazaheri, Parsa, et al.
Veröffentlicht: (2026)
von: Mazaheri, Parsa, et al.
Veröffentlicht: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
Automating Domain-Driven Design: Experience with a Prompting Framework
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2026)
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2026)
Leveraging Large Language Models for Use Case Model Generation from Software Requirements
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2025)
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2025)
Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development
von: Tran, Hung, et al.
Veröffentlicht: (2026)
von: Tran, Hung, et al.
Veröffentlicht: (2026)
TCProF: Time-Complexity Prediction SSL Framework
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones
von: Rai, Daking, et al.
Veröffentlicht: (2025)
von: Rai, Daking, et al.
Veröffentlicht: (2025)
MEC$^3$O: Multi-Expert Consensus for Code Time Complexity Prediction
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
von: Lim, Soohan, et al.
Veröffentlicht: (2025)
von: Lim, Soohan, et al.
Veröffentlicht: (2025)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
A Serverless Architecture for Real-Time Stock Analysis using Large Language Models: An Iterative Development and Debugging Case Study
von: Ashraf, Taniv
Veröffentlicht: (2025)
von: Ashraf, Taniv
Veröffentlicht: (2025)
Smaller Models, Smarter Rewards: A Two-Sided Approach to Process and Outcome Rewards
von: Groeneveld, Jan Niklas, et al.
Veröffentlicht: (2025)
von: Groeneveld, Jan Niklas, et al.
Veröffentlicht: (2025)
Narrow Transformer: StarCoder-Based Java-LM For Desktop
von: Rathinasamy, Kamalkumar, et al.
Veröffentlicht: (2024)
von: Rathinasamy, Kamalkumar, et al.
Veröffentlicht: (2024)
Mechanistic Understanding of Language Models in Syntactic Code Completion
von: Miller, Samuel, et al.
Veröffentlicht: (2025)
von: Miller, Samuel, et al.
Veröffentlicht: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
Plan with Code: Comparing approaches for robust NL to DSL generation
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025) -
A Framework for Testing and Adapting REST APIs as LLM Tools
von: Bandlamudi, Jayachandu, et al.
Veröffentlicht: (2025) -
ContractBench: Can LLM Agents Preserve Observation Contracts?
von: Wang, Jicheng, et al.
Veröffentlicht: (2026) -
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026) -
Achieving Tool Calling Functionality in LLMs Using Only Prompt Engineering Without Fine-Tuning
von: He, Shengtao
Veröffentlicht: (2024)