Saved in:
| Main Authors: | Davvetas, Athanasios, Ziouvelou, Xenia, Dami, Ypatia, Kaponis, Alexios, Giouvanopoulou, Konstantina, Papademas, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.17514 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI Act Evaluation Benchmark: An Open, Transparent, and Reproducible Evaluation Dataset for NLP and RAG Systems
by: Davvetas, Athanasios, et al.
Published: (2026)
by: Davvetas, Athanasios, et al.
Published: (2026)
Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems
by: Papademas, Michael, et al.
Published: (2025)
by: Papademas, Michael, et al.
Published: (2025)
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
by: Kokane, Shirley, et al.
Published: (2024)
by: Kokane, Shirley, et al.
Published: (2024)
Query Routing for Homogeneous Tools: An Instantiation in the RAG Scenario
by: Mu, Feiteng, et al.
Published: (2024)
by: Mu, Feiteng, et al.
Published: (2024)
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
by: Glentis, Athanasios, et al.
Published: (2025)
by: Glentis, Athanasios, et al.
Published: (2025)
Multimodal Proposal for an AI-Based Tool to Increase Cross-Assessment of Messages
by: Castro, Alejandro Álvarez, et al.
Published: (2025)
by: Castro, Alejandro Álvarez, et al.
Published: (2025)
Responsible AI Question Bank: A Comprehensive Tool for AI Risk Assessment
by: Lee, Sung Une, et al.
Published: (2024)
by: Lee, Sung Une, et al.
Published: (2024)
ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation
by: Zhou, Jingqi, et al.
Published: (2026)
by: Zhou, Jingqi, et al.
Published: (2026)
Resources for Automated Evaluation of Assistive RAG Systems that Help Readers with News Trustworthiness Assessment
by: Zhang, Dake, et al.
Published: (2026)
by: Zhang, Dake, et al.
Published: (2026)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
by: Tsay, Jason, et al.
Published: (2025)
by: Tsay, Jason, et al.
Published: (2025)
Empowering Affected Individuals to Shape AI Fairness Assessments: Processes, Criteria, and Tools
by: Luo, Lin, et al.
Published: (2026)
by: Luo, Lin, et al.
Published: (2026)
iMedic: Towards Smartphone-based Self-Auscultation Tool for AI-Powered Pediatric Respiratory Assessment
by: Jeong, Seung Gyu, et al.
Published: (2025)
by: Jeong, Seung Gyu, et al.
Published: (2025)
The Global AI Vibrancy Tool
by: Fattorini, Loredana, et al.
Published: (2024)
by: Fattorini, Loredana, et al.
Published: (2024)
TAI3: Testing Agent Integrity in Interpreting User Intent
by: Feng, Shiwei, et al.
Published: (2025)
by: Feng, Shiwei, et al.
Published: (2025)
ToolACE-DEV: Self-Improving Tool Learning via Decomposition and EVolution
by: Huang, Xu, et al.
Published: (2025)
by: Huang, Xu, et al.
Published: (2025)
Design of AI-Powered Tool for Self-Regulation Support in Programming Education
by: Li, Huiyong, et al.
Published: (2025)
by: Li, Huiyong, et al.
Published: (2025)
AquiLLM: a RAG Tool for Capturing Tacit Knowledge in Research Groups
by: Campbell, Chandler, et al.
Published: (2025)
by: Campbell, Chandler, et al.
Published: (2025)
GTM: Simulating the World of Tools for AI Agents
by: Ren, Zhenzhen, et al.
Published: (2025)
by: Ren, Zhenzhen, et al.
Published: (2025)
AgenticRAG: Tool-Augmented Foundation Models for Zero-Shot Explainable Recommender Systems
by: Ma, Bo, et al.
Published: (2025)
by: Ma, Bo, et al.
Published: (2025)
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking
by: Akbar, Muhammad Sajjad
Published: (2025)
by: Akbar, Muhammad Sajjad
Published: (2025)
AI-Driven Tools in Modern Software Quality Assurance: An Assessment of Benefits, Challenges, and Future Directions
by: Pysmennyi, Ihor, et al.
Published: (2025)
by: Pysmennyi, Ihor, et al.
Published: (2025)
Model-diff: A Tool for Comparative Study of Language Models in the Input Space
by: Liu, Weitang, et al.
Published: (2024)
by: Liu, Weitang, et al.
Published: (2024)
Graph-Based Self-Healing Tool Routing for Cost-Efficient LLM Agents
by: Bholani, Neeraj
Published: (2026)
by: Bholani, Neeraj
Published: (2026)
SeCon-RAG: A Two-Stage Semantic Filtering and Conflict-Free Framework for Trustworthy RAG
by: Si, Xiaonan, et al.
Published: (2025)
by: Si, Xiaonan, et al.
Published: (2025)
Democratizing AI scientists using ToolUniverse
by: Gao, Shanghua, et al.
Published: (2025)
by: Gao, Shanghua, et al.
Published: (2025)
Bridging the AI Trustworthiness Gap between Functions and Norms
by: Di Scala, Daan, et al.
Published: (2025)
by: Di Scala, Daan, et al.
Published: (2025)
OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
by: Liang, Jiaxuan, et al.
Published: (2025)
by: Liang, Jiaxuan, et al.
Published: (2025)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
by: Gan, Tiantian, et al.
Published: (2025)
by: Gan, Tiantian, et al.
Published: (2025)
A Review of Prominent Paradigms for LLM-Based Agents: Tool Use (Including RAG), Planning, and Feedback Learning
by: Li, Xinzhe
Published: (2024)
by: Li, Xinzhe
Published: (2024)
STEM Agent: A Self-Adapting, Tool-Enabled, Extensible Architecture for Multi-Protocol AI Agent Systems
by: Shen, Alfred, et al.
Published: (2026)
by: Shen, Alfred, et al.
Published: (2026)
ToolRLA: Multiplicative Reward Decomposition for Tool-Integrated Agents
by: Liu, Pengbo
Published: (2026)
by: Liu, Pengbo
Published: (2026)
CharTool: Tool-Integrated Visual Reasoning for Chart Understanding
by: Zhang, Situo, et al.
Published: (2026)
by: Zhang, Situo, et al.
Published: (2026)
Developer Insights into Designing AI-Based Computer Perception Tools
by: Guhan, Maya, et al.
Published: (2025)
by: Guhan, Maya, et al.
Published: (2025)
EvoTool: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
ToolFuzz -- Automated Agent Tool Testing
by: Milev, Ivan, et al.
Published: (2025)
by: Milev, Ivan, et al.
Published: (2025)
ToolNet: Connecting Large Language Models with Massive Tools via Tool Graph
by: Liu, Xukun, et al.
Published: (2024)
by: Liu, Xukun, et al.
Published: (2024)
ToolBrain: A Flexible Reinforcement Learning Framework for Agentic Tools
by: Le, Quy Minh, et al.
Published: (2025)
by: Le, Quy Minh, et al.
Published: (2025)
ToolCritic: Detecting and Correcting Tool-Use Errors in Dialogue Systems
by: Hamad, Hassan, et al.
Published: (2025)
by: Hamad, Hassan, et al.
Published: (2025)
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
by: Guo, Ruocheng, et al.
Published: (2026)
by: Guo, Ruocheng, et al.
Published: (2026)
AutoTool: Efficient Tool Selection for Large Language Model Agents
by: Jia, Jingyi, et al.
Published: (2025)
by: Jia, Jingyi, et al.
Published: (2025)
Similar Items
-
AI Act Evaluation Benchmark: An Open, Transparent, and Reproducible Evaluation Dataset for NLP and RAG Systems
by: Davvetas, Athanasios, et al.
Published: (2026) -
Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems
by: Papademas, Michael, et al.
Published: (2025) -
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
by: Kokane, Shirley, et al.
Published: (2024) -
Query Routing for Homogeneous Tools: An Instantiation in the RAG Scenario
by: Mu, Feiteng, et al.
Published: (2024) -
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
by: Glentis, Athanasios, et al.
Published: (2025)