TAI Scan Tool: A RAG-Based Tool With Minimalistic Input for Trustworthy AI Self-Assessment
Fuente:
arXiv
Salvato in:
| Autori principali: | Davvetas, Athanasios, Ziouvelou, Xenia, Dami, Ypatia, Kaponis, Alexios, Giouvanopoulou, Konstantina, Papademas, Michael |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AI Act Evaluation Benchmark: An Open, Transparent, and Reproducible Evaluation Dataset for NLP and RAG Systems
di: Davvetas, Athanasios, et al.
Pubblicazione: (2026)
di: Davvetas, Athanasios, et al.
Pubblicazione: (2026)
Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems
di: Papademas, Michael, et al.
Pubblicazione: (2025)
di: Papademas, Michael, et al.
Pubblicazione: (2025)
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
di: Kokane, Shirley, et al.
Pubblicazione: (2024)
di: Kokane, Shirley, et al.
Pubblicazione: (2024)
Query Routing for Homogeneous Tools: An Instantiation in the RAG Scenario
di: Mu, Feiteng, et al.
Pubblicazione: (2024)
di: Mu, Feiteng, et al.
Pubblicazione: (2024)
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
di: Glentis, Athanasios, et al.
Pubblicazione: (2025)
di: Glentis, Athanasios, et al.
Pubblicazione: (2025)
Multimodal Proposal for an AI-Based Tool to Increase Cross-Assessment of Messages
di: Castro, Alejandro Álvarez, et al.
Pubblicazione: (2025)
di: Castro, Alejandro Álvarez, et al.
Pubblicazione: (2025)
ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation
di: Zhou, Jingqi, et al.
Pubblicazione: (2026)
di: Zhou, Jingqi, et al.
Pubblicazione: (2026)
Responsible AI Question Bank: A Comprehensive Tool for AI Risk Assessment
di: Lee, Sung Une, et al.
Pubblicazione: (2024)
di: Lee, Sung Une, et al.
Pubblicazione: (2024)
Resources for Automated Evaluation of Assistive RAG Systems that Help Readers with News Trustworthiness Assessment
di: Zhang, Dake, et al.
Pubblicazione: (2026)
di: Zhang, Dake, et al.
Pubblicazione: (2026)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
di: Tsay, Jason, et al.
Pubblicazione: (2025)
di: Tsay, Jason, et al.
Pubblicazione: (2025)
Empowering Affected Individuals to Shape AI Fairness Assessments: Processes, Criteria, and Tools
di: Luo, Lin, et al.
Pubblicazione: (2026)
di: Luo, Lin, et al.
Pubblicazione: (2026)
The Global AI Vibrancy Tool
di: Fattorini, Loredana, et al.
Pubblicazione: (2024)
di: Fattorini, Loredana, et al.
Pubblicazione: (2024)
ToolACE-DEV: Self-Improving Tool Learning via Decomposition and EVolution
di: Huang, Xu, et al.
Pubblicazione: (2025)
di: Huang, Xu, et al.
Pubblicazione: (2025)
iMedic: Towards Smartphone-based Self-Auscultation Tool for AI-Powered Pediatric Respiratory Assessment
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
di: Jeong, Seung Gyu, et al.
Pubblicazione: (2025)
GTM: Simulating the World of Tools for AI Agents
di: Ren, Zhenzhen, et al.
Pubblicazione: (2025)
di: Ren, Zhenzhen, et al.
Pubblicazione: (2025)
Design of AI-Powered Tool for Self-Regulation Support in Programming Education
di: Li, Huiyong, et al.
Pubblicazione: (2025)
di: Li, Huiyong, et al.
Pubblicazione: (2025)
AquiLLM: a RAG Tool for Capturing Tacit Knowledge in Research Groups
di: Campbell, Chandler, et al.
Pubblicazione: (2025)
di: Campbell, Chandler, et al.
Pubblicazione: (2025)
AgenticRAG: Tool-Augmented Foundation Models for Zero-Shot Explainable Recommender Systems
di: Ma, Bo, et al.
Pubblicazione: (2025)
di: Ma, Bo, et al.
Pubblicazione: (2025)
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking
di: Akbar, Muhammad Sajjad
Pubblicazione: (2025)
di: Akbar, Muhammad Sajjad
Pubblicazione: (2025)
AI-Driven Tools in Modern Software Quality Assurance: An Assessment of Benefits, Challenges, and Future Directions
di: Pysmennyi, Ihor, et al.
Pubblicazione: (2025)
di: Pysmennyi, Ihor, et al.
Pubblicazione: (2025)
Bridging the AI Trustworthiness Gap between Functions and Norms
di: Di Scala, Daan, et al.
Pubblicazione: (2025)
di: Di Scala, Daan, et al.
Pubblicazione: (2025)
STEM Agent: A Self-Adapting, Tool-Enabled, Extensible Architecture for Multi-Protocol AI Agent Systems
di: Shen, Alfred, et al.
Pubblicazione: (2026)
di: Shen, Alfred, et al.
Pubblicazione: (2026)
Graph-Based Self-Healing Tool Routing for Cost-Efficient LLM Agents
di: Bholani, Neeraj
Pubblicazione: (2026)
di: Bholani, Neeraj
Pubblicazione: (2026)
TAI3: Testing Agent Integrity in Interpreting User Intent
di: Feng, Shiwei, et al.
Pubblicazione: (2025)
di: Feng, Shiwei, et al.
Pubblicazione: (2025)
SeCon-RAG: A Two-Stage Semantic Filtering and Conflict-Free Framework for Trustworthy RAG
di: Si, Xiaonan, et al.
Pubblicazione: (2025)
di: Si, Xiaonan, et al.
Pubblicazione: (2025)
Democratizing AI scientists using ToolUniverse
di: Gao, Shanghua, et al.
Pubblicazione: (2025)
di: Gao, Shanghua, et al.
Pubblicazione: (2025)
ToolRLA: Multiplicative Reward Decomposition for Tool-Integrated Agents
di: Liu, Pengbo
Pubblicazione: (2026)
di: Liu, Pengbo
Pubblicazione: (2026)
CharTool: Tool-Integrated Visual Reasoning for Chart Understanding
di: Zhang, Situo, et al.
Pubblicazione: (2026)
di: Zhang, Situo, et al.
Pubblicazione: (2026)
EvoTool: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection
di: Yang, Shuo, et al.
Pubblicazione: (2026)
di: Yang, Shuo, et al.
Pubblicazione: (2026)
OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
di: Liang, Jiaxuan, et al.
Pubblicazione: (2025)
di: Liang, Jiaxuan, et al.
Pubblicazione: (2025)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
di: Gan, Tiantian, et al.
Pubblicazione: (2025)
di: Gan, Tiantian, et al.
Pubblicazione: (2025)
Model-diff: A Tool for Comparative Study of Language Models in the Input Space
di: Liu, Weitang, et al.
Pubblicazione: (2024)
di: Liu, Weitang, et al.
Pubblicazione: (2024)
A Review of Prominent Paradigms for LLM-Based Agents: Tool Use (Including RAG), Planning, and Feedback Learning
di: Li, Xinzhe
Pubblicazione: (2024)
di: Li, Xinzhe
Pubblicazione: (2024)
AI-Compass: A Comprehensive and Effective Multi-module Testing Tool for AI Systems
di: Zhu, Zhiyu, et al.
Pubblicazione: (2024)
di: Zhu, Zhiyu, et al.
Pubblicazione: (2024)
ToolBrain: A Flexible Reinforcement Learning Framework for Agentic Tools
di: Le, Quy Minh, et al.
Pubblicazione: (2025)
di: Le, Quy Minh, et al.
Pubblicazione: (2025)
ToolCritic: Detecting and Correcting Tool-Use Errors in Dialogue Systems
di: Hamad, Hassan, et al.
Pubblicazione: (2025)
di: Hamad, Hassan, et al.
Pubblicazione: (2025)
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
di: Guo, Ruocheng, et al.
Pubblicazione: (2026)
di: Guo, Ruocheng, et al.
Pubblicazione: (2026)
AutoTool: Efficient Tool Selection for Large Language Model Agents
di: Jia, Jingyi, et al.
Pubblicazione: (2025)
di: Jia, Jingyi, et al.
Pubblicazione: (2025)
Developer Insights into Designing AI-Based Computer Perception Tools
di: Guhan, Maya, et al.
Pubblicazione: (2025)
di: Guhan, Maya, et al.
Pubblicazione: (2025)
Toward Maturity-Based Certification of Embodied AI: Quantifying Trustworthiness Through Measurement Mechanisms
di: Darling, Michael C., et al.
Pubblicazione: (2026)
di: Darling, Michael C., et al.
Pubblicazione: (2026)
Documenti analoghi
-
AI Act Evaluation Benchmark: An Open, Transparent, and Reproducible Evaluation Dataset for NLP and RAG Systems
di: Davvetas, Athanasios, et al.
Pubblicazione: (2026) -
Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems
di: Papademas, Michael, et al.
Pubblicazione: (2025) -
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
di: Kokane, Shirley, et al.
Pubblicazione: (2024) -
Query Routing for Homogeneous Tools: An Instantiation in the RAG Scenario
di: Mu, Feiteng, et al.
Pubblicazione: (2024) -
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
di: Glentis, Athanasios, et al.
Pubblicazione: (2025)