Human-Centered Evaluation of RAG outputs: a framework and questionnaire for human-AI collaboration
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mangold, Aline, Hoffmann, Kiran |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Useful for Exploration, Risky for Precision: Evaluating AI Tools in Academic Research
par: Dathe, Anthea, et autres
Publié: (2026)
par: Dathe, Anthea, et autres
Publié: (2026)
On the Design and Evaluation of Human-centered Explainable AI Systems: A Systematic Review and Taxonomy
par: Mangold, Aline, et autres
Publié: (2025)
par: Mangold, Aline, et autres
Publié: (2025)
Human + AI for Accelerating Ad Localization Evaluation
par: Rajgarhia, Harshit, et autres
Publié: (2025)
par: Rajgarhia, Harshit, et autres
Publié: (2025)
Human-Centered Human-AI Interaction (HC-HAII): A Human-Centered AI Perspective
par: Xu, Wei
Publié: (2025)
par: Xu, Wei
Publié: (2025)
Evaluating Identity Leakage in Speaker De-Identification Systems
par: Seo, Seungmin, et autres
Publié: (2025)
par: Seo, Seungmin, et autres
Publié: (2025)
AI Harmonics: a human-centric and harms severity-adaptive AI risk assessment framework
par: Vei, Sofia, et autres
Publié: (2025)
par: Vei, Sofia, et autres
Publié: (2025)
Human-AI collaboration is not very collaborative yet: A taxonomy of interaction patterns in AI-assisted decision making from a systematic review
par: Gomez, Catalina, et autres
Publié: (2023)
par: Gomez, Catalina, et autres
Publié: (2023)
Human and AI collaboration in Fitness Education:A Longitudinal Study with a Pilates Instructor
par: Huang, Qian, et autres
Publié: (2025)
par: Huang, Qian, et autres
Publié: (2025)
Beyond Inefficiency: Systemic Costs of Incivility in Multi-Agent Monte Carlo Simulations
par: Moldovan-Mauer, Alison, et autres
Publié: (2026)
par: Moldovan-Mauer, Alison, et autres
Publié: (2026)
AI Act Evaluation Benchmark: An Open, Transparent, and Reproducible Evaluation Dataset for NLP and RAG Systems
par: Davvetas, Athanasios, et autres
Publié: (2026)
par: Davvetas, Athanasios, et autres
Publié: (2026)
Human-AI collaboration or obedient and often clueless AI in instruct, serve, repeat dynamics?
par: Saqr, Mohammed, et autres
Publié: (2025)
par: Saqr, Mohammed, et autres
Publié: (2025)
Human-AI collaborative autonomous synthesis with pulsed laser deposition for remote epitaxy
par: Haque, Asraful, et autres
Publié: (2025)
par: Haque, Asraful, et autres
Publié: (2025)
The High Cost of Incivility: Quantifying Interaction Inefficiency via Multi-Agent Monte Carlo Simulations
par: Mangold, Benedikt
Publié: (2025)
par: Mangold, Benedikt
Publié: (2025)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
par: Shen, Hua, et autres
Publié: (2025)
par: Shen, Hua, et autres
Publié: (2025)
Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance
par: Ruffle, James K, et autres
Publié: (2025)
par: Ruffle, James K, et autres
Publié: (2025)
Toward Human-Centered Readability Evaluation
par: İlgen, Bahar, et autres
Publié: (2025)
par: İlgen, Bahar, et autres
Publié: (2025)
RAGalyst: Automated Human-Aligned Agentic Evaluation for Domain-Specific RAG
par: Gao, Joshua, et autres
Publié: (2025)
par: Gao, Joshua, et autres
Publié: (2025)
How Human-Centered Explainable AI Interface Are Designed and Evaluated: A Systematic Survey
par: Nguyen, Thu, et autres
Publié: (2024)
par: Nguyen, Thu, et autres
Publié: (2024)
DMA: Online RAG Alignment with Human Feedback
par: Bai, Yu, et autres
Publié: (2025)
par: Bai, Yu, et autres
Publié: (2025)
ChatGPTest: opportunities and cautionary tales of utilizing AI for questionnaire pretesting
par: Olivos, Francisco, et autres
Publié: (2024)
par: Olivos, Francisco, et autres
Publié: (2024)
Understanding the Process of Human-AI Value Alignment
par: McKinlay, Jack, et autres
Publié: (2025)
par: McKinlay, Jack, et autres
Publié: (2025)
AISSISTANT: Human-AI Collaborative Review and Perspective Research Workflows in Data Science
par: Gaddipati, Sasi Kiran, et autres
Publié: (2025)
par: Gaddipati, Sasi Kiran, et autres
Publié: (2025)
Human-Centered Human-AI Collaboration (HCHAC)
par: Gao, Qi, et autres
Publié: (2025)
par: Gao, Qi, et autres
Publié: (2025)
Comprehensive AI Literacy: The Case for Centering Human Agency
par: Tadimalla, Sri Yash, et autres
Publié: (2025)
par: Tadimalla, Sri Yash, et autres
Publié: (2025)
Aetheria: A multimodal interpretable content safety framework based on multi-agent debate and collaboration
par: He, Yuxiang, et autres
Publié: (2025)
par: He, Yuxiang, et autres
Publié: (2025)
The Budget AI Researcher and the Power of RAG Chains
par: Lee, Franklin, et autres
Publié: (2025)
par: Lee, Franklin, et autres
Publié: (2025)
Human-Centered AI and Autonomy in Robotics: Insights from a Bibliometric Study
par: Casini, Simona, et autres
Publié: (2025)
par: Casini, Simona, et autres
Publié: (2025)
GOLF: Goal-Oriented Long-term liFe tasks supported by human-AI collaboration
par: Wang, Ben
Publié: (2024)
par: Wang, Ben
Publié: (2024)
Deepchecks: Evaluating Retrieval-Augmented Generation (RAG)
par: Gerner, Assaf, et autres
Publié: (2026)
par: Gerner, Assaf, et autres
Publié: (2026)
DAO-enabled decentralized physical AI: A new paradigm for human-machine collaboration
par: Ballandies, Mark C., et autres
Publié: (2026)
par: Ballandies, Mark C., et autres
Publié: (2026)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
par: Tsay, Jason, et autres
Publié: (2025)
par: Tsay, Jason, et autres
Publié: (2025)
Human Centered AI for Indian Legal Text Analytics
par: Ghosh, Sudipto, et autres
Publié: (2024)
par: Ghosh, Sudipto, et autres
Publié: (2024)
Toward a Human-Centered AI-assisted Colonoscopy System in Australia
par: Chen, Hsiang-Ting, et autres
Publié: (2025)
par: Chen, Hsiang-Ting, et autres
Publié: (2025)
EcphoryRAG: Re-Imagining Knowledge-Graph RAG via Human Associative Memory
par: Liao, Zirui
Publié: (2025)
par: Liao, Zirui
Publié: (2025)
EmoRAG: Evaluating RAG Robustness to Symbolic Perturbations
par: Zhou, Xinyun, et autres
Publié: (2025)
par: Zhou, Xinyun, et autres
Publié: (2025)
A Unifying Human-Centered AI Fairness Framework
par: Rahman, Munshi Mahbubur, et autres
Publié: (2025)
par: Rahman, Munshi Mahbubur, et autres
Publié: (2025)
Human-Centered Evaluation of XAI Methods
par: Dawoud, Karam, et autres
Publié: (2023)
par: Dawoud, Karam, et autres
Publié: (2023)
Designing and Evaluating Malinowski's Lens: An AI-Native Educational Game for Ethnographic Learning
par: Hoffmann, Michael, et autres
Publié: (2025)
par: Hoffmann, Michael, et autres
Publié: (2025)
Retrieval Augmented Generation (RAG) for Fintech: Agentic Design and Evaluation
par: Cook, Thomas, et autres
Publié: (2025)
par: Cook, Thomas, et autres
Publié: (2025)
From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering
par: Dong, Tao, et autres
Publié: (2025)
par: Dong, Tao, et autres
Publié: (2025)
Documents similaires
-
Useful for Exploration, Risky for Precision: Evaluating AI Tools in Academic Research
par: Dathe, Anthea, et autres
Publié: (2026) -
On the Design and Evaluation of Human-centered Explainable AI Systems: A Systematic Review and Taxonomy
par: Mangold, Aline, et autres
Publié: (2025) -
Human + AI for Accelerating Ad Localization Evaluation
par: Rajgarhia, Harshit, et autres
Publié: (2025) -
Human-Centered Human-AI Interaction (HC-HAII): A Human-Centered AI Perspective
par: Xu, Wei
Publié: (2025) -
Evaluating Identity Leakage in Speaker De-Identification Systems
par: Seo, Seungmin, et autres
Publié: (2025)