Principles and Guidelines for Randomized Controlled Trials in AI Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Kelly, Christopher, Chowdhury, Angelica, Campili, Alexandra, Ayoola, Bimpe, Barbour, Devin, Dawson, Thomas Chen, Chin, Ze Shen, Gipiškis, Rokas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Generative AI in the Lab: Methodological Challenges and Guidelines
by: Park, Hyerim, et al.
Published: (2026)
by: Park, Hyerim, et al.
Published: (2026)
Investigating Students' Preferences for AI Roles in Mathematical Modelling: Evidence from a Randomized Controlled Trial
by: Zhu, Wangda, et al.
Published: (2025)
by: Zhu, Wangda, et al.
Published: (2025)
RAI Guidelines: Method for Generating Responsible AI Guidelines Grounded in Regulations and Usable by (Non-)Technical Roles
by: Constantinides, Marios, et al.
Published: (2023)
by: Constantinides, Marios, et al.
Published: (2023)
PreCare: Designing AI Assistants for Advance Care Planning (ACP) to Enhance Personal Value Exploration, Patient Knowledge, and Decisional Confidence
by: Hsu, Yu Lun, et al.
Published: (2025)
by: Hsu, Yu Lun, et al.
Published: (2025)
Guidelines for Integrating Value Sensitive Design in Responsible AI Toolkits
by: Sadek, Malak, et al.
Published: (2024)
by: Sadek, Malak, et al.
Published: (2024)
From Guidelines to Practice: Evaluating the Reproducibility of Methods in Computational Social Science
by: Momeni, Fakhri, et al.
Published: (2026)
by: Momeni, Fakhri, et al.
Published: (2026)
TrialView: An AI-powered Visual Analytics System for Temporal Event Data in Clinical Trials
by: Li, Zuotian, et al.
Published: (2023)
by: Li, Zuotian, et al.
Published: (2023)
Case Law Grounding: Using Precedents to Align Decision-Making for Humans and AI
by: Chen, Quan Ze, et al.
Published: (2023)
by: Chen, Quan Ze, et al.
Published: (2023)
Principles of Safe AI Companions for Youth: Parent and Expert Perspectives
by: Yu, Yaman, et al.
Published: (2025)
by: Yu, Yaman, et al.
Published: (2025)
pixelLOG: Logging of Online Gameplay for Cognitive Research
by: Lu, Zeyu, et al.
Published: (2026)
by: Lu, Zeyu, et al.
Published: (2026)
Stories That Teach: Eastern Wisdom for Human-AI Creative Partnerships
by: Nie, Kexin, et al.
Published: (2025)
by: Nie, Kexin, et al.
Published: (2025)
The User-first Approach to AI Ethics: Preferences for Ethical Principles in AI Systems across Cultures and Contexts
by: Carroll, Benjamin J., et al.
Published: (2025)
by: Carroll, Benjamin J., et al.
Published: (2025)
Analyzing Security and Privacy Challenges in Generative AI Usage Guidelines for Higher Education
by: Ng, Bei Yi, et al.
Published: (2025)
by: Ng, Bei Yi, et al.
Published: (2025)
VisPubs Games: Joyful Discovery of Visualization Research(ers)
by: Lange, Devin, et al.
Published: (2025)
by: Lange, Devin, et al.
Published: (2025)
AI for Proactive Mental Health: A Multi-Institutional, Longitudinal, Randomized Controlled Trial
by: Cachia, Julie Y. A., et al.
Published: (2025)
by: Cachia, Julie Y. A., et al.
Published: (2025)
A Tool to Facilitate Web-Browsing
by: Kelly, Christopher, et al.
Published: (2024)
by: Kelly, Christopher, et al.
Published: (2024)
GuidelineExplorer -- Navigating through the Forrest of Actionable Guidelines on Node-Link Graph Visualization
by: Guckes, Kathrin, et al.
Published: (2024)
by: Guckes, Kathrin, et al.
Published: (2024)
Conversational AI for Automated Patient Questionnaire Completion: Development Insights and Design Principles
by: Navarro, David Fraile, et al.
Published: (2026)
by: Navarro, David Fraile, et al.
Published: (2026)
Creating an African American-Sounding TTS: Guidelines, Technical Challenges,and Surprising Evaluations
by: Pinhanez, Claudio, et al.
Published: (2024)
by: Pinhanez, Claudio, et al.
Published: (2024)
`Socheton': A Culturally Appropriate AI Tool to Support Reproductive Well-being
by: Sultana, Sharifa, et al.
Published: (2025)
by: Sultana, Sharifa, et al.
Published: (2025)
Auxilio and Beyond: Comparative Evaluation, Usability, and Design Guidelines for Head Movement-based Assistive Mouse Controllers
by: Kabir, Mohammad Ridwan, et al.
Published: (2022)
by: Kabir, Mohammad Ridwan, et al.
Published: (2022)
Disability-First AI Dataset Annotation: Co-designing Stuttered Speech Annotation Guidelines with People Who Stutter
by: Tang, Xinru, et al.
Published: (2026)
by: Tang, Xinru, et al.
Published: (2026)
Can Conversational AI Counsel for Change? A Theory-Driven Approach to Supporting Dietary Intentions in Ambivalent Individuals
by: Bak, Michelle, et al.
Published: (2025)
by: Bak, Michelle, et al.
Published: (2025)
IntentFlow: Investigating Fluid Dynamics of Intent Communication in Generative AI
by: Kim, Yoonsu, et al.
Published: (2025)
by: Kim, Yoonsu, et al.
Published: (2025)
Engineering Trustworthy Automation: Design Principles and Evaluation for AutoML Tools for Novices
by: Thys, Jarne, et al.
Published: (2025)
by: Thys, Jarne, et al.
Published: (2025)
Towards AI as Colleagues: Multi-Agent System Improves Structured Ideation Processes
by: Quan, Kexin, et al.
Published: (2025)
by: Quan, Kexin, et al.
Published: (2025)
"Pragmatic Tools or Empowering Friends?" Discovering and Co-Designing Personality-Aligned AI Writing Companions
by: Wu, Mengke, et al.
Published: (2025)
by: Wu, Mengke, et al.
Published: (2025)
GPT Chatbots for Alleviating Anxiety and Depression: A Pilot Randomized Controlled Trial with Afghan Women
by: Sahab, Sofia, et al.
Published: (2025)
by: Sahab, Sofia, et al.
Published: (2025)
Understanding Attitudes and Trust of Generative AI Chatbots for Social Anxiety Support
by: Wang, Yimeng, et al.
Published: (2025)
by: Wang, Yimeng, et al.
Published: (2025)
Risk of Harm in VR Dating from the Perspective of Women and LGBTQIA+ Stakeholders
by: Tebbe, Devin, et al.
Published: (2024)
by: Tebbe, Devin, et al.
Published: (2024)
What Makes an AI Writing Companion a Good Fit? A Personality-Informed Co-Design Study
by: Wu, Mengke, et al.
Published: (2026)
by: Wu, Mengke, et al.
Published: (2026)
Evaluating Actionability in Explainable AI
by: Mansi, Gennie, et al.
Published: (2026)
by: Mansi, Gennie, et al.
Published: (2026)
Computational Cognitive Modeling to understand the effects of Racializing AI on Human-AI cooperation with PigChase Task
by: Dulam, Swapnika, et al.
Published: (2025)
by: Dulam, Swapnika, et al.
Published: (2025)
The Effects of Generative AI on Design Fixation and Divergent Thinking
by: Wadinambiarachchi, Samangi, et al.
Published: (2024)
by: Wadinambiarachchi, Samangi, et al.
Published: (2024)
Design Principles for Generative AI Applications
by: Weisz, Justin D., et al.
Published: (2024)
by: Weisz, Justin D., et al.
Published: (2024)
Guidelines para Desenvolvimento de Jogos Mobile Inclusivos
by: Zorzo, Gabriela Panta, et al.
Published: (2024)
by: Zorzo, Gabriela Panta, et al.
Published: (2024)
Understanding the Research-Practice Gap in Visualization Design Guidelines
by: Kim, Nam Wook, et al.
Published: (2023)
by: Kim, Nam Wook, et al.
Published: (2023)
AI-Mediated Feedback Improves Student Revisions: A Randomized Trial with FeedbackWriter in a Large Undergraduate Course
by: Lu, Xinyi, et al.
Published: (2026)
by: Lu, Xinyi, et al.
Published: (2026)
Troubling Taxonomies in GenAI Evaluation
by: Berman, Glen, et al.
Published: (2024)
by: Berman, Glen, et al.
Published: (2024)
An Exploratory Study on AI-driven Visualisation Techniques on Decision Making in Extended Reality
by: Dong, Ze, et al.
Published: (2025)
by: Dong, Ze, et al.
Published: (2025)
Similar Items
-
Evaluating Generative AI in the Lab: Methodological Challenges and Guidelines
by: Park, Hyerim, et al.
Published: (2026) -
Investigating Students' Preferences for AI Roles in Mathematical Modelling: Evidence from a Randomized Controlled Trial
by: Zhu, Wangda, et al.
Published: (2025) -
RAI Guidelines: Method for Generating Responsible AI Guidelines Grounded in Regulations and Usable by (Non-)Technical Roles
by: Constantinides, Marios, et al.
Published: (2023) -
PreCare: Designing AI Assistants for Advance Care Planning (ACP) to Enhance Personal Value Exploration, Patient Knowledge, and Decisional Confidence
by: Hsu, Yu Lun, et al.
Published: (2025) -
Guidelines for Integrating Value Sensitive Design in Responsible AI Toolkits
by: Sadek, Malak, et al.
Published: (2024)