PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yu, Kim, Ig-Jae, Yoon, Dongwook |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The AI Genie Phenomenon and Three Types of AI Chatbot Addiction: Escapist Roleplays, Pseudosocial Companions, and Epistemic Rabbit Holes
by: Shen, M. Karen, et al.
Published: (2026)
by: Shen, M. Karen, et al.
Published: (2026)
Social Media Clones: Exploring the Impact of Social Delegation with AI Clones through a Design Workbook Study
by: Liu, Jackie, et al.
Published: (2025)
by: Liu, Jackie, et al.
Published: (2025)
Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2025)
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2025)
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task
by: Yoon, Sion, et al.
Published: (2024)
by: Yoon, Sion, et al.
Published: (2024)
Cloning the Self for Mental Well-Being: A Framework for Designing Safe and Therapeutic Self-Clone Chatbots
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2026)
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2026)
Toward Scalable Audio Description Quality Control: A Workflow for Evaluating Human and VLM Raters
by: Do, Lana, et al.
Published: (2026)
by: Do, Lana, et al.
Published: (2026)
An Empirical Examination of the Evaluative AI Framework
by: Kornowicz, Jaroslaw
Published: (2024)
by: Kornowicz, Jaroslaw
Published: (2024)
Beyond Compliance: How AI Could Help Creative Writers by Refusing Them
by: Qin, Hua Xuan, et al.
Published: (2026)
by: Qin, Hua Xuan, et al.
Published: (2026)
A Scalable Framework for Evaluating Health Language Models
by: Mallinar, Neil, et al.
Published: (2025)
by: Mallinar, Neil, et al.
Published: (2025)
Beyond Compliance: A Resistance-Informed Motivation Reasoning Framework for Challenging Psychological Client Simulation
by: Liu, Danni, et al.
Published: (2026)
by: Liu, Danni, et al.
Published: (2026)
Transforming GenAI Policy to Prompting Instruction: An RCT of Scalable Prompting Interventions in a CS1 Course
by: Xiao, Ruiwei, et al.
Published: (2026)
by: Xiao, Ruiwei, et al.
Published: (2026)
Evaluating Human-AI Collaboration: A Review and Methodological Framework
by: Fragiadakis, George, et al.
Published: (2024)
by: Fragiadakis, George, et al.
Published: (2024)
When Should an AI Act? A Human-Centered Model of Scene, Context, and Behavior for Agentic AI Design
by: Jung, Soyoung, et al.
Published: (2026)
by: Jung, Soyoung, et al.
Published: (2026)
Authors' Values and Attitudes Towards AI-bridged Scalable Personalization of Creative Language Arts
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Children's Mental Models of AI Reasoning: Implications for AI Literacy Education
by: Dangol, Aayushi, et al.
Published: (2025)
by: Dangol, Aayushi, et al.
Published: (2025)
DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models
by: Kim, Eugenia, et al.
Published: (2026)
by: Kim, Eugenia, et al.
Published: (2026)
Towards a Comprehensive Human-Centred Evaluation Framework for Explainable AI
by: Donoso-Guzmán, Ivania, et al.
Published: (2023)
by: Donoso-Guzmán, Ivania, et al.
Published: (2023)
"My Grade is Wrong!": A Contestable AI Framework for Interactive Feedback in Evaluating Student Essays
by: Hong, Shengxin, et al.
Published: (2024)
by: Hong, Shengxin, et al.
Published: (2024)
Evaluating the Effects of AI Directors for Quest Selection
by: Yu, Kristen K., et al.
Published: (2024)
by: Yu, Kristen K., et al.
Published: (2024)
Do Consumers Accept AIs as Moral Compliance Agents?
by: Nyilasy, Greg, et al.
Published: (2026)
by: Nyilasy, Greg, et al.
Published: (2026)
A Multi-Layered Research Framework for Human-Centered AI: Defining the Path to Explainability and Trust
by: De Silva, Chameera, et al.
Published: (2025)
by: De Silva, Chameera, et al.
Published: (2025)
SPHERE: An Evaluation Card for Human-AI Systems
by: Ma, Qianou, et al.
Published: (2025)
by: Ma, Qianou, et al.
Published: (2025)
A Scalable Approach to Benchmarking the In-Conversation Differential Diagnostic Accuracy of a Health AI
by: Bhatt, Deep, et al.
Published: (2024)
by: Bhatt, Deep, et al.
Published: (2024)
Learning Complementary Policies for Human-AI Teams
by: Gao, Ruijiang, et al.
Published: (2023)
by: Gao, Ruijiang, et al.
Published: (2023)
Evaluating Machine Expertise: How Graduate Students Develop Frameworks for Assessing GenAI Content
by: Chen, Celia, et al.
Published: (2025)
by: Chen, Celia, et al.
Published: (2025)
SHARPIE: A Modular Framework for Reinforcement Learning and Human-AI Interaction Experiments
by: Aydın, Hüseyin, et al.
Published: (2025)
by: Aydın, Hüseyin, et al.
Published: (2025)
Toward an AI-Powered Computational Testbed for Workforce Policy
by: Vaid, Sumer S., et al.
Published: (2026)
by: Vaid, Sumer S., et al.
Published: (2026)
Wikibench: Community-Driven Data Curation for AI Evaluation on Wikipedia
by: Kuo, Tzu-Sheng, et al.
Published: (2024)
by: Kuo, Tzu-Sheng, et al.
Published: (2024)
When Generative AI Meets Extended Reality: Enabling Scalable and Natural Interactions
by: Zhu, Mingyu, et al.
Published: (2026)
by: Zhu, Mingyu, et al.
Published: (2026)
A Framework for Longitudinal Health AI Agents
by: Lin, Georgianna, et al.
Published: (2026)
by: Lin, Georgianna, et al.
Published: (2026)
CAIM: Development and Evaluation of a Cognitive AI Memory Framework for Long-Term Interaction with Intelligent Agents
by: Westhäußer, Rebecca, et al.
Published: (2025)
by: Westhäußer, Rebecca, et al.
Published: (2025)
StreetDesignAI: Broadening Designer Perspectives Through Multi-Persona Evaluation of Cycling Infrastructure
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
Balancing Usability and Compliance in AI Smart Devices: A Privacy-by-Design Audit of Google Home, Alexa, and Siri
by: De Clark, Trevor, et al.
Published: (2026)
by: De Clark, Trevor, et al.
Published: (2026)
Beyond Procedural Compliance: Human Oversight as a Dimension of Well-being Efficacy in AI Governance
by: Xie, Yao, et al.
Published: (2025)
by: Xie, Yao, et al.
Published: (2025)
Toward AI Matching Policies in Homeless Services: A Qualitative Study with Policymakers
by: Johnston, Caroline M., et al.
Published: (2025)
by: Johnston, Caroline M., et al.
Published: (2025)
A Multi-Agent Large Language Model Framework for Automated Qualitative Analysis
by: Xu, Qidi, et al.
Published: (2025)
by: Xu, Qidi, et al.
Published: (2025)
A LLM-based Controllable, Scalable, Human-Involved User Simulator Framework for Conversational Recommender Systems
by: Zhu, Lixi, et al.
Published: (2024)
by: Zhu, Lixi, et al.
Published: (2024)
A Decision Theoretic Framework for Measuring AI Reliance
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
Pensieve Discuss: Scalable Small-Group CS Tutoring System with AI
by: Yang, Yoonseok, et al.
Published: (2024)
by: Yang, Yoonseok, et al.
Published: (2024)
InnerPond: Fostering Inter-Self Dialogue with a Multi-Agent Approach for Introspection
by: Jeon, Hayeon, et al.
Published: (2026)
by: Jeon, Hayeon, et al.
Published: (2026)
Similar Items
-
The AI Genie Phenomenon and Three Types of AI Chatbot Addiction: Escapist Roleplays, Pseudosocial Companions, and Epistemic Rabbit Holes
by: Shen, M. Karen, et al.
Published: (2026) -
Social Media Clones: Exploring the Impact of Social Delegation with AI Clones through a Design Workbook Study
by: Liu, Jackie, et al.
Published: (2025) -
Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2025) -
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task
by: Yoon, Sion, et al.
Published: (2024) -
Cloning the Self for Mental Well-Being: A Framework for Designing Safe and Therapeutic Self-Clone Chatbots
by: Shirvani, Mehrnoosh Sadat, et al.
Published: (2026)