EUDAIMONIA: Evaluating Undesirable Dynamics in AI
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Jun Rui, Zhu, Wang Bill, Liu, Ziyi, Fast, Nathanael, Iyer, Ravi, Jia, Robin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
Improving Interactive Diagnostic Ability of a Large Language Model Agent Through Clinical Experience Learning
by: Sun, Zhoujian, et al.
Published: (2025)
by: Sun, Zhoujian, et al.
Published: (2025)
Human Bias in the Face of AI: Examining Human Judgment Against Text Labeled as AI Generated
by: Zhu, Tiffany, et al.
Published: (2024)
by: Zhu, Tiffany, et al.
Published: (2024)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
by: Oh, Juhyun, et al.
Published: (2024)
by: Oh, Juhyun, et al.
Published: (2024)
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
by: Clark, Nicholas, et al.
Published: (2025)
by: Clark, Nicholas, et al.
Published: (2025)
Evaluation and Incident Prevention in an Enterprise AI Assistant
by: Maharaj, Akash V., et al.
Published: (2025)
by: Maharaj, Akash V., et al.
Published: (2025)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
VizTrust: A Visual Analytics Tool for Capturing User Trust Dynamics in Human-AI Communication
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
A Risk Ontology for Evaluating AI-Powered Psychotherapy Virtual Agents
by: Steenstra, Ian, et al.
Published: (2025)
by: Steenstra, Ian, et al.
Published: (2025)
BLIP: Facilitating the Exploration of Undesirable Consequences of Digital Technologies
by: Pang, Rock Yuren, et al.
Published: (2024)
by: Pang, Rock Yuren, et al.
Published: (2024)
Awaking the Slides: A Tuning-free and Knowledge-regulated AI Tutoring System via Language Model Coordination
by: Zhang-Li, Daniel, et al.
Published: (2024)
by: Zhang-Li, Daniel, et al.
Published: (2024)
The Good, The Bad, and Why: Unveiling Emotions in Generative AI
by: Li, Cheng, et al.
Published: (2023)
by: Li, Cheng, et al.
Published: (2023)
CulturalTeaming: AI-Assisted Interactive Red-Teaming for Challenging LLMs' (Lack of) Multicultural Knowledge
by: Chiu, Yu Ying, et al.
Published: (2024)
by: Chiu, Yu Ying, et al.
Published: (2024)
VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents
by: Lee, Sam Yu-Te, et al.
Published: (2025)
by: Lee, Sam Yu-Te, et al.
Published: (2025)
Ensembling Large Language Models to Characterize Affective Dynamics in Student-AI Tutor Dialogues
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
Exploring how EFL students talk to and through AI to develop texts
by: Woo, David James, et al.
Published: (2026)
by: Woo, David James, et al.
Published: (2026)
Incentives shape how humans co-create with generative AI
by: Jo, Nathanael, et al.
Published: (2026)
by: Jo, Nathanael, et al.
Published: (2026)
Direct Advantage Regression: Aligning LLMs with Online AI Reward
by: He, Li, et al.
Published: (2025)
by: He, Li, et al.
Published: (2025)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
by: Liu, Hongtao, et al.
Published: (2025)
by: Liu, Hongtao, et al.
Published: (2025)
PsychBench: A comprehensive and professional benchmark for evaluating the performance of LLM-assisted psychiatric clinical practice
by: Liu, Shuyu, et al.
Published: (2025)
by: Liu, Shuyu, et al.
Published: (2025)
Human-Centered AI in Multidisciplinary Medical Discussions: Evaluating the Feasibility of a Chat-Based Approach to Case Assessment
by: Sawano, Shinnosuke, et al.
Published: (2025)
by: Sawano, Shinnosuke, et al.
Published: (2025)
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task
by: Yoon, Sion, et al.
Published: (2024)
by: Yoon, Sion, et al.
Published: (2024)
MimiTalk: Revolutionizing Qualitative Research with Dual-Agent AI
by: Liu, Fengming, et al.
Published: (2025)
by: Liu, Fengming, et al.
Published: (2025)
Evaluating the Application of ChatGPT in Outpatient Triage Guidance: A Comparative Study
by: Liu, Dou, et al.
Published: (2024)
by: Liu, Dou, et al.
Published: (2024)
CoCo Matrix: Taxonomy of Cognitive Contributions in Co-writing with Intelligent Agents
by: Wan, Ruyuan, et al.
Published: (2024)
by: Wan, Ruyuan, et al.
Published: (2024)
Collaborative Evaluation of Deepfake Text with Deliberation-Enhancing Dialogue Systems
by: Lee, Jooyoung, et al.
Published: (2025)
by: Lee, Jooyoung, et al.
Published: (2025)
ViSP: A PPO-Driven Framework for Sarcasm Generation with Contrastive Learning
by: Wang, Changli, et al.
Published: (2025)
by: Wang, Changli, et al.
Published: (2025)
The Great AI Witch Hunt: Reviewers Perception and (Mis)Conception of Generative AI in Research Writing
by: Hadan, Hilda, et al.
Published: (2024)
by: Hadan, Hilda, et al.
Published: (2024)
Inclusion Arena: An Open Platform for Evaluating Large Foundation Models with Real-World Apps
by: Wang, Kangyu, et al.
Published: (2025)
by: Wang, Kangyu, et al.
Published: (2025)
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents
by: Zhao, Qinlin, et al.
Published: (2023)
by: Zhao, Qinlin, et al.
Published: (2023)
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
by: Wang, Zora Zhiruo, et al.
Published: (2025)
by: Wang, Zora Zhiruo, et al.
Published: (2025)
Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts
by: Tirumala, Shreyas, et al.
Published: (2025)
by: Tirumala, Shreyas, et al.
Published: (2025)
WordDecipher: Enhancing Digital Workspace Communication with Explainable AI for Non-native English Speakers
by: Chen, Yuexi, et al.
Published: (2024)
by: Chen, Yuexi, et al.
Published: (2024)
Representation Bias of Adolescents in AI: A Bilingual, Bicultural Study
by: Wolfe, Robert, et al.
Published: (2024)
by: Wolfe, Robert, et al.
Published: (2024)
BADGE: BADminton report Generation and Evaluation with LLM
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
Direct Language Model Alignment from Online AI Feedback
by: Guo, Shangmin, et al.
Published: (2024)
by: Guo, Shangmin, et al.
Published: (2024)
PleaSQLarify: Visual Pragmatic Repair for Natural Language Database Querying
by: Chan, Robin Shing Moon, et al.
Published: (2026)
by: Chan, Robin Shing Moon, et al.
Published: (2026)
A Survey on Human-AI Collaboration with Large Foundation Models
by: Vats, Vanshika, et al.
Published: (2024)
by: Vats, Vanshika, et al.
Published: (2024)
Game Plot Design with an LLM-powered Assistant: An Empirical Study with Game Designers
by: Alavi, Seyed Hossein, et al.
Published: (2024)
by: Alavi, Seyed Hossein, et al.
Published: (2024)
Similar Items
-
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses
by: Lin, Jionghao, et al.
Published: (2024) -
Improving Interactive Diagnostic Ability of a Large Language Model Agent Through Clinical Experience Learning
by: Sun, Zhoujian, et al.
Published: (2025) -
Human Bias in the Face of AI: Examining Human Judgment Against Text Labeled as AI Generated
by: Zhu, Tiffany, et al.
Published: (2024) -
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
by: Oh, Juhyun, et al.
Published: (2024) -
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
by: Clark, Nicholas, et al.
Published: (2025)