Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Gurkan, Necdet, Njoki, Kimathi, Suchow, Jordan W. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Evaluating Large Language Models on the GMAT: Implications for the Future of Business Education
par: Ashrafimoghari, Vahid, et autres
Publié: (2024)
par: Ashrafimoghari, Vahid, et autres
Publié: (2024)
Exploring Public Opinion on Responsible AI Through The Lens of Cultural Consensus Theory
par: Gurkan, Necdet, et autres
Publié: (2024)
par: Gurkan, Necdet, et autres
Publié: (2024)
Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4
par: Uhler, Lydia, et autres
Publié: (2024)
par: Uhler, Lydia, et autres
Publié: (2024)
PaperBench: Evaluating AI's Ability to Replicate AI Research
par: Starace, Giulio, et autres
Publié: (2025)
par: Starace, Giulio, et autres
Publié: (2025)
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences
par: Nguyen, Bang, et autres
Publié: (2026)
par: Nguyen, Bang, et autres
Publié: (2026)
Classifying Human-Generated and AI-Generated Election Claims in Social Media
par: Dmonte, Alphaeus, et autres
Publié: (2024)
par: Dmonte, Alphaeus, et autres
Publié: (2024)
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions
par: Vaugrante, Laurène, et autres
Publié: (2024)
par: Vaugrante, Laurène, et autres
Publié: (2024)
TextMineX: Data, Evaluation Framework and Ontology-guided LLM Pipeline for Humanitarian Mine Action
par: Zhou, Chenyue, et autres
Publié: (2025)
par: Zhou, Chenyue, et autres
Publié: (2025)
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs
par: Colakoglu, Gaye, et autres
Publié: (2025)
par: Colakoglu, Gaye, et autres
Publié: (2025)
ConSiDERS-The-Human Evaluation Framework: Rethinking Human Evaluation for Generative Large Language Models
par: Elangovan, Aparna, et autres
Publié: (2024)
par: Elangovan, Aparna, et autres
Publié: (2024)
FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean
par: Meadows, Jordan, et autres
Publié: (2026)
par: Meadows, Jordan, et autres
Publié: (2026)
Using Large Language Models to Create AI Personas for Replication, Generalization and Prediction of Media Effects: An Empirical Test of 133 Published Experimental Research Findings
par: Yeykelis, Leo, et autres
Publié: (2024)
par: Yeykelis, Leo, et autres
Publié: (2024)
Exploring Human Perceptions of AI Responses: Insights from a Mixed-Methods Study on Risk Mitigation in Generative Models
par: Candello, Heloisa, et autres
Publié: (2025)
par: Candello, Heloisa, et autres
Publié: (2025)
Evaluating Creative Short Story Generation in Humans and Large Language Models
par: Ismayilzada, Mete, et autres
Publié: (2024)
par: Ismayilzada, Mete, et autres
Publié: (2024)
QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI
par: Veysi, Marjan, et autres
Publié: (2026)
par: Veysi, Marjan, et autres
Publié: (2026)
Measuring Human and AI Values Based on Generative Psychometrics with Large Language Models
par: Ye, Haoran, et autres
Publié: (2024)
par: Ye, Haoran, et autres
Publié: (2024)
Can we Debias Social Stereotypes in AI-Generated Images? Examining Text-to-Image Outputs and User Perceptions
par: Barve, Saharsh, et autres
Publié: (2025)
par: Barve, Saharsh, et autres
Publié: (2025)
Evaluating Neural Language Models as Cognitive Models of Language Acquisition
par: Martínez, Héctor Javier Vázquez, et autres
Publié: (2023)
par: Martínez, Héctor Javier Vázquez, et autres
Publié: (2023)
Do AI Models Perform Human-like Abstract Reasoning Across Modalities?
par: Beger, Claas, et autres
Publié: (2025)
par: Beger, Claas, et autres
Publié: (2025)
Generative AI Perceptions: A Survey to Measure the Perceptions of Faculty, Staff, and Students on Generative AI Tools in Academia
par: Amani, Sara, et autres
Publié: (2023)
par: Amani, Sara, et autres
Publié: (2023)
Perceptions of Linguistic Uncertainty by Language Models and Humans
par: Belem, Catarina G, et autres
Publié: (2024)
par: Belem, Catarina G, et autres
Publié: (2024)
Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
par: Chiang, Wei-Lin, et autres
Publié: (2024)
par: Chiang, Wei-Lin, et autres
Publié: (2024)
FinTrace: Holistic Trajectory-Level Evaluation of LLM Tool Calling for Long-Horizon Financial Tasks
par: Cao, Yupeng, et autres
Publié: (2026)
par: Cao, Yupeng, et autres
Publié: (2026)
Style over Substance: Distilled Language Models Reason Via Stylistic Replication
par: Lippmann, Philip, et autres
Publié: (2025)
par: Lippmann, Philip, et autres
Publié: (2025)
Detecting AI-Generated Sentences in Human-AI Collaborative Hybrid Texts: Challenges, Strategies, and Insights
par: Zeng, Zijie, et autres
Publié: (2024)
par: Zeng, Zijie, et autres
Publié: (2024)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
par: Garg, Madhav Krishan, et autres
Publié: (2025)
par: Garg, Madhav Krishan, et autres
Publié: (2025)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
par: Shen, Hua, et autres
Publié: (2025)
par: Shen, Hua, et autres
Publié: (2025)
Humanity in AI: Detecting the Personality of Large Language Models
par: Zhan, Baohua, et autres
Publié: (2024)
par: Zhan, Baohua, et autres
Publié: (2024)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
par: Ryan, Michael J., et autres
Publié: (2025)
par: Ryan, Michael J., et autres
Publié: (2025)
Dialogue You Can Trust: Human and AI Perspectives on Generated Conversations
par: Ebubechukwu, Ike, et autres
Publié: (2024)
par: Ebubechukwu, Ike, et autres
Publié: (2024)
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
par: Schaaff, Kristina, et autres
Publié: (2023)
par: Schaaff, Kristina, et autres
Publié: (2023)
The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations
par: Arriaga, Carlos, et autres
Publié: (2025)
par: Arriaga, Carlos, et autres
Publié: (2025)
Social Bias Benchmark for Generation: A Comparison of Generation and QA-Based Evaluations
par: Jin, Jiho, et autres
Publié: (2025)
par: Jin, Jiho, et autres
Publié: (2025)
Evaluating LLMs with Multiple Problems at once
par: Wang, Zhengxiang, et autres
Publié: (2024)
par: Wang, Zhengxiang, et autres
Publié: (2024)
Interpretable Predictability-Based AI Text Detection: A Replication Study
par: Skurla, Adam, et autres
Publié: (2026)
par: Skurla, Adam, et autres
Publié: (2026)
Measuring Human Contribution in AI-Assisted Content Generation
par: Xie, Yueqi, et autres
Publié: (2024)
par: Xie, Yueqi, et autres
Publié: (2024)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
par: Fayyaz, Mohsen, et autres
Publié: (2024)
par: Fayyaz, Mohsen, et autres
Publié: (2024)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
par: Raza, Shaina, et autres
Publié: (2024)
par: Raza, Shaina, et autres
Publié: (2024)
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
par: AlDakheel, Abdulaziz, et autres
Publié: (2026)
par: AlDakheel, Abdulaziz, et autres
Publié: (2026)
Detecting Machine-Generated Texts: Not Just "AI vs Humans" and Explainability is Complicated
par: Ji, Jiazhou, et autres
Publié: (2024)
par: Ji, Jiazhou, et autres
Publié: (2024)
Documents similaires
-
Evaluating Large Language Models on the GMAT: Implications for the Future of Business Education
par: Ashrafimoghari, Vahid, et autres
Publié: (2024) -
Exploring Public Opinion on Responsible AI Through The Lens of Cultural Consensus Theory
par: Gurkan, Necdet, et autres
Publié: (2024) -
Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4
par: Uhler, Lydia, et autres
Publié: (2024) -
PaperBench: Evaluating AI's Ability to Replicate AI Research
par: Starace, Giulio, et autres
Publié: (2025) -
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences
par: Nguyen, Bang, et autres
Publié: (2026)