Theory Trace Card: Theory-Driven Socio-Cognitive Evaluation of LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Karimi-Malekabadi, Farzan, Abdurahman, Suhaib, Sourati, Zhivar, Trager, Jackson, Dehghani, Morteza |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Realistic threat perception drives intergroup conflict: A causal, dynamic analysis using generative-agent simulations
par: Abdurahman, Suhaib, et autres
Publié: (2025)
par: Abdurahman, Suhaib, et autres
Publié: (2025)
Tracing Moral Foundations in Large Language Models
par: Yu, Chenxiao, et autres
Publié: (2026)
par: Yu, Chenxiao, et autres
Publié: (2026)
Reasoning on a Spectrum: Aligning LLMs to System 1 and System 2 Thinking
par: Ziabari, Alireza S., et autres
Publié: (2025)
par: Ziabari, Alireza S., et autres
Publié: (2025)
The Shrinking Landscape of Linguistic Diversity in the Age of Large Language Models
par: Sourati, Zhivar, et autres
Publié: (2025)
par: Sourati, Zhivar, et autres
Publié: (2025)
The Homogenizing Effect of Large Language Models on Human Expression and Thought
par: Sourati, Zhivar, et autres
Publié: (2025)
par: Sourati, Zhivar, et autres
Publié: (2025)
Scaling Item-to-Standard Alignment with Large Language Models: Accuracy, Limits, and Solutions
par: Karimi-Malekabadi, Farzan, et autres
Publié: (2025)
par: Karimi-Malekabadi, Farzan, et autres
Publié: (2025)
The Limits of Goal-Setting Theory in LLM-Driven Assessment
par: Kumar, Mrityunjay
Publié: (2025)
par: Kumar, Mrityunjay
Publié: (2025)
EvalCards: A Framework for Standardized Evaluation Reporting
par: Dhar, Ruchira, et autres
Publié: (2025)
par: Dhar, Ruchira, et autres
Publié: (2025)
Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training
par: Rahmati, Elnaz, et autres
Publié: (2026)
par: Rahmati, Elnaz, et autres
Publié: (2026)
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs
par: Wachter, Jasmin, et autres
Publié: (2025)
par: Wachter, Jasmin, et autres
Publié: (2025)
Red Teaming LLMs as Socio-Technical Practice: From Exploration and Data Creation to Evaluation
par: Garcia, Adriana Alvarado, et autres
Publié: (2026)
par: Garcia, Adriana Alvarado, et autres
Publié: (2026)
AI and Social Theory
par: Mokander, Jakob, et autres
Publié: (2024)
par: Mokander, Jakob, et autres
Publié: (2024)
Evaluation Cards for XAI Metrics
par: Gipiškis, Rokas, et autres
Publié: (2026)
par: Gipiškis, Rokas, et autres
Publié: (2026)
Analyzing User Characteristics of Hate Speech Spreaders on Social Media
par: Geissler, Dominique, et autres
Publié: (2023)
par: Geissler, Dominique, et autres
Publié: (2023)
From Vision to Validation: A Theory- and Data-Driven Construction of a GCC-Specific AI Adoption Index
par: Albous, Mohammad Rashed, et autres
Publié: (2025)
par: Albous, Mohammad Rashed, et autres
Publié: (2025)
Leveraging Pedagogical Theories to Understand Student Learning Process with Graph-based Reasonable Knowledge Tracing
par: Cui, Jiajun, et autres
Publié: (2024)
par: Cui, Jiajun, et autres
Publié: (2024)
Industrial AI Robustness Card for Time Series Models
par: Windmann, Alexander, et autres
Publié: (2025)
par: Windmann, Alexander, et autres
Publié: (2025)
Mapping Social Choice Theory to RLHF
par: Dai, Jessica, et autres
Publié: (2024)
par: Dai, Jessica, et autres
Publié: (2024)
Can A Cognitive Architecture Fundamentally Enhance LLMs? Or Vice Versa?
par: Sun, Ron
Publié: (2024)
par: Sun, Ron
Publié: (2024)
Position: AI Evaluations Should be Grounded on a Theory of Capability
par: Jo, Nathanael, et autres
Publié: (2025)
par: Jo, Nathanael, et autres
Publié: (2025)
Towards Sustainability Model Cards
par: Jouneaux, Gwendal, et autres
Publié: (2025)
par: Jouneaux, Gwendal, et autres
Publié: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
par: Bucknall, Ben, et autres
Publié: (2025)
par: Bucknall, Ben, et autres
Publié: (2025)
Assessing Cognitive Biases in LLMs for Judicial Decision Support: Virtuous Victim and Halo Effects
par: Liu, Sierra S.
Publié: (2026)
par: Liu, Sierra S.
Publié: (2026)
Dialogue with the Machine and Dialogue with the Art World: Evaluating Generative AI for Culturally-Situated Creativity
par: Qadri, Rida, et autres
Publié: (2024)
par: Qadri, Rida, et autres
Publié: (2024)
Survey on AI Ethics: A Socio-technical Perspective
par: Mbiazi, Dave, et autres
Publié: (2023)
par: Mbiazi, Dave, et autres
Publié: (2023)
The Moral Foundations Reddit Corpus
par: Trager, Jackson, et autres
Publié: (2022)
par: Trager, Jackson, et autres
Publié: (2022)
From Perceptions to Decisions: Wildfire Evacuation Decision Prediction with Behavioral Theory-informed LLMs
par: Chen, Ruxiao, et autres
Publié: (2025)
par: Chen, Ruxiao, et autres
Publié: (2025)
AI Literacy for All: Adjustable Interdisciplinary Socio-technical Curriculum
par: Tadimalla, Sri Yash, et autres
Publié: (2024)
par: Tadimalla, Sri Yash, et autres
Publié: (2024)
Homoglyph-based Adversarial Perturbation of Introductory Computer Science Theory Problems
par: Alexander, Aidan, et autres
Publié: (2026)
par: Alexander, Aidan, et autres
Publié: (2026)
Reconciling Different Theories of Learning with an Agent-based Model of Procedural Learning
par: Rismanchian, Sina, et autres
Publié: (2024)
par: Rismanchian, Sina, et autres
Publié: (2024)
Using AI Alignment Theory to understand the potential pitfalls of regulatory frameworks
par: Tlaie, Alejandro
Publié: (2024)
par: Tlaie, Alejandro
Publié: (2024)
Simulating Generative Social Agents via Theory-Informed Workflow Design
par: Yan, Yuwei, et autres
Publié: (2025)
par: Yan, Yuwei, et autres
Publié: (2025)
LLM-Driven Personalized Answer Generation and Evaluation
par: Molavi, Mohammadreza, et autres
Publié: (2025)
par: Molavi, Mohammadreza, et autres
Publié: (2025)
ZPD-SCA: Unveiling the Blind Spots of LLMs in Assessing Students' Cognitive Abilities
par: Dong, Wenhan, et autres
Publié: (2025)
par: Dong, Wenhan, et autres
Publié: (2025)
Exploring Public Opinion on Responsible AI Through The Lens of Cultural Consensus Theory
par: Gurkan, Necdet, et autres
Publié: (2024)
par: Gurkan, Necdet, et autres
Publié: (2024)
AI Cards: Towards an Applied Framework for Machine-Readable AI and Risk Documentation Inspired by the EU AI Act
par: Golpayegani, Delaram, et autres
Publié: (2024)
par: Golpayegani, Delaram, et autres
Publié: (2024)
Navigating Pitfalls: Evaluating LLMs in Machine Learning Programming Education
par: Kumar, Smitha, et autres
Publié: (2025)
par: Kumar, Smitha, et autres
Publié: (2025)
Evaluating Code Generation of LLMs in Advanced Computer Science Problems
par: Catir, Emir, et autres
Publié: (2025)
par: Catir, Emir, et autres
Publié: (2025)
AI Application Operations -- A Socio-Technical Framework for Data-driven Organizations
par: Jönsson, Daniel, et autres
Publié: (2025)
par: Jönsson, Daniel, et autres
Publié: (2025)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
par: Caputo, Nicholas A.
Publié: (2024)
par: Caputo, Nicholas A.
Publié: (2024)
Documents similaires
-
Realistic threat perception drives intergroup conflict: A causal, dynamic analysis using generative-agent simulations
par: Abdurahman, Suhaib, et autres
Publié: (2025) -
Tracing Moral Foundations in Large Language Models
par: Yu, Chenxiao, et autres
Publié: (2026) -
Reasoning on a Spectrum: Aligning LLMs to System 1 and System 2 Thinking
par: Ziabari, Alireza S., et autres
Publié: (2025) -
The Shrinking Landscape of Linguistic Diversity in the Age of Large Language Models
par: Sourati, Zhivar, et autres
Publié: (2025) -
The Homogenizing Effect of Large Language Models on Human Expression and Thought
par: Sourati, Zhivar, et autres
Publié: (2025)