IDEAlign: Comparing Large Language Models to Human Experts in Open-ended Interpretive Annotations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nam, Hyunji, Langlois, Lucia, Malamut, James, Tan, Mei, Demszky, Dorottya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EduCoder: An Open-Source Annotation System for Education Transcript Data
von: Ashraf, Saad, et al.
Veröffentlicht: (2025)
von: Ashraf, Saad, et al.
Veröffentlicht: (2025)
Mitigating LLM biases toward spurious social contexts using direct preference optimization
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
Marked Pedagogies: Examining Linguistic Biases in Personalized Automated Writing Feedback
von: Tan, Mei, et al.
Veröffentlicht: (2026)
von: Tan, Mei, et al.
Veröffentlicht: (2026)
Mapping the Methodological Space of Classroom Interaction Research: Scale, Duration, and Modality in an Age of AI
von: Demszky, Dorottya, et al.
Veröffentlicht: (2026)
von: Demszky, Dorottya, et al.
Veröffentlicht: (2026)
Edu-ConvoKit: An Open-Source Library for Education Conversation Data
von: Wang, Rose E., et al.
Veröffentlicht: (2024)
von: Wang, Rose E., et al.
Veröffentlicht: (2024)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
Measuring Large Language Models Capacity to Annotate Journalistic Sourcing
von: Vincent, Subramaniam, et al.
Veröffentlicht: (2024)
von: Vincent, Subramaniam, et al.
Veröffentlicht: (2024)
TeachLM: Post-Training LLMs for Education Using Authentic Learning Data
von: Perczel, Janos, et al.
Veröffentlicht: (2025)
von: Perczel, Janos, et al.
Veröffentlicht: (2025)
Bridging the Novice-Expert Gap via Models of Decision-Making: A Case Study on Remediating Math Mistakes
von: Wang, Rose E., et al.
Veröffentlicht: (2023)
von: Wang, Rose E., et al.
Veröffentlicht: (2023)
Open Models, Closed Minds? On Agents Capabilities in Mimicking Human Personalities through Open Large Language Models
von: La Cava, Lucio, et al.
Veröffentlicht: (2024)
von: La Cava, Lucio, et al.
Veröffentlicht: (2024)
LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
von: Faiz, Ahmad, et al.
Veröffentlicht: (2023)
von: Faiz, Ahmad, et al.
Veröffentlicht: (2023)
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
von: Lin, Hao, et al.
Veröffentlicht: (2025)
von: Lin, Hao, et al.
Veröffentlicht: (2025)
Hypothesis Generation with Large Language Models
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
Better Call GPT, Comparing Large Language Models Against Lawyers
von: Martin, Lauren, et al.
Veröffentlicht: (2024)
von: Martin, Lauren, et al.
Veröffentlicht: (2024)
Tell, Don't Show: Leveraging Language Models' Abstractive Retellings to Model Literary Themes
von: Lucy, Li, et al.
Veröffentlicht: (2025)
von: Lucy, Li, et al.
Veröffentlicht: (2025)
Human vs. Machine: Behavioral Differences Between Expert Humans and Language Models in Wargame Simulations
von: Lamparth, Max, et al.
Veröffentlicht: (2024)
von: Lamparth, Max, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Predictive Analysis of Human Misery
von: Seal, Bishanka, et al.
Veröffentlicht: (2025)
von: Seal, Bishanka, et al.
Veröffentlicht: (2025)
Large Language Models Approach Expert Pedagogical Quality in Math Tutoring but Differ in Instructional and Linguistic Profiles
von: Abdulsalam, Ramatu Oiza, et al.
Veröffentlicht: (2025)
von: Abdulsalam, Ramatu Oiza, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models Against Human Annotators in Latent Content Analysis: Sentiment, Political Leaning, Emotional Intensity, and Sarcasm
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2025)
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2025)
Open-Ended Wargames with Large Language Models
von: Hogan, Daniel P., et al.
Veröffentlicht: (2024)
von: Hogan, Daniel P., et al.
Veröffentlicht: (2024)
Evaluating Digital Inclusiveness of Digital Agri-Food Tools Using Large Language Models: A Comparative Analysis Between Human and AI-Based Evaluations
von: Pewinya, Githma, et al.
Veröffentlicht: (2026)
von: Pewinya, Githma, et al.
Veröffentlicht: (2026)
Strategic Insights in Human and Large Language Model Tactics at Word Guessing Games
von: Rikters, Matīss, et al.
Veröffentlicht: (2024)
von: Rikters, Matīss, et al.
Veröffentlicht: (2024)
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
Prompt Selection Matters: Enhancing Text Annotations for Social Sciences with Large Language Models
von: Abraham, Louis, et al.
Veröffentlicht: (2024)
von: Abraham, Louis, et al.
Veröffentlicht: (2024)
Using LLMs for Knowledge Component-level Correctness Labeling in Open-ended Coding Problems
von: Duan, Zhangqi, et al.
Veröffentlicht: (2026)
von: Duan, Zhangqi, et al.
Veröffentlicht: (2026)
AuditWen:An Open-Source Large Language Model for Audit
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
Large Language Models' Accuracy in Emulating Human Experts' Evaluation of Public Sentiments about Heated Tobacco Products on Social Media
von: Kim, Kwanho, et al.
Veröffentlicht: (2025)
von: Kim, Kwanho, et al.
Veröffentlicht: (2025)
Assessing the Impact of Conspiracy Theories Using Large Language Models
von: Jiang, Bohan, et al.
Veröffentlicht: (2024)
von: Jiang, Bohan, et al.
Veröffentlicht: (2024)
Human-Level and Beyond: Benchmarking Large Language Models Against Clinical Pharmacists in Prescription Review
von: Yang, Yan, et al.
Veröffentlicht: (2025)
von: Yang, Yan, et al.
Veröffentlicht: (2025)
HerO at AVeriTeC: The Herd of Open Large Language Models for Verifying Real-World Claims
von: Yoon, Yejun, et al.
Veröffentlicht: (2024)
von: Yoon, Yejun, et al.
Veröffentlicht: (2024)
The Human Condition as Reflected in Contemporary Large Language Models
von: Neuman, W. Russell
Veröffentlicht: (2026)
von: Neuman, W. Russell
Veröffentlicht: (2026)
CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China
von: Sun, Bolun, et al.
Veröffentlicht: (2025)
von: Sun, Bolun, et al.
Veröffentlicht: (2025)
Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation
von: Nowak, Sebastian, et al.
Veröffentlicht: (2026)
von: Nowak, Sebastian, et al.
Veröffentlicht: (2026)
A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses
von: Jiang, Yan, et al.
Veröffentlicht: (2026)
von: Jiang, Yan, et al.
Veröffentlicht: (2026)
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models
von: Chanenson, Jake, et al.
Veröffentlicht: (2023)
von: Chanenson, Jake, et al.
Veröffentlicht: (2023)
Motivation in Large Language Models
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
Test Case-Informed Knowledge Tracing for Open-ended Coding Tasks
von: Duan, Zhangqi, et al.
Veröffentlicht: (2024)
von: Duan, Zhangqi, et al.
Veröffentlicht: (2024)
Interpreting Public Sentiment in Diplomacy Events: A Counterfactual Analysis Framework Using Large Language Models
von: Ouyang, Leyi
Veröffentlicht: (2025)
von: Ouyang, Leyi
Veröffentlicht: (2025)
Language of Thought Shapes Output Diversity in Large Language Models
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
Evaluating Large Language Models as Expert Annotators
von: Tseng, Yu-Min, et al.
Veröffentlicht: (2025)
von: Tseng, Yu-Min, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EduCoder: An Open-Source Annotation System for Education Transcript Data
von: Ashraf, Saad, et al.
Veröffentlicht: (2025) -
Mitigating LLM biases toward spurious social contexts using direct preference optimization
von: Nam, Hyunji, et al.
Veröffentlicht: (2026) -
Marked Pedagogies: Examining Linguistic Biases in Personalized Automated Writing Feedback
von: Tan, Mei, et al.
Veröffentlicht: (2026) -
Mapping the Methodological Space of Classroom Interaction Research: Scale, Duration, and Modality in an Age of AI
von: Demszky, Dorottya, et al.
Veröffentlicht: (2026) -
Edu-ConvoKit: An Open-Source Library for Education Conversation Data
von: Wang, Rose E., et al.
Veröffentlicht: (2024)