Text Annotation via Inductive Coding: Comparing Human Experts to LLMs in Qualitative Data Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Parfenova, Angelina, Marfurt, Andreas, Denzler, Alexander, Pfeffer, Juergen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Emergent Convergence in Multi-Agent LLM Annotation
by: Parfenova, Angelina, et al.
Published: (2025)
by: Parfenova, Angelina, et al.
Published: (2025)
From Quotes to Concepts: Axial Coding of Political Debates with Ensemble LMs
by: Parfenova, Angelina, et al.
Published: (2026)
by: Parfenova, Angelina, et al.
Published: (2026)
GROUNDEDKG-RAG: Grounded Knowledge Graph Index for Long-document Question Answering
by: Zhang, Tianyi, et al.
Published: (2026)
by: Zhang, Tianyi, et al.
Published: (2026)
Prompts Matter: Comparing ML/GAI Approaches for Generating Inductive Qualitative Coding Results
by: Chen, John, et al.
Published: (2024)
by: Chen, John, et al.
Published: (2024)
HICode: Hierarchical Inductive Coding with LLMs
by: Zhong, Mian, et al.
Published: (2025)
by: Zhong, Mian, et al.
Published: (2025)
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data
by: Nemkova, Poli Apollinaire, et al.
Published: (2025)
by: Nemkova, Poli Apollinaire, et al.
Published: (2025)
Unveiling Divergent Inductive Biases of LLMs on Temporal Data
by: Kishore, Sindhu, et al.
Published: (2024)
by: Kishore, Sindhu, et al.
Published: (2024)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Automating the Information Extraction from Semi-Structured Interview Transcripts
by: Parfenova, Angelina
Published: (2024)
by: Parfenova, Angelina
Published: (2024)
Scalable Qualitative Coding with LLMs: Chain-of-Thought Reasoning Matches Human Performance in Some Hermeneutic Tasks
by: Dunivin, Zackary Okun
Published: (2024)
by: Dunivin, Zackary Okun
Published: (2024)
COMMENTATOR: A Code-mixed Multilingual Text Annotation Framework
by: Sheth, Rajvee, et al.
Published: (2024)
by: Sheth, Rajvee, et al.
Published: (2024)
A Comparative Analysis of Instruction Fine-Tuning LLMs for Financial Text Classification
by: Fatemi, Sorouralsadat, et al.
Published: (2024)
by: Fatemi, Sorouralsadat, et al.
Published: (2024)
Conditioning LLMs to Generate Code-Switched Text
by: Heredia, Maite, et al.
Published: (2025)
by: Heredia, Maite, et al.
Published: (2025)
The Effectiveness of LLMs as Annotators: A Comparative Overview and Empirical Analysis of Direct Representation
by: Pavlovic, Maja, et al.
Published: (2024)
by: Pavlovic, Maja, et al.
Published: (2024)
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
by: Sheth, Rajvee, et al.
Published: (2025)
by: Sheth, Rajvee, et al.
Published: (2025)
Risk prediction of pathological gambling on social media
by: Parfenova, Angelina, et al.
Published: (2024)
by: Parfenova, Angelina, et al.
Published: (2024)
Inductive Entity Representations from Text via Link Prediction
by: Daza, Daniel, et al.
Published: (2020)
by: Daza, Daniel, et al.
Published: (2020)
An Expert is Worth One Token: Synergizing Multiple Expert LLMs as Generalist via Expert Token Routing
by: Chai, Ziwei, et al.
Published: (2024)
by: Chai, Ziwei, et al.
Published: (2024)
From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?
by: Amin, Hasan, et al.
Published: (2026)
by: Amin, Hasan, et al.
Published: (2026)
Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection
by: Hakimi, Ahmad Dawar, et al.
Published: (2026)
by: Hakimi, Ahmad Dawar, et al.
Published: (2026)
LLMs for Science: Usage for Code Generation and Data Analysis
by: Nejjar, Mohamed, et al.
Published: (2023)
by: Nejjar, Mohamed, et al.
Published: (2023)
From Variance to Invariance: Qualitative Content Analysis for Narrative Graph Annotation
by: Huang, Junbo, et al.
Published: (2026)
by: Huang, Junbo, et al.
Published: (2026)
The AI Co-Ethnographer: How Far Can Automation Take Qualitative Research?
by: Retkowski, Fabian, et al.
Published: (2025)
by: Retkowski, Fabian, et al.
Published: (2025)
Are Chatbots Reliable Text Annotators? Sometimes
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
Thematic Analysis with Open-Source Generative AI and Machine Learning: A New Method for Inductive Qualitative Codebook Development
by: Katz, Andrew, et al.
Published: (2024)
by: Katz, Andrew, et al.
Published: (2024)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
by: Calderon, Nitay, et al.
Published: (2025)
by: Calderon, Nitay, et al.
Published: (2025)
Are LLMs Ready to Replace Bangla Annotators?
by: Hasan, Md. Najib, et al.
Published: (2026)
by: Hasan, Md. Najib, et al.
Published: (2026)
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective
by: Xing, Jinming, et al.
Published: (2024)
by: Xing, Jinming, et al.
Published: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
by: Niu, Tianyi, et al.
Published: (2026)
by: Niu, Tianyi, et al.
Published: (2026)
Inductive Learning of Logical Theories with LLMs: An Expressivity-Graded Analysis
by: Gandarela, João Pedro, et al.
Published: (2024)
by: Gandarela, João Pedro, et al.
Published: (2024)
Multi-task Code LLMs: Data Mix or Model Merge?
by: Zhu, Mingzhi, et al.
Published: (2026)
by: Zhu, Mingzhi, et al.
Published: (2026)
A Comparative Analysis of Counterfactual Explanation Methods for Text Classifiers
by: McAleese, Stephen, et al.
Published: (2024)
by: McAleese, Stephen, et al.
Published: (2024)
Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators
by: Gu, Feng, et al.
Published: (2025)
by: Gu, Feng, et al.
Published: (2025)
Conversational Process Modeling: Can Generative AI Empower Domain Experts in Creating and Redesigning Process Models?
by: Klievtsova, Nataliia, et al.
Published: (2023)
by: Klievtsova, Nataliia, et al.
Published: (2023)
From Human Annotation to Automation: LLM-in-the-Loop Active Learning for Arabic Sentiment Analysis
by: Refai, Dania, et al.
Published: (2025)
by: Refai, Dania, et al.
Published: (2025)
DACO: Towards Application-Driven and Comprehensive Data Analysis via Code Generation
by: Wu, Xueqing, et al.
Published: (2024)
by: Wu, Xueqing, et al.
Published: (2024)
Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
by: Baumann, Joachim, et al.
Published: (2025)
by: Baumann, Joachim, et al.
Published: (2025)
Stylometry Analysis of Human and Machine Text for Academic Integrity
by: Albaqami, Hezam, et al.
Published: (2026)
by: Albaqami, Hezam, et al.
Published: (2026)
Dr. SoW: Density Ratio of Strong-over-weak LLMs for Reducing the Cost of Human Annotation in Preference Tuning
by: Xu, Guangxuan, et al.
Published: (2024)
by: Xu, Guangxuan, et al.
Published: (2024)
Human and LLM Biases in Hate Speech Annotations: A Socio-Demographic Analysis of Annotators and Targets
by: Giorgi, Tommaso, et al.
Published: (2024)
by: Giorgi, Tommaso, et al.
Published: (2024)
Similar Items
-
Emergent Convergence in Multi-Agent LLM Annotation
by: Parfenova, Angelina, et al.
Published: (2025) -
From Quotes to Concepts: Axial Coding of Political Debates with Ensemble LMs
by: Parfenova, Angelina, et al.
Published: (2026) -
GROUNDEDKG-RAG: Grounded Knowledge Graph Index for Long-document Question Answering
by: Zhang, Tianyi, et al.
Published: (2026) -
Prompts Matter: Comparing ML/GAI Approaches for Generating Inductive Qualitative Coding Results
by: Chen, John, et al.
Published: (2024) -
HICode: Hierarchical Inductive Coding with LLMs
by: Zhong, Mian, et al.
Published: (2025)