Large language models struggle with ethnographic text annotation
Fuente:
arXiv
Saved in:
| Main Authors: | Goodall, Leonardo S., Shilton, Dor, Mullins, Daniel A., Whitehouse, Harvey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are generative AI text annotations systematically biased?
by: Stolwijk, Sjoerd B., et al.
Published: (2025)
by: Stolwijk, Sjoerd B., et al.
Published: (2025)
A thorough benchmark of automatic text classification: From traditional approaches to large language models
by: Cunha, Washington, et al.
Published: (2025)
by: Cunha, Washington, et al.
Published: (2025)
ARC-Encoder: learning compressed text representations for large language models
by: Pilchen, Hippolyte, et al.
Published: (2025)
by: Pilchen, Hippolyte, et al.
Published: (2025)
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
by: Górski, Franciszek, et al.
Published: (2025)
by: Górski, Franciszek, et al.
Published: (2025)
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
by: Nédellec, Claire, et al.
Published: (2024)
by: Nédellec, Claire, et al.
Published: (2024)
Differentially-private text generation degrades output language quality
by: Çano, Erion, et al.
Published: (2025)
by: Çano, Erion, et al.
Published: (2025)
Large language models and linguistic intentionality
by: Grindrod, Jumbly
Published: (2024)
by: Grindrod, Jumbly
Published: (2024)
Large language models in medicine: the potentials and pitfalls
by: Omiye, Jesutofunmi A., et al.
Published: (2023)
by: Omiye, Jesutofunmi A., et al.
Published: (2023)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
ks-lit-3m: A 3.1 million word kashmiri text dataset for large language model pretraining
by: Malik, Haq Nawaz
Published: (2026)
by: Malik, Haq Nawaz
Published: (2026)
Large language model empowered participatory urban planning
by: Zhou, Zhilun, et al.
Published: (2024)
by: Zhou, Zhilun, et al.
Published: (2024)
LLM_annotate: A Python package for annotating and analyzing fiction characters
by: Rosenbusch, Hannes
Published: (2025)
by: Rosenbusch, Hannes
Published: (2025)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
by: Lin, Zuhong, et al.
Published: (2025)
by: Lin, Zuhong, et al.
Published: (2025)
Large language model for Bible sentiment analysis: Sermon on the Mount
by: Vora, Mahek, et al.
Published: (2024)
by: Vora, Mahek, et al.
Published: (2024)
Large language models management of medications: three performance analyses
by: Henry, Kelli, et al.
Published: (2025)
by: Henry, Kelli, et al.
Published: (2025)
Large language models in healthcare and medical domain: A review
by: Nazi, Zabir Al, et al.
Published: (2023)
by: Nazi, Zabir Al, et al.
Published: (2023)
Augmenting emotion features in irony detection with Large language modeling
by: Lin, Yucheng, et al.
Published: (2024)
by: Lin, Yucheng, et al.
Published: (2024)
A survey of textual cyber abuse detection using cutting-edge language models and large language models
by: Diaz-Garcia, Jose A., et al.
Published: (2025)
by: Diaz-Garcia, Jose A., et al.
Published: (2025)
PLACID: Privacy-preserving Large language models for Acronym Clinical Inference and Disambiguation
by: Aithal, Manjushree B., et al.
Published: (2026)
by: Aithal, Manjushree B., et al.
Published: (2026)
Large language models show fragile cognitive reasoning about human emotions
by: Bhattacharyya, Sree, et al.
Published: (2025)
by: Bhattacharyya, Sree, et al.
Published: (2025)
Tokens, the oft-overlooked appetizer: Large language models, the distributional hypothesis, and meaning
by: Zimmerman, Julia Witte, et al.
Published: (2024)
by: Zimmerman, Julia Witte, et al.
Published: (2024)
Meta-aware Learning in text-to-SQL Large Language Model
by: Zhang, Wenda
Published: (2025)
by: Zhang, Wenda
Published: (2025)
Segment-Level Diffusion: A Framework for Controllable Long-Form Generation with Diffusion Language Models
by: Zhu, Xiaochen, et al.
Published: (2024)
by: Zhu, Xiaochen, et al.
Published: (2024)
Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators
by: Cao, Yang Trista, et al.
Published: (2023)
by: Cao, Yang Trista, et al.
Published: (2023)
Large language models for folktale type automation based on motifs: Cinderella case study
by: Arčon, Tjaša, et al.
Published: (2025)
by: Arčon, Tjaša, et al.
Published: (2025)
Social preferences with unstable interactive reasoning: Large language models in economic trust games
by: Jiamin, Ou, et al.
Published: (2025)
by: Jiamin, Ou, et al.
Published: (2025)
Outraged AI: Large language models prioritise emotion over cost in fairness enforcement
by: Liu, Hao, et al.
Published: (2025)
by: Liu, Hao, et al.
Published: (2025)
Fluent dreaming for language models
by: Thompson, T. Ben, et al.
Published: (2024)
by: Thompson, T. Ben, et al.
Published: (2024)
Algorithmic progress in language models
by: Ho, Anson, et al.
Published: (2024)
by: Ho, Anson, et al.
Published: (2024)
chDzDT: Word-level morphology-aware language model for Algerian social media text
by: Aries, Abdelkrime
Published: (2025)
by: Aries, Abdelkrime
Published: (2025)
Feedback Forensics: A Toolkit to Measure AI Personality
by: Findeis, Arduin, et al.
Published: (2025)
by: Findeis, Arduin, et al.
Published: (2025)
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System
by: Zhao, Yan, et al.
Published: (2024)
by: Zhao, Yan, et al.
Published: (2024)
Multilingual transformer and BERTopic for short text topic modeling: The case of Serbian
by: Medvecki, Darija, et al.
Published: (2024)
by: Medvecki, Darija, et al.
Published: (2024)
Scalable multilingual PII annotation for responsible AI in LLMs
by: Meena, Bharti, et al.
Published: (2025)
by: Meena, Bharti, et al.
Published: (2025)
On the attribution of confidence to large language models
by: Keeling, Geoff, et al.
Published: (2024)
by: Keeling, Geoff, et al.
Published: (2024)
Attentive Reasoning Queries: A Systematic Method for Optimizing Instruction-Following in Large Language Models
by: Karov, Bar, et al.
Published: (2025)
by: Karov, Bar, et al.
Published: (2025)
Do language models practice what they preach? Examining language ideologies about gendered language reform encoded in LLMs
by: Watson, Julia, et al.
Published: (2024)
by: Watson, Julia, et al.
Published: (2024)
Controllable and explainable personality sliders for LLMs at inference time
by: Hoppe, Florian, et al.
Published: (2026)
by: Hoppe, Florian, et al.
Published: (2026)
Superhuman performance of a large language model on the reasoning tasks of a physician
by: Brodeur, Peter G., et al.
Published: (2024)
by: Brodeur, Peter G., et al.
Published: (2024)
Linguistic traces of stochastic empathy in language models
by: Kleinberg, Bennett, et al.
Published: (2024)
by: Kleinberg, Bennett, et al.
Published: (2024)
Similar Items
-
Are generative AI text annotations systematically biased?
by: Stolwijk, Sjoerd B., et al.
Published: (2025) -
A thorough benchmark of automatic text classification: From traditional approaches to large language models
by: Cunha, Washington, et al.
Published: (2025) -
ARC-Encoder: learning compressed text representations for large language models
by: Pilchen, Hippolyte, et al.
Published: (2025) -
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
by: Górski, Franciszek, et al.
Published: (2025) -
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
by: Nédellec, Claire, et al.
Published: (2024)