Are generative AI text annotations systematically biased?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Stolwijk, Sjoerd B., Boukes, Mark, Trilling, Damian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Large language models struggle with ethnographic text annotation
von: Goodall, Leonardo S., et al.
Veröffentlicht: (2026)
von: Goodall, Leonardo S., et al.
Veröffentlicht: (2026)
Humans can learn to detect AI-generated texts, or at least learn when they can't
von: Milička, Jiří, et al.
Veröffentlicht: (2025)
von: Milička, Jiří, et al.
Veröffentlicht: (2025)
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
von: Miralles-González, Pablo, et al.
Veröffentlicht: (2025)
von: Miralles-González, Pablo, et al.
Veröffentlicht: (2025)
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
von: Nédellec, Claire, et al.
Veröffentlicht: (2024)
von: Nédellec, Claire, et al.
Veröffentlicht: (2024)
Benchmark of stylistic variation in LLM-generated texts
von: Milička, Jiří, et al.
Veröffentlicht: (2025)
von: Milička, Jiří, et al.
Veröffentlicht: (2025)
LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2026)
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2026)
Scalable multilingual PII annotation for responsible AI in LLMs
von: Meena, Bharti, et al.
Veröffentlicht: (2025)
von: Meena, Bharti, et al.
Veröffentlicht: (2025)
Can professional translators identify machine-generated text?
von: Farrell, Michael
Veröffentlicht: (2026)
von: Farrell, Michael
Veröffentlicht: (2026)
People who frequently use ChatGPT for writing tasks are accurate and robust detectors of AI-generated text
von: Russell, Jenna, et al.
Veröffentlicht: (2025)
von: Russell, Jenna, et al.
Veröffentlicht: (2025)
Kill two birds with one stone: generalized and robust AI-generated text detection via dynamic perturbations
von: Zhou, Yinghan, et al.
Veröffentlicht: (2025)
von: Zhou, Yinghan, et al.
Veröffentlicht: (2025)
Can postgraduate translation students identify machine-generated text?
von: Farrell, Michael
Veröffentlicht: (2025)
von: Farrell, Michael
Veröffentlicht: (2025)
Differentially-private text generation degrades output language quality
von: Çano, Erion, et al.
Veröffentlicht: (2025)
von: Çano, Erion, et al.
Veröffentlicht: (2025)
LLM_annotate: A Python package for annotating and analyzing fiction characters
von: Rosenbusch, Hannes
Veröffentlicht: (2025)
von: Rosenbusch, Hannes
Veröffentlicht: (2025)
ReadCtrl: Personalizing text generation with readability-controlled instruction learning
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
Reducing annotator bias by belief elicitation
von: Jakobsen, Terne Sasha Thorn, et al.
Veröffentlicht: (2024)
von: Jakobsen, Terne Sasha Thorn, et al.
Veröffentlicht: (2024)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
Beyond checkmate: exploring the creative chokepoints in AI text
von: Tripto, Nafis Irtiza, et al.
Veröffentlicht: (2025)
von: Tripto, Nafis Irtiza, et al.
Veröffentlicht: (2025)
MedReadCtrl: Personalizing medical text generation with readability-controlled instruction learning
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
From text to multimodal: a survey of adversarial example generation in question answering systems
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
Stylometry recognizes human and LLM-generated texts in short samples
von: Przystalski, Karol, et al.
Veröffentlicht: (2025)
von: Przystalski, Karol, et al.
Veröffentlicht: (2025)
<think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
Defining bias in AI-systems: Biased models are fair models
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
Autocorrect for Estonian texts: final report from project EKTB25
von: Luhtaru, Agnes, et al.
Veröffentlicht: (2024)
von: Luhtaru, Agnes, et al.
Veröffentlicht: (2024)
A systematic framework for generating novel experimental hypotheses from language models
von: Misra, Kanishka, et al.
Veröffentlicht: (2024)
von: Misra, Kanishka, et al.
Veröffentlicht: (2024)
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
von: Górski, Franciszek, et al.
Veröffentlicht: (2025)
von: Górski, Franciszek, et al.
Veröffentlicht: (2025)
CARMA: Comprehensive Automatically-annotated Reddit Mental Health Dataset for Arabic
von: Mankarious, Saad, et al.
Veröffentlicht: (2025)
von: Mankarious, Saad, et al.
Veröffentlicht: (2025)
Exploring LLM biases to manipulate AI search overview
von: Smirnov, Roman
Veröffentlicht: (2026)
von: Smirnov, Roman
Veröffentlicht: (2026)
The power of text similarity in identifying AI-LLM paraphrased documents: The case of BBC news articles and ChatGPT
von: Xylogiannopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
von: Xylogiannopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
Exploring how EFL students talk to and through AI to develop texts
von: Woo, David James, et al.
Veröffentlicht: (2026)
von: Woo, David James, et al.
Veröffentlicht: (2026)
Abusive text transformation using LLMs
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics
von: Humblot-Renaux, Galadrielle, et al.
Veröffentlicht: (2026)
von: Humblot-Renaux, Galadrielle, et al.
Veröffentlicht: (2026)
AI-generated Text Detection with a GLTR-based Approach
von: Wu, Lucía Yan, et al.
Veröffentlicht: (2025)
von: Wu, Lucía Yan, et al.
Veröffentlicht: (2025)
PAGE: Prompt Augmentation for text Generation Enhancement
von: Pacchiotti, Mauro Jose, et al.
Veröffentlicht: (2025)
von: Pacchiotti, Mauro Jose, et al.
Veröffentlicht: (2025)
Serialized EHR make for good text representations
von: Chou, Zhirong, et al.
Veröffentlicht: (2025)
von: Chou, Zhirong, et al.
Veröffentlicht: (2025)
Evaluating how LLM annotations represent diverse views on contentious topics
von: Brown, Megan A., et al.
Veröffentlicht: (2025)
von: Brown, Megan A., et al.
Veröffentlicht: (2025)
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology
von: Yu, Danni, et al.
Veröffentlicht: (2023)
von: Yu, Danni, et al.
Veröffentlicht: (2023)
To Bias or Not to Bias: Detecting bias in News with bias-detector
von: Ghosh, Himel, et al.
Veröffentlicht: (2025)
von: Ghosh, Himel, et al.
Veröffentlicht: (2025)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
On the locality bias and results in the Long Range Arena
von: Miralles-González, Pablo, et al.
Veröffentlicht: (2025)
von: Miralles-González, Pablo, et al.
Veröffentlicht: (2025)
FRACCO: A gold-standard annotated corpus of oncological entities with ICD-O-3.1 normalisation
von: Pignat, Johann, et al.
Veröffentlicht: (2025)
von: Pignat, Johann, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Large language models struggle with ethnographic text annotation
von: Goodall, Leonardo S., et al.
Veröffentlicht: (2026) -
Humans can learn to detect AI-generated texts, or at least learn when they can't
von: Milička, Jiří, et al.
Veröffentlicht: (2025) -
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
von: Miralles-González, Pablo, et al.
Veröffentlicht: (2025) -
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
von: Nédellec, Claire, et al.
Veröffentlicht: (2024) -
Benchmark of stylistic variation in LLM-generated texts
von: Milička, Jiří, et al.
Veröffentlicht: (2025)