Undesirable Biases in NLP: Addressing Challenges of Measurement
Fuente:
arXiv
Saved in:
| Main Authors: | van der Wal, Oskar, Bachmann, Dominik, Leidinger, Alina, van Maanen, Leendert, Zuidema, Willem, Schulz, Katrin |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are LLMs classical or nonmonotonic reasoners? Lessons from generics
by: Leidinger, Alina, et al.
Published: (2024)
by: Leidinger, Alina, et al.
Published: (2024)
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
Facilitating Opinion Diversity through Hybrid NLP Approaches
by: van der Meer, Michiel
Published: (2024)
by: van der Meer, Michiel
Published: (2024)
In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads
by: Pouw, Charlotte, et al.
Published: (2026)
by: Pouw, Charlotte, et al.
Published: (2026)
CIVICS: Building a Dataset for Examining Culturally-Informed Values in Large Language Models
by: Pistilli, Giada, et al.
Published: (2024)
by: Pistilli, Giada, et al.
Published: (2024)
Tokenization and Representation Biases in Multilingual Models on Dialectal NLP Tasks
by: Kanjirangat, Vani, et al.
Published: (2025)
by: Kanjirangat, Vani, et al.
Published: (2025)
Undesirable Memorization in Large Language Models: A Survey
by: Satvaty, Ali, et al.
Published: (2024)
by: Satvaty, Ali, et al.
Published: (2024)
Misaligned by Reward: Socially Undesirable Preferences in LLMs
by: Ghazaryan, Gayane, et al.
Published: (2026)
by: Ghazaryan, Gayane, et al.
Published: (2026)
EUDAIMONIA: Evaluating Undesirable Dynamics in AI
by: Huang, Jun Rui, et al.
Published: (2026)
by: Huang, Jun Rui, et al.
Published: (2026)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
Creativity in AI: Progresses and Challenges
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Sequence models for by-trial decoding of cognitive strategies from neural data
by: Otter, Rick den, et al.
Published: (2025)
by: Otter, Rick den, et al.
Published: (2025)
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
by: van der Wal, Oskar, et al.
Published: (2025)
by: van der Wal, Oskar, et al.
Published: (2025)
We Need to Measure Data Diversity in NLP -- Better and Broader
by: Nguyen, Dong, et al.
Published: (2025)
by: Nguyen, Dong, et al.
Published: (2025)
NLP Methods May Actually Be Better Than Professors at Estimating Question Difficulty
by: Zotos, Leonidas, et al.
Published: (2025)
by: Zotos, Leonidas, et al.
Published: (2025)
Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference
by: Zhao, Stephen, et al.
Published: (2025)
by: Zhao, Stephen, et al.
Published: (2025)
Can we Evaluate RAGs with Synthetic Data?
by: van Elburg, Jonas, et al.
Published: (2025)
by: van Elburg, Jonas, et al.
Published: (2025)
How Are LLMs Mitigating Stereotyping Harms? Learning from Search Engine Studies
by: Leidinger, Alina, et al.
Published: (2024)
by: Leidinger, Alina, et al.
Published: (2024)
Addressing Longstanding Challenges in Cognitive Science with Language Models
by: Wulff, Dirk U., et al.
Published: (2025)
by: Wulff, Dirk U., et al.
Published: (2025)
Investigating the Role of Prompting and External Tools in Hallucination Rates of Large Language Models
by: Barkley, Liam, et al.
Published: (2024)
by: Barkley, Liam, et al.
Published: (2024)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
by: Fanconi, Claudio, et al.
Published: (2025)
by: Fanconi, Claudio, et al.
Published: (2025)
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge
by: Vadlapati, Praneeth
Published: (2024)
by: Vadlapati, Praneeth
Published: (2024)
An exploration of the effect of quantisation on energy consumption and inference time of StarCoder2
by: de Reus, Pepijn, et al.
Published: (2024)
by: de Reus, Pepijn, et al.
Published: (2024)
Frame of Reference: Addressing the Challenges of Common Ground Representation in Situational Dialogs
by: Mohapatra, Biswesh, et al.
Published: (2026)
by: Mohapatra, Biswesh, et al.
Published: (2026)
Tracking the emergence of linguistic structure in self-supervised models learning from speech
by: Kloots, Marianne de Heer, et al.
Published: (2026)
by: Kloots, Marianne de Heer, et al.
Published: (2026)
Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient Tracing
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025)
by: Lent, Heather, et al.
Published: (2025)
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
State of NLP in Kenya: A Survey
by: Amol, Cynthia Jayne, et al.
Published: (2024)
by: Amol, Cynthia Jayne, et al.
Published: (2024)
What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training
by: Kloots, Marianne de Heer, et al.
Published: (2025)
by: Kloots, Marianne de Heer, et al.
Published: (2025)
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
Multi-Stage Balanced Distillation: Addressing Long-Tail Challenges in Sequence-Level Knowledge Distillation
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
Scaling few-shot spoken word classification with generative meta-continual learning
by: Beyers, Louise, et al.
Published: (2026)
by: Beyers, Louise, et al.
Published: (2026)
SocialNLP Fake-EmoReact 2021 Challenge Overview: Predicting Fake Tweets from Their Replies and GIFs
by: Huang, Chien-Kun, et al.
Published: (2024)
by: Huang, Chien-Kun, et al.
Published: (2024)
RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
by: Yuan, Zhuowen, et al.
Published: (2024)
by: Yuan, Zhuowen, et al.
Published: (2024)
Evaluation Metrics for Text Data Augmentation in NLP
by: Amadeus, Marcellus, et al.
Published: (2024)
by: Amadeus, Marcellus, et al.
Published: (2024)
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026)
by: Purificato, Antonio, et al.
Published: (2026)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
by: Manna, Supriya, et al.
Published: (2024)
by: Manna, Supriya, et al.
Published: (2024)
Similar Items
-
Are LLMs classical or nonmonotonic reasoners? Lessons from generics
by: Leidinger, Alina, et al.
Published: (2024) -
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024) -
Facilitating Opinion Diversity through Hybrid NLP Approaches
by: van der Meer, Michiel
Published: (2024) -
In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads
by: Pouw, Charlotte, et al.
Published: (2026) -
CIVICS: Building a Dataset for Examining Culturally-Informed Values in Large Language Models
by: Pistilli, Giada, et al.
Published: (2024)