ABCD: All Biases Come Disguised
Fuente:
arXiv
Saved in:
| Main Authors: | Nowak, Mateusz, Cadet, Xavier, Chin, Peter |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RADAR: Relative Angular Divergence Across Representations
by: Cadet, Xavier, et al.
Published: (2026)
by: Cadet, Xavier, et al.
Published: (2026)
Chain-of-Thought Unfaithfulness as Disguised Accuracy
by: Bentham, Oliver, et al.
Published: (2024)
by: Bentham, Oliver, et al.
Published: (2024)
VoD-3DGS: View-opacity-Dependent 3D Gaussian Splatting
by: Nowak, Mateusz, et al.
Published: (2025)
by: Nowak, Mateusz, et al.
Published: (2025)
ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links
by: Basch, Serwar, et al.
Published: (2025)
by: Basch, Serwar, et al.
Published: (2025)
$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks
by: Chowdhary, Pratim, et al.
Published: (2025)
by: Chowdhary, Pratim, et al.
Published: (2025)
Reward-RAG: Enhancing RAG with Reward Driven Supervision
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
EqualizeIR: Mitigating Linguistic Biases in Retrieval Models
by: Cheng, Jiali, et al.
Published: (2025)
by: Cheng, Jiali, et al.
Published: (2025)
Language Models Need Inductive Biases to Count Inductively
by: Chang, Yingshan, et al.
Published: (2024)
by: Chang, Yingshan, et al.
Published: (2024)
Silenced Biases: The Dark Side LLMs Learned to Refuse
by: Himelstein, Rom, et al.
Published: (2025)
by: Himelstein, Rom, et al.
Published: (2025)
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
by: Huber, Christian, et al.
Published: (2025)
by: Huber, Christian, et al.
Published: (2025)
Controlled LLM Decoding via Discrete Auto-regressive Biasing
by: Pynadath, Patrick, et al.
Published: (2025)
by: Pynadath, Patrick, et al.
Published: (2025)
Understanding Intrinsic Socioeconomic Biases in Large Language Models
by: Arzaghi, Mina, et al.
Published: (2024)
by: Arzaghi, Mina, et al.
Published: (2024)
Problems with Chinchilla Approach 2: Systematic Biases in IsoFLOP Parabola Fits
by: Czech, Eric, et al.
Published: (2026)
by: Czech, Eric, et al.
Published: (2026)
Biased or Flawed? Mitigating Stereotypes in Generative Language Models by Addressing Task-Specific Flaws
by: Jha, Akshita, et al.
Published: (2024)
by: Jha, Akshita, et al.
Published: (2024)
On the Representational Capacity of Recurrent Neural Language Models
by: Nowak, Franz, et al.
Published: (2023)
by: Nowak, Franz, et al.
Published: (2023)
eFontes. Part of Speech Tagging and Lemmatization of Medieval Latin Texts.A Cross-Genre Survey
by: Nowak, Krzysztof, et al.
Published: (2024)
by: Nowak, Krzysztof, et al.
Published: (2024)
ASTE Transformer Modelling Dependencies in Aspect-Sentiment Triplet Extraction
by: Naglik, Iwo, et al.
Published: (2024)
by: Naglik, Iwo, et al.
Published: (2024)
Relative Value Biases in Large Language Models
by: Hayes, William M., et al.
Published: (2024)
by: Hayes, William M., et al.
Published: (2024)
Large Language Models are Biased Reinforcement Learners
by: Hayes, William M., et al.
Published: (2024)
by: Hayes, William M., et al.
Published: (2024)
Toward Automated Detection of Biased Social Signals from the Content of Clinical Conversations
by: Chen, Feng, et al.
Published: (2024)
by: Chen, Feng, et al.
Published: (2024)
A Toolbox for Surfacing Health Equity Harms and Biases in Large Language Models
by: Pfohl, Stephen R., et al.
Published: (2024)
by: Pfohl, Stephen R., et al.
Published: (2024)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
by: Tyukin, Georgy, et al.
Published: (2024)
by: Tyukin, Georgy, et al.
Published: (2024)
Large Language Models are Geographically Biased
by: Manvi, Rohin, et al.
Published: (2024)
by: Manvi, Rohin, et al.
Published: (2024)
Machine Learning Approaches for Mental Illness Detection on Social Media: A Systematic Review of Biases and Methodological Challenges
by: Cao, Yuchen, et al.
Published: (2024)
by: Cao, Yuchen, et al.
Published: (2024)
GPT-4o as the Gold Standard: A Scalable and General Purpose Approach to Filter Language Model Pretraining Data
by: Zhang, Jifan, et al.
Published: (2024)
by: Zhang, Jifan, et al.
Published: (2024)
An Algebraic View of the Expressivity of Recurrent Language Models
by: Nowak, Franz, et al.
Published: (2026)
by: Nowak, Franz, et al.
Published: (2026)
FairPair: A Robust Evaluation of Biases in Language Models through Paired Perturbations
by: Dwivedi-Yu, Jane, et al.
Published: (2024)
by: Dwivedi-Yu, Jane, et al.
Published: (2024)
Language verY Rare for All
by: Merad, Ibrahim, et al.
Published: (2024)
by: Merad, Ibrahim, et al.
Published: (2024)
On the Hidden Objective Biases of Group-based Reinforcement Learning
by: Fontana, Aleksandar, et al.
Published: (2026)
by: Fontana, Aleksandar, et al.
Published: (2026)
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
by: Yang, Xikang, et al.
Published: (2025)
by: Yang, Xikang, et al.
Published: (2025)
Self-Speculative Biased Decoding for Faster Re-Translation
by: Zeng, Linxiao, et al.
Published: (2025)
by: Zeng, Linxiao, et al.
Published: (2025)
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
by: Maheri, Mohammad M, et al.
Published: (2025)
by: Maheri, Mohammad M, et al.
Published: (2025)
Explore Reinforced: Equilibrium Approximation with Reinforcement Learning
by: Yu, Ryan, et al.
Published: (2024)
by: Yu, Ryan, et al.
Published: (2024)
Putting It All into Context: Simplifying Agents with LCLMs
by: Jiang, Mingjian, et al.
Published: (2025)
by: Jiang, Mingjian, et al.
Published: (2025)
Proofread: Fixes All Errors with One Tap
by: Liu, Renjie, et al.
Published: (2024)
by: Liu, Renjie, et al.
Published: (2024)
Heuristics and Biases in AI Decision-Making: Implications for Responsible AGI
by: Saeedi, Payam, et al.
Published: (2024)
by: Saeedi, Payam, et al.
Published: (2024)
FairFlow: Mitigating Dataset Biases through Undecided Learning
by: Cheng, Jiali, et al.
Published: (2025)
by: Cheng, Jiali, et al.
Published: (2025)
Do Large Language Models Show Biases in Causal Learning?
by: Carro, Maria Victoria, et al.
Published: (2024)
by: Carro, Maria Victoria, et al.
Published: (2024)
Are Smaller Open-Weight LLMs Closing the Gap to Proprietary Models for Biomedical Question Answering?
by: Stachura, Damian, et al.
Published: (2025)
by: Stachura, Damian, et al.
Published: (2025)
Reward Models Inherit Value Biases from Pretraining
by: Christian, Brian, et al.
Published: (2026)
by: Christian, Brian, et al.
Published: (2026)
Similar Items
-
RADAR: Relative Angular Divergence Across Representations
by: Cadet, Xavier, et al.
Published: (2026) -
Chain-of-Thought Unfaithfulness as Disguised Accuracy
by: Bentham, Oliver, et al.
Published: (2024) -
VoD-3DGS: View-opacity-Dependent 3D Gaussian Splatting
by: Nowak, Mateusz, et al.
Published: (2025) -
ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links
by: Basch, Serwar, et al.
Published: (2025) -
$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks
by: Chowdhary, Pratim, et al.
Published: (2025)