Breaking the Benchmark: Revealing LLM Bias via Minimal Contextual Augmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Miandoab, Kaveh Eskandari, Kamruzzaman, Mahammed, Gharooni, Arshia, Kim, Gene Louis, Sarathy, Vasanth, Mehrabi, Ninareh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
"Let's Argue Both Sides": Argument Generation Can Force Small Models to Utilize Previously Inaccessible Reasoning Capabilities
by: Miandoab, Kaveh Eskandari, et al.
Published: (2024)
by: Miandoab, Kaveh Eskandari, et al.
Published: (2024)
The Impact of Disability Disclosure on Fairness and Bias in LLM-Driven Candidate Selection
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
Where Norms and References Collide: Evaluating LLMs on Normative Reasoning
by: Abrams, Mitchell, et al.
Published: (2026)
by: Abrams, Mitchell, et al.
Published: (2026)
Exploring Changes in Nation Perception with Nationality-Assigned Personas in LLMs
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
Efficient Sentiment Analysis: A Resource-Aware Evaluation of Feature Extraction Techniques, Ensembling, and Deep Learning Models
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
IntelliProof: An Argumentation Network-based Conversational Helper for Organized Reflection
by: Miandoab, Kaveh Eskandari, et al.
Published: (2025)
by: Miandoab, Kaveh Eskandari, et al.
Published: (2025)
"Global is Good, Local is Bad?": Understanding Brand Bias in LLMs
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
Investigating Subtler Biases in LLMs: Ageism, Beauty, Institutional, and Nationality Bias in Generative Models
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools
by: Lymperopoulos, Panagiotis, et al.
Published: (2025)
by: Lymperopoulos, Panagiotis, et al.
Published: (2025)
Noise Injection Systemically Degrades Large Language Model Safety Guardrails
by: Shahani, Prithviraj Singh, et al.
Published: (2025)
by: Shahani, Prithviraj Singh, et al.
Published: (2025)
"A Woman is More Culturally Knowledgeable than A Man?": The Effect of Personas on Cultural Norm Interpretation in LLMs
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
From Anger to Joy: How Nationality Personas Shape Emotion Attribution in Large Language Models
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
BanStereoSet: A Dataset to Measure Stereotypical Social Biases in LLMs for Bangla
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
by: Parekh, Tanmay, et al.
Published: (2025)
by: Parekh, Tanmay, et al.
Published: (2025)
Large Language Models Know What To Say But Not When To Speak
by: Umair, Muhammad, et al.
Published: (2024)
by: Umair, Muhammad, et al.
Published: (2024)
Robust Persona-Aware Toxicity Detection with Prompt Optimization and Learned Ensembling
by: Atil, Berk, et al.
Published: (2026)
by: Atil, Berk, et al.
Published: (2026)
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Prompt Perturbation Consistency Learning for Robust Language Models
by: Qiang, Yao, et al.
Published: (2024)
by: Qiang, Yao, et al.
Published: (2024)
FERRET: Framework for Expansion Reliant Red Teaming
by: Mehrabi, Ninareh, et al.
Published: (2026)
by: Mehrabi, Ninareh, et al.
Published: (2026)
Mitigating Gender Bias in Contextual Word Embeddings
by: Yarrabelly, Navya, et al.
Published: (2024)
by: Yarrabelly, Navya, et al.
Published: (2024)
K-Edit: Language Model Editing with Contextual Knowledge Awareness
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
Tokenization Matters: Navigating Data-Scarce Tokenization for Gender Inclusive Language Technologies
by: Ovalle, Anaelia, et al.
Published: (2023)
by: Ovalle, Anaelia, et al.
Published: (2023)
RAP: Retrieval-Augmented Planning with Contextual Memory for Multimodal LLM Agents
by: Kagaya, Tomoyuki, et al.
Published: (2024)
by: Kagaya, Tomoyuki, et al.
Published: (2024)
Analogical Reasoning Within a Conceptual Hyperspace
by: Goldowsky, Howard, et al.
Published: (2024)
by: Goldowsky, Howard, et al.
Published: (2024)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
by: Li, Miaomiao, et al.
Published: (2025)
by: Li, Miaomiao, et al.
Published: (2025)
Benchmark Inflation: Revealing LLM Performance Gaps Using Retro-Holdouts
by: Haimes, Jacob, et al.
Published: (2024)
by: Haimes, Jacob, et al.
Published: (2024)
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG
by: Kermani, Arshia, et al.
Published: (2025)
by: Kermani, Arshia, et al.
Published: (2025)
Estimating Privacy Leakage of Augmented Contextual Knowledge in Language Models
by: Flemings, James, et al.
Published: (2024)
by: Flemings, James, et al.
Published: (2024)
Single Character Perturbations Break LLM Alignment
by: Lin, Leon, et al.
Published: (2024)
by: Lin, Leon, et al.
Published: (2024)
PARCO: Phoneme-Augmented Robust Contextual ASR via Contrastive Entity Disambiguation
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
User-LLM: Efficient LLM Contextualization with User Embeddings
by: Ning, Lin, et al.
Published: (2024)
by: Ning, Lin, et al.
Published: (2024)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
by: Pal, Soumyadeep, et al.
Published: (2025)
by: Pal, Soumyadeep, et al.
Published: (2025)
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
by: Ma, Sibo, et al.
Published: (2025)
by: Ma, Sibo, et al.
Published: (2025)
Break the Sequential Dependency of LLM Inference Using Lookahead Decoding
by: Fu, Yichao, et al.
Published: (2024)
by: Fu, Yichao, et al.
Published: (2024)
Does Context Matter? ContextualJudgeBench for Evaluating LLM-based Judges in Contextual Settings
by: Xu, Austin, et al.
Published: (2025)
by: Xu, Austin, et al.
Published: (2025)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
by: Li, Huihan, et al.
Published: (2025)
by: Li, Huihan, et al.
Published: (2025)
Explaining Length Bias in LLM-Based Preference Evaluations
by: Hu, Zhengyu, et al.
Published: (2024)
by: Hu, Zhengyu, et al.
Published: (2024)
Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling
by: Xiao, Tim Z., et al.
Published: (2025)
by: Xiao, Tim Z., et al.
Published: (2025)
Contextually Entangled Gradient Mapping for Optimized LLM Comprehension
by: Sisate, Colin, et al.
Published: (2025)
by: Sisate, Colin, et al.
Published: (2025)
Similar Items
-
"Let's Argue Both Sides": Argument Generation Can Force Small Models to Utilize Previously Inaccessible Reasoning Capabilities
by: Miandoab, Kaveh Eskandari, et al.
Published: (2024) -
The Impact of Disability Disclosure on Fairness and Bias in LLM-Driven Candidate Selection
by: Kamruzzaman, Mahammed, et al.
Published: (2025) -
Where Norms and References Collide: Evaluating LLMs on Normative Reasoning
by: Abrams, Mitchell, et al.
Published: (2026) -
Exploring Changes in Nation Perception with Nationality-Assigned Personas in LLMs
by: Kamruzzaman, Mahammed, et al.
Published: (2024) -
Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes
by: Kamruzzaman, Mahammed, et al.
Published: (2024)