Saved in:
| Main Authors: | Dong, Wenchao, Zhunis, Assem, Chin, Hyojin, Han, Jiyoung, Cha, Meeyoung |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.10436 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Persona Setting Pitfall: Persistent Outgroup Biases in Large Language Models Arising from Social Identity Adoption
by: Dong, Wenchao, et al.
Published: (2024)
by: Dong, Wenchao, et al.
Published: (2024)
Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions
by: Kim, Jiseon, et al.
Published: (2026)
by: Kim, Jiseon, et al.
Published: (2026)
Adversarial Style Augmentation via Large Language Model for Robust Fake News Detection
by: Park, Sungwon, et al.
Published: (2024)
by: Park, Sungwon, et al.
Published: (2024)
EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models
by: Lee, Donggyu, et al.
Published: (2025)
by: Lee, Donggyu, et al.
Published: (2025)
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
by: Kim, Minsung, et al.
Published: (2025)
by: Kim, Minsung, et al.
Published: (2025)
How You Ask Matters! Adaptive RAG Robustness to Query Variations
by: Jang, Yunah, et al.
Published: (2026)
by: Jang, Yunah, et al.
Published: (2026)
Generalization Bias in Large Language Model Summarization of Scientific Research
by: Peters, Uwe, et al.
Published: (2025)
by: Peters, Uwe, et al.
Published: (2025)
Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
by: Kwon, Jea, et al.
Published: (2025)
by: Kwon, Jea, et al.
Published: (2025)
Bias after Prompting: Persistent Discrimination in Large Language Models
by: Sivakumar, Nivedha, et al.
Published: (2025)
by: Sivakumar, Nivedha, et al.
Published: (2025)
Enhancing Contextual Understanding in Large Language Models through Contrastive Decoding
by: Zhao, Zheng, et al.
Published: (2024)
by: Zhao, Zheng, et al.
Published: (2024)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
by: Kim, Jiseon, et al.
Published: (2025)
by: Kim, Jiseon, et al.
Published: (2025)
I Am Aligned, But With Whom? MENA Values Benchmark for Evaluating Cultural Alignment and Multilingual Bias in LLMs
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
by: Xu, Yuzheng, et al.
Published: (2026)
by: Xu, Yuzheng, et al.
Published: (2026)
Am I Blue or Is My Hobby Counting Teardrops? Expression Leakage in Large Language Models as a Symptom of Irrelevancy Disruption
by: Köprü, Berkay, et al.
Published: (2025)
by: Köprü, Berkay, et al.
Published: (2025)
DROID: Dual Representation for Out-of-Scope Intent Detection
by: Rashwan, Wael, et al.
Published: (2025)
by: Rashwan, Wael, et al.
Published: (2025)
Improved Out-of-Scope Intent Classification with Dual Encoding and Threshold-based Re-Classification
by: Zawbaa, Hossam M., et al.
Published: (2024)
by: Zawbaa, Hossam M., et al.
Published: (2024)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
by: Sakhawat, Adib, et al.
Published: (2026)
by: Sakhawat, Adib, et al.
Published: (2026)
LTD-Bench: Evaluating Large Language Models by Letting Them Draw
by: Lin, Liuhao, et al.
Published: (2025)
by: Lin, Liuhao, et al.
Published: (2025)
Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts
by: Elsafoury, Fatma, et al.
Published: (2025)
by: Elsafoury, Fatma, et al.
Published: (2025)
Characterizing AI Manipulation Risks in Brazilian YouTube Climate Discourse
by: Dong, Wenchao, et al.
Published: (2025)
by: Dong, Wenchao, et al.
Published: (2025)
Bias in, Bias out: Annotation Bias in Multilingual Large Language Models
by: Cui, Xia, et al.
Published: (2025)
by: Cui, Xia, et al.
Published: (2025)
Acquiescence Bias in Large Language Models
by: Braun, Daniel
Published: (2025)
by: Braun, Daniel
Published: (2025)
Benchmarking Gender and Political Bias in Large Language Models
by: Yang, Jinrui, et al.
Published: (2025)
by: Yang, Jinrui, et al.
Published: (2025)
Factuality Challenges in the Era of Large Language Models
by: Augenstein, Isabelle, et al.
Published: (2023)
by: Augenstein, Isabelle, et al.
Published: (2023)
Mitigate Position Bias in Large Language Models via Scaling a Single Dimension
by: Yu, Yijiong, et al.
Published: (2024)
by: Yu, Yijiong, et al.
Published: (2024)
Persistent Topological Features in Large Language Models
by: Gardinazzi, Yuri, et al.
Published: (2024)
by: Gardinazzi, Yuri, et al.
Published: (2024)
Am I eligible? Natural Language Inference for Clinical Trial Patient Recruitment: the Patient's Point of View
by: Aguiar, Mathilde, et al.
Published: (2025)
by: Aguiar, Mathilde, et al.
Published: (2025)
Evaluating Gender Bias in Large Language Models
by: Döll, Michael, et al.
Published: (2024)
by: Döll, Michael, et al.
Published: (2024)
Mitigating the Bias of Large Language Model Evaluation
by: Zhou, Hongli, et al.
Published: (2024)
by: Zhou, Hongli, et al.
Published: (2024)
Language Models Predict Empathy Gaps Between Social In-groups and Out-groups
by: Hou, Yu, et al.
Published: (2025)
by: Hou, Yu, et al.
Published: (2025)
SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models
by: Su, Binxian, et al.
Published: (2026)
by: Su, Binxian, et al.
Published: (2026)
Constructions Are So Difficult That Even Large Language Models Get Them Right for the Wrong Reasons
by: Zhou, Shijia, et al.
Published: (2024)
by: Zhou, Shijia, et al.
Published: (2024)
A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models
by: Shi, Zeru, et al.
Published: (2026)
by: Shi, Zeru, et al.
Published: (2026)
What Am I Missing? Question-Answering as Hidden State Probing
by: Luo, Chu Fei, et al.
Published: (2026)
by: Luo, Chu Fei, et al.
Published: (2026)
Conversations: Love Them, Hate Them, Steer Them
by: Chebrolu, Niranjan, et al.
Published: (2025)
by: Chebrolu, Niranjan, et al.
Published: (2025)
Regional Bias in Large Language Models
by: Gopinadh, M P V S, et al.
Published: (2026)
by: Gopinadh, M P V S, et al.
Published: (2026)
JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models
by: Yanaka, Hitomi, et al.
Published: (2024)
by: Yanaka, Hitomi, et al.
Published: (2024)
Out-of-Context Reasoning in Large Language Models
by: Shaki, Jonathan, et al.
Published: (2025)
by: Shaki, Jonathan, et al.
Published: (2025)
Gender Bias in Large Language Models across Multiple Languages
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
Do They Understand Them? An Updated Evaluation on Nonbinary Pronoun Handling in Large Language Models
by: Tang, Xushuo, et al.
Published: (2025)
by: Tang, Xushuo, et al.
Published: (2025)
Similar Items
-
Persona Setting Pitfall: Persistent Outgroup Biases in Large Language Models Arising from Social Identity Adoption
by: Dong, Wenchao, et al.
Published: (2024) -
Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions
by: Kim, Jiseon, et al.
Published: (2026) -
Adversarial Style Augmentation via Large Language Model for Robust Fake News Detection
by: Park, Sungwon, et al.
Published: (2024) -
EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models
by: Lee, Donggyu, et al.
Published: (2025) -
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
by: Kim, Minsung, et al.
Published: (2025)