Deontological Keyword Bias: The Impact of Modal Expressions on Normative Judgments of Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Park, Bumjin, Lee, Jinsil, Choi, Jaesik |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Memorizing Documents with Guidance in Large Language Models
par: Park, Bumjin, et autres
Publié: (2024)
par: Park, Bumjin, et autres
Publié: (2024)
Identifying the Source of Generation for Large Language Models
par: Park, Bumjin, et autres
Publié: (2024)
par: Park, Bumjin, et autres
Publié: (2024)
Are Language Models Consequentialist or Deontological Moral Reasoners?
par: Samway, Keenan, et autres
Publié: (2025)
par: Samway, Keenan, et autres
Publié: (2025)
Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling
par: Liu, Shuliang, et autres
Publié: (2025)
par: Liu, Shuliang, et autres
Publié: (2025)
Optimal path for Biomedical Text Summarization Using Pointer GPT
par: Han, Hyunkyung, et autres
Publié: (2024)
par: Han, Hyunkyung, et autres
Publié: (2024)
Inertia in Moral and Value Judgments of Large Language Models
par: Lee, Bruce W., et autres
Publié: (2024)
par: Lee, Bruce W., et autres
Publié: (2024)
Systematic Bias in Large Language Models: Discrepant Response Patterns in Binary vs. Continuous Judgment Tasks
par: Lu, Yi-Long, et autres
Publié: (2025)
par: Lu, Yi-Long, et autres
Publié: (2025)
Normative Reasoning in Large Language Models: A Comparative Benchmark from Logical and Modal Perspectives
par: Ozeki, Kentaro, et autres
Publié: (2025)
par: Ozeki, Kentaro, et autres
Publié: (2025)
Chaos with Keywords: Exposing Large Language Models Sycophantic Hallucination to Misleading Keywords and Evaluating Defense Strategies
par: RRV, Aswin, et autres
Publié: (2024)
par: RRV, Aswin, et autres
Publié: (2024)
Assessing Modality Bias in Video Question Answering Benchmarks with Multimodal Large Language Models
par: Park, Jean, et autres
Publié: (2024)
par: Park, Jean, et autres
Publié: (2024)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
par: Wehnert, Sabine, et autres
Publié: (2025)
par: Wehnert, Sabine, et autres
Publié: (2025)
Exploring Automated Keyword Mnemonics Generation with Large Language Models via Overgenerate-and-Rank
par: Lee, Jaewook, et autres
Publié: (2024)
par: Lee, Jaewook, et autres
Publié: (2024)
KiC: Keyword-inspired Cascade for Cost-Efficient Text Generation with LLMs
par: Kim, Woo-Chan, et autres
Publié: (2025)
par: Kim, Woo-Chan, et autres
Publié: (2025)
Reasons to Reject? Aligning Language Models with Judgments
par: Xu, Weiwen, et autres
Publié: (2023)
par: Xu, Weiwen, et autres
Publié: (2023)
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
par: Hahm, Sungeun, et autres
Publié: (2025)
par: Hahm, Sungeun, et autres
Publié: (2025)
Large Language Models Still Exhibit Bias in Long Text
par: Jeung, Wonje, et autres
Publié: (2024)
par: Jeung, Wonje, et autres
Publié: (2024)
Detection of Conspiracy Theories Beyond Keyword Bias in German-Language Telegram Using Large Language Models
par: Pustet, Milena, et autres
Publié: (2024)
par: Pustet, Milena, et autres
Publié: (2024)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
par: Fernandes, Gustavo Lúcius, et autres
Publié: (2026)
par: Fernandes, Gustavo Lúcius, et autres
Publié: (2026)
Aligning Large Language Models by On-Policy Self-Judgment
par: Lee, Sangkyu, et autres
Publié: (2024)
par: Lee, Sangkyu, et autres
Publié: (2024)
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
par: Elsafoury, Fatma, et autres
Publié: (2023)
par: Elsafoury, Fatma, et autres
Publié: (2023)
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings
par: Lee, Jonghyun, et autres
Publié: (2025)
par: Lee, Jonghyun, et autres
Publié: (2025)
Bias Attribution in Filipino Language Models: Extending a Bias Interpretability Metric for Application on Agglutinative Languages
par: Gamboa, Lance Calvin Lim, et autres
Publié: (2025)
par: Gamboa, Lance Calvin Lim, et autres
Publié: (2025)
What is Your Favorite Gender, MLM? Gender Bias Evaluation in Multilingual Masked Language Models
par: Yu, Jeongrok, et autres
Publié: (2024)
par: Yu, Jeongrok, et autres
Publié: (2024)
A Unified Representation Underlying the Judgment of Large Language Models
par: Lu, Yi-Long, et autres
Publié: (2025)
par: Lu, Yi-Long, et autres
Publié: (2025)
Incoherent Probability Judgments in Large Language Models
par: Zhu, Jian-Qiao, et autres
Publié: (2024)
par: Zhu, Jian-Qiao, et autres
Publié: (2024)
Pretraining Exposure Explains Popularity Judgments in Large Language Models
par: Mozafari, Jamshid, et autres
Publié: (2026)
par: Mozafari, Jamshid, et autres
Publié: (2026)
Do Emotions Influence Moral Judgment in Large Language Models?
par: Saim, Mohammad, et autres
Publié: (2026)
par: Saim, Mohammad, et autres
Publié: (2026)
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models
par: Park, Seong-Il, et autres
Publié: (2024)
par: Park, Seong-Il, et autres
Publié: (2024)
Soft Inductive Bias Approach via Explicit Reasoning Perspectives in Inappropriate Utterance Detection Using Large Language Models
par: Kim, Ju-Young, et autres
Publié: (2025)
par: Kim, Ju-Young, et autres
Publié: (2025)
Smoothie-Qwen: Post-Hoc Smoothing to Reduce Language Bias in Multilingual LLMs
par: Ji, SeungWon, et autres
Publié: (2025)
par: Ji, SeungWon, et autres
Publié: (2025)
Improving LLM-as-a-Judge Inference with the Judgment Distribution
par: Wang, Victor, et autres
Publié: (2025)
par: Wang, Victor, et autres
Publié: (2025)
Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models
par: Yu, Seunguk, et autres
Publié: (2025)
par: Yu, Seunguk, et autres
Publié: (2025)
Humane Speech Synthesis through Zero-Shot Emotion and Disfluency Generation
par: Chaudhury, Rohan, et autres
Publié: (2024)
par: Chaudhury, Rohan, et autres
Publié: (2024)
If Probable, Then Acceptable? Understanding Conditional Acceptability Judgments in Large Language Models
par: Orth, Jasmin, et autres
Publié: (2025)
par: Orth, Jasmin, et autres
Publié: (2025)
Grammaticality Judgments in Humans and Language Models: Revisiting Generative Grammar with LLMs
par: Johnsen, Lars G. B.
Publié: (2025)
par: Johnsen, Lars G. B.
Publié: (2025)
Social Bias in Multilingual Language Models: A Survey
par: Gamboa, Lance Calvin Lim, et autres
Publié: (2025)
par: Gamboa, Lance Calvin Lim, et autres
Publié: (2025)
KLAAD: Refining Attention Mechanisms to Reduce Societal Bias in Generative Language Models
par: Kim, Seorin, et autres
Publié: (2025)
par: Kim, Seorin, et autres
Publié: (2025)
Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models
par: Kim, Mingyeong, et autres
Publié: (2026)
par: Kim, Mingyeong, et autres
Publié: (2026)
Beyond Keywords: Evaluating Large Language Model Classification of Nuanced Ableism
par: Rizvi, Naba, et autres
Publié: (2025)
par: Rizvi, Naba, et autres
Publié: (2025)
Language-Universal Speech Attributes Modeling for Zero-Shot Multilingual Spoken Keyword Recognition
par: Yen, Hao, et autres
Publié: (2024)
par: Yen, Hao, et autres
Publié: (2024)
Documents similaires
-
Memorizing Documents with Guidance in Large Language Models
par: Park, Bumjin, et autres
Publié: (2024) -
Identifying the Source of Generation for Large Language Models
par: Park, Bumjin, et autres
Publié: (2024) -
Are Language Models Consequentialist or Deontological Moral Reasoners?
par: Samway, Keenan, et autres
Publié: (2025) -
Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling
par: Liu, Shuliang, et autres
Publié: (2025) -
Optimal path for Biomedical Text Summarization Using Pointer GPT
par: Han, Hyunkyung, et autres
Publié: (2024)