Where Fact Ends and Fairness Begins: Redefining AI Bias Evaluation through Cognitive Biases
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Jen-tse, Yan, Yuhang, Liu, Linqi, Wan, Yixin, Wang, Wenxuan, Chang, Kai-Wei, Lyu, Michael R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Where Δμ Begins, Gödel Ends
di: Ednyashev, Sanal
Pubblicazione: (2025)
di: Ednyashev, Sanal
Pubblicazione: (2025)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
di: Du, Yongkang, et al.
Pubblicazione: (2025)
di: Du, Yongkang, et al.
Pubblicazione: (2025)
Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation
di: Huang, Jen-tse, et al.
Pubblicazione: (2026)
di: Huang, Jen-tse, et al.
Pubblicazione: (2026)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
AI Sees Your Location, But With A Bias Toward The Wealthy World
di: Huang, Jingyuan, et al.
Pubblicazione: (2025)
di: Huang, Jingyuan, et al.
Pubblicazione: (2025)
The Male CEO and the Female Assistant: Evaluation and Mitigation of Gender Biases in Text-To-Image Generation of Dual Subjects
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
New Job, New Gender? Measuring the Social Bias in Image Generation Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs
di: Kukreja, Dikshant, et al.
Pubblicazione: (2026)
di: Kukreja, Dikshant, et al.
Pubblicazione: (2026)
Follow‐Up After Preterm Birth: Where Evidence Ends and Uncertainty Begins
di: Ilari Kuitunen
Pubblicazione: (2026)
di: Ilari Kuitunen
Pubblicazione: (2026)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
BIASINSPECTOR: Detecting Bias in Structured Data through LLM Agents
di: Li, Haoxuan, et al.
Pubblicazione: (2025)
di: Li, Haoxuan, et al.
Pubblicazione: (2025)
On the Shortcut Learning in Multilingual Neural Machine Translation
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
di: Wan, Yuxuan, et al.
Pubblicazione: (2024)
di: Wan, Yuxuan, et al.
Pubblicazione: (2024)
On the Failure of Latent State Persistence in Large Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
di: Wan, Yixin, et al.
Pubblicazione: (2025)
di: Wan, Yixin, et al.
Pubblicazione: (2025)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
All Languages Matter: On the Multilingual Safety of Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
FAIRGAMER: Evaluating Social Biases in LLM-Based Video Game NPCs
di: Shi, Bingkang, et al.
Pubblicazione: (2025)
di: Shi, Bingkang, et al.
Pubblicazione: (2025)
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
Where It all Begins
di: Turner, Dorothy B.
Pubblicazione: (1973)
di: Turner, Dorothy B.
Pubblicazione: (1973)
ComboBench: Can LLMs Manipulate Physical Devices to Play Virtual Reality Games?
di: Li, Shuqing, et al.
Pubblicazione: (2025)
di: Li, Shuqing, et al.
Pubblicazione: (2025)
Revisiting the Reliability of Psychological Scales on Large Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2023)
di: Huang, Jen-tse, et al.
Pubblicazione: (2023)
Emotionally Numb or Empathetic? Evaluating How LLMs Feel Using EmotionBench
di: Huang, Jen-tse, et al.
Pubblicazione: (2023)
di: Huang, Jen-tse, et al.
Pubblicazione: (2023)
CompAlign: Improving Compositional Text-to-Image Generation with a Complex Benchmark and Fine-Grained Feedback
di: Wan, Yixin, et al.
Pubblicazione: (2025)
di: Wan, Yixin, et al.
Pubblicazione: (2025)
Bias Begins with Data: The FairGround Corpus for Robust and Reproducible Research on Algorithmic Fairness
di: Simson, Jan, et al.
Pubblicazione: (2025)
di: Simson, Jan, et al.
Pubblicazione: (2025)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
di: Lam, Man Ho, et al.
Pubblicazione: (2025)
di: Lam, Man Ho, et al.
Pubblicazione: (2025)
Collaboration: Where Does It Begin?
di: Small, Ruth V.
Pubblicazione: (2002)
di: Small, Ruth V.
Pubblicazione: (2002)
Towards Evaluating Proactive Risk Awareness of Multimodal Language Models
di: Yuan, Youliang, et al.
Pubblicazione: (2025)
di: Yuan, Youliang, et al.
Pubblicazione: (2025)
Survey of Bias In Text-to-Image Generation: Definition, Evaluation, and Mitigation
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
Hallucination Begins Where Saliency Drops
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2026)
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2026)
Technology--Where Do We Begin?
di: Kostecki, Sister Gladys
Pubblicazione: (1973)
di: Kostecki, Sister Gladys
Pubblicazione: (1973)
Fairness at Risk: Where Bias Emerges in Machine Learning
di: Otavio de Paula Albuquerque, et al.
Pubblicazione: (2026)
di: Otavio de Paula Albuquerque, et al.
Pubblicazione: (2026)
Where Should I Study? Biased Language Models Decide! Evaluating Fairness in LMs for Academic Recommendations
di: Shailya, Krithi, et al.
Pubblicazione: (2025)
di: Shailya, Krithi, et al.
Pubblicazione: (2025)
Are Bias Evaluation Methods Biased ?
di: Berrayana, Lina, et al.
Pubblicazione: (2025)
di: Berrayana, Lina, et al.
Pubblicazione: (2025)
GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
di: Yuan, Youliang, et al.
Pubblicazione: (2023)
di: Yuan, Youliang, et al.
Pubblicazione: (2023)
Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
di: Liu, Xiaoyuan, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyuan, et al.
Pubblicazione: (2024)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval
di: Wu, Di, et al.
Pubblicazione: (2025)
di: Wu, Di, et al.
Pubblicazione: (2025)
The Hrunting of AI: Where and How to Improve English Dialectal Fairness
di: Li, Wei, et al.
Pubblicazione: (2026)
di: Li, Wei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Where Δμ Begins, Gödel Ends
di: Ednyashev, Sanal
Pubblicazione: (2025) -
FairCoder: Evaluating Social Bias of LLMs in Code Generation
di: Du, Yongkang, et al.
Pubblicazione: (2025) -
Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation
di: Huang, Jen-tse, et al.
Pubblicazione: (2026) -
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2025) -
AI Sees Your Location, But With A Bias Toward The Wealthy World
di: Huang, Jingyuan, et al.
Pubblicazione: (2025)