A Comprehensive Evaluation of Cognitive Biases in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Malberg, Simon, Poletukhin, Roman, Schuster, Carolin M., Groh, Georg |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Framework of Thoughts: A Foundation Framework for Dynamic and Optimized Reasoning based on Chains, Trees, and Graphs
by: Fricke, Felix, et al.
Published: (2026)
by: Fricke, Felix, et al.
Published: (2026)
From Roots to Rewards: Dynamic Tree Reasoning with Reinforcement Learning
by: Bahloul, Ahmed, et al.
Published: (2025)
by: Bahloul, Ahmed, et al.
Published: (2025)
Adapter-based Approaches to Knowledge-enhanced Language Models -- A Survey
by: Fichtl, Alexander, et al.
Published: (2024)
by: Fichtl, Alexander, et al.
Published: (2024)
Cross-lingual Text Classification Transfer: The Case of Ukrainian
by: Dementieva, Daryna, et al.
Published: (2024)
by: Dementieva, Daryna, et al.
Published: (2024)
DIALECTIC: A Multi-Agent System for Startup Evaluation
by: Bae, Jae Yoon, et al.
Published: (2026)
by: Bae, Jae Yoon, et al.
Published: (2026)
Profiling Bias in LLMs: Stereotype Dimensions in Contextual Word Embeddings
by: Schuster, Carolin M., et al.
Published: (2024)
by: Schuster, Carolin M., et al.
Published: (2024)
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
by: Yang, Xikang, et al.
Published: (2025)
by: Yang, Xikang, et al.
Published: (2025)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
by: Ghosh, Rajarshi, et al.
Published: (2025)
by: Ghosh, Rajarshi, et al.
Published: (2025)
TUM-MiKaNi at SemEval-2025 Task 3: Towards Multilingual and Knowledge-Aware Non-factual Hallucination Identification
by: Anschütz, Miriam, et al.
Published: (2025)
by: Anschütz, Miriam, et al.
Published: (2025)
To Bias or Not to Bias: Detecting bias in News with bias-detector
by: Ghosh, Himel, et al.
Published: (2025)
by: Ghosh, Himel, et al.
Published: (2025)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
by: Itzhak, Itay, et al.
Published: (2025)
by: Itzhak, Itay, et al.
Published: (2025)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
by: Nguyen, Van Bach, et al.
Published: (2024)
by: Nguyen, Van Bach, et al.
Published: (2024)
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
by: Jin, Keyan, et al.
Published: (2025)
by: Jin, Keyan, et al.
Published: (2025)
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments
by: Sumita, Yasuaki, et al.
Published: (2024)
by: Sumita, Yasuaki, et al.
Published: (2024)
Are Bias Evaluation Methods Biased ?
by: Berrayana, Lina, et al.
Published: (2025)
by: Berrayana, Lina, et al.
Published: (2025)
Unveiling Divergent Inductive Biases of LLMs on Temporal Data
by: Kishore, Sindhu, et al.
Published: (2024)
by: Kishore, Sindhu, et al.
Published: (2024)
Semantic Component Analysis: Introducing Multi-Topic Distributions to Clustering-Based Topic Modeling
by: Eichin, Florian, et al.
Published: (2024)
by: Eichin, Florian, et al.
Published: (2024)
EasyJudge: an Easy-to-use Tool for Comprehensive Response Evaluation of LLMs
by: Li, Yijie, et al.
Published: (2024)
by: Li, Yijie, et al.
Published: (2024)
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
by: Sturm, Jakob, et al.
Published: (2026)
by: Sturm, Jakob, et al.
Published: (2026)
Benchmarking Cognitive Biases in Large Language Models as Evaluators
by: Koo, Ryan, et al.
Published: (2023)
by: Koo, Ryan, et al.
Published: (2023)
Tuning Into Bias: A Computational Study of Gender Bias in Song Lyrics
by: Chen, Danqing, et al.
Published: (2024)
by: Chen, Danqing, et al.
Published: (2024)
Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs
by: Mor-Lan, Guy, et al.
Published: (2026)
by: Mor-Lan, Guy, et al.
Published: (2026)
(Ir)rationality and Cognitive Biases in Large Language Models
by: Macmillan-Scott, Olivia, et al.
Published: (2024)
by: Macmillan-Scott, Olivia, et al.
Published: (2024)
A Comprehensive Evaluation framework of Alignment Techniques for LLMs
by: Azmat, Muneeza, et al.
Published: (2025)
by: Azmat, Muneeza, et al.
Published: (2025)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
by: Wei, Jianhui, et al.
Published: (2025)
by: Wei, Jianhui, et al.
Published: (2025)
Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs
by: de Landa, Joseba Fernandez, et al.
Published: (2026)
by: de Landa, Joseba Fernandez, et al.
Published: (2026)
Cognitive Biases in Large Language Models for News Recommendation
by: Lyu, Yougang, et al.
Published: (2024)
by: Lyu, Yougang, et al.
Published: (2024)
Depth-Recurrent Attention Mixtures: Giving Latent Reasoning the Attention it Deserves
by: Knupp, Jonas, et al.
Published: (2026)
by: Knupp, Jonas, et al.
Published: (2026)
Identifying High-Confidence Social Biases in LLMs for Trustworthy Conversational Tutoring Agents
by: Alvarez, Aitor Arronte, et al.
Published: (2026)
by: Alvarez, Aitor Arronte, et al.
Published: (2026)
Balancing Rigor and Utility: Mitigating Cognitive Biases in Large Language Models for Multiple-Choice Questions
by: Zhong, Hanyang, et al.
Published: (2024)
by: Zhong, Hanyang, et al.
Published: (2024)
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs
by: Liu, Songyang, et al.
Published: (2025)
by: Liu, Songyang, et al.
Published: (2025)
LLMs for Explainable AI: A Comprehensive Survey
by: Bilal, Ahsan, et al.
Published: (2025)
by: Bilal, Ahsan, et al.
Published: (2025)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
by: Wu, Chengwei, et al.
Published: (2025)
by: Wu, Chengwei, et al.
Published: (2025)
Split and Merge: Aligning Position Biases in LLM-based Evaluators
by: Li, Zongjie, et al.
Published: (2023)
by: Li, Zongjie, et al.
Published: (2023)
Cognitive Bias in Decision-Making with LLMs
by: Echterhoff, Jessica, et al.
Published: (2024)
by: Echterhoff, Jessica, et al.
Published: (2024)
Cannot See the Forest for the Trees: Invoking Heuristics and Biases to Elicit Irrational Choices of LLMs
by: Yang, Haoming, et al.
Published: (2025)
by: Yang, Haoming, et al.
Published: (2025)
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment
by: Chen, Nuo, et al.
Published: (2024)
by: Chen, Nuo, et al.
Published: (2024)
CCR-Bench: A Comprehensive Benchmark for Evaluating LLMs on Complex Constraints, Control Flows, and Real-World Cases
by: Xue, Xiaona, et al.
Published: (2026)
by: Xue, Xiaona, et al.
Published: (2026)
Empowering LLMs with Logical Reasoning: A Comprehensive Survey
by: Cheng, Fengxiang, et al.
Published: (2025)
by: Cheng, Fengxiang, et al.
Published: (2025)
Discrete Tokenization for Multimodal LLMs: A Comprehensive Survey
by: Li, Jindong, et al.
Published: (2025)
by: Li, Jindong, et al.
Published: (2025)
Similar Items
-
Framework of Thoughts: A Foundation Framework for Dynamic and Optimized Reasoning based on Chains, Trees, and Graphs
by: Fricke, Felix, et al.
Published: (2026) -
From Roots to Rewards: Dynamic Tree Reasoning with Reinforcement Learning
by: Bahloul, Ahmed, et al.
Published: (2025) -
Adapter-based Approaches to Knowledge-enhanced Language Models -- A Survey
by: Fichtl, Alexander, et al.
Published: (2024) -
Cross-lingual Text Classification Transfer: The Case of Ukrainian
by: Dementieva, Daryna, et al.
Published: (2024) -
DIALECTIC: A Multi-Agent System for Startup Evaluation
by: Bae, Jae Yoon, et al.
Published: (2026)