Fin-Bias: Comprehensive Evaluation for LLM Decision-Making under human bias in Finance Domain
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Xiaoyu, Zhao, Jinman |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FinTrust: A Comprehensive Benchmark of Trustworthiness Evaluation in Finance Domain
by: Hu, Tiansheng, et al.
Published: (2025)
by: Hu, Tiansheng, et al.
Published: (2025)
Evaluating the Bias in LLMs for Surveying Opinion and Decision Making in Healthcare
by: Khaokaew, Yonchanok, et al.
Published: (2025)
by: Khaokaew, Yonchanok, et al.
Published: (2025)
FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents
by: Jia, Haoxuan, et al.
Published: (2026)
by: Jia, Haoxuan, et al.
Published: (2026)
'Finance Wizard' at the FinLLM Challenge Task: Financial Text Summarization
by: Lee, Meisin, et al.
Published: (2024)
by: Lee, Meisin, et al.
Published: (2024)
FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Making
by: Yu, Yangyang, et al.
Published: (2024)
by: Yu, Yangyang, et al.
Published: (2024)
Gender Bias in Large Language Models across Multiple Languages
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
Cognitive Bias in Decision-Making with LLMs
by: Echterhoff, Jessica, et al.
Published: (2024)
by: Echterhoff, Jessica, et al.
Published: (2024)
FinNuE: Exposing the Risks of Using BERTScore for Numerical Semantic Evaluation in Finance
by: Huang, Yu-Shiang, et al.
Published: (2025)
by: Huang, Yu-Shiang, et al.
Published: (2025)
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains
by: Zhao, Yilun, et al.
Published: (2023)
by: Zhao, Yilun, et al.
Published: (2023)
Exploring the Limitations of Large Language Models in Compositional Relation Reasoning
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
STRUX: An LLM for Decision-Making with Structured Explanations
by: Lu, Yiming, et al.
Published: (2024)
by: Lu, Yiming, et al.
Published: (2024)
To Bias or Not to Bias: Detecting bias in News with bias-detector
by: Ghosh, Himel, et al.
Published: (2025)
by: Ghosh, Himel, et al.
Published: (2025)
Evaluating Scoring Bias in LLM-as-a-Judge
by: Li, Qingquan, et al.
Published: (2025)
by: Li, Qingquan, et al.
Published: (2025)
No LLM is Free From Bias: A Comprehensive Study of Bias Evaluation in Large Language Models
by: Kumar, Charaka Vinayak, et al.
Published: (2025)
by: Kumar, Charaka Vinayak, et al.
Published: (2025)
MMREC: LLM Based Multi-Modal Recommender System
by: Tian, Jiahao, et al.
Published: (2024)
by: Tian, Jiahao, et al.
Published: (2024)
FinGen: A Dataset for Argument Generation in Finance
by: Chen, Chung-Chi, et al.
Published: (2024)
by: Chen, Chung-Chi, et al.
Published: (2024)
FinMTEB: Finance Massive Text Embedding Benchmark
by: Tang, Yixuan, et al.
Published: (2025)
by: Tang, Yixuan, et al.
Published: (2025)
FinXABSA: Explainable Finance through Aspect-Based Sentiment Analysis
by: Ong, Keane, et al.
Published: (2023)
by: Ong, Keane, et al.
Published: (2023)
LLMs Meet Finance: Fine-Tuning Foundation Models for the Open FinLLM Leaderboard
by: Rao, Varun, et al.
Published: (2025)
by: Rao, Varun, et al.
Published: (2025)
MedVoiceBias: A Controlled Study of Audio LLM Behavior in Clinical Decision-Making
by: Tam, Zhi Rui, et al.
Published: (2025)
by: Tam, Zhi Rui, et al.
Published: (2025)
An Electoral Approach to Diversify LLM-based Multi-Agent Collective Decision-Making
by: Zhao, Xiutian, et al.
Published: (2024)
by: Zhao, Xiutian, et al.
Published: (2024)
Pretraining on the Test Set Is No Longer All You Need: A Debate-Driven Approach to QA Benchmarks
by: Cao, Linbo, et al.
Published: (2025)
by: Cao, Linbo, et al.
Published: (2025)
Discrimination by LLMs: Cross-lingual Bias Assessment and Mitigation in Decision-Making and Summarisation
by: Huijzer, Willem, et al.
Published: (2025)
by: Huijzer, Willem, et al.
Published: (2025)
FinLFQA: Evaluating Attributed Text Generation of LLMs in Financial Long-Form Question Answering
by: Long, Yitao, et al.
Published: (2025)
by: Long, Yitao, et al.
Published: (2025)
A Survey of Large Language Models in Finance (FinLLMs)
by: Lee, Jean, et al.
Published: (2024)
by: Lee, Jean, et al.
Published: (2024)
FinEval-KR: A Financial Domain Evaluation Framework for Large Language Models' Knowledge and Reasoning
by: Dou, Shaoyu, et al.
Published: (2025)
by: Dou, Shaoyu, et al.
Published: (2025)
Evaluating Gender Bias of LLMs in Making Morality Judgements
by: Bajaj, Divij, et al.
Published: (2024)
by: Bajaj, Divij, et al.
Published: (2024)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
by: Chen, Yupeng, et al.
Published: (2025)
by: Chen, Yupeng, et al.
Published: (2025)
Implicit Causality-biases in humans and LLMs as a tool for benchmarking LLM discourse capabilities
by: Kankowski, Florian, et al.
Published: (2025)
by: Kankowski, Florian, et al.
Published: (2025)
Gender Bias in Decision-Making with Large Language Models: A Study of Relationship Conflicts
by: Levy, Sharon, et al.
Published: (2024)
by: Levy, Sharon, et al.
Published: (2024)
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases
by: Mohanty, Suvendu
Published: (2025)
by: Mohanty, Suvendu
Published: (2025)
FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios
by: Hou, Yutao, et al.
Published: (2026)
by: Hou, Yutao, et al.
Published: (2026)
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models
by: Guo, Xin, et al.
Published: (2023)
by: Guo, Xin, et al.
Published: (2023)
FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain
by: Zhao, Suifeng, et al.
Published: (2025)
by: Zhao, Suifeng, et al.
Published: (2025)
THaLLE-ThaiLLM: Domain-Specialized Small LLMs for Finance and Thai -- Technical Report
by: Labs, KBTG, et al.
Published: (2026)
by: Labs, KBTG, et al.
Published: (2026)
Learning to Contextualize Web Pages for Enhanced Decision Making by LLM Agents
by: Lee, Dongjun, et al.
Published: (2025)
by: Lee, Dongjun, et al.
Published: (2025)
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models
by: Qin, Zhanyue, et al.
Published: (2024)
by: Qin, Zhanyue, et al.
Published: (2024)
VIVA+: Human-Centered Situational Decision-Making
by: Hu, Zhe, et al.
Published: (2025)
by: Hu, Zhe, et al.
Published: (2025)
Explaining Length Bias in LLM-Based Preference Evaluations
by: Hu, Zhengyu, et al.
Published: (2024)
by: Hu, Zhengyu, et al.
Published: (2024)
Similar Items
-
FinTrust: A Comprehensive Benchmark of Trustworthiness Evaluation in Finance Domain
by: Hu, Tiansheng, et al.
Published: (2025) -
Evaluating the Bias in LLMs for Surveying Opinion and Decision Making in Healthcare
by: Khaokaew, Yonchanok, et al.
Published: (2025) -
FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents
by: Jia, Haoxuan, et al.
Published: (2026) -
'Finance Wizard' at the FinLLM Challenge Task: Financial Text Summarization
by: Lee, Meisin, et al.
Published: (2024) -
FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Making
by: Yu, Yangyang, et al.
Published: (2024)