UCFE: A User-Centric Financial Expertise Benchmark for Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Yuzhe, Zhang, Yifei, Hu, Yan, Guo, Yilin, Gan, Ruoli, He, Yueru, Lei, Mingcong, Zhang, Xiao, Wang, Haining, Xie, Qianqian, Huang, Jimin, Yu, Honghai, Wang, Benyou |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TwinMarket: A Scalable Behavioral and Social Simulation for Financial Markets
par: Yang, Yuzhe, et autres
Publié: (2025)
par: Yang, Yuzhe, et autres
Publié: (2025)
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
par: Huang, Jimin, et autres
Publié: (2024)
par: Huang, Jimin, et autres
Publié: (2024)
INVESTORBENCH: A Benchmark for Financial Decision-Making Tasks with LLM-based Agent
par: Li, Haohang, et autres
Publié: (2024)
par: Li, Haohang, et autres
Publié: (2024)
No Language is an Island: Unifying Chinese and English in Financial Large Language Models, Instruction Data, and Benchmarks
par: Hu, Gang, et autres
Publié: (2024)
par: Hu, Gang, et autres
Publié: (2024)
All That Glisters Is Not Gold: A Benchmark for Reference-Free Counterfactual Financial Misinformation Detection
par: Jiang, Yuechen, et autres
Publié: (2026)
par: Jiang, Yuechen, et autres
Publié: (2026)
AI-Trader: Benchmarking Autonomous Agents in Real-Time Financial Markets
par: Fan, Tianyu, et autres
Publié: (2025)
par: Fan, Tianyu, et autres
Publié: (2025)
FinBen: A Holistic Financial Benchmark for Large Language Models
par: Xie, Qianqian, et autres
Publié: (2024)
par: Xie, Qianqian, et autres
Publié: (2024)
Machine Learning Methods for Pricing Financial Derivatives
par: Fan, Lei, et autres
Publié: (2024)
par: Fan, Lei, et autres
Publié: (2024)
FinAudio: A Benchmark for Audio Large Language Models in Financial Applications
par: Cao, Yupeng, et autres
Publié: (2025)
par: Cao, Yupeng, et autres
Publié: (2025)
Open FinLLM Leaderboard: Towards Financial AI Readiness
par: Lin, Shengyuan Colin, et autres
Publié: (2025)
par: Lin, Shengyuan Colin, et autres
Publié: (2025)
SusGen-GPT: A Data-Centric LLM for Financial NLP and Sustainability Report Generation
par: Wu, Qilong, et autres
Publié: (2024)
par: Wu, Qilong, et autres
Publié: (2024)
Conv-FinRe: A Conversational and Longitudinal Benchmark for Utility-Grounded Financial Recommendation
par: Wang, Yan, et autres
Publié: (2026)
par: Wang, Yan, et autres
Publié: (2026)
Modeling of Measurement Error in Financial Returns Data
par: Jasra, Ajay, et autres
Publié: (2024)
par: Jasra, Ajay, et autres
Publié: (2024)
FinTagging: Benchmarking LLMs for Extracting and Structuring Financial Information
par: Wang, Yan, et autres
Publié: (2025)
par: Wang, Yan, et autres
Publié: (2025)
MMFCTUB: Multi-Modal Financial Credit Table Understanding Benchmark
par: Yakun, Cui, et autres
Publié: (2026)
par: Yakun, Cui, et autres
Publié: (2026)
Identifying and Quantifying Financial Bubbles with the Hyped Log-Periodic Power Law Model
par: Cao, Zheng, et autres
Publié: (2025)
par: Cao, Zheng, et autres
Publié: (2025)
Evaluation and Benchmarking Suite for Financial Large Language Models and Agents
par: Lin, Shengyuan, et autres
Publié: (2026)
par: Lin, Shengyuan, et autres
Publié: (2026)
From Volatility to Variance: A Skew-Enhanced SABR Model and Its Empirical Study in the Chinese Financial Options Market
par: Zhang, Wenxuan, et autres
Publié: (2026)
par: Zhang, Wenxuan, et autres
Publié: (2026)
FinHEAR: Human Expertise and Adaptive Risk-Aware Temporal Reasoning for Financial Decision-Making
par: Chen, Jiaxiang, et autres
Publié: (2025)
par: Chen, Jiaxiang, et autres
Publié: (2025)
Modeling News Interactions and Influence for Financial Market Prediction
par: Wang, Mengyu, et autres
Publié: (2024)
par: Wang, Mengyu, et autres
Publié: (2024)
RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?
par: Dai, Yuyang, et autres
Publié: (2026)
par: Dai, Yuyang, et autres
Publié: (2026)
FinMTM: A Multi-Turn Multimodal Benchmark for Financial Reasoning and Agent Evaluation
par: Zhang, Chenxi, et autres
Publié: (2026)
par: Zhang, Chenxi, et autres
Publié: (2026)
The Statistical Significance of the Inclusion of Graph Neural Networks in the Financial Time Series Forecasting Problem
par: Gregnanin, Marco, et autres
Publié: (2026)
par: Gregnanin, Marco, et autres
Publié: (2026)
Construction of a Japanese Financial Benchmark for Large Language Models
par: Hirano, Masanori
Publié: (2024)
par: Hirano, Masanori
Publié: (2024)
FinRL Contests: Benchmarking Data-driven Financial Reinforcement Learning Agents
par: Wang, Keyi, et autres
Publié: (2025)
par: Wang, Keyi, et autres
Publié: (2025)
Beyond the Numbers: Causal Effects of Financial Report Sentiment on Bank Profitability
par: Neupane, Krishna, et autres
Publié: (2026)
par: Neupane, Krishna, et autres
Publié: (2026)
Long-Range Dependence in Financial Markets: Empirical Evidence and Generative Modeling Challenges
par: He, Yifan, et autres
Publié: (2025)
par: He, Yifan, et autres
Publié: (2025)
R&D-Agent-Quant: A Multi-Agent Framework for Data-Centric Factors and Model Joint Optimization
par: Li, Yuante, et autres
Publié: (2025)
par: Li, Yuante, et autres
Publié: (2025)
MFMDQwen: Multilingual Financial Misinformation Detection Based on Large Language Model
par: Liu, Zhiwei, et autres
Publié: (2026)
par: Liu, Zhiwei, et autres
Publié: (2026)
Exploiting Distributional Value Functions for Financial Market Valuation, Enhanced Feature Creation and Improvement of Trading Algorithms
par: Grab, Colin D.
Publié: (2024)
par: Grab, Colin D.
Publié: (2024)
From Scores to Skills: A Cognitive Diagnosis Framework for Evaluating Financial Large Language Models
par: Kuang, Ziyan, et autres
Publié: (2025)
par: Kuang, Ziyan, et autres
Publié: (2025)
QuantBench: Benchmarking AI Methods for Quantitative Investment
par: Wang, Saizhuo, et autres
Publié: (2025)
par: Wang, Saizhuo, et autres
Publié: (2025)
AMA-LSTM: Pioneering Robust and Fair Financial Audio Analysis for Stock Volatility Prediction
par: Wang, Shengkun, et autres
Publié: (2024)
par: Wang, Shengkun, et autres
Publié: (2024)
Quantifying Semantic Shift in Financial NLP: Robust Metrics for Market Prediction Stability
par: Sun, Zhongtian, et autres
Publié: (2025)
par: Sun, Zhongtian, et autres
Publié: (2025)
FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs
par: Wang, Yan, et autres
Publié: (2025)
par: Wang, Yan, et autres
Publié: (2025)
Neural Term Structure of Additive Process for Option Pricing
par: Lin, Jimin, et autres
Publié: (2024)
par: Lin, Jimin, et autres
Publié: (2024)
SeQwen at the Financial Misinformation Detection Challenge Task: Sequential Learning for Claim Verification and Explanation Generation in Financial Domains
par: Purbey, Jebish, et autres
Publié: (2024)
par: Purbey, Jebish, et autres
Publié: (2024)
Shallow Representation of Option Implied Information
par: Lin, Jimin
Publié: (2026)
par: Lin, Jimin
Publié: (2026)
Financial Wind Tunnel: A Retrieval-Augmented Market Simulator
par: Cao, Bokai, et autres
Publié: (2025)
par: Cao, Bokai, et autres
Publié: (2025)
FinReflectKG -- HalluBench: GraphRAG Hallucination Benchmark for Financial Question Answering Systems
par: Kumar, Mahesh, et autres
Publié: (2026)
par: Kumar, Mahesh, et autres
Publié: (2026)
Documents similaires
-
TwinMarket: A Scalable Behavioral and Social Simulation for Financial Markets
par: Yang, Yuzhe, et autres
Publié: (2025) -
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
par: Huang, Jimin, et autres
Publié: (2024) -
INVESTORBENCH: A Benchmark for Financial Decision-Making Tasks with LLM-based Agent
par: Li, Haohang, et autres
Publié: (2024) -
No Language is an Island: Unifying Chinese and English in Financial Large Language Models, Instruction Data, and Benchmarks
par: Hu, Gang, et autres
Publié: (2024) -
All That Glisters Is Not Gold: A Benchmark for Reference-Free Counterfactual Financial Misinformation Detection
par: Jiang, Yuechen, et autres
Publié: (2026)