Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Benhenda, Mostapha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
YC Bench: a Live Benchmark for Forecasting Startup Outperformance in Y Combinator Batches
von: Benhenda, Mostapha
Veröffentlicht: (2026)
von: Benhenda, Mostapha
Veröffentlicht: (2026)
Assessing Consistency and Reproducibility in the Outputs of Large Language Models: Evidence Across Diverse Finance and Accounting Tasks
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk
von: Chen, Zichen, et al.
Veröffentlicht: (2025)
von: Chen, Zichen, et al.
Veröffentlicht: (2025)
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models
von: Li, Weixian Waylon, et al.
Veröffentlicht: (2026)
von: Li, Weixian Waylon, et al.
Veröffentlicht: (2026)
Evaluating LLMs in Finance Requires Explicit Bias Consideration
von: Kong, Yaxuan, et al.
Veröffentlicht: (2026)
von: Kong, Yaxuan, et al.
Veröffentlicht: (2026)
Ideological Bias in LLMs' Economic Causal Reasoning
von: Lee, Donggyu, et al.
Veröffentlicht: (2026)
von: Lee, Donggyu, et al.
Veröffentlicht: (2026)
A Scoping Review of ChatGPT Research in Accounting and Finance
von: Dong, Mengming Michael, et al.
Veröffentlicht: (2024)
von: Dong, Mengming Michael, et al.
Veröffentlicht: (2024)
Large Language Models in Finance: A Survey
von: Li, Yinheng, et al.
Veröffentlicht: (2023)
von: Li, Yinheng, et al.
Veröffentlicht: (2023)
StakeBench: Evaluating Language Understanding Grounded in Market Commitment
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
Chat Bankman-Fried: an Exploration of LLM Alignment in Finance
von: Biancotti, Claudia, et al.
Veröffentlicht: (2024)
von: Biancotti, Claudia, et al.
Veröffentlicht: (2024)
Emoji Driven Crypto Assets Market Reactions
von: Zuo, Xiaorui, et al.
Veröffentlicht: (2024)
von: Zuo, Xiaorui, et al.
Veröffentlicht: (2024)
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
Words That Unite The World: A Unified Framework for Deciphering Central Bank Communications Globally
von: Shah, Agam, et al.
Veröffentlicht: (2025)
von: Shah, Agam, et al.
Veröffentlicht: (2025)
When Single Answer Is Not Enough: Rethinking Single-Step Retrosynthesis Benchmarks for LLMs
von: Zagribelnyy, Bogdan, et al.
Veröffentlicht: (2026)
von: Zagribelnyy, Bogdan, et al.
Veröffentlicht: (2026)
FinMaster: A Holistic Benchmark for Mastering Full-Pipeline Financial Workflows with LLMs
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
A Survey of Large Language Models in Finance (FinLLMs)
von: Lee, Jean, et al.
Veröffentlicht: (2024)
von: Lee, Jean, et al.
Veröffentlicht: (2024)
PolyBench: Benchmarking LLM Forecasting and Trading Capabilities on Live Prediction Market Data
von: Cheng, Pu, et al.
Veröffentlicht: (2026)
von: Cheng, Pu, et al.
Veröffentlicht: (2026)
PARROT: Persuasion and Agreement Robustness Rating of Output Truth -- A Sycophancy Robustness Benchmark for LLMs
von: Çelebi, Yusuf, et al.
Veröffentlicht: (2025)
von: Çelebi, Yusuf, et al.
Veröffentlicht: (2025)
FiMI: A Domain-Specific Language Model for Indian Finance Ecosystem
von: Kathar, Aboli, et al.
Veröffentlicht: (2026)
von: Kathar, Aboli, et al.
Veröffentlicht: (2026)
Leveraging Large Language Models to Democratize Access to Costly Datasets for Academic Research
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2024)
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2024)
Can AI Help with Your Personal Finances?
von: Hean, Oudom, et al.
Veröffentlicht: (2024)
von: Hean, Oudom, et al.
Veröffentlicht: (2024)
UniFinEval: Towards Unified Evaluation of Financial Multimodal Models across Text, Images and Videos
von: Yang, Zhi, et al.
Veröffentlicht: (2026)
von: Yang, Zhi, et al.
Veröffentlicht: (2026)
Reasoning Models Ace the CFA Exams
von: Patel, Jaisal, et al.
Veröffentlicht: (2025)
von: Patel, Jaisal, et al.
Veröffentlicht: (2025)
MetaBench: A Multi-task Benchmark for Assessing LLMs in Metabolomics
von: Lu, Yuxing, et al.
Veröffentlicht: (2025)
von: Lu, Yuxing, et al.
Veröffentlicht: (2025)
Baichuan4-Finance Technical Report
von: Zhang, Hanyu, et al.
Veröffentlicht: (2024)
von: Zhang, Hanyu, et al.
Veröffentlicht: (2024)
Reasoning on Time-Series for Financial Technical Analysis
von: Koa, Kelvin J. L., et al.
Veröffentlicht: (2025)
von: Koa, Kelvin J. L., et al.
Veröffentlicht: (2025)
EDINET-Bench: Evaluating LLMs on Complex Financial Tasks using Japanese Financial Statements
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
Financial Statement Analysis with Large Language Models
von: Kim, Alex, et al.
Veröffentlicht: (2024)
von: Kim, Alex, et al.
Veröffentlicht: (2024)
NumLLM: Numeric-Sensitive Large Language Model for Chinese Finance
von: Su, Huan-Yi, et al.
Veröffentlicht: (2024)
von: Su, Huan-Yi, et al.
Veröffentlicht: (2024)
A Survey of Large Language Models for Financial Applications: Progress, Prospects and Challenges
von: Nie, Yuqi, et al.
Veröffentlicht: (2024)
von: Nie, Yuqi, et al.
Veröffentlicht: (2024)
ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition
von: Liu, Yujie, et al.
Veröffentlicht: (2025)
von: Liu, Yujie, et al.
Veröffentlicht: (2025)
Forecasting Future Language: Context Design for Mention Markets
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
Topic Identification in LLM Input-Output Pairs through the Lens of Information Bottleneck
von: Halperin, Igor
Veröffentlicht: (2025)
von: Halperin, Igor
Veröffentlicht: (2025)
Multi-Dimensional Behavioral Evaluation of Agentic Stock Prediction Systems Using Large Language Model Judges with Closed-Loop Reinforcement Learning Feedback
von: Ridhawi, Mohammad Al, et al.
Veröffentlicht: (2026)
von: Ridhawi, Mohammad Al, et al.
Veröffentlicht: (2026)
Exploring the Synergy of Quantitative Factors and Newsflow Representations from Large Language Models for Stock Return Prediction
von: Guo, Tian, et al.
Veröffentlicht: (2025)
von: Guo, Tian, et al.
Veröffentlicht: (2025)
Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
von: Halperin, Igor
Veröffentlicht: (2025)
von: Halperin, Igor
Veröffentlicht: (2025)
CTBench: Cryptocurrency Time Series Generation Benchmark
von: Ang, Yihao, et al.
Veröffentlicht: (2025)
von: Ang, Yihao, et al.
Veröffentlicht: (2025)
Demystifying Domain-adaptive Post-training for Financial LLMs
von: Ke, Zixuan, et al.
Veröffentlicht: (2025)
von: Ke, Zixuan, et al.
Veröffentlicht: (2025)
BizFinBench: A Business-Driven Real-World Financial Benchmark for Evaluating LLMs
von: Lu, Guilong, et al.
Veröffentlicht: (2025)
von: Lu, Guilong, et al.
Veröffentlicht: (2025)
The Agentic Regulator: Risks for AI in Finance and a Proposed Agent-based Framework for Governance
von: Kurshan, Eren, et al.
Veröffentlicht: (2025)
von: Kurshan, Eren, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
YC Bench: a Live Benchmark for Forecasting Startup Outperformance in Y Combinator Batches
von: Benhenda, Mostapha
Veröffentlicht: (2026) -
Assessing Consistency and Reproducibility in the Outputs of Large Language Models: Evidence Across Diverse Finance and Accounting Tasks
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025) -
Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk
von: Chen, Zichen, et al.
Veröffentlicht: (2025) -
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models
von: Li, Weixian Waylon, et al.
Veröffentlicht: (2026) -
Evaluating LLMs in Finance Requires Explicit Bias Consideration
von: Kong, Yaxuan, et al.
Veröffentlicht: (2026)