LemonadeBench: Evaluating the Economic Intuition of Large Language Models in Simple Markets
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Vyas, Aidan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StakeBench: Evaluating Language Understanding Grounded in Market Commitment
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
Integrating Large Language Models in Financial Investments and Market Analysis: A Survey
von: Mahdavi, Sedigheh, et al.
Veröffentlicht: (2025)
von: Mahdavi, Sedigheh, et al.
Veröffentlicht: (2025)
A Survey of Large Language Models for Financial Applications: Progress, Prospects and Challenges
von: Nie, Yuqi, et al.
Veröffentlicht: (2024)
von: Nie, Yuqi, et al.
Veröffentlicht: (2024)
Large Language Models in Finance: A Survey
von: Li, Yinheng, et al.
Veröffentlicht: (2023)
von: Li, Yinheng, et al.
Veröffentlicht: (2023)
Financial Statement Analysis with Large Language Models
von: Kim, Alex, et al.
Veröffentlicht: (2024)
von: Kim, Alex, et al.
Veröffentlicht: (2024)
Dissecting AI Trading: Behavioral Finance and Market Bubbles
von: Ouyang, Shumiao, et al.
Veröffentlicht: (2026)
von: Ouyang, Shumiao, et al.
Veröffentlicht: (2026)
Do LLM Personas Dream of Bull Markets? Comparing Human and AI Investment Strategies Through the Lens of the Five-Factor Model
von: Borman, Harris, et al.
Veröffentlicht: (2024)
von: Borman, Harris, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models to Democratize Access to Costly Datasets for Academic Research
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2024)
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2024)
Leveraging Natural Language and Item Response Theory Models for ESG Scoring
von: Soares, César Pedrosa
Veröffentlicht: (2024)
von: Soares, César Pedrosa
Veröffentlicht: (2024)
Artificial Intelligence and Systemic Risk: A Unified Model of Performative Prediction, Algorithmic Herding, and Cognitive Dependency in Financial Markets
von: Meng, Shuchen, et al.
Veröffentlicht: (2026)
von: Meng, Shuchen, et al.
Veröffentlicht: (2026)
Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
von: Benhenda, Mostapha
Veröffentlicht: (2026)
von: Benhenda, Mostapha
Veröffentlicht: (2026)
UniFinEval: Towards Unified Evaluation of Financial Multimodal Models across Text, Images and Videos
von: Yang, Zhi, et al.
Veröffentlicht: (2026)
von: Yang, Zhi, et al.
Veröffentlicht: (2026)
Generative AI for Analysts
von: Xue, Jian, et al.
Veröffentlicht: (2025)
von: Xue, Jian, et al.
Veröffentlicht: (2025)
NoLBERT: A No Lookahead(back) Foundational Language Model
von: Kakhbod, Ali, et al.
Veröffentlicht: (2025)
von: Kakhbod, Ali, et al.
Veröffentlicht: (2025)
Assessing Consistency and Reproducibility in the Outputs of Large Language Models: Evidence Across Diverse Finance and Accounting Tasks
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
An Algorithmic Framework for Systematic Literature Reviews: A Case Study for Financial Narratives
von: Taibi, Gabin, et al.
Veröffentlicht: (2026)
von: Taibi, Gabin, et al.
Veröffentlicht: (2026)
Seeing the Goal, Missing the Truth: Human Accountability for AI Bias
von: Cao, Sean, et al.
Veröffentlicht: (2026)
von: Cao, Sean, et al.
Veröffentlicht: (2026)
The Sleeping Beauty Problem: Sleeping Kelly is a Thirder
von: Abramowitz, Ben
Veröffentlicht: (2025)
von: Abramowitz, Ben
Veröffentlicht: (2025)
Explaining the Unexplainable: A Systematic Review of Explainable AI in Finance
von: Mohsin, Md Talha, et al.
Veröffentlicht: (2025)
von: Mohsin, Md Talha, et al.
Veröffentlicht: (2025)
Know Your Intent: An Autonomous Multi-Perspective LLM Agent Framework for DeFi User Transaction Intent Mining
von: Mao, Qian'ang, et al.
Veröffentlicht: (2025)
von: Mao, Qian'ang, et al.
Veröffentlicht: (2025)
FedSight AI: Multi-Agent System Architecture for Federal Funds Target Rate Prediction
von: Hou, Yuhan, et al.
Veröffentlicht: (2025)
von: Hou, Yuhan, et al.
Veröffentlicht: (2025)
Towards Data-Centric Automatic R&D
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
Reasoning Models Ace the CFA Exams
von: Patel, Jaisal, et al.
Veröffentlicht: (2025)
von: Patel, Jaisal, et al.
Veröffentlicht: (2025)
DeFi TrustBoost: Blockchain and AI for Trustworthy Decentralized Financial Decisions
von: Sachan, Swati, et al.
Veröffentlicht: (2025)
von: Sachan, Swati, et al.
Veröffentlicht: (2025)
Personalized Chain-of-Thought Summarization of Financial News for Investor Decision Support
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Bloated Disclosures: Can ChatGPT Help Investors Process Information?
von: Kim, Alex, et al.
Veröffentlicht: (2023)
von: Kim, Alex, et al.
Veröffentlicht: (2023)
AI-Driven Alpha Decay: Algorithmic Homogenization, Reflexive Signal Erosion, and the Paradox of Intelligent Markets
von: Meng, Shuchen, et al.
Veröffentlicht: (2026)
von: Meng, Shuchen, et al.
Veröffentlicht: (2026)
A Financial Brain Scan of the LLM
von: Chen, Hui, et al.
Veröffentlicht: (2025)
von: Chen, Hui, et al.
Veröffentlicht: (2025)
PredictionMarketBench: A SWE-bench-Style Framework for Backtesting Trading Agents on Prediction Markets
von: Arora, Avi, et al.
Veröffentlicht: (2026)
von: Arora, Avi, et al.
Veröffentlicht: (2026)
Deep Learning for Art Market Valuation
von: Mei, Jianping, et al.
Veröffentlicht: (2025)
von: Mei, Jianping, et al.
Veröffentlicht: (2025)
DrafterBench: Benchmarking Large Language Models for Tasks Automation in Civil Engineering
von: Li, Yinsheng, et al.
Veröffentlicht: (2025)
von: Li, Yinsheng, et al.
Veröffentlicht: (2025)
MarketBench: Evaluating AI Agents as Market Participants
von: Fradkin, Andrey, et al.
Veröffentlicht: (2026)
von: Fradkin, Andrey, et al.
Veröffentlicht: (2026)
Can Large Language Models Improve Venture Capital Exit Timing After IPO?
von: Rashidi, Mohammadhossien
Veröffentlicht: (2025)
von: Rashidi, Mohammadhossien
Veröffentlicht: (2025)
Can Large Language Models Trade? Testing Financial Theories with LLM Agents in Market Simulations
von: Lopez-Lira, Alejandro
Veröffentlicht: (2025)
von: Lopez-Lira, Alejandro
Veröffentlicht: (2025)
FinMaster: A Holistic Benchmark for Mastering Full-Pipeline Financial Workflows with LLMs
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
FinRobot: Generative Business Process AI Agents for Enterprise Resource Planning in Finance
von: Yang, Hongyang, et al.
Veröffentlicht: (2025)
von: Yang, Hongyang, et al.
Veröffentlicht: (2025)
Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk
von: Chen, Zichen, et al.
Veröffentlicht: (2025)
von: Chen, Zichen, et al.
Veröffentlicht: (2025)
Evaluating Investment Risks in LATAM AI Startups: Ranking of Investment Potential and Framework for Valuation
von: Ramos-Torres, Abraham, et al.
Veröffentlicht: (2024)
von: Ramos-Torres, Abraham, et al.
Veröffentlicht: (2024)
A Survey of Sustainability in Large Language Models: Applications, Economics, and Challenges
von: Singh, Aditi, et al.
Veröffentlicht: (2024)
von: Singh, Aditi, et al.
Veröffentlicht: (2024)
A Scoping Review of ChatGPT Research in Accounting and Finance
von: Dong, Mengming Michael, et al.
Veröffentlicht: (2024)
von: Dong, Mengming Michael, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
StakeBench: Evaluating Language Understanding Grounded in Market Commitment
von: Pei, Yunhua, et al.
Veröffentlicht: (2026) -
Integrating Large Language Models in Financial Investments and Market Analysis: A Survey
von: Mahdavi, Sedigheh, et al.
Veröffentlicht: (2025) -
A Survey of Large Language Models for Financial Applications: Progress, Prospects and Challenges
von: Nie, Yuqi, et al.
Veröffentlicht: (2024) -
Large Language Models in Finance: A Survey
von: Li, Yinheng, et al.
Veröffentlicht: (2023) -
Financial Statement Analysis with Large Language Models
von: Kim, Alex, et al.
Veröffentlicht: (2024)