When Reasoning Fails: Evaluating 'Thinking' LLMs for Stock Prediction
Fuente:
arXiv
Saved in:
| Main Author: | Sodha, Rakeshkumar H |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
by: Hashemi, Helia, et al.
Published: (2024)
by: Hashemi, Helia, et al.
Published: (2024)
The Arrival of AGI? When Expert Personas Exceed Expert Benchmarks
by: Mullens, Drake, et al.
Published: (2026)
by: Mullens, Drake, et al.
Published: (2026)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025)
by: Petrov, Egor, et al.
Published: (2025)
Counterfactual Causal Inference in Natural Language with Large Language Models
by: Gendron, Gaël, et al.
Published: (2024)
by: Gendron, Gaël, et al.
Published: (2024)
Causal Cartographer: From Mapping to Reasoning Over Counterfactual Worlds
by: Gendron, Gaël, et al.
Published: (2025)
by: Gendron, Gaël, et al.
Published: (2025)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
by: Sauter, Andreas, et al.
Published: (2026)
by: Sauter, Andreas, et al.
Published: (2026)
Leveraging Large Language Models to Extract and Translate Medical Information in Doctors' Notes for Health Records and Diagnostic Billing Codes
by: Hartnett, Peter, et al.
Published: (2026)
by: Hartnett, Peter, et al.
Published: (2026)
Can LLMs Identify Tax Abuse?
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting
by: Lettner, Jan Niklas, et al.
Published: (2026)
by: Lettner, Jan Niklas, et al.
Published: (2026)
BioAlchemy: Distilling Biological Literature into Reasoning-Ready Reinforcement Learning Training Data
by: Hsu, Brian, et al.
Published: (2026)
by: Hsu, Brian, et al.
Published: (2026)
LLMs as Architects and Critics for Multi-Source Opinion Summarization
by: Attri, Anuj, et al.
Published: (2025)
by: Attri, Anuj, et al.
Published: (2025)
The Reasoning-Creativity Trade-off: Toward Creativity-Driven Problem Solving
by: Luyten, Max Ruiz, et al.
Published: (2026)
by: Luyten, Max Ruiz, et al.
Published: (2026)
Deep Policy Iteration with Integer Programming for Inventory Management
by: Harsha, Pavithra, et al.
Published: (2021)
by: Harsha, Pavithra, et al.
Published: (2021)
Training Language Models to Use Prolog as a Tool
by: Mellgren, Niklas, et al.
Published: (2025)
by: Mellgren, Niklas, et al.
Published: (2025)
Survey Transfer Learning: Recycling Data with Silicon Responses
by: Amini, Ali
Published: (2025)
by: Amini, Ali
Published: (2025)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
by: Zhang, Xue
Published: (2025)
by: Zhang, Xue
Published: (2025)
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
by: Yost, Alexandra, et al.
Published: (2025)
by: Yost, Alexandra, et al.
Published: (2025)
Shallow Robustness, Deep Vulnerabilities: Multi-Turn Evaluation of Medical LLMs
by: Manczak, Blazej, et al.
Published: (2025)
by: Manczak, Blazej, et al.
Published: (2025)
Model selection meets clinical semantics: Optimizing ICD-10-CM prediction via LLM-as-Judge evaluation, redundancy-aware sampling, and section-aware fine-tuning
by: Dai, Hong-Jie, et al.
Published: (2025)
by: Dai, Hong-Jie, et al.
Published: (2025)
Ask WhAI:Probing Belief Formation in Role-Primed LLM Agents
by: Moore, Keith, et al.
Published: (2025)
by: Moore, Keith, et al.
Published: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
A Lightweight Multi-Expert Generative Language Model System for Engineering Information and Knowledge Extraction
by: Bogachov, Bogdan, et al.
Published: (2025)
by: Bogachov, Bogdan, et al.
Published: (2025)
Thinking Machines: Mathematical Reasoning in the Age of LLMs
by: Asperti, Andrea, et al.
Published: (2025)
by: Asperti, Andrea, et al.
Published: (2025)
Improving ML Training Data with Gold-Standard Quality Metrics
by: Barrett, Leslie, et al.
Published: (2025)
by: Barrett, Leslie, et al.
Published: (2025)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
by: Zhang, Yiqing, et al.
Published: (2026)
by: Zhang, Yiqing, et al.
Published: (2026)
A Graph-based RAG for Energy Efficiency Question Answering
by: Campi, Riccardo, et al.
Published: (2025)
by: Campi, Riccardo, et al.
Published: (2025)
Contrastive Similarity Learning for Market Forecasting: The ContraSim Framework
by: Vinden, Nicholas, et al.
Published: (2025)
by: Vinden, Nicholas, et al.
Published: (2025)
IntelliCode: A Multi-Agent LLM Tutoring System with Centralized Learner Modeling
by: David, Jones, et al.
Published: (2025)
by: David, Jones, et al.
Published: (2025)
Why We Feel What We Feel: Joint Detection of Emotions and Their Opinion Triggers in E-commerce
by: Attri, Arnav, et al.
Published: (2025)
by: Attri, Arnav, et al.
Published: (2025)
Hierarchical Pooling and Explainability in Graph Neural Networks for Tumor and Tissue-of-Origin Classification Using RNA-seq Data
by: Fontanari, Thomas Vaitses, et al.
Published: (2026)
by: Fontanari, Thomas Vaitses, et al.
Published: (2026)
Cryptogenic stroke and migraine: using probabilistic independence and machine learning to uncover latent sources of disease from the electronic health record
by: Betts, Joshua W., et al.
Published: (2025)
by: Betts, Joshua W., et al.
Published: (2025)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
by: van der Meer, Virgill, et al.
Published: (2026)
by: van der Meer, Virgill, et al.
Published: (2026)
Hierarchical Dual-Head Model for Suicide Risk Assessment via MentalRoBERTa
by: Yang, Chang, et al.
Published: (2025)
by: Yang, Chang, et al.
Published: (2025)
CS-Guide: Leveraging LLMs and Student Reflections to Provide Frequent, Scalable Academic Monitoring Feedback to Computer Science Students
by: Chacko, Samuel Jacob, et al.
Published: (2025)
by: Chacko, Samuel Jacob, et al.
Published: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
by: Pawar, Tejas, et al.
Published: (2025)
by: Pawar, Tejas, et al.
Published: (2025)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
by: Sun, Yuhui, et al.
Published: (2025)
by: Sun, Yuhui, et al.
Published: (2025)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
by: Reddy, Sandeep, et al.
Published: (2025)
by: Reddy, Sandeep, et al.
Published: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
Evaluating Large Language Models for IUCN Red List Species Information
by: Uryu, Shinya
Published: (2025)
by: Uryu, Shinya
Published: (2025)
Steering Conceptual Bias via Transformer Latent-Subspace Activation
by: Sharma, Vansh, et al.
Published: (2025)
by: Sharma, Vansh, et al.
Published: (2025)
Similar Items
-
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
by: Hashemi, Helia, et al.
Published: (2024) -
The Arrival of AGI? When Expert Personas Exceed Expert Benchmarks
by: Mullens, Drake, et al.
Published: (2026) -
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025) -
Counterfactual Causal Inference in Natural Language with Large Language Models
by: Gendron, Gaël, et al.
Published: (2024) -
Causal Cartographer: From Mapping to Reasoning Over Counterfactual Worlds
by: Gendron, Gaël, et al.
Published: (2025)