LLM-as-a-Prophet: Understanding Predictive Intelligence with Prophet Arena
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Qingchuan, Mahns, Simon, Li, Sida, Gu, Anri, Wu, Jibang, Xu, Haifeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TopicProphet: Prophesies on Temporal Topic Trends and Stocks
por: Kim, Olivia
Publicado: (2025)
por: Kim, Olivia
Publicado: (2025)
Measuring all the noises of LLM Evals
por: Wang, Sida
Publicado: (2025)
por: Wang, Sida
Publicado: (2025)
AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting
por: Wu, Jibang, et al.
Publicado: (2025)
por: Wu, Jibang, et al.
Publicado: (2025)
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
por: Yang, Chenghao, et al.
Publicado: (2025)
por: Yang, Chenghao, et al.
Publicado: (2025)
Stock Market Price Prediction using Neural Prophet with Deep Neural Network
por: Chhibber, Navin, et al.
Publicado: (2026)
por: Chhibber, Navin, et al.
Publicado: (2026)
SpreadsheetArena: Decomposing Preference in LLM Generation of Spreadsheet Workbooks
por: Kundurthy, Srivatsa, et al.
Publicado: (2026)
por: Kundurthy, Srivatsa, et al.
Publicado: (2026)
Evaluating LLM Understanding via Structured Tabular Decision Simulations
por: Li, Sichao, et al.
Publicado: (2025)
por: Li, Sichao, et al.
Publicado: (2025)
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena
por: Luo, Haipeng, et al.
Publicado: (2024)
por: Luo, Haipeng, et al.
Publicado: (2024)
Understanding LLM Embeddings for Regression
por: Tang, Eric, et al.
Publicado: (2024)
por: Tang, Eric, et al.
Publicado: (2024)
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
por: Younesian, Sharareh, et al.
Publicado: (2026)
por: Younesian, Sharareh, et al.
Publicado: (2026)
TextArena
por: Guertler, Leon, et al.
Publicado: (2025)
por: Guertler, Leon, et al.
Publicado: (2025)
Agent Trading Arena: A Study on Numerical Understanding in LLM-Based Agents
por: Ma, Tianmi, et al.
Publicado: (2025)
por: Ma, Tianmi, et al.
Publicado: (2025)
Contractual Reinforcement Learning: Pulling Arms with Invisible Hands
por: Wu, Jibang, et al.
Publicado: (2024)
por: Wu, Jibang, et al.
Publicado: (2024)
Understanding the planning of LLM agents: A survey
por: Huang, Xu, et al.
Publicado: (2024)
por: Huang, Xu, et al.
Publicado: (2024)
Bridging Human and LLM Judgments: Understanding and Narrowing the Gap
por: Polo, Felipe Maia, et al.
Publicado: (2025)
por: Polo, Felipe Maia, et al.
Publicado: (2025)
Eureka: Intelligent Feature Engineering for Enterprise AI Cloud Resource Demand Prediction
por: Li, Hangxuan, et al.
Publicado: (2026)
por: Li, Hangxuan, et al.
Publicado: (2026)
From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
por: Li, Tianle, et al.
Publicado: (2024)
por: Li, Tianle, et al.
Publicado: (2024)
ClawArena: Benchmarking AI Agents in Evolving Information Environments
por: Ji, Haonian, et al.
Publicado: (2026)
por: Ji, Haonian, et al.
Publicado: (2026)
BPO: Staying Close to the Behavior LLM Creates Better Online LLM Alignment
por: Xu, Wenda, et al.
Publicado: (2024)
por: Xu, Wenda, et al.
Publicado: (2024)
RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
por: Wang, Zihan, et al.
Publicado: (2025)
por: Wang, Zihan, et al.
Publicado: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
por: Wang, Ganghua, et al.
Publicado: (2025)
por: Wang, Ganghua, et al.
Publicado: (2025)
FedProphet: Memory-Efficient Federated Adversarial Training via Robust and Consistent Cascade Learning
por: Tang, Minxue, et al.
Publicado: (2024)
por: Tang, Minxue, et al.
Publicado: (2024)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
por: Li, Miaomiao, et al.
Publicado: (2025)
por: Li, Miaomiao, et al.
Publicado: (2025)
Fine-Tuning Improves Information Conveyance in Language Models
por: Cheng, Yuwei, et al.
Publicado: (2026)
por: Cheng, Yuwei, et al.
Publicado: (2026)
WebArena: A Realistic Web Environment for Building Autonomous Agents
por: Zhou, Shuyan, et al.
Publicado: (2023)
por: Zhou, Shuyan, et al.
Publicado: (2023)
Understanding World or Predicting Future? A Comprehensive Survey of World Models
por: Ding, Jingtao, et al.
Publicado: (2024)
por: Ding, Jingtao, et al.
Publicado: (2024)
SafeArena: Evaluating the Safety of Autonomous Web Agents
por: Tur, Ada Defne, et al.
Publicado: (2025)
por: Tur, Ada Defne, et al.
Publicado: (2025)
Understanding and Mitigating Dataset Corruption in LLM Steering
por: Anderson, Cullen, et al.
Publicado: (2026)
por: Anderson, Cullen, et al.
Publicado: (2026)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
por: Kirk, Robert, et al.
Publicado: (2023)
por: Kirk, Robert, et al.
Publicado: (2023)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
LLM-Guided Indoor Navigation with Multimodal Map Understanding
por: Coffrini, Alberto, et al.
Publicado: (2025)
por: Coffrini, Alberto, et al.
Publicado: (2025)
ArenaBencher: Automatic Benchmark Evolution via Multi-Model Competitive Evaluation
por: Liu, Qin, et al.
Publicado: (2025)
por: Liu, Qin, et al.
Publicado: (2025)
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
por: Choudhury, Deepro, et al.
Publicado: (2025)
por: Choudhury, Deepro, et al.
Publicado: (2025)
OnPrem.LLM: A Privacy-Conscious Document Intelligence Toolkit
por: Maiya, Arun S.
Publicado: (2025)
por: Maiya, Arun S.
Publicado: (2025)
Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression
por: Park, Jungsoo, et al.
Publicado: (2026)
por: Park, Jungsoo, et al.
Publicado: (2026)
Health-LLM: Large Language Models for Health Prediction via Wearable Sensor Data
por: Kim, Yubin, et al.
Publicado: (2024)
por: Kim, Yubin, et al.
Publicado: (2024)
PredictaBoard: Benchmarking LLM Score Predictability
por: Pacchiardi, Lorenzo, et al.
Publicado: (2025)
por: Pacchiardi, Lorenzo, et al.
Publicado: (2025)
WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks
por: Miyai, Atsuyuki, et al.
Publicado: (2025)
por: Miyai, Atsuyuki, et al.
Publicado: (2025)
On the Design of KL-Regularized Policy Gradient Algorithms for LLM Reasoning
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
por: Xu, Ran, et al.
Publicado: (2025)
por: Xu, Ran, et al.
Publicado: (2025)
Ejemplares similares
-
TopicProphet: Prophesies on Temporal Topic Trends and Stocks
por: Kim, Olivia
Publicado: (2025) -
Measuring all the noises of LLM Evals
por: Wang, Sida
Publicado: (2025) -
AI Realtor: Towards Grounded Persuasive Language Generation for Automated Copywriting
por: Wu, Jibang, et al.
Publicado: (2025) -
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
por: Yang, Chenghao, et al.
Publicado: (2025) -
Stock Market Price Prediction using Neural Prophet with Deep Neural Network
por: Chhibber, Navin, et al.
Publicado: (2026)