XForecast: Evaluating Natural Language Explanations for Time Series Forecasting
Fuente:
arXiv
Guardado en:
| Autores principales: | Aksu, Taha, Liu, Chenghao, Saha, Amrita, Tan, Sarah, Xiong, Caiming, Sahoo, Doyen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation
por: Aksu, Taha, et al.
Publicado: (2024)
por: Aksu, Taha, et al.
Publicado: (2024)
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
por: Zhao, Zirui, et al.
Publicado: (2024)
por: Zhao, Zirui, et al.
Publicado: (2024)
Moirai 2.0: When Less Is More for Time Series Forecasting
por: Liu, Chenghao, et al.
Publicado: (2025)
por: Liu, Chenghao, et al.
Publicado: (2025)
Unified Training of Universal Time Series Forecasting Transformers
por: Woo, Gerald, et al.
Publicado: (2024)
por: Woo, Gerald, et al.
Publicado: (2024)
Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
por: Xu, Yiheng, et al.
Publicado: (2024)
por: Xu, Yiheng, et al.
Publicado: (2024)
MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
por: Luo, Ziyang, et al.
Publicado: (2025)
por: Luo, Ziyang, et al.
Publicado: (2025)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
por: Wang, Lei, et al.
Publicado: (2024)
por: Wang, Lei, et al.
Publicado: (2024)
ThinK: Thinner Key Cache by Query-Driven Pruning
por: Xu, Yuhui, et al.
Publicado: (2024)
por: Xu, Yuhui, et al.
Publicado: (2024)
CodeTree: Agent-guided Tree Search for Code Generation with Large Language Models
por: Li, Jierui, et al.
Publicado: (2024)
por: Li, Jierui, et al.
Publicado: (2024)
UniTST: Effectively Modeling Inter-Series and Intra-Series Dependencies for Multivariate Time Series Forecasting
por: Liu, Juncheng, et al.
Publicado: (2024)
por: Liu, Juncheng, et al.
Publicado: (2024)
CodeChain: Towards Modular Code Generation Through Chain of Self-revisions with Representative Sub-modules
por: Le, Hung, et al.
Publicado: (2023)
por: Le, Hung, et al.
Publicado: (2023)
Moirai-MoE: Empowering Time Series Foundation Models with Sparse Mixture of Experts
por: Liu, Xu, et al.
Publicado: (2024)
por: Liu, Xu, et al.
Publicado: (2024)
INDICT: Code Generation with Internal Dialogues of Critiques for Both Security and Helpfulness
por: Le, Hung, et al.
Publicado: (2024)
por: Le, Hung, et al.
Publicado: (2024)
Empowering Time Series Analysis with Synthetic Data: A Survey and Outlook in the Era of Foundation Models
por: Liu, Xu, et al.
Publicado: (2025)
por: Liu, Xu, et al.
Publicado: (2025)
Entropy-Based Block Pruning for Efficient Large Language Models
por: Yang, Liangwei, et al.
Publicado: (2025)
por: Yang, Liangwei, et al.
Publicado: (2025)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
por: Peng, Yun, et al.
Publicado: (2024)
por: Peng, Yun, et al.
Publicado: (2024)
Scalable Chain of Thoughts via Elastic Reasoning
por: Xu, Yuhui, et al.
Publicado: (2025)
por: Xu, Yuhui, et al.
Publicado: (2025)
Fractured Chain-of-Thought Reasoning
por: Liao, Baohao, et al.
Publicado: (2025)
por: Liao, Baohao, et al.
Publicado: (2025)
Granular Change Accuracy: A More Accurate Performance Metric for Dialogue State Tracking
por: Aksu, Taha, et al.
Publicado: (2024)
por: Aksu, Taha, et al.
Publicado: (2024)
Reward-Guided Speculative Decoding for Efficient LLM Reasoning
por: Liao, Baohao, et al.
Publicado: (2025)
por: Liao, Baohao, et al.
Publicado: (2025)
Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models
por: Hu, Zhiyuan, et al.
Publicado: (2025)
por: Hu, Zhiyuan, et al.
Publicado: (2025)
Man Made Language Models? Evaluating LLMs' Perpetuation of Masculine Generics Bias
por: Doyen, Enzo, et al.
Publicado: (2025)
por: Doyen, Enzo, et al.
Publicado: (2025)
LExT: Towards Evaluating Trustworthiness of Natural Language Explanations
por: Shailya, Krithi, et al.
Publicado: (2025)
por: Shailya, Krithi, et al.
Publicado: (2025)
RLHF Workflow: From Reward Modeling to Online RLHF
por: Dong, Hanze, et al.
Publicado: (2024)
por: Dong, Hanze, et al.
Publicado: (2024)
Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators
por: Zhou, Yilun, et al.
Publicado: (2025)
por: Zhou, Yilun, et al.
Publicado: (2025)
Situated Natural Language Explanations
por: Zhu, Zining, et al.
Publicado: (2023)
por: Zhu, Zining, et al.
Publicado: (2023)
On the Importance and Evaluation of Narrativity in Natural Language AI Explanations
por: Cedro, Mateusz, et al.
Publicado: (2026)
por: Cedro, Mateusz, et al.
Publicado: (2026)
A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
por: Xiong, Wei, et al.
Publicado: (2025)
por: Xiong, Wei, et al.
Publicado: (2025)
Fusing Large Language Models with Temporal Transformers for Time Series Forecasting
por: Su, Chen, et al.
Publicado: (2025)
por: Su, Chen, et al.
Publicado: (2025)
Explanation based Bias Decoupling Regularization for Natural Language Inference
por: Zang, Jianxiang, et al.
Publicado: (2024)
por: Zang, Jianxiang, et al.
Publicado: (2024)
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models
por: Liu, Yong, et al.
Publicado: (2024)
por: Liu, Yong, et al.
Publicado: (2024)
Reasoning with Natural Language Explanations
por: Valentino, Marco, et al.
Publicado: (2024)
por: Valentino, Marco, et al.
Publicado: (2024)
Sonar-TS: Search-Then-Verify Natural Language Querying for Time Series Databases
por: Tan, Zhao, et al.
Publicado: (2026)
por: Tan, Zhao, et al.
Publicado: (2026)
Retrieval-augmented Large Language Models for Financial Time Series Forecasting
por: Xiao, Mengxi, et al.
Publicado: (2025)
por: Xiao, Mengxi, et al.
Publicado: (2025)
A Taxonomy for Design and Evaluation of Prompt-Based Natural Language Explanations
por: Nejadgholi, Isar, et al.
Publicado: (2025)
por: Nejadgholi, Isar, et al.
Publicado: (2025)
Unanswerability Evaluation for Retrieval Augmented Generation
por: Peng, Xiangyu, et al.
Publicado: (2024)
por: Peng, Xiangyu, et al.
Publicado: (2024)
Retrieving Time-Series Differences Using Natural Language Queries
por: Dohi, Kota, et al.
Publicado: (2025)
por: Dohi, Kota, et al.
Publicado: (2025)
Lemur: Harmonizing Natural Language and Code for Language Agents
por: Xu, Yiheng, et al.
Publicado: (2023)
por: Xu, Yiheng, et al.
Publicado: (2023)
LLM-as-a-Judge for Time Series Explanations
por: Sivalingam, Preetham, et al.
Publicado: (2026)
por: Sivalingam, Preetham, et al.
Publicado: (2026)
DiffNator: Generating Structured Explanations of Time-Series Differences
por: Dohi, Kota, et al.
Publicado: (2025)
por: Dohi, Kota, et al.
Publicado: (2025)
Ejemplares similares
-
GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation
por: Aksu, Taha, et al.
Publicado: (2024) -
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
por: Zhao, Zirui, et al.
Publicado: (2024) -
Moirai 2.0: When Less Is More for Time Series Forecasting
por: Liu, Chenghao, et al.
Publicado: (2025) -
Unified Training of Universal Time Series Forecasting Transformers
por: Woo, Gerald, et al.
Publicado: (2024) -
Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
por: Xu, Yiheng, et al.
Publicado: (2024)