Evaluating Large Language Models on Time Series Feature Understanding: A Comprehensive Taxonomy and Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Fons, Elizabeth, Kaur, Rachneet, Palande, Soham, Zeng, Zhen, Balch, Tucker, Veloso, Manuela, Vyetrenko, Svitlana |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TADACap: Time-series Adaptive Domain-Aware Captioning
by: Fons, Elizabeth, et al.
Published: (2025)
by: Fons, Elizabeth, et al.
Published: (2025)
AI Analyst: Framework and Comprehensive Evaluation of Large Language Models for Financial Time Series Report Generation
by: Fons, Elizabeth, et al.
Published: (2025)
by: Fons, Elizabeth, et al.
Published: (2025)
LETS-C: Leveraging Text Embedding for Time Series Classification
by: Kaur, Rachneet, et al.
Published: (2024)
by: Kaur, Rachneet, et al.
Published: (2024)
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
Multi-Modal Financial Time-Series Retrieval Through Latent Space Projections
by: Bamford, Tom, et al.
Published: (2023)
by: Bamford, Tom, et al.
Published: (2023)
AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
by: Verma, Gaurav, et al.
Published: (2024)
by: Verma, Gaurav, et al.
Published: (2024)
A Language Model-Guided Framework for Mining Time Series with Distributional Shifts
by: Zhu, Haibei, et al.
Published: (2024)
by: Zhu, Haibei, et al.
Published: (2024)
ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering
by: Kaur, Rachneet, et al.
Published: (2025)
by: Kaur, Rachneet, et al.
Published: (2025)
TS-Agent: Understanding and Reasoning Over Raw Time Series via Iterative Insight Gathering
by: Liu, Penghang, et al.
Published: (2025)
by: Liu, Penghang, et al.
Published: (2025)
LSCD: Lomb-Scargle Conditioned Diffusion for Time series Imputation
by: Fons, Elizabeth, et al.
Published: (2025)
by: Fons, Elizabeth, et al.
Published: (2025)
Towards Interpretable Time Series Foundation Models
by: Boileau, Matthieu, et al.
Published: (2025)
by: Boileau, Matthieu, et al.
Published: (2025)
LAW: Legal Agentic Workflows for Custody and Fund Services Contracts
by: Watson, William, et al.
Published: (2024)
by: Watson, William, et al.
Published: (2024)
SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding
by: Jin, Yiqiao, et al.
Published: (2025)
by: Jin, Yiqiao, et al.
Published: (2025)
HiddenTables & PyQTax: A Cooperative Game and Dataset For TableQA to Ensure Scale and Data Privacy Across a Myriad of Taxonomies
by: Watson, William, et al.
Published: (2024)
by: Watson, William, et al.
Published: (2024)
Empirical Equilibria in Agent-based Economic systems with Learning agents
by: Dwarakanath, Kshama, et al.
Published: (2024)
by: Dwarakanath, Kshama, et al.
Published: (2024)
ABIDES-Economist: Agent-Based Simulator of Economic Systems with Learning Agents
by: Dwarakanath, Kshama, et al.
Published: (2024)
by: Dwarakanath, Kshama, et al.
Published: (2024)
Ensemble Methods for Sequence Classification with Hidden Markov Models
by: Kawawa-Beaudan, Maxime, et al.
Published: (2024)
by: Kawawa-Beaudan, Maxime, et al.
Published: (2024)
Behavioral Sequence Modeling with Ensemble Learning
by: Kawawa-Beaudan, Maxime, et al.
Published: (2024)
by: Kawawa-Beaudan, Maxime, et al.
Published: (2024)
Augment on Manifold: Mixup Regularization with UMAP
by: El-Laham, Yousef, et al.
Published: (2023)
by: El-Laham, Yousef, et al.
Published: (2023)
Transparency as Delayed Observability in Multi-Agent Systems
by: Dwarakanath, Kshama, et al.
Published: (2024)
by: Dwarakanath, Kshama, et al.
Published: (2024)
Limited or Biased: Modeling Sub-Rational Human Investors in Financial Markets
by: Liu, Penghang, et al.
Published: (2022)
by: Liu, Penghang, et al.
Published: (2022)
FlowMind: Automatic Workflow Generation with LLMs
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
Privacy-Aware Time Series Synthesis via Public Knowledge Distillation
by: Liu, Penghang, et al.
Published: (2025)
by: Liu, Penghang, et al.
Published: (2025)
LLM-driven Imitation of Subrational Behavior : Illusion or Reality?
by: Coletta, Andrea, et al.
Published: (2024)
by: Coletta, Andrea, et al.
Published: (2024)
COLE: a Comprehensive Benchmark for French Language Understanding Evaluation
by: Beauchemin, David, et al.
Published: (2025)
by: Beauchemin, David, et al.
Published: (2025)
Mixup Regularization: A Probabilistic Perspective
by: El-Laham, Yousef, et al.
Published: (2025)
by: El-Laham, Yousef, et al.
Published: (2025)
Are Emergent Abilities in Large Language Models just In-Context Learning?
by: Lu, Sheng, et al.
Published: (2023)
by: Lu, Sheng, et al.
Published: (2023)
Variational Neural Stochastic Differential Equations with Change Points
by: El-Laham, Yousef, et al.
Published: (2024)
by: El-Laham, Yousef, et al.
Published: (2024)
SinhalaMMLU: A Comprehensive Benchmark for Evaluating Multitask Language Understanding in Sinhala
by: Pramodya, Ashmari, et al.
Published: (2025)
by: Pramodya, Ashmari, et al.
Published: (2025)
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design
by: Neehal, Nafis, et al.
Published: (2024)
by: Neehal, Nafis, et al.
Published: (2024)
HKCanto-Eval: A Benchmark for Evaluating Cantonese Language Understanding and Cultural Comprehension in LLMs
by: Cheng, Tsz Chung, et al.
Published: (2025)
by: Cheng, Tsz Chung, et al.
Published: (2025)
Once Burned, Twice Shy? The Effect of Stock Market Bubbles on Traders that Learn by Experience
by: Zhu, Haibei, et al.
Published: (2023)
by: Zhu, Haibei, et al.
Published: (2023)
Taxonomy-based CheckList for Large Language Model Evaluation
by: Zhang, Damin
Published: (2023)
by: Zhang, Damin
Published: (2023)
C$^{3}$Bench: A Comprehensive Classical Chinese Understanding Benchmark for Large Language Models
by: Cao, Jiahuan, et al.
Published: (2024)
by: Cao, Jiahuan, et al.
Published: (2024)
Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems
by: Cui, Tianyu, et al.
Published: (2024)
by: Cui, Tianyu, et al.
Published: (2024)
Rating Multi-Modal Time-Series Forecasting Models (MM-TSFM) for Robustness Through a Causal Lens
by: Lakkaraju, Kausik, et al.
Published: (2024)
by: Lakkaraju, Kausik, et al.
Published: (2024)
Towards Comprehensive Stage-wise Benchmarking of Large Language Models in Fact-Checking
by: Lin, Hongzhan, et al.
Published: (2026)
by: Lin, Hongzhan, et al.
Published: (2026)
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions
by: Sachdeva, Rachneet, et al.
Published: (2025)
by: Sachdeva, Rachneet, et al.
Published: (2025)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
by: Sachdeva, Rachneet, et al.
Published: (2023)
by: Sachdeva, Rachneet, et al.
Published: (2023)
SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
by: Xia, Haotian, et al.
Published: (2024)
by: Xia, Haotian, et al.
Published: (2024)
Similar Items
-
TADACap: Time-series Adaptive Domain-Aware Captioning
by: Fons, Elizabeth, et al.
Published: (2025) -
AI Analyst: Framework and Comprehensive Evaluation of Large Language Models for Financial Time Series Report Generation
by: Fons, Elizabeth, et al.
Published: (2025) -
LETS-C: Leveraging Text Embedding for Time Series Classification
by: Kaur, Rachneet, et al.
Published: (2024) -
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting
by: Zeng, Zhen, et al.
Published: (2024) -
Multi-Modal Financial Time-Series Retrieval Through Latent Space Projections
by: Bamford, Tom, et al.
Published: (2023)