BEDTime: A Unified Benchmark for Automatically Describing Time Series
Fuente:
arXiv
Saved in:
| Main Authors: | Sen, Medhasweta, Gottesman, Zachary, Qiu, Jiaxing, Bruss, C. Bayan, Nguyen, Nam, Hartvigsen, Tom |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DynaGuard: A Dynamic Guardian Model With User-Defined Policies
by: Hoover, Monte, et al.
Published: (2025)
by: Hoover, Monte, et al.
Published: (2025)
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2025)
by: Agrawal, Aakriti, et al.
Published: (2025)
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
by: Thede, Lukas, et al.
Published: (2025)
by: Thede, Lukas, et al.
Published: (2025)
Instruction-based Time Series Editing
by: Qiu, Jiaxing, et al.
Published: (2025)
by: Qiu, Jiaxing, et al.
Published: (2025)
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
by: Jayawardhana, Mayuka, et al.
Published: (2026)
by: Jayawardhana, Mayuka, et al.
Published: (2026)
Domain-Independent Automatic Generation of Descriptive Texts for Time-Series Data
by: Dohi, Kota, et al.
Published: (2024)
by: Dohi, Kota, et al.
Published: (2024)
ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
by: Wang, Chengsen, et al.
Published: (2024)
by: Wang, Chengsen, et al.
Published: (2024)
VITRO: Vocabulary Inversion for Time-series Representation Optimization
by: Bellos, Filippos, et al.
Published: (2024)
by: Bellos, Filippos, et al.
Published: (2024)
Explainable Automatic Grading with Neural Additive Models
by: Condor, Aubrey, et al.
Published: (2024)
by: Condor, Aubrey, et al.
Published: (2024)
In-Context Fine-Tuning for Time-Series Foundation Models
by: Das, Abhimanyu, et al.
Published: (2024)
by: Das, Abhimanyu, et al.
Published: (2024)
Time-IMM: A Dataset and Benchmark for Irregular Multimodal Multivariate Time Series
by: Chang, Ching, et al.
Published: (2025)
by: Chang, Ching, et al.
Published: (2025)
The Few Govern the Many:Unveiling Few-Layer Dominance for Time Series Models
by: Qiu, Xin, et al.
Published: (2025)
by: Qiu, Xin, et al.
Published: (2025)
Estimating Semantic Alphabet Size for LLM Uncertainty Quantification
by: McCabe, Lucas H., et al.
Published: (2025)
by: McCabe, Lucas H., et al.
Published: (2025)
Constrained Discrete Diffusion
by: Cardei, Michael, et al.
Published: (2025)
by: Cardei, Michael, et al.
Published: (2025)
Automatic Prompt Selection for Large Language Models
by: Do, Viet-Tung, et al.
Published: (2024)
by: Do, Viet-Tung, et al.
Published: (2024)
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens
by: Ouyang, Xu, et al.
Published: (2024)
by: Ouyang, Xu, et al.
Published: (2024)
OWLViz: An Open-World Benchmark for Visual Question Answering
by: Nguyen, Thuy, et al.
Published: (2025)
by: Nguyen, Thuy, et al.
Published: (2025)
Integrating Sequential and Relational Modeling for User Events: Datasets and Prediction Tasks
by: Fathony, Rizal, et al.
Published: (2025)
by: Fathony, Rizal, et al.
Published: (2025)
Automatically Identifying Local and Global Circuits with Linear Computation Graphs
by: Ge, Xuyang, et al.
Published: (2024)
by: Ge, Xuyang, et al.
Published: (2024)
Sparse Transformer with Local and Seasonal Adaptation for Multivariate Time Series Forecasting
by: Zhang, Yifan, et al.
Published: (2023)
by: Zhang, Yifan, et al.
Published: (2023)
Time-MMD: Multi-Domain Multimodal Dataset for Time Series Analysis
by: Liu, Haoxin, et al.
Published: (2024)
by: Liu, Haoxin, et al.
Published: (2024)
AntiLeakBench: Preventing Data Contamination by Automatically Constructing Benchmarks with Updated Real-World Knowledge
by: Wu, Xiaobao, et al.
Published: (2024)
by: Wu, Xiaobao, et al.
Published: (2024)
Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization
by: Du, He, et al.
Published: (2026)
by: Du, He, et al.
Published: (2026)
A Simple Baseline for Predicting Events with Auto-Regressive Tabular Transformers
by: Stein, Alex, et al.
Published: (2024)
by: Stein, Alex, et al.
Published: (2024)
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models
by: Liu, Yong, et al.
Published: (2024)
by: Liu, Yong, et al.
Published: (2024)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
by: Xiao, Yuxin, et al.
Published: (2025)
by: Xiao, Yuxin, et al.
Published: (2025)
Time Series Language Model for Descriptive Caption Generation
by: Trabelsi, Mohamed, et al.
Published: (2025)
by: Trabelsi, Mohamed, et al.
Published: (2025)
Metadata Matters for Time Series: Informative Forecasting with Transformers
by: Dong, Jiaxiang, et al.
Published: (2024)
by: Dong, Jiaxiang, et al.
Published: (2024)
Unifying Linear-Time Attention via Latent Probabilistic Modelling
by: Dolga, Rares, et al.
Published: (2024)
by: Dolga, Rares, et al.
Published: (2024)
Temporal Tokenization Strategies for Event Sequence Modeling with Large Language Models
by: Liu, Zefang, et al.
Published: (2025)
by: Liu, Zefang, et al.
Published: (2025)
AI versus AI in Financial Crimes and Detection: GenAI Crime Waves to Co-Evolutionary AI
by: Kurshan, Eren, et al.
Published: (2024)
by: Kurshan, Eren, et al.
Published: (2024)
Can Multimodal LLMs Perform Time Series Anomaly Detection?
by: Xu, Xiongxiao, et al.
Published: (2025)
by: Xu, Xiongxiao, et al.
Published: (2025)
Using Pre-trained LLMs for Multivariate Time Series Forecasting
by: Wolff, Malcolm L., et al.
Published: (2025)
by: Wolff, Malcolm L., et al.
Published: (2025)
LLM-Mixer: Multiscale Mixing in LLMs for Time Series Forecasting
by: Kowsher, Md, et al.
Published: (2024)
by: Kowsher, Md, et al.
Published: (2024)
UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models
by: Nguyen, Nam V., et al.
Published: (2024)
by: Nguyen, Nam V., et al.
Published: (2024)
Automatically Labeling Clinical Trial Outcomes: A Large-Scale Benchmark for Drug Development
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
A Review on Generative AI Models for Synthetic Medical Text, Time Series, and Longitudinal Data
by: Loni, Mohammad, et al.
Published: (2024)
by: Loni, Mohammad, et al.
Published: (2024)
Benchmarking ChatGPT on Algorithmic Reasoning
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
Sparse Autoencoder Features for Classifications and Transferability
by: Gallifant, Jack, et al.
Published: (2025)
by: Gallifant, Jack, et al.
Published: (2025)
Similar Items
-
DynaGuard: A Dynamic Guardian Model With User-Defined Policies
by: Hoover, Monte, et al.
Published: (2025) -
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2025) -
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
by: Thede, Lukas, et al.
Published: (2025) -
Instruction-based Time Series Editing
by: Qiu, Jiaxing, et al.
Published: (2025) -
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
by: Jayawardhana, Mayuka, et al.
Published: (2026)