CaTS-Bench: Can Language Models Describe Time Series?
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Luca, Yashwante, Pratham, Fisher, Marshall, Sampieri, Alessio, Zhou, Zihao, Galasso, Fabio, Yu, Rose |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time Series, Vision, and Language: Exploring the Limits of Alignment in Contrastive Representation Spaces
by: Yashwante, Pratham, et al.
Published: (2026)
by: Yashwante, Pratham, et al.
Published: (2026)
Length-Aware Motion Synthesis via Latent Diffusion
by: Sampieri, Alessio, et al.
Published: (2024)
by: Sampieri, Alessio, et al.
Published: (2024)
Social EgoMesh Estimation
by: Scofano, Luca, et al.
Published: (2024)
by: Scofano, Luca, et al.
Published: (2024)
How Do Inpainting Artifacts Propagate to Language?
by: Yashwante, Pratham, et al.
Published: (2026)
by: Yashwante, Pratham, et al.
Published: (2026)
About latent roles in forecasting players in team sports
by: Scofano, Luca, et al.
Published: (2023)
by: Scofano, Luca, et al.
Published: (2023)
Human Motion Unlearning
by: De Matteis, Edoardo, et al.
Published: (2025)
by: De Matteis, Edoardo, et al.
Published: (2025)
Following the Human Thread in Social Navigation
by: Scofano, Luca, et al.
Published: (2024)
by: Scofano, Luca, et al.
Published: (2024)
Contracting Skeletal Kinematics for Human-Related Video Anomaly Detection
by: Flaborea, Alessandro, et al.
Published: (2023)
by: Flaborea, Alessandro, et al.
Published: (2023)
Video Unlearning via Low-Rank Refusal Vector
by: Facchiano, Simone, et al.
Published: (2025)
by: Facchiano, Simone, et al.
Published: (2025)
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
by: Collorone, Luca, et al.
Published: (2025)
by: Collorone, Luca, et al.
Published: (2025)
Can LLMs Understand Time Series Anomalies?
by: Zhou, Zihao, et al.
Published: (2024)
by: Zhou, Zihao, et al.
Published: (2024)
MoDiPO: text-to-motion alignment via AI-feedback-driven Direct Preference Optimization
by: Pappa, Massimiliano, et al.
Published: (2024)
by: Pappa, Massimiliano, et al.
Published: (2024)
MLLM4TS: Leveraging Vision and Multimodal Language Models for General Time-Series Analysis
by: Liu, Qinghua, et al.
Published: (2025)
by: Liu, Qinghua, et al.
Published: (2025)
SeRpEnt: Selective Resampling for Expressive State Space Models
by: Rando, Stefano, et al.
Published: (2025)
by: Rando, Stefano, et al.
Published: (2025)
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
by: Bao, Han, et al.
Published: (2024)
by: Bao, Han, et al.
Published: (2024)
TriTS: Time Series Forecasting from a Multimodal Perspective
by: Ao, Xiang
Published: (2026)
by: Ao, Xiang
Published: (2026)
MonSTeR: a Unified Model for Motion, Scene, Text Retrieval
by: Collorone, Luca, et al.
Published: (2025)
by: Collorone, Luca, et al.
Published: (2025)
TS-SatFire: A Multi-Task Satellite Image Time-Series Dataset for Wildfire Detection and Prediction
by: Zhao, Yu, et al.
Published: (2024)
by: Zhao, Yu, et al.
Published: (2024)
VisionTS++: Cross-Modal Time Series Foundation Model with Continual Pre-trained Vision Backbones
by: Shen, Lefei, et al.
Published: (2025)
by: Shen, Lefei, et al.
Published: (2025)
LongCodeBench: Evaluating Coding LLMs at 1M Context Windows
by: Rando, Stefano, et al.
Published: (2025)
by: Rando, Stefano, et al.
Published: (2025)
Modal Aphasia: Can Unified Multimodal Models Describe Images From Memory?
by: Aerni, Michael, et al.
Published: (2025)
by: Aerni, Michael, et al.
Published: (2025)
AgriBench: A Hierarchical Agriculture Benchmark for Multimodal Large Language Models
by: Zhou, Yutong, et al.
Published: (2024)
by: Zhou, Yutong, et al.
Published: (2024)
TS-P$^2$CL: Plug-and-Play Dual Contrastive Learning for Vision-Guided Medical Time Series Classification
by: Xu, Qi'ao, et al.
Published: (2025)
by: Xu, Qi'ao, et al.
Published: (2025)
Describe Anything Anywhere At Any Moment
by: Gorlo, Nicolas, et al.
Published: (2025)
by: Gorlo, Nicolas, et al.
Published: (2025)
Hyperbolic Active Learning for Semantic Segmentation under Domain Shift
by: Franco, Luca, et al.
Published: (2023)
by: Franco, Luca, et al.
Published: (2023)
DescribeEarth: Describe Anything for Remote Sensing Images
by: Li, Kaiyu, et al.
Published: (2025)
by: Li, Kaiyu, et al.
Published: (2025)
TS3IM: Unveiling Structural Similarity in Time Series through Image Similarity Assessment Insights
by: Liu, Yuhan, et al.
Published: (2024)
by: Liu, Yuhan, et al.
Published: (2024)
StegaVision: Enhancing Steganography with Attention Mechanism
by: Kumar, Abhinav, et al.
Published: (2024)
by: Kumar, Abhinav, et al.
Published: (2024)
TS-SAM: Fine-Tuning Segment-Anything Model for Downstream Tasks
by: Yu, Yang, et al.
Published: (2024)
by: Yu, Yang, et al.
Published: (2024)
VisualActBench: Can VLMs See and Act like a Human?
by: Zhang, Daoan, et al.
Published: (2025)
by: Zhang, Daoan, et al.
Published: (2025)
TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models
by: Zhang, Junyi, et al.
Published: (2025)
by: Zhang, Junyi, et al.
Published: (2025)
VARS: Vision-based Assessment of Risk in Security Systems
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
VisionTS: Visual Masked Autoencoders Are Free-Lunch Zero-Shot Time Series Forecasters
by: Chen, Mouxiang, et al.
Published: (2024)
by: Chen, Mouxiang, et al.
Published: (2024)
Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval
by: Wang, Zhichuan, et al.
Published: (2025)
by: Wang, Zhichuan, et al.
Published: (2025)
DiffuSyn Bench: Evaluating Vision-Language Models on Real-World Complexities with Diffusion-Generated Synthetic Benchmarks
by: Zhou, Haokun, et al.
Published: (2024)
by: Zhou, Haokun, et al.
Published: (2024)
ANTHROPOS-V: benchmarking the novel task of Crowd Volume Estimation
by: Collorone, Luca, et al.
Published: (2025)
by: Collorone, Luca, et al.
Published: (2025)
VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?
by: Liu, Yuanxin, et al.
Published: (2025)
by: Liu, Yuanxin, et al.
Published: (2025)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
by: Pal, Avik, et al.
Published: (2024)
by: Pal, Avik, et al.
Published: (2024)
Give me a hint: Can LLMs take a hint to solve math problems?
by: Agrawal, Vansh, et al.
Published: (2024)
by: Agrawal, Vansh, et al.
Published: (2024)
Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent
by: Ci, En, et al.
Published: (2025)
by: Ci, En, et al.
Published: (2025)
Similar Items
-
Time Series, Vision, and Language: Exploring the Limits of Alignment in Contrastive Representation Spaces
by: Yashwante, Pratham, et al.
Published: (2026) -
Length-Aware Motion Synthesis via Latent Diffusion
by: Sampieri, Alessio, et al.
Published: (2024) -
Social EgoMesh Estimation
by: Scofano, Luca, et al.
Published: (2024) -
How Do Inpainting Artifacts Propagate to Language?
by: Yashwante, Pratham, et al.
Published: (2026) -
About latent roles in forecasting players in team sports
by: Scofano, Luca, et al.
Published: (2023)