Teaching Time Series to See and Speak: Forecasting with Aligned Visual and Textual Perspectives
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Sixun, Fan, Wei, Wu, Teresa, Fu, Yanjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens
von: Qin, Yiming, et al.
Veröffentlicht: (2025)
von: Qin, Yiming, et al.
Veröffentlicht: (2025)
Textualize Visual Prompt for Image Editing via Diffusion Bridge
von: Xu, Pengcheng, et al.
Veröffentlicht: (2025)
von: Xu, Pengcheng, et al.
Veröffentlicht: (2025)
TimePre: Bridging Accuracy, Efficiency, and Stability in Probabilistic Time-Series Forecasting
von: Jiang, Lingyu, et al.
Veröffentlicht: (2025)
von: Jiang, Lingyu, et al.
Veröffentlicht: (2025)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
von: Zhong, Siru, et al.
Veröffentlicht: (2025)
von: Zhong, Siru, et al.
Veröffentlicht: (2025)
ViTime: Foundation Model for Time Series Forecasting Powered by Vision Intelligence
von: Yang, Luoxiao, et al.
Veröffentlicht: (2024)
von: Yang, Luoxiao, et al.
Veröffentlicht: (2024)
Learning Temporal Saliency for Time Series Forecasting with Cross-Scale Attention
von: Delibasoglu, Ibrahim, et al.
Veröffentlicht: (2025)
von: Delibasoglu, Ibrahim, et al.
Veröffentlicht: (2025)
Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
TagFog: Textual Anchor Guidance and Fake Outlier Generation for Visual Out-of-Distribution Detection
von: Chen, Jiankang, et al.
Veröffentlicht: (2024)
von: Chen, Jiankang, et al.
Veröffentlicht: (2024)
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
OccamVTS: Distilling Vision Models to 1% Parameters for Time Series Forecasting
von: Lyu, Sisuo, et al.
Veröffentlicht: (2025)
von: Lyu, Sisuo, et al.
Veröffentlicht: (2025)
Probabilistic NDVI Forecasting from Sparse Satellite Time Series and Weather Covariates
von: Iele, Irene, et al.
Veröffentlicht: (2026)
von: Iele, Irene, et al.
Veröffentlicht: (2026)
VisionTS: Visual Masked Autoencoders Are Free-Lunch Zero-Shot Time Series Forecasters
von: Chen, Mouxiang, et al.
Veröffentlicht: (2024)
von: Chen, Mouxiang, et al.
Veröffentlicht: (2024)
EEO-TFV: Escape-Explore Optimizer for Web-Scale Time-Series Forecasting and Vision Analysis
von: Wang, Hua, et al.
Veröffentlicht: (2026)
von: Wang, Hua, et al.
Veröffentlicht: (2026)
ZOTTA: Test-Time Adaptation with Gradient-Free Zeroth-Order Optimization
von: Zhang, Ronghao, et al.
Veröffentlicht: (2026)
von: Zhang, Ronghao, et al.
Veröffentlicht: (2026)
Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models
von: Gan, Woody Haosheng, et al.
Veröffentlicht: (2025)
von: Gan, Woody Haosheng, et al.
Veröffentlicht: (2025)
When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models
von: Sun, Zhengyang, et al.
Veröffentlicht: (2026)
von: Sun, Zhengyang, et al.
Veröffentlicht: (2026)
VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
Mitigating Data Redundancy to Revitalize Transformer-based Long-Term Time Series Forecasting System
von: Li, Mingjie, et al.
Veröffentlicht: (2022)
von: Li, Mingjie, et al.
Veröffentlicht: (2022)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
MTS-DMAE: Dual-Masked Autoencoder for Unsupervised Multivariate Time Series Representation Learning
von: Xu, Yi, et al.
Veröffentlicht: (2025)
von: Xu, Yi, et al.
Veröffentlicht: (2025)
Prompting Forgetting: Unlearning in GANs via Textual Guidance
von: Nagasubramaniam, Piyush, et al.
Veröffentlicht: (2025)
von: Nagasubramaniam, Piyush, et al.
Veröffentlicht: (2025)
Forecasting as Rendering: A 2D Gaussian Splatting Framework for Time Series Forecasting
von: Wang, Yixin, et al.
Veröffentlicht: (2026)
von: Wang, Yixin, et al.
Veröffentlicht: (2026)
Learning to See Before Seeing: Demystifying LLM Visual Priors from Language Pre-training
von: Han, Junlin, et al.
Veröffentlicht: (2025)
von: Han, Junlin, et al.
Veröffentlicht: (2025)
Differential-Integral Neural Operator for Long-Term Turbulence Forecasting
von: Wu, Hao, et al.
Veröffentlicht: (2025)
von: Wu, Hao, et al.
Veröffentlicht: (2025)
Seeing and Knowing in the Wild: Open-domain Visual Entity Recognition with Large-scale Knowledge Graphs via Contrastive Learning
von: Zhou, Hongkuan, et al.
Veröffentlicht: (2025)
von: Zhou, Hongkuan, et al.
Veröffentlicht: (2025)
Aligning Visual Contrastive learning models via Preference Optimization
von: Afzali, Amirabbas, et al.
Veröffentlicht: (2024)
von: Afzali, Amirabbas, et al.
Veröffentlicht: (2024)
CLIPErase: Efficient Unlearning of Visual-Textual Associations in CLIP
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
MMTok: Multimodal Coverage Maximization for Efficient Inference of VLMs
von: Dong, Sixun, et al.
Veröffentlicht: (2025)
von: Dong, Sixun, et al.
Veröffentlicht: (2025)
Human-Aligned Image Models Improve Visual Decoding from the Brain
von: Rajabi, Nona, et al.
Veröffentlicht: (2025)
von: Rajabi, Nona, et al.
Veröffentlicht: (2025)
Symmetrical Visual Contrastive Optimization: Aligning Vision-Language Models with Minimal Contrastive Images
von: Wu, Shengguang, et al.
Veröffentlicht: (2025)
von: Wu, Shengguang, et al.
Veröffentlicht: (2025)
Improving Position Encoding of Transformers for Multivariate Time Series Classification
von: Foumani, Navid Mohammadi, et al.
Veröffentlicht: (2023)
von: Foumani, Navid Mohammadi, et al.
Veröffentlicht: (2023)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2024)
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2024)
Aligning Human Knowledge with Visual Concepts Towards Explainable Medical Image Classification
von: Gao, Yunhe, et al.
Veröffentlicht: (2024)
von: Gao, Yunhe, et al.
Veröffentlicht: (2024)
SimpleOCR: Rendering Visualized Questions to Teach MLLMs to Read
von: Peng, Yibo, et al.
Veröffentlicht: (2026)
von: Peng, Yibo, et al.
Veröffentlicht: (2026)
EmbodiTTA: Resource-Efficient Test-Time Adaptation for Embodied Visual Systems
von: Ma, Xiao, et al.
Veröffentlicht: (2025)
von: Ma, Xiao, et al.
Veröffentlicht: (2025)
Don't Get Me Wrong: How to Apply Deep Visual Interpretations to Time Series
von: Loeffler, Christoffer, et al.
Veröffentlicht: (2022)
von: Loeffler, Christoffer, et al.
Veröffentlicht: (2022)
Align Your Flow: Scaling Continuous-Time Flow Map Distillation
von: Sabour, Amirmojtaba, et al.
Veröffentlicht: (2025)
von: Sabour, Amirmojtaba, et al.
Veröffentlicht: (2025)
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
von: Hu, Zhangyi, et al.
Veröffentlicht: (2025)
von: Hu, Zhangyi, et al.
Veröffentlicht: (2025)
Highlight Every Step: Knowledge Distillation via Collaborative Teaching
von: Zhao, Haoran, et al.
Veröffentlicht: (2019)
von: Zhao, Haoran, et al.
Veröffentlicht: (2019)
Ähnliche Einträge
-
Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens
von: Qin, Yiming, et al.
Veröffentlicht: (2025) -
Textualize Visual Prompt for Image Editing via Diffusion Bridge
von: Xu, Pengcheng, et al.
Veröffentlicht: (2025) -
TimePre: Bridging Accuracy, Efficiency, and Stability in Probabilistic Time-Series Forecasting
von: Jiang, Lingyu, et al.
Veröffentlicht: (2025) -
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
von: Góral, Gracjan, et al.
Veröffentlicht: (2024) -
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
von: Zhong, Siru, et al.
Veröffentlicht: (2025)