Vision-Enhanced Time Series Forecasting via Latent Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ruan, Weilin, Zhong, Siru, Wen, Haomin, Liang, Yuxuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross Space and Time: A Spatio-Temporal Unitized Model for Traffic Flow Forecasting
by: Ruan, Weilin, et al.
Published: (2024)
by: Ruan, Weilin, et al.
Published: (2024)
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
by: Zhong, Siru, et al.
Published: (2025)
by: Zhong, Siru, et al.
Published: (2025)
OccamVTS: Distilling Vision Models to 1% Parameters for Time Series Forecasting
by: Lyu, Sisuo, et al.
Published: (2025)
by: Lyu, Sisuo, et al.
Published: (2025)
UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction
by: Hao, Xixuan, et al.
Published: (2024)
by: Hao, Xixuan, et al.
Published: (2024)
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
by: Zhong, Siru, et al.
Published: (2024)
by: Zhong, Siru, et al.
Published: (2024)
SSDA: Bridging Spectral and Structural Gaps via Dual Adaptation for Vision-Based Time Series Forecasting
by: Zhang, Mingrui, et al.
Published: (2026)
by: Zhang, Mingrui, et al.
Published: (2026)
ViTime: Foundation Model for Time Series Forecasting Powered by Vision Intelligence
by: Yang, Luoxiao, et al.
Published: (2024)
by: Yang, Luoxiao, et al.
Published: (2024)
Enhancing Spatiotemporal Disease Progression Models via Latent Diffusion and Prior Knowledge
by: Puglisi, Lemuel, et al.
Published: (2024)
by: Puglisi, Lemuel, et al.
Published: (2024)
MM-ISTS: Cooperating Irregularly Sampled Time Series Forecasting with Multimodal Vision-Text LLMs
by: Lei, Zhi, et al.
Published: (2026)
by: Lei, Zhi, et al.
Published: (2026)
The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents
by: Sun, Yuwei, et al.
Published: (2026)
by: Sun, Yuwei, et al.
Published: (2026)
Text-to-CT Generation via 3D Latent Diffusion Model with Contrastive Vision-Language Pretraining
by: Molino, Daniele, et al.
Published: (2025)
by: Molino, Daniele, et al.
Published: (2025)
StructXLIP: Enhancing Vision-language Models with Multimodal Structural Cues
by: Ruan, Zanxi, et al.
Published: (2026)
by: Ruan, Zanxi, et al.
Published: (2026)
Sword: Style-Robust World Models as Simulators via Dynamic Latent Bootstrapping for VLA Policy Post-Training
by: Gao, Jiaxuan, et al.
Published: (2026)
by: Gao, Jiaxuan, et al.
Published: (2026)
LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model
by: Mei, Xiaodong, et al.
Published: (2026)
by: Mei, Xiaodong, et al.
Published: (2026)
DDT: Decoupled Diffusion Transformer
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
TAMMs: Change Understanding and Forecasting in Satellite Image Time Series with Temporal-Aware Multimodal Models
by: Guo, Zhongbin, et al.
Published: (2025)
by: Guo, Zhongbin, et al.
Published: (2025)
EfficientFSL: Enhancing Few-Shot Classification via Query-Only Tuning in Vision Transformers
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
by: Liu, Gongye, et al.
Published: (2026)
by: Liu, Gongye, et al.
Published: (2026)
TriTS: Time Series Forecasting from a Multimodal Perspective
by: Ao, Xiang
Published: (2026)
by: Ao, Xiang
Published: (2026)
EarthCrafter: Scalable 3D Earth Generation via Dual-Sparse Latent Diffusion
by: Liu, Shang, et al.
Published: (2025)
by: Liu, Shang, et al.
Published: (2025)
A Multi-Modal Knowledge-Enhanced Framework for Vessel Trajectory Prediction
by: Yu, Haomin, et al.
Published: (2025)
by: Yu, Haomin, et al.
Published: (2025)
Latent Guidance in Diffusion Models for Perceptual Evaluations
by: Saini, Shreshth, et al.
Published: (2025)
by: Saini, Shreshth, et al.
Published: (2025)
Latent Diffusion Model without Variational Autoencoder
by: Shi, Minglei, et al.
Published: (2025)
by: Shi, Minglei, et al.
Published: (2025)
Fuse Your Latents: Video Editing with Multi-source Latent Diffusion Models
by: Lu, Tianyi, et al.
Published: (2023)
by: Lu, Tianyi, et al.
Published: (2023)
Forecasting When to Forecast: Accelerating Diffusion Models with Confidence-Gated Taylor
by: Guan, Xiaoliu, et al.
Published: (2025)
by: Guan, Xiaoliu, et al.
Published: (2025)
From Images to Signals: Are Large Vision Models Useful for Time Series Analysis?
by: Zhao, Ziming, et al.
Published: (2025)
by: Zhao, Ziming, et al.
Published: (2025)
LDEdit: Towards Generalized Text Guided Image Manipulation via Latent Diffusion Models
by: Chandramouli, Paramanand, et al.
Published: (2022)
by: Chandramouli, Paramanand, et al.
Published: (2022)
VisionTS: Visual Masked Autoencoders Are Free-Lunch Zero-Shot Time Series Forecasters
by: Chen, Mouxiang, et al.
Published: (2024)
by: Chen, Mouxiang, et al.
Published: (2024)
EEO-TFV: Escape-Explore Optimizer for Web-Scale Time-Series Forecasting and Vision Analysis
by: Wang, Hua, et al.
Published: (2026)
by: Wang, Hua, et al.
Published: (2026)
TS-P$^2$CL: Plug-and-Play Dual Contrastive Learning for Vision-Guided Medical Time Series Classification
by: Xu, Qi'ao, et al.
Published: (2025)
by: Xu, Qi'ao, et al.
Published: (2025)
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
WF-VAE: Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Model
by: Li, Zongjian, et al.
Published: (2024)
by: Li, Zongjian, et al.
Published: (2024)
Color encoding in Latent Space of Stable Diffusion Models
by: Arias, Guillem, et al.
Published: (2025)
by: Arias, Guillem, et al.
Published: (2025)
Latent Diffusion Models for Attribute-Preserving Image Anonymization
by: Piano, Luca, et al.
Published: (2024)
by: Piano, Luca, et al.
Published: (2024)
Latent-based Diffusion Model for Long-tailed Recognition
by: Han, Pengxiao, et al.
Published: (2024)
by: Han, Pengxiao, et al.
Published: (2024)
Latent Feature-Guided Diffusion Models for Shadow Removal
by: Mei, Kangfu, et al.
Published: (2023)
by: Mei, Kangfu, et al.
Published: (2023)
Latent-Compressed Variational Autoencoder for Video Diffusion Models
by: Guan, Jiarui, et al.
Published: (2026)
by: Guan, Jiarui, et al.
Published: (2026)
Reprogramming Vision Foundation Models for Spatio-Temporal Forecasting
by: Chen, Changlu, et al.
Published: (2025)
by: Chen, Changlu, et al.
Published: (2025)
Image Restoration via Diffusion Models with Dynamic Resolution
by: Zheng, Yang, et al.
Published: (2026)
by: Zheng, Yang, et al.
Published: (2026)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
by: Zhong, Weihong, et al.
Published: (2024)
by: Zhong, Weihong, et al.
Published: (2024)
Similar Items
-
Cross Space and Time: A Spatio-Temporal Unitized Model for Traffic Flow Forecasting
by: Ruan, Weilin, et al.
Published: (2024) -
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
by: Zhong, Siru, et al.
Published: (2025) -
OccamVTS: Distilling Vision Models to 1% Parameters for Time Series Forecasting
by: Lyu, Sisuo, et al.
Published: (2025) -
UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction
by: Hao, Xixuan, et al.
Published: (2024) -
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
by: Zhong, Siru, et al.
Published: (2024)