Latent Knowledge-Guided Video Diffusion for Scientific Phenomena Generation from a Single Initial Frame
Fuente:
arXiv
Saved in:
| Main Authors: | Cao, Qinglong, Li, Xirui, Wang, Ding, Ma, Chao, Chen, Yuntian, Yang, Xiaokang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vision-Informed Flow Image Super-Resolution with Quaternion Spatial Modeling and Dynamic Flow Convolution
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Learning Domain Knowledge in Multimodal Large Language Models through Reinforcement Fine-Tuning
by: Cao, Qinglong, et al.
Published: (2026)
by: Cao, Qinglong, et al.
Published: (2026)
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Open-Vocabulary Remote Sensing Image Semantic Segmentation
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Dynamic Training-Free Fusion of Subject and Style LoRAs
by: Cao, Qinglong, et al.
Published: (2026)
by: Cao, Qinglong, et al.
Published: (2026)
Latent Image and Video Resolution Prediction using Convolutional Neural Networks
by: Kansabanik, Rittwika, et al.
Published: (2024)
by: Kansabanik, Rittwika, et al.
Published: (2024)
UAVDB: Point-Guided Masks for UAV Detection and Segmentation
by: Chen, Yu-Hsi
Published: (2024)
by: Chen, Yu-Hsi
Published: (2024)
Scalable Vision-Guided Crop Yield Estimation
by: Li, Harrison H., et al.
Published: (2025)
by: Li, Harrison H., et al.
Published: (2025)
StomaD2: An All-in-One System for Intelligent Stomatal Phenotype Analysis via Diffusion-Based Restoration Detection Network
by: Zhao, Quanling, et al.
Published: (2026)
by: Zhao, Quanling, et al.
Published: (2026)
Auto-Regressive Moving Diffusion Models for Time Series Forecasting
by: Gao, Jiaxin, et al.
Published: (2024)
by: Gao, Jiaxin, et al.
Published: (2024)
Exploring the Magnitude-Shape Plot Framework for Anomaly Detection in Crowded Video Scenes
by: Wang, Zuzheng, et al.
Published: (2024)
by: Wang, Zuzheng, et al.
Published: (2024)
Dynamic Atomic Column Detection in Transmission Electron Microscopy Videos via Ridge Estimation
by: Xu, Yuchen, et al.
Published: (2023)
by: Xu, Yuchen, et al.
Published: (2023)
Preliminary Study on Space Utilization and Emergent Behaviors of Group vs. Single Pedestrians in Real-World Trajectories
by: Sanjjamts, Amartaivan, et al.
Published: (2025)
by: Sanjjamts, Amartaivan, et al.
Published: (2025)
Convolutional Unscented Kalman Filter for Multi-Object Tracking with Outliers
by: Liu, Shiqi, et al.
Published: (2024)
by: Liu, Shiqi, et al.
Published: (2024)
Motion-aware Latent Diffusion Models for Video Frame Interpolation
by: Huang, Zhilin, et al.
Published: (2024)
by: Huang, Zhilin, et al.
Published: (2024)
VACT: A Video Automatic Causal Testing System and a Benchmark
by: Yang, Haotong, et al.
Published: (2025)
by: Yang, Haotong, et al.
Published: (2025)
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate
by: Yuan, Zhihang, et al.
Published: (2025)
by: Yuan, Zhihang, et al.
Published: (2025)
QuantCache: Adaptive Importance-Guided Quantization with Hierarchical Latent and Layer Caching for Video Generation
by: Wu, Junyi, et al.
Published: (2025)
by: Wu, Junyi, et al.
Published: (2025)
Generation of synthetic gait data: application to multiple sclerosis patients' gait patterns
by: Gall, Klervi Le, et al.
Published: (2024)
by: Gall, Klervi Le, et al.
Published: (2024)
Significance and Stability Analysis of Gene-Environment Interaction using RGxEStat
by: Qin, Meng'en, et al.
Published: (2026)
by: Qin, Meng'en, et al.
Published: (2026)
TLB-VFI: Temporal-Aware Latent Brownian Bridge Diffusion for Video Frame Interpolation
by: Lyu, Zonglin, et al.
Published: (2025)
by: Lyu, Zonglin, et al.
Published: (2025)
Latte: Latent Diffusion Transformer for Video Generation
by: Ma, Xin, et al.
Published: (2024)
by: Ma, Xin, et al.
Published: (2024)
Multimodal Prototyping for cancer survival prediction
by: Song, Andrew H., et al.
Published: (2024)
by: Song, Andrew H., et al.
Published: (2024)
Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks
by: Tinoco, Daniel, et al.
Published: (2026)
by: Tinoco, Daniel, et al.
Published: (2026)
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
by: Lan, Tian, et al.
Published: (2025)
by: Lan, Tian, et al.
Published: (2025)
AI-driven 3D Spatial Transcriptomics
by: Almagro-Pérez, Cristina, et al.
Published: (2025)
by: Almagro-Pérez, Cristina, et al.
Published: (2025)
Towards Explainable Industrial Anomaly Detection via Knowledge-Guided Latent Reasoning
by: Chen, Peng, et al.
Published: (2026)
by: Chen, Peng, et al.
Published: (2026)
Motion-Guided Latent Diffusion for Temporally Consistent Real-world Video Super-resolution
by: Yang, Xi, et al.
Published: (2023)
by: Yang, Xi, et al.
Published: (2023)
When No-Reference Image Quality Models Meet MAP Estimation in Diffusion Latents
by: Zhang, Weixia, et al.
Published: (2024)
by: Zhang, Weixia, et al.
Published: (2024)
Morphological Prototyping for Unsupervised Slide Representation Learning in Computational Pathology
by: Song, Andrew H., et al.
Published: (2024)
by: Song, Andrew H., et al.
Published: (2024)
AI Pose Analysis and Kinematic Profiling of Range-of-Motion Variations in Resistance Training
by: Diamant, Adam
Published: (2025)
by: Diamant, Adam
Published: (2025)
An Explainable Anomaly Detection Framework for Monitoring Depression and Anxiety Using Consumer Wearable Devices
by: Zhang, Yuezhou, et al.
Published: (2025)
by: Zhang, Yuezhou, et al.
Published: (2025)
A statistical method for crack pre-detection in 3D concrete images
by: Makogin, Vitalii, et al.
Published: (2024)
by: Makogin, Vitalii, et al.
Published: (2024)
Unsupervised cell segmentation by fast Gaussian Processes
by: Baracaldo, Laura, et al.
Published: (2025)
by: Baracaldo, Laura, et al.
Published: (2025)
Deep Clustering of Remote Sensing Scenes through Heterogeneous Transfer Learning
by: Ray, Isaac, et al.
Published: (2024)
by: Ray, Isaac, et al.
Published: (2024)
A Stochastic-Geometrical Framework for Object Pose Estimation based on Mixture Models Avoiding the Correspondence Problem
by: Hoegele, Wolfgang
Published: (2023)
by: Hoegele, Wolfgang
Published: (2023)
Estimating the Impact of COVID-19 on Travel Demand in Houston Area Using Deep Learning and Satellite Imagery
by: Pachika, Alekhya, et al.
Published: (2026)
by: Pachika, Alekhya, et al.
Published: (2026)
To use or not to use proprietary street view images in (health and place) research? That is the question
by: Helbich, Marco, et al.
Published: (2024)
by: Helbich, Marco, et al.
Published: (2024)
SIGN: A Statistically-Informed Gaze Network for Gaze Time Prediction
by: Ye, Jianping, et al.
Published: (2025)
by: Ye, Jianping, et al.
Published: (2025)
Efficient Surgical Tool Recognition via HMM-Stabilized Deep Learning
by: Wang, Haifeng, et al.
Published: (2024)
by: Wang, Haifeng, et al.
Published: (2024)
Similar Items
-
Vision-Informed Flow Image Super-Resolution with Quaternion Spatial Modeling and Dynamic Flow Convolution
by: Cao, Qinglong, et al.
Published: (2024) -
Learning Domain Knowledge in Multimodal Large Language Models through Reinforcement Fine-Tuning
by: Cao, Qinglong, et al.
Published: (2026) -
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
by: Cao, Qinglong, et al.
Published: (2024) -
Open-Vocabulary Remote Sensing Image Semantic Segmentation
by: Cao, Qinglong, et al.
Published: (2024) -
Dynamic Training-Free Fusion of Subject and Style LoRAs
by: Cao, Qinglong, et al.
Published: (2026)