CV-VAE: A Compatible Video VAE for Latent Generative Video Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Sijie, Zhang, Yong, Cun, Xiaodong, Yang, Shaoshu, Niu, Muyao, Li, Xiaoyu, Hu, Wenbo, Shan, Ying |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improved Video VAE for Latent Video Diffusion Model
by: Wu, Pingyu, et al.
Published: (2024)
by: Wu, Pingyu, et al.
Published: (2024)
LeanVAE: An Ultra-Efficient Reconstruction VAE for Video Diffusion Models
by: Cheng, Yu, et al.
Published: (2025)
by: Cheng, Yu, et al.
Published: (2025)
OD-VAE: An Omni-dimensional Video Compressor for Improving Latent Video Diffusion Model
by: Chen, Liuhan, et al.
Published: (2024)
by: Chen, Liuhan, et al.
Published: (2024)
LDPM: Towards undersampled MRI reconstruction with MR-VAE and Latent Diffusion Prior
by: Tang, Xingjian, et al.
Published: (2024)
by: Tang, Xingjian, et al.
Published: (2024)
Semi-Supervised Contrastive VAE for Disentanglement of Digital Pathology Images
by: Hasan, Mahmudul, et al.
Published: (2024)
by: Hasan, Mahmudul, et al.
Published: (2024)
BrightVAE: Luminosity Enhancement in Underexposed Endoscopic Images
by: Koohestani, Farzaneh, et al.
Published: (2024)
by: Koohestani, Farzaneh, et al.
Published: (2024)
Epsilon-VAE: Denoising as Visual Decoding
by: Zhao, Long, et al.
Published: (2024)
by: Zhao, Long, et al.
Published: (2024)
CE-VAE: Capsule Enhanced Variational AutoEncoder for Underwater Image Enhancement
by: Pucci, Rita, et al.
Published: (2024)
by: Pucci, Rita, et al.
Published: (2024)
Generative Latent Video Compression
by: Guo, Zongyu, et al.
Published: (2025)
by: Guo, Zongyu, et al.
Published: (2025)
D-CNN and VQ-VAE Autoencoders for Compression and Denoising of Industrial X-ray Computed Tomography Images
by: Hejazi, Bardia, et al.
Published: (2025)
by: Hejazi, Bardia, et al.
Published: (2025)
Prompt-Guided Patch UNet-VAE with Adversarial Supervision for Adrenal Gland Segmentation in Computed Tomography Medical Images
by: Ghouse, Hania, et al.
Published: (2025)
by: Ghouse, Hania, et al.
Published: (2025)
Variational Autoencoder Framework for Hyperspectral Retrievals (Hyper-VAE) of Phytoplankton Absorption and Chlorophyll a in Coastal Waters for NASA's EMIT and PACE Missions
by: Lou, Jiadong, et al.
Published: (2025)
by: Lou, Jiadong, et al.
Published: (2025)
Predicting Pulmonary Hypertension in Newborns: A Multi-view VAE Approach
by: Erlacher, Lucas, et al.
Published: (2025)
by: Erlacher, Lucas, et al.
Published: (2025)
GCVAMD: A Modified CausalVAE Model for Causal Age-related Macular Degeneration Risk Factor Detection and Prediction
by: Kim, Daeyoung
Published: (2025)
by: Kim, Daeyoung
Published: (2025)
VesselVAE: Recursive Variational Autoencoders for 3D Blood Vessel Synthesis
by: Feldman, Paula, et al.
Published: (2023)
by: Feldman, Paula, et al.
Published: (2023)
Generative Latent Coding for Ultra-Low Bitrate Image and Video Compression
by: Qi, Linfeng, et al.
Published: (2025)
by: Qi, Linfeng, et al.
Published: (2025)
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders
by: Varma, Maya, et al.
Published: (2025)
by: Varma, Maya, et al.
Published: (2025)
Ultrasound Image-to-Video Synthesis via Latent Dynamic Diffusion Models
by: Chen, Tingxiu, et al.
Published: (2025)
by: Chen, Tingxiu, et al.
Published: (2025)
LDMVFI: Video Frame Interpolation with Latent Diffusion Models
by: Danier, Duolikun, et al.
Published: (2023)
by: Danier, Duolikun, et al.
Published: (2023)
Semantic and Temporal Integration in Latent Diffusion Space for High-Fidelity Video Super-Resolution
by: Wang, Yiwen, et al.
Published: (2025)
by: Wang, Yiwen, et al.
Published: (2025)
Latent Motion Profiling for Annotation-free Cardiac Phase Detection in Adult and Fetal Echocardiography Videos
by: Yang, Yingyu, et al.
Published: (2025)
by: Yang, Yingyu, et al.
Published: (2025)
BVI-UGC: A Video Quality Database for User-Generated Content Transcoding
by: Qi, Zihao, et al.
Published: (2024)
by: Qi, Zihao, et al.
Published: (2024)
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023)
by: Liang, Jingyun, et al.
Published: (2023)
Extreme Video Compression with Pre-trained Diffusion Models
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Binarized Low-light Raw Video Enhancement
by: Zhang, Gengchen, et al.
Published: (2024)
by: Zhang, Gengchen, et al.
Published: (2024)
MRI Parameter Mapping via Gaussian Mixture VAE: Breaking the Assumption of Independent Pixels
by: Xu, Moucheng, et al.
Published: (2024)
by: Xu, Moucheng, et al.
Published: (2024)
Taming Diffusion Transformer for Efficient Mobile Video Generation in Seconds
by: Wu, Yushu, et al.
Published: (2025)
by: Wu, Yushu, et al.
Published: (2025)
VEnhancer: Generative Space-Time Enhancement for Video Generation
by: He, Jingwen, et al.
Published: (2024)
by: He, Jingwen, et al.
Published: (2024)
U-Motion: Learned Point Cloud Video Compression with U-Structured Temporal Context Generation
by: Fan, Tingyu, et al.
Published: (2024)
by: Fan, Tingyu, et al.
Published: (2024)
Bora: Biomedical Generalist Video Generation Model
by: Sun, Weixiang, et al.
Published: (2024)
by: Sun, Weixiang, et al.
Published: (2024)
GIViC: Generative Implicit Video Compression
by: Gao, Ge, et al.
Published: (2025)
by: Gao, Ge, et al.
Published: (2025)
Research on Audio-Visual Quality Assessment Dataset and Method for User-Generated Omnidirectional Video
by: Zhao, Fei, et al.
Published: (2025)
by: Zhao, Fei, et al.
Published: (2025)
Joint Reference Frame Synthesis and Post Filter Enhancement for Versatile Video Coding
by: Bao, Weijie, et al.
Published: (2024)
by: Bao, Weijie, et al.
Published: (2024)
Online Video Quality Enhancement with Spatial-Temporal Look-up Tables
by: Qu, Zefan, et al.
Published: (2023)
by: Qu, Zefan, et al.
Published: (2023)
SF-V: Single Forward Video Generation Model
by: Zhang, Zhixing, et al.
Published: (2024)
by: Zhang, Zhixing, et al.
Published: (2024)
EndoGen: Conditional Autoregressive Endoscopic Video Generation
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
UnDIVE: Generalized Underwater Video Enhancement Using Generative Priors
by: Srinath, Suhas, et al.
Published: (2024)
by: Srinath, Suhas, et al.
Published: (2024)
Embedding Compression Distortion in Video Coding for Machines
by: Sun, Yuxiao, et al.
Published: (2025)
by: Sun, Yuxiao, et al.
Published: (2025)
KVQ: Kwai Video Quality Assessment for Short-form Videos
by: Lu, Yiting, et al.
Published: (2024)
by: Lu, Yiting, et al.
Published: (2024)
3D MedDiffusion: A 3D Medical Latent Diffusion Model for Controllable and High-quality Medical Image Generation
by: Wang, Haoshen, et al.
Published: (2024)
by: Wang, Haoshen, et al.
Published: (2024)
Similar Items
-
Improved Video VAE for Latent Video Diffusion Model
by: Wu, Pingyu, et al.
Published: (2024) -
LeanVAE: An Ultra-Efficient Reconstruction VAE for Video Diffusion Models
by: Cheng, Yu, et al.
Published: (2025) -
OD-VAE: An Omni-dimensional Video Compressor for Improving Latent Video Diffusion Model
by: Chen, Liuhan, et al.
Published: (2024) -
LDPM: Towards undersampled MRI reconstruction with MR-VAE and Latent Diffusion Prior
by: Tang, Xingjian, et al.
Published: (2024) -
Semi-Supervised Contrastive VAE for Disentanglement of Digital Pathology Images
by: Hasan, Mahmudul, et al.
Published: (2024)