TRIP: Temporal Residual Learning with Image Noise Prior for Image-to-Video Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zhongwei, Long, Fuchen, Pan, Yingwei, Qiu, Zhaofan, Yao, Ting, Cao, Yang, Mei, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MotionPro: A Precise Motion Controller for Image-to-Video Generation
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
Learning Spatial Adaptation and Temporal Coherence in Diffusion Models for Video Super-Resolution
von: Chen, Zhikai, et al.
Veröffentlicht: (2024)
von: Chen, Zhikai, et al.
Veröffentlicht: (2024)
Region-Constraint In-Context Generation for Instructional Video Editing
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
Creatively Upscaling Images with Global-Regional Priors
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
von: Luo, Yang, et al.
Veröffentlicht: (2024)
von: Luo, Yang, et al.
Veröffentlicht: (2024)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2025)
von: Cai, Qi, et al.
Veröffentlicht: (2025)
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
Improving Virtual Try-On with Garment-focused Diffusion Models
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024)
von: Yao, Ting, et al.
Veröffentlicht: (2024)
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
von: Li, Dong, et al.
Veröffentlicht: (2025)
von: Li, Dong, et al.
Veröffentlicht: (2025)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Visual Autoregressive Modeling for Instruction-Guided Image Editing
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
VideoStudio: Generating Consistent-Content and Multi-Scene Videos
von: Long, Fuchen, et al.
Veröffentlicht: (2024)
von: Long, Fuchen, et al.
Veröffentlicht: (2024)
Boosting Diffusion Models with Moving Average Sampling in Frequency Domain
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2026)
von: Cai, Qi, et al.
Veröffentlicht: (2026)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
Improving Text-guided Object Inpainting with Semantic Pre-inpainting
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion
von: Chen, Jingyuan, et al.
Veröffentlicht: (2025)
von: Chen, Jingyuan, et al.
Veröffentlicht: (2025)
Residual Prior-driven Frequency-aware Network for Image Fusion
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
VideoZeroBench: Probing the Limits of Video MLLMs with Spatio-Temporal Evidence Verification
von: Meng, Jiahao, et al.
Veröffentlicht: (2026)
von: Meng, Jiahao, et al.
Veröffentlicht: (2026)
Exposure Completing for Temporally Consistent Neural High Dynamic Range Video Rendering
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
DeepSPG: Exploring Deep Semantic Prior Guidance for Low-light Image Enhancement with Multimodal Learning
von: Lu, Jialang, et al.
Veröffentlicht: (2025)
von: Lu, Jialang, et al.
Veröffentlicht: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
Disparity-based Stereo Image Compression with Aligned Cross-View Priors
von: Zhai, Yongqi, et al.
Veröffentlicht: (2022)
von: Zhai, Yongqi, et al.
Veröffentlicht: (2022)
ChatVTG: Video Temporal Grounding via Chat with Video Dialogue Large Language Models
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
StableDub: Taming Diffusion Prior for Generalized and Efficient Visual Dubbing
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
Towards Efficient Low-rate Image Compression with Frequency-aware Diffusion Prior Refinement
von: Xia, Yichong, et al.
Veröffentlicht: (2026)
von: Xia, Yichong, et al.
Veröffentlicht: (2026)
3D-LMVIC: Learning-based Multi-View Image Coding with 3D Gaussian Geometric Priors
von: Huang, Yujun, et al.
Veröffentlicht: (2024)
von: Huang, Yujun, et al.
Veröffentlicht: (2024)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
Exploring Phrase-Level Grounding with Text-to-Image Diffusion Model
von: Yang, Danni, et al.
Veröffentlicht: (2024)
von: Yang, Danni, et al.
Veröffentlicht: (2024)
Vision-Language Models Learn Super Images for Efficient Partially Relevant Video Retrieval
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits
von: Pan, Minzhou, et al.
Veröffentlicht: (2024)
von: Pan, Minzhou, et al.
Veröffentlicht: (2024)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
von: Zhang, Long, et al.
Veröffentlicht: (2025)
von: Zhang, Long, et al.
Veröffentlicht: (2025)
Synthetic Perception: Can Generated Images Unlock Latent Visual Prior for Text-Centric Reasoning?
von: Huang, Yuesheng, et al.
Veröffentlicht: (2025)
von: Huang, Yuesheng, et al.
Veröffentlicht: (2025)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
von: Cao, Pu, et al.
Veröffentlicht: (2023)
von: Cao, Pu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
MotionPro: A Precise Motion Controller for Image-to-Video Generation
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025) -
Learning Spatial Adaptation and Temporal Coherence in Diffusion Models for Video Super-Resolution
von: Chen, Zhikai, et al.
Veröffentlicht: (2024) -
Region-Constraint In-Context Generation for Instructional Video Editing
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025) -
Creatively Upscaling Images with Global-Regional Priors
von: Qian, Yurui, et al.
Veröffentlicht: (2025) -
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
von: Luo, Yang, et al.
Veröffentlicht: (2024)