DreamVideo: High-Fidelity Image-to-Video Generation with Image Retention and Text Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Cong, Gu, Jiaxi, Hu, Panwen, Xu, Songcen, Xu, Hang, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
von: Wang, Cong, et al.
Veröffentlicht: (2024)
von: Wang, Cong, et al.
Veröffentlicht: (2024)
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023)
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023)
BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation
von: Hu, Panwen, et al.
Veröffentlicht: (2025)
von: Hu, Panwen, et al.
Veröffentlicht: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
AtomoVideo: High Fidelity Image-to-Video Generation
von: Gong, Litong, et al.
Veröffentlicht: (2024)
von: Gong, Litong, et al.
Veröffentlicht: (2024)
MagDiff: Multi-Alignment Diffusion for High-Fidelity Video Generation and Editing
von: Zhao, Haoyu, et al.
Veröffentlicht: (2023)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2023)
Fuse Your Latents: Video Editing with Multi-source Latent Diffusion Models
von: Lu, Tianyi, et al.
Veröffentlicht: (2023)
von: Lu, Tianyi, et al.
Veröffentlicht: (2023)
LayerDiff: Exploring Text-guided Multi-layered Composable Image Synthesis via Layer-Collaborative Diffusion Model
von: Huang, Runhui, et al.
Veröffentlicht: (2024)
von: Huang, Runhui, et al.
Veröffentlicht: (2024)
AutoTVG: A New Vision-language Pre-training Paradigm for Temporal Video Grounding
von: Zhang, Xing, et al.
Veröffentlicht: (2024)
von: Zhang, Xing, et al.
Veröffentlicht: (2024)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
DirectSwap: Mask-Free Cross-Identity Training and Benchmarking for Expression-Consistent Video Head Swapping
von: Wang, Yanan, et al.
Veröffentlicht: (2025)
von: Wang, Yanan, et al.
Veröffentlicht: (2025)
DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer
von: Guo, Xu, et al.
Veröffentlicht: (2026)
von: Guo, Xu, et al.
Veröffentlicht: (2026)
DreamText: High Fidelity Scene Text Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance
von: Gao, Danyi
Veröffentlicht: (2025)
von: Gao, Danyi
Veröffentlicht: (2025)
Counting Guidance for High Fidelity Text-to-Image Synthesis
von: Kang, Wonjun, et al.
Veröffentlicht: (2023)
von: Kang, Wonjun, et al.
Veröffentlicht: (2023)
DualDiff+: Dual-Branch Diffusion for High-Fidelity Video Generation with Reward Guidance
von: Yang, Zhao, et al.
Veröffentlicht: (2025)
von: Yang, Zhao, et al.
Veröffentlicht: (2025)
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
von: Li, Weijie, et al.
Veröffentlicht: (2024)
von: Li, Weijie, et al.
Veröffentlicht: (2024)
StoryAgent: Customized Storytelling Video Generation via Multi-Agent Collaboration
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation
von: Song, Wenhui, et al.
Veröffentlicht: (2025)
von: Song, Wenhui, et al.
Veröffentlicht: (2025)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
Lynx: Towards High-Fidelity Personalized Video Generation
von: Sang, Shen, et al.
Veröffentlicht: (2025)
von: Sang, Shen, et al.
Veröffentlicht: (2025)
Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis
von: Su, Tongtong, et al.
Veröffentlicht: (2025)
von: Su, Tongtong, et al.
Veröffentlicht: (2025)
AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
DreamVE: Unified Instruction-based Image and Video Editing
von: Xia, Bin, et al.
Veröffentlicht: (2025)
von: Xia, Bin, et al.
Veröffentlicht: (2025)
Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes
von: Chen, Jianqi, et al.
Veröffentlicht: (2024)
von: Chen, Jianqi, et al.
Veröffentlicht: (2024)
Bring Your Dreams to Life: Continual Text-to-Video Customization
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
Repeating Words for Video-Language Retrieval with Coarse-to-Fine Objectives
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidance
von: Zheng, Jun, et al.
Veröffentlicht: (2026)
von: Zheng, Jun, et al.
Veröffentlicht: (2026)
VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
von: Liao, Xinyao, et al.
Veröffentlicht: (2026)
Conditional Text-to-Image Generation with Reference Guidance
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
DreamMapping: High-Fidelity Text-to-3D Generation via Variational Distribution Mapping
von: Cai, Zeyu, et al.
Veröffentlicht: (2024)
von: Cai, Zeyu, et al.
Veröffentlicht: (2024)
DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers
von: Wang, Lizhen, et al.
Veröffentlicht: (2025)
von: Wang, Lizhen, et al.
Veröffentlicht: (2025)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
Geometric-Aware Low-Light Image and Video Enhancement via Depth Guidance
von: Lin, Yingqi, et al.
Veröffentlicht: (2023)
von: Lin, Yingqi, et al.
Veröffentlicht: (2023)
DreamVAR: Taming Reinforced Visual Autoregressive Model for High-Fidelity Subject-Driven Image Generation
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
von: Shen, Guibao, et al.
Veröffentlicht: (2024)
von: Shen, Guibao, et al.
Veröffentlicht: (2024)
Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance
von: Bompai, Stelio, et al.
Veröffentlicht: (2026)
von: Bompai, Stelio, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
von: Wang, Cong, et al.
Veröffentlicht: (2024) -
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023) -
BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation
von: Hu, Panwen, et al.
Veröffentlicht: (2025) -
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
von: Wei, Yujie, et al.
Veröffentlicht: (2024) -
DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning
von: Wei, Yujie, et al.
Veröffentlicht: (2026)