Gespeichert in:
| Hauptverfasser: | Liu, Zichen, Meng, Yihao, Ouyang, Hao, Yu, Yue, Zhao, Bolin, Cohen-Or, Daniel, Qu, Huamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2404.11614 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
Bring Your Dreams to Life: Continual Text-to-Video Customization
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
Animated Stickers: Bringing Stickers to Life with Video Diffusion
von: Yan, David, et al.
Veröffentlicht: (2024)
von: Yan, David, et al.
Veröffentlicht: (2024)
Kinetic Typography Diffusion Model
von: Park, Seonmi, et al.
Veröffentlicht: (2024)
von: Park, Seonmi, et al.
Veröffentlicht: (2024)
AniDoc: Animation Creation Made Easier
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
WordCon: Word-level Typography Control in Scene Text Rendering
von: Shi, Wenda, et al.
Veröffentlicht: (2025)
von: Shi, Wenda, et al.
Veröffentlicht: (2025)
Generative Neural Video Compression via Video Diffusion Prior
von: Mao, Qi, et al.
Veröffentlicht: (2025)
von: Mao, Qi, et al.
Veröffentlicht: (2025)
Diffusion Models Need Visual Priors for Image Generation
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2024)
FonTS: Text Rendering with Typography and Style Controls
von: Shi, Wenda, et al.
Veröffentlicht: (2024)
von: Shi, Wenda, et al.
Veröffentlicht: (2024)
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
von: Feng, Kailai, et al.
Veröffentlicht: (2024)
von: Feng, Kailai, et al.
Veröffentlicht: (2024)
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
FaceShot: Bring Any Character into Life
von: Gao, Junyao, et al.
Veröffentlicht: (2025)
von: Gao, Junyao, et al.
Veröffentlicht: (2025)
GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and Generation
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
Bring the Power of Diffusion Model to Defect Detection
von: Yu, Xuyi
Veröffentlicht: (2024)
von: Yu, Xuyi
Veröffentlicht: (2024)
XPSR: Cross-modal Priors for Diffusion-based Image Super-Resolution
von: Qu, Yunpeng, et al.
Veröffentlicht: (2024)
von: Qu, Yunpeng, et al.
Veröffentlicht: (2024)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
von: Shao, Jiahao, et al.
Veröffentlicht: (2024)
von: Shao, Jiahao, et al.
Veröffentlicht: (2024)
Reading $\neq$ Seeing: Diagnosing and Closing the Typography Gap in Vision-Language Models
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
Prior-Enhanced Gaussian Splatting for Dynamic Scene Reconstruction from Casual Video
von: Shih, Meng-Li, et al.
Veröffentlicht: (2025)
von: Shih, Meng-Li, et al.
Veröffentlicht: (2025)
The Dynamic Prior: Understanding 3D Structures for Casual Dynamic Videos
von: Wu, Zhuoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhuoyuan, et al.
Veröffentlicht: (2025)
Calligrapher: Freestyle Text Image Customization
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
Boosting Visual Recognition in Real-world Degradations via Unsupervised Feature Enhancement Module with Deep Channel Prior
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
CameraCtrl II: Dynamic Scene Exploration via Camera-controlled Video Diffusion Models
von: He, Hao, et al.
Veröffentlicht: (2025)
von: He, Hao, et al.
Veröffentlicht: (2025)
UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors
von: Chen, Houyuan, et al.
Veröffentlicht: (2026)
von: Chen, Houyuan, et al.
Veröffentlicht: (2026)
Video Deblurring by Sharpness Prior Detection and Edge Information
von: Tian, Yang, et al.
Veröffentlicht: (2025)
von: Tian, Yang, et al.
Veröffentlicht: (2025)
Towards Transferable Attacks Against Vision-LLMs in Autonomous Driving with Typography
von: Chung, Nhat, et al.
Veröffentlicht: (2024)
von: Chung, Nhat, et al.
Veröffentlicht: (2024)
WordCraft: Interactive Artistic Typography with Attention Awareness and Noise Blending
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
Typography-Based Monocular Distance Estimation Framework for Vehicle Safety Systems
von: Reddy, Manognya Lokesh, et al.
Veröffentlicht: (2026)
von: Reddy, Manognya Lokesh, et al.
Veröffentlicht: (2026)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
Unfolding Videos Dynamics via Taylor Expansion
von: Chen, Siyi, et al.
Veröffentlicht: (2024)
von: Chen, Siyi, et al.
Veröffentlicht: (2024)
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
4Dynamic: Text-to-4D Generation with Hybrid Priors
von: Yuan, Yu-Jie, et al.
Veröffentlicht: (2024)
von: Yuan, Yu-Jie, et al.
Veröffentlicht: (2024)
End-to-End Training for Autoregressive Video Diffusion via Self-Resampling
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
Single-Shot HDR Recovery via a Video Diffusion Prior
von: Talegaonkar, Chinmay, et al.
Veröffentlicht: (2026)
von: Talegaonkar, Chinmay, et al.
Veröffentlicht: (2026)
Dysen-VDM: Empowering Dynamics-aware Text-to-Video Diffusion with LLMs
von: Fei, Hao, et al.
Veröffentlicht: (2023)
von: Fei, Hao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2026) -
Bring Your Dreams to Life: Continual Text-to-Video Customization
von: Dong, Jiahua, et al.
Veröffentlicht: (2025) -
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2025) -
Animated Stickers: Bringing Stickers to Life with Video Diffusion
von: Yan, David, et al.
Veröffentlicht: (2024) -
Kinetic Typography Diffusion Model
von: Park, Seonmi, et al.
Veröffentlicht: (2024)