Vivid-ZOO: Multi-View Video Generation with Diffusion Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Bing, Zheng, Cheng, Zhu, Wenxuan, Mai, Jinjie, Zhang, Biao, Wonka, Peter, Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Video Diffusion Model Reconstruct 4D Geometry?
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
TrackNeRF: Bundle Adjusting NeRF from Sparse and Noisy Views via Feature Tracks
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
4D-Bench: Benchmarking Multi-modal Large Language Models for 4D Object Understanding
von: Zhu, Wenxuan, et al.
Veröffentlicht: (2025)
von: Zhu, Wenxuan, et al.
Veröffentlicht: (2025)
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2025)
PoseGAM: Robust Unseen Object Pose Estimation via Geometry-Aware Multi-View Reasoning
von: Chen, Jianqi, et al.
Veröffentlicht: (2025)
von: Chen, Jianqi, et al.
Veröffentlicht: (2025)
Generative Human Geometry Distribution
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
efunc: An Efficient Function Representation without Neural Networks
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
Hybrid Structure-from-Motion and Camera Relocalization for Enhanced Egocentric Localization
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
Human Geometry Distribution for 3D Animation Generation
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
NearID: Identity Representation Learning via Near-identity Distractors
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2026)
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2026)
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
von: Li, Yan, et al.
Veröffentlicht: (2024)
von: Li, Yan, et al.
Veröffentlicht: (2024)
GenView: Enhancing View Quality with Pretrained Generative Model for Self-Supervised Learning
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2023)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
Geometry Distributions
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
VividFace: High-Quality and Efficient One-Step Diffusion For Video Face Enhancement
von: Zhang, Shulian, et al.
Veröffentlicht: (2025)
von: Zhang, Shulian, et al.
Veröffentlicht: (2025)
Lazy Layers to Make Fine-Tuned Diffusion Models More Traceable
von: Liu, Haozhe, et al.
Veröffentlicht: (2024)
von: Liu, Haozhe, et al.
Veröffentlicht: (2024)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
No Mesh, No Problem: Estimating Coral Volume and Surface from Sparse Multi-View Images
von: Farchione, Diego Eustachio, et al.
Veröffentlicht: (2025)
von: Farchione, Diego Eustachio, et al.
Veröffentlicht: (2025)
HairDiffusion: Vivid Multi-Colored Hair Editing via Latent Diffusion
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
von: Fang, Zixun, et al.
Veröffentlicht: (2025)
von: Fang, Zixun, et al.
Veröffentlicht: (2025)
SVP: Style-Enhanced Vivid Portrait Talking Head Diffusion Model
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
LaRI: Layered Ray Intersections for Single-view 3D Geometric Reasoning
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration
von: Bai, Haoran, et al.
Veröffentlicht: (2025)
von: Bai, Haoran, et al.
Veröffentlicht: (2025)
iFlame: Interleaving Full and Linear Attention for Efficient Mesh Generation
von: Wang, Hanxiao, et al.
Veröffentlicht: (2025)
von: Wang, Hanxiao, et al.
Veröffentlicht: (2025)
PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion Models
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2025)
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2025)
ImmersePro: End-to-End Stereo Video Synthesis Via Implicit Disparity Learning
von: Shi, Jian, et al.
Veröffentlicht: (2024)
von: Shi, Jian, et al.
Veröffentlicht: (2024)
Any Resolution Any Geometry: From Multi-View To Multi-Patch
von: Cui, Wenqing, et al.
Veröffentlicht: (2026)
von: Cui, Wenqing, et al.
Veröffentlicht: (2026)
MatCLIP: Light- and Shape-Insensitive Assignment of PBR Material Models
von: Birsak, Michael, et al.
Veröffentlicht: (2025)
von: Birsak, Michael, et al.
Veröffentlicht: (2025)
LumiX: Structured and Coherent Text-to-Intrinsic Generation
von: Han, Xu, et al.
Veröffentlicht: (2025)
von: Han, Xu, et al.
Veröffentlicht: (2025)
BrepGPT: Autoregressive B-rep Generation with Voronoi Half-Patch
von: Li, Pu, et al.
Veröffentlicht: (2025)
von: Li, Pu, et al.
Veröffentlicht: (2025)
Autoregressive Generation of Static and Growing Trees
von: Wang, Hanxiao, et al.
Veröffentlicht: (2025)
von: Wang, Hanxiao, et al.
Veröffentlicht: (2025)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
von: Zhu, Ruishu, et al.
Veröffentlicht: (2025)
von: Zhu, Ruishu, et al.
Veröffentlicht: (2025)
Video Self-Stitching Graph Network for Temporal Action Localization
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
Dynamically Masked Discriminator for Generative Adversarial Networks
von: Zhang, Wentian, et al.
Veröffentlicht: (2023)
von: Zhang, Wentian, et al.
Veröffentlicht: (2023)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can Video Diffusion Model Reconstruct 4D Geometry?
von: Mai, Jinjie, et al.
Veröffentlicht: (2025) -
TrackNeRF: Bundle Adjusting NeRF from Sparse and Noisy Views via Feature Tracks
von: Mai, Jinjie, et al.
Veröffentlicht: (2024) -
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
von: Zhang, Biao, et al.
Veröffentlicht: (2024) -
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025) -
4D-Bench: Benchmarking Multi-modal Large Language Models for 4D Object Understanding
von: Zhu, Wenxuan, et al.
Veröffentlicht: (2025)