EchoShot: Multi-Shot Portrait Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Jiahao, Sheng, Hualian, Cai, Sijia, Zhang, Weizhan, Yan, Caixia, Feng, Yachuang, Deng, Bing, Ye, Jieping |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References
by: Wang, Jiahao, et al.
Published: (2026)
by: Wang, Jiahao, et al.
Published: (2026)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
by: Yang, Yuxiao, et al.
Published: (2025)
by: Yang, Yuxiao, et al.
Published: (2025)
CT3D++: Improving 3D Object Detection with Keypoint-induced Channel-wise Transformer
by: Sheng, Hualian, et al.
Published: (2024)
by: Sheng, Hualian, et al.
Published: (2024)
PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models
by: Zhang, Jinhua, et al.
Published: (2024)
by: Zhang, Jinhua, et al.
Published: (2024)
RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception
by: Zhu, Xiaosu, et al.
Published: (2024)
by: Zhu, Xiaosu, et al.
Published: (2024)
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability
by: Hu, Xirui, et al.
Published: (2025)
by: Hu, Xirui, et al.
Published: (2025)
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
by: Kara, Ozgur, et al.
Published: (2025)
by: Kara, Ozgur, et al.
Published: (2025)
MVGSR: Multi-View Consistent 3D Gaussian Super-Resolution via Epipolar Guidance
by: Zhang, Kaizhe, et al.
Published: (2025)
by: Zhang, Kaizhe, et al.
Published: (2025)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
by: Wang, Qinghe, et al.
Published: (2025)
by: Wang, Qinghe, et al.
Published: (2025)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
by: Luo, Yawen, et al.
Published: (2026)
by: Luo, Yawen, et al.
Published: (2026)
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
OneActor: Consistent Character Generation via Cluster-Conditioned Guidance
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
ShotDirector: Directorially Controllable Multi-Shot Video Generation with Cinematographic Transitions
by: Wu, Xiaoxue, et al.
Published: (2025)
by: Wu, Xiaoxue, et al.
Published: (2025)
LLaFS: When Large Language Models Meet Few-Shot Segmentation
by: Zhu, Lanyun, et al.
Published: (2023)
by: Zhu, Lanyun, et al.
Published: (2023)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
by: Zhang, Zhihao, et al.
Published: (2023)
by: Zhang, Zhihao, et al.
Published: (2023)
DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis
by: Gu, Yuming, et al.
Published: (2023)
by: Gu, Yuming, et al.
Published: (2023)
MSVBench: Towards Human-Level Evaluation of Multi-Shot Video Generation
by: Shi, Haoyuan, et al.
Published: (2026)
by: Shi, Haoyuan, et al.
Published: (2026)
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
by: Liu, Jin, et al.
Published: (2024)
by: Liu, Jin, et al.
Published: (2024)
Data-Free Generalized Zero-Shot Learning
by: Tang, Bowen, et al.
Published: (2024)
by: Tang, Bowen, et al.
Published: (2024)
Generalized Few-Shot Out-of-Distribution Detection
by: Li, Pinxuan, et al.
Published: (2025)
by: Li, Pinxuan, et al.
Published: (2025)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
by: Nazarieh, Fatemeh, et al.
Published: (2024)
by: Nazarieh, Fatemeh, et al.
Published: (2024)
ShotVerse: Advancing Cinematic Camera Control for Text-Driven Multi-Shot Video Creation
by: Yang, Songlin, et al.
Published: (2026)
by: Yang, Songlin, et al.
Published: (2026)
OneStory: Coherent Multi-Shot Video Generation with Adaptive Memory
by: An, Zhaochong, et al.
Published: (2025)
by: An, Zhaochong, et al.
Published: (2025)
Portrait4D: Learning One-Shot 4D Head Avatar Synthesis using Synthetic Data
by: Deng, Yu, et al.
Published: (2023)
by: Deng, Yu, et al.
Published: (2023)
MindShot: Multi-Shot Video Reconstruction from fMRI with LLM Decoding
by: Zeng, Wenwen, et al.
Published: (2025)
by: Zeng, Wenwen, et al.
Published: (2025)
DreaMontage: Arbitrary Frame-Guided One-Shot Video Generation
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation
by: Wei, Yujie, et al.
Published: (2026)
by: Wei, Yujie, et al.
Published: (2026)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
by: He, Xuanhua, et al.
Published: (2024)
by: He, Xuanhua, et al.
Published: (2024)
MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation
by: Zhang, Haojie, et al.
Published: (2026)
by: Zhang, Haojie, et al.
Published: (2026)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
by: Wei, Yujie, et al.
Published: (2024)
by: Wei, Yujie, et al.
Published: (2024)
Scaling Zero-Shot Reference-to-Video Generation
by: Zhou, Zijian, et al.
Published: (2025)
by: Zhou, Zijian, et al.
Published: (2025)
High-Discriminative Attribute Feature Learning for Generalized Zero-Shot Learning
by: Lei, Yu, et al.
Published: (2024)
by: Lei, Yu, et al.
Published: (2024)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
by: Meng, Yihao, et al.
Published: (2025)
by: Meng, Yihao, et al.
Published: (2025)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
by: Meng, Yihao, et al.
Published: (2026)
by: Meng, Yihao, et al.
Published: (2026)
DepthArb: Training-Free Depth-Arbitrated Generation for Occlusion-Robust Image Synthesis
by: Niu, Hongjin, et al.
Published: (2026)
by: Niu, Hongjin, et al.
Published: (2026)
Few-Shot Generative Model Adaption via Identity Injection and Preservation
by: He, Yeqi, et al.
Published: (2026)
by: He, Yeqi, et al.
Published: (2026)
MUSE: Harnessing Precise and Diverse Semantics for Few-Shot Whole Slide Image Classification
by: Xu, Jiahao, et al.
Published: (2026)
by: Xu, Jiahao, et al.
Published: (2026)
MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation
by: Xing, Jinbo, et al.
Published: (2025)
by: Xing, Jinbo, et al.
Published: (2025)
Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation
by: Huang, Binyuan, et al.
Published: (2026)
by: Huang, Binyuan, et al.
Published: (2026)
SKALD: Learning-Based Shot Assembly for Coherent Multi-Shot Video Creation
by: Lu, Chen Yi, et al.
Published: (2025)
by: Lu, Chen Yi, et al.
Published: (2025)
Similar Items
-
AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References
by: Wang, Jiahao, et al.
Published: (2026) -
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
by: Yang, Yuxiao, et al.
Published: (2025) -
CT3D++: Improving 3D Object Detection with Keypoint-induced Channel-wise Transformer
by: Sheng, Hualian, et al.
Published: (2024) -
PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models
by: Zhang, Jinhua, et al.
Published: (2024) -
RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception
by: Zhu, Xiaosu, et al.
Published: (2024)