A Self-supervised Motion Representation for Portrait Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Qiyuan, Wu, Chenyu, Sun, Wenzhang, Liu, Huaize, Di, Donglin, Chen, Wei, Zou, Changqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
DH-FaceVid-1K: A Large-Scale High-Quality Dataset for Face Video Generation
von: Di, Donglin, et al.
Veröffentlicht: (2024)
von: Di, Donglin, et al.
Veröffentlicht: (2024)
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
von: Feng, He, et al.
Veröffentlicht: (2026)
von: Feng, He, et al.
Veröffentlicht: (2026)
Collaboratively Self-supervised Video Representation Learning for Action Recognition
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
ChronoTailor: Harnessing Attention Guidance for Fine-Grained Video Virtual Try-On
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion Representations
von: Shi, Yuxiang, et al.
Veröffentlicht: (2025)
von: Shi, Yuxiang, et al.
Veröffentlicht: (2025)
Leveraging SAM for Single-Source Domain Generalization in Medical Image Segmentation
von: Wang, Hanhui, et al.
Veröffentlicht: (2024)
von: Wang, Hanhui, et al.
Veröffentlicht: (2024)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
PhysRVG: Physics-Aware Unified Reinforcement Learning for Video Generative Models
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
ExpPortrait: Expressive Portrait Generation via Personalized Representation
von: Wang, Junyi, et al.
Veröffentlicht: (2026)
von: Wang, Junyi, et al.
Veröffentlicht: (2026)
Portrait3D: Text-Guided High-Quality 3D Portrait Generation Using Pyramid Representation and GANs Prior
von: Wu, Yiqian, et al.
Veröffentlicht: (2024)
von: Wu, Yiqian, et al.
Veröffentlicht: (2024)
TV-3DG: Mastering Text-to-3D Customized Generation with Visual Prompt
von: Yang, Jiahui, et al.
Veröffentlicht: (2024)
von: Yang, Jiahui, et al.
Veröffentlicht: (2024)
Masked Modeling for Self-supervised Representation Learning on Vision and Beyond
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
von: Feng, He, et al.
Veröffentlicht: (2025)
von: Feng, He, et al.
Veröffentlicht: (2025)
Real Face Video Animation Platform
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
PRISM: Streaming Human Motion Generation with Per-Joint Latent Decomposition
von: Ling, Zeyu, et al.
Veröffentlicht: (2026)
von: Ling, Zeyu, et al.
Veröffentlicht: (2026)
Towards Imbalanced Motion: Part-Decoupling Network for Video Portrait Segmentation
von: Yu, Tianshu, et al.
Veröffentlicht: (2023)
von: Yu, Tianshu, et al.
Veröffentlicht: (2023)
Semantic-guided Adversarial Diffusion Model for Self-supervised Shadow Removal
von: Zeng, Ziqi, et al.
Veröffentlicht: (2024)
von: Zeng, Ziqi, et al.
Veröffentlicht: (2024)
LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control
von: Guo, Jianzhu, et al.
Veröffentlicht: (2024)
von: Guo, Jianzhu, et al.
Veröffentlicht: (2024)
LumiSculpt: Enabling Consistent Portrait Lighting in Video Generation
von: Zhang, Yuxin, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxin, et al.
Veröffentlicht: (2024)
DO3D: Self-supervised Learning of Decomposed Object-aware 3D Motion and Depth from Monocular Videos
von: Wu, Xiuzhe, et al.
Veröffentlicht: (2024)
von: Wu, Xiuzhe, et al.
Veröffentlicht: (2024)
ScrollScape: Unlocking 32K Image Generation With Video Diffusion Priors
von: Yu, Haodong, et al.
Veröffentlicht: (2026)
von: Yu, Haodong, et al.
Veröffentlicht: (2026)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
MVP: Enhancing Video Large Language Models via Self-supervised Masked Video Prediction
von: Sun, Xiaokun, et al.
Veröffentlicht: (2026)
von: Sun, Xiaokun, et al.
Veröffentlicht: (2026)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
EchoShot: Multi-Shot Portrait Video Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
Masked Diffusion as Self-supervised Representation Learner
von: Pan, Zixuan, et al.
Veröffentlicht: (2023)
von: Pan, Zixuan, et al.
Veröffentlicht: (2023)
FPGA: Flexible Portrait Generation Approach
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Stable Video Portraits
von: Ostrek, Mirela, et al.
Veröffentlicht: (2024)
von: Ostrek, Mirela, et al.
Veröffentlicht: (2024)
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
von: Chefer, Hila, et al.
Veröffentlicht: (2025)
von: Chefer, Hila, et al.
Veröffentlicht: (2025)
Return of Unconditional Generation: A Self-supervised Representation Generation Method
von: Li, Tianhong, et al.
Veröffentlicht: (2023)
von: Li, Tianhong, et al.
Veröffentlicht: (2023)
Learning Human Motion from Monocular Videos via Cross-Modal Manifold Alignment
von: Hou, Shuaiying, et al.
Veröffentlicht: (2024)
von: Hou, Shuaiying, et al.
Veröffentlicht: (2024)
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
VersatileMotion: A Unified Framework for Motion Synthesis and Comprehension
von: Ling, Zeyu, et al.
Veröffentlicht: (2024)
von: Ling, Zeyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion
von: Liu, Huaize, et al.
Veröffentlicht: (2025) -
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
von: Liu, Huaize, et al.
Veröffentlicht: (2025) -
UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024) -
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025) -
DH-FaceVid-1K: A Large-Scale High-Quality Dataset for Face Video Generation
von: Di, Donglin, et al.
Veröffentlicht: (2024)