Exploring Timeline Control for Facial Motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yifeng, Qi, Jinwei, Ji, Chaonan, Zhang, Peng, Zhang, Bang, Deng, Zhidong, Bo, Liefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Controllable and Expressive One-Shot Video Head Swapping
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment
von: Ji, Chaonan, et al.
Veröffentlicht: (2026)
von: Ji, Chaonan, et al.
Veröffentlicht: (2026)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
AnyText2: Visual Text Generation and Editing With Customizable Attributes
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
von: Hu, Li, et al.
Veröffentlicht: (2023)
von: Hu, Li, et al.
Veröffentlicht: (2023)
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024)
von: He, Junjie, et al.
Veröffentlicht: (2024)
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation
von: Chen, Yingjie, et al.
Veröffentlicht: (2025)
von: Chen, Yingjie, et al.
Veröffentlicht: (2025)
Knowledge-Guided Prompt Learning for Deepfake Facial Image Detection
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
von: Song, Yafei, et al.
Veröffentlicht: (2025)
von: Song, Yafei, et al.
Veröffentlicht: (2025)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
von: He, Junjie, et al.
Veröffentlicht: (2025)
von: He, Junjie, et al.
Veröffentlicht: (2025)
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
von: Petrovich, Mathis, et al.
Veröffentlicht: (2024)
von: Petrovich, Mathis, et al.
Veröffentlicht: (2024)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
von: Hu, Li, et al.
Veröffentlicht: (2025)
von: Hu, Li, et al.
Veröffentlicht: (2025)
MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
von: Men, Yifang, et al.
Veröffentlicht: (2024)
von: Men, Yifang, et al.
Veröffentlicht: (2024)
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
von: Yuan, Weihao, et al.
Veröffentlicht: (2024)
von: Yuan, Weihao, et al.
Veröffentlicht: (2024)
TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
Wan-S2V: Audio-Driven Cinematic Video Generation
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
von: Sun, Ke, et al.
Veröffentlicht: (2024)
von: Sun, Ke, et al.
Veröffentlicht: (2024)
MoSAM: Motion-Guided Segment Anything Model with Spatial-Temporal Memory Selection
von: Yang, Qiushi, et al.
Veröffentlicht: (2025)
von: Yang, Qiushi, et al.
Veröffentlicht: (2025)
VQTalker: Towards Multilingual Talking Avatars through Facial Motion Tokenization
von: Liu, Tao, et al.
Veröffentlicht: (2024)
von: Liu, Tao, et al.
Veröffentlicht: (2024)
EmojiDiff: Advanced Facial Expression Control with High Identity Preservation in Portrait Generation
von: Jiang, Liangwei, et al.
Veröffentlicht: (2024)
von: Jiang, Liangwei, et al.
Veröffentlicht: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
I4VGen: Image as Free Stepping Stone for Text-to-Video Generation
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
MaTe3D: Mask-guided Text-based 3D-aware Portrait Editing
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
BeyondFacial: Identity-Preserving Personalized Generation Beyond Facial Close-ups
von: Zhang, Songsong, et al.
Veröffentlicht: (2025)
von: Zhang, Songsong, et al.
Veröffentlicht: (2025)
GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
HyperMotionX: The Dataset and Benchmark with DiT-Based Pose-Guided Human Image Animation of Complex Motions
von: Xu, Shuolin, et al.
Veröffentlicht: (2025)
von: Xu, Shuolin, et al.
Veröffentlicht: (2025)
MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
CoGenAV: Versatile Audio-Visual Representation Learning via Contrastive-Generative Synchronization
von: Bai, Detao, et al.
Veröffentlicht: (2025)
von: Bai, Detao, et al.
Veröffentlicht: (2025)
Event-Based Motion Magnification
von: Chen, Yutian, et al.
Veröffentlicht: (2024)
von: Chen, Yutian, et al.
Veröffentlicht: (2024)
Capturing the Unseen: Vision-Free Facial Motion Capture Using Inertial Measurement Units
von: Wang, Youjia, et al.
Veröffentlicht: (2024)
von: Wang, Youjia, et al.
Veröffentlicht: (2024)
DiffuEraser: A Diffusion Model for Video Inpainting
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Controllable and Expressive One-Shot Video Head Swapping
von: Ji, Chaonan, et al.
Veröffentlicht: (2025) -
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
von: Qi, Jinwei, et al.
Veröffentlicht: (2025) -
PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment
von: Ji, Chaonan, et al.
Veröffentlicht: (2026) -
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025) -
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)