Lightning Fast Caching-based Parallel Denoising Prediction for Accelerating Talking Head Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Long, Jianzhi, Sun, Wenhao, Tu, Rongcheng, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Turbo4DGen: Ultra-Fast Acceleration for 4D Generation
by: Man, Yuanbin, et al.
Published: (2026)
by: Man, Yuanbin, et al.
Published: (2026)
Seeing Fast and Slow: Learning the Flow of Time in Videos
by: Wu, Yen-Siang, et al.
Published: (2026)
by: Wu, Yen-Siang, et al.
Published: (2026)
PersonaTalk: Bring Attention to Your Persona in Visual Dubbing
by: Zhang, Longhao, et al.
Published: (2024)
by: Zhang, Longhao, et al.
Published: (2024)
VOODOO XP: Expressive One-Shot Head Reenactment for VR Telepresence
by: Tran, Phong, et al.
Published: (2024)
by: Tran, Phong, et al.
Published: (2024)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
by: Mao, Yuxiang, et al.
Published: (2025)
by: Mao, Yuxiang, et al.
Published: (2025)
EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation
by: Zhou, Wenyang, et al.
Published: (2023)
by: Zhou, Wenyang, et al.
Published: (2023)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
by: Aneja, Shivangi, et al.
Published: (2023)
by: Aneja, Shivangi, et al.
Published: (2023)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
by: Rakesh, Vineet Kumar, et al.
Published: (2025)
by: Rakesh, Vineet Kumar, et al.
Published: (2025)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
by: Sung-Bin, Kim, et al.
Published: (2024)
by: Sung-Bin, Kim, et al.
Published: (2024)
TexAvatars : Hybrid Texel-3D Representations for Stable Rigging of Photorealistic Gaussian Head Avatars
by: Lee, Jaeseong, et al.
Published: (2025)
by: Lee, Jaeseong, et al.
Published: (2025)
SurFhead: Affine Rig Blending for Geometrically Accurate 2D Gaussian Surfel Head Avatars
by: Lee, Jaeseong, et al.
Published: (2024)
by: Lee, Jaeseong, et al.
Published: (2024)
Morse: Dual-Sampling for Lossless Acceleration of Diffusion Models
by: Li, Chao, et al.
Published: (2025)
by: Li, Chao, et al.
Published: (2025)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
by: Wang, Xinmu, et al.
Published: (2025)
by: Wang, Xinmu, et al.
Published: (2025)
GeneOH Diffusion: Towards Generalizable Hand-Object Interaction Denoising via Denoising Diffusion
by: Liu, Xueyi, et al.
Published: (2024)
by: Liu, Xueyi, et al.
Published: (2024)
StyGazeTalk: Learning Stylized Generation of Gaze and Head Dynamics
by: Shi, Chengwei, et al.
Published: (2025)
by: Shi, Chengwei, et al.
Published: (2025)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
by: Li, Xinyang, et al.
Published: (2025)
by: Li, Xinyang, et al.
Published: (2025)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
by: Cheng, Shenggan, et al.
Published: (2025)
by: Cheng, Shenggan, et al.
Published: (2025)
SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
by: Liu, Yuan, et al.
Published: (2023)
by: Liu, Yuan, et al.
Published: (2023)
GazeFusion: Saliency-Guided Image Generation
by: Zhang, Yunxiang, et al.
Published: (2024)
by: Zhang, Yunxiang, et al.
Published: (2024)
DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos
by: Hu, Wenbo, et al.
Published: (2024)
by: Hu, Wenbo, et al.
Published: (2024)
Vector Grimoire: Codebook-based Shape Generation under Raster Image Supervision
by: Feuerpfeil, Moritz, et al.
Published: (2024)
by: Feuerpfeil, Moritz, et al.
Published: (2024)
SpaRP: Fast 3D Object Reconstruction and Pose Estimation from Sparse Views
by: Xu, Chao, et al.
Published: (2024)
by: Xu, Chao, et al.
Published: (2024)
FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models
by: Zhang, Zhanwei, et al.
Published: (2024)
by: Zhang, Zhanwei, et al.
Published: (2024)
SeqTex: Generate Mesh Textures in Video Sequence
by: Yuan, Ze, et al.
Published: (2025)
by: Yuan, Ze, et al.
Published: (2025)
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024)
by: Jiang, Liyao, et al.
Published: (2024)
DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
End-to-End Training for Unified Tokenization and Latent Denoising
by: Duggal, Shivam, et al.
Published: (2026)
by: Duggal, Shivam, et al.
Published: (2026)
Res2NetFuse: A Novel Res2Net-based Fusion Method for Infrared and Visible Images
by: Song, Xu, et al.
Published: (2021)
by: Song, Xu, et al.
Published: (2021)
Frankenstein: Generating Semantic-Compositional 3D Scenes in One Tri-Plane
by: Yan, Han, et al.
Published: (2024)
by: Yan, Han, et al.
Published: (2024)
Pandora3D: A Comprehensive Framework for High-Quality 3D Shape and Texture Generation
by: Yang, Jiayu, et al.
Published: (2025)
by: Yang, Jiayu, et al.
Published: (2025)
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
by: Wu, Zhennan, et al.
Published: (2024)
by: Wu, Zhennan, et al.
Published: (2024)
Generating by Understanding: Neural Visual Generation with Logical Symbol Groundings
by: Peng, Yifei, et al.
Published: (2023)
by: Peng, Yifei, et al.
Published: (2023)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
Freehand Sketch Generation from Mechanical Components
by: Liao, Zhichao, et al.
Published: (2024)
by: Liao, Zhichao, et al.
Published: (2024)
Annotated Hands for Generative Models
by: Yang, Yue, et al.
Published: (2024)
by: Yang, Yue, et al.
Published: (2024)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
by: Huang, Yihuan, et al.
Published: (2025)
by: Huang, Yihuan, et al.
Published: (2025)
Transcending Dimensions using Generative AI: Real-Time 3D Model Generation in Augmented Reality
by: Behravan, Majid, et al.
Published: (2025)
by: Behravan, Majid, et al.
Published: (2025)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
by: Sun, Zhiyao, et al.
Published: (2023)
by: Sun, Zhiyao, et al.
Published: (2023)
Physical Simulator In-the-Loop Video Generation
by: Foo, Lin Geng, et al.
Published: (2026)
by: Foo, Lin Geng, et al.
Published: (2026)
AirSketch: Generative Motion to Sketch
by: Lim, Hui Xian Grace, et al.
Published: (2024)
by: Lim, Hui Xian Grace, et al.
Published: (2024)
Similar Items
-
Turbo4DGen: Ultra-Fast Acceleration for 4D Generation
by: Man, Yuanbin, et al.
Published: (2026) -
Seeing Fast and Slow: Learning the Flow of Time in Videos
by: Wu, Yen-Siang, et al.
Published: (2026) -
PersonaTalk: Bring Attention to Your Persona in Visual Dubbing
by: Zhang, Longhao, et al.
Published: (2024) -
VOODOO XP: Expressive One-Shot Head Reenactment for VR Telepresence
by: Tran, Phong, et al.
Published: (2024) -
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
by: Mao, Yuxiang, et al.
Published: (2025)