Gespeichert in:
| Hauptverfasser: | Li, Chaochao, Wang, Ruikui, Zhou, Liangbo, Feng, Jinheng, Luo, Huaishao, Zhang, Huan, Wu, Youzheng, He, Xiaodong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2512.11423 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
JoyStreamer: Unlocking Highly Expressive Avatars via Harmonized Text-Audio Conditioning
von: Wang, Ruikui, et al.
Veröffentlicht: (2026)
von: Wang, Ruikui, et al.
Veröffentlicht: (2026)
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
SoulX-FlashTalk: Real-Time Infinite Streaming of Audio-Driven Avatars via Self-Correcting Bidirectional Distillation
von: Shen, Le, et al.
Veröffentlicht: (2025)
von: Shen, Le, et al.
Veröffentlicht: (2025)
MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space
von: Xiao, Lixing, et al.
Veröffentlicht: (2025)
von: Xiao, Lixing, et al.
Veröffentlicht: (2025)
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation
von: Zhou, Junkang, et al.
Veröffentlicht: (2026)
von: Zhou, Junkang, et al.
Veröffentlicht: (2026)
SoulX-FlashHead: Oracle-guided Generation of Infinite Real-time Streaming Talking Heads
von: Yu, Tan, et al.
Veröffentlicht: (2026)
von: Yu, Tan, et al.
Veröffentlicht: (2026)
Infinite Gaze Generation for Videos with Autoregressive Diffusion
von: Kang, Jenna, et al.
Veröffentlicht: (2026)
von: Kang, Jenna, et al.
Veröffentlicht: (2026)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation
von: Cao, Xuyang, et al.
Veröffentlicht: (2024)
von: Cao, Xuyang, et al.
Veröffentlicht: (2024)
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
von: Cai, Aichen, et al.
Veröffentlicht: (2026)
von: Cai, Aichen, et al.
Veröffentlicht: (2026)
OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
A Unit Enhancement and Guidance Framework for Audio-Driven Avatar Video Generation
von: Zhou, S. Z., et al.
Veröffentlicht: (2025)
von: Zhou, S. Z., et al.
Veröffentlicht: (2025)
Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
InfiniteAudio: Infinite-Length Audio Generation with Consistency
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
von: Sun, Zhiyao, et al.
Veröffentlicht: (2025)
von: Sun, Zhiyao, et al.
Veröffentlicht: (2025)
Audio-Driven Universal Gaussian Head Avatars
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
GaussianSpeech: Audio-Driven Gaussian Avatars
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
FlashAvatar: High-fidelity Head Avatar with Efficient Gaussian Embedding
von: Xiang, Jun, et al.
Veröffentlicht: (2023)
von: Xiang, Jun, et al.
Veröffentlicht: (2023)
TalkingMachines: Real-Time Audio-Driven FaceTime-Style Video via Autoregressive Diffusion Models
von: Low, Chetwin, et al.
Veröffentlicht: (2025)
von: Low, Chetwin, et al.
Veröffentlicht: (2025)
FlashAudio: Rectified Flows for Fast and High-Fidelity Text-to-Audio Generation
von: Liu, Huadai, et al.
Veröffentlicht: (2024)
von: Liu, Huadai, et al.
Veröffentlicht: (2024)
Nods of Agreement: Webcam-Driven Avatars Improve Meeting Outcomes and Avatar Satisfaction Over Audio-Driven or Static Avatars in All-Avatar Work Videoconferencing
von: Ma, Fang, et al.
Veröffentlicht: (2024)
von: Ma, Fang, et al.
Veröffentlicht: (2024)
UME: Upcycling Mixture-of-Experts for Scalable and Efficient Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2024)
von: Fu, Li, et al.
Veröffentlicht: (2024)
PAC: Pronunciation-Aware Contextualized Large Language Model-based Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2025)
von: Fu, Li, et al.
Veröffentlicht: (2025)
XHand: Real-time Expressive Hand Avatar
von: Gan, Qijun, et al.
Veröffentlicht: (2024)
von: Gan, Qijun, et al.
Veröffentlicht: (2024)
Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer
von: Lei, Ke, et al.
Veröffentlicht: (2026)
von: Lei, Ke, et al.
Veröffentlicht: (2026)
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing
von: Yang, Shaoshu, et al.
Veröffentlicht: (2025)
von: Yang, Shaoshu, et al.
Veröffentlicht: (2025)
JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing
von: Wang, Qili, et al.
Veröffentlicht: (2025)
von: Wang, Qili, et al.
Veröffentlicht: (2025)
Computational Model for Photoionization in Pure SF6 Streamer at 1-15 atm
von: Feng, Zihao, et al.
Veröffentlicht: (2025)
von: Feng, Zihao, et al.
Veröffentlicht: (2025)
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
von: Chen, Yi, et al.
Veröffentlicht: (2025)
von: Chen, Yi, et al.
Veröffentlicht: (2025)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
Differences in Text Generated by Diffusion and Autoregressive Language Models
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
AudCast: Audio-Driven Human Video Generation by Cascaded Diffusion Transformers
von: Guan, Jiazhi, et al.
Veröffentlicht: (2025)
von: Guan, Jiazhi, et al.
Veröffentlicht: (2025)
Real-Time Motion-Controllable Autoregressive Video Diffusion
von: Zhao, Kesen, et al.
Veröffentlicht: (2025)
von: Zhao, Kesen, et al.
Veröffentlicht: (2025)
LINA: Learning INterventions Adaptively for Physical Alignment and Generalization in Diffusion Models
von: Yu, Shu, et al.
Veröffentlicht: (2025)
von: Yu, Shu, et al.
Veröffentlicht: (2025)
Neighboring Autoregressive Modeling for Efficient Visual Generation
von: He, Yefei, et al.
Veröffentlicht: (2025)
von: He, Yefei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
JoyStreamer: Unlocking Highly Expressive Avatars via Harmonized Text-Audio Conditioning
von: Wang, Ruikui, et al.
Veröffentlicht: (2026) -
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length
von: Huang, Yubo, et al.
Veröffentlicht: (2025) -
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025) -
SoulX-FlashTalk: Real-Time Infinite Streaming of Audio-Driven Avatars via Self-Correcting Bidirectional Distillation
von: Shen, Le, et al.
Veröffentlicht: (2025) -
MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space
von: Xiao, Lixing, et al.
Veröffentlicht: (2025)