Wan-S2V: Audio-Driven Cinematic Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Xin, Hu, Li, Hu, Siqi, Huang, Mingyang, Ji, Chaonan, Meng, Dechao, Qi, Jinwei, Qiao, Penchong, Shen, Zhen, Song, Yafei, Sun, Ke, Tian, Linrui, Wang, Guangyuan, Wang, Qi, Wang, Zhongjian, Xiao, Jiayu, Xu, Sheng, Zhang, Bang, Zhang, Peng, Zhang, Xindi, Zhang, Zhe, Zhou, Jingren, Zhuo, Lian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Wan-Animate: Unified Character Animation and Replacement with Holistic Replication
por: Cheng, Gang, et al.
Publicado: (2025)
por: Cheng, Gang, et al.
Publicado: (2025)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
por: Wang, Zhongjian, et al.
Publicado: (2025)
por: Wang, Zhongjian, et al.
Publicado: (2025)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
por: Meng, Dechao, et al.
Publicado: (2025)
por: Meng, Dechao, et al.
Publicado: (2025)
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
por: Tian, Linrui, et al.
Publicado: (2025)
por: Tian, Linrui, et al.
Publicado: (2025)
SyncAnyone: Implicit Disentanglement via Progressive Self-Correction for Lip-Syncing in the wild
por: Zhang, Xindi, et al.
Publicado: (2025)
por: Zhang, Xindi, et al.
Publicado: (2025)
Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation
por: Xiao, Steven, et al.
Publicado: (2025)
por: Xiao, Steven, et al.
Publicado: (2025)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
por: Tian, Linrui, et al.
Publicado: (2024)
por: Tian, Linrui, et al.
Publicado: (2024)
PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment
por: Ji, Chaonan, et al.
Publicado: (2026)
por: Ji, Chaonan, et al.
Publicado: (2026)
Controllable and Expressive One-Shot Video Head Swapping
por: Ji, Chaonan, et al.
Publicado: (2025)
por: Ji, Chaonan, et al.
Publicado: (2025)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
por: Hu, Li, et al.
Publicado: (2025)
por: Hu, Li, et al.
Publicado: (2025)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
por: Sun, Ke, et al.
Publicado: (2024)
por: Sun, Ke, et al.
Publicado: (2024)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
por: Qi, Jinwei, et al.
Publicado: (2025)
por: Qi, Jinwei, et al.
Publicado: (2025)
Exploring Timeline Control for Facial Motion Generation
por: Ma, Yifeng, et al.
Publicado: (2025)
por: Ma, Yifeng, et al.
Publicado: (2025)
A new proof of Carlitz-Wan conjecture on exceptional polynomials
por: Hu, Yilong, et al.
Publicado: (2026)
por: Hu, Yilong, et al.
Publicado: (2026)
A Review on the Current Research Status of Key Areas in Wireless Capsule Endoscopy
por: Guangyuan Wang, et al.
Publicado: (2025)
por: Guangyuan Wang, et al.
Publicado: (2025)
UCM: Unifying Camera Control and Memory with Time-aware Positional Encoding Warping for World Models
por: Xu, Tianxing, et al.
Publicado: (2026)
por: Xu, Tianxing, et al.
Publicado: (2026)
Cinematic Audio Source Separation Using Visual Cues
por: Zhang, Kang, et al.
Publicado: (2026)
por: Zhang, Kang, et al.
Publicado: (2026)
Wan-R1: Verifiable-Reinforcement Learning for Video Reasoning
por: Liu, Ming, et al.
Publicado: (2026)
por: Liu, Ming, et al.
Publicado: (2026)
A mesh-free method for interface problems using the deep learning approach
por: Wang, Zhongjian, et al.
Publicado: (2019)
por: Wang, Zhongjian, et al.
Publicado: (2019)
A Novel Stochastic Particle-Field Algorithm for a Reaction-Diffusion-Advection Cancer Invasion Model
por: Hu, Jingyuan, et al.
Publicado: (2026)
por: Hu, Jingyuan, et al.
Publicado: (2026)
A Stochastic Interacting Particle-Field Algorithm for a Haptotaxis Advection-Diffusion System Modeling Cancer Cell Invasion
por: Hu, Boyi, et al.
Publicado: (2024)
por: Hu, Boyi, et al.
Publicado: (2024)
A Stochastic Genetic Interacting Particle Method for Reaction-Diffusion-Advection Equations
por: Hu, Boyi, et al.
Publicado: (2025)
por: Hu, Boyi, et al.
Publicado: (2025)
Convergence Analysis of a Stochastic Interacting Particle-Field Algorithm for 3D Parabolic-Parabolic Keller-Segel Systems
por: Hu, Boyi, et al.
Publicado: (2025)
por: Hu, Boyi, et al.
Publicado: (2025)
A fast stochastic interacting particle-field method for 3D parabolic parabolic Chemotaxis systems: numerical algorithms and error analysis
por: Hu, Jingyuan, et al.
Publicado: (2025)
por: Hu, Jingyuan, et al.
Publicado: (2025)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
por: Song, Yafei, et al.
Publicado: (2025)
por: Song, Yafei, et al.
Publicado: (2025)
Projection Head is Secretly an Information Bottleneck
por: Ouyang, Zhuo, et al.
Publicado: (2025)
por: Ouyang, Zhuo, et al.
Publicado: (2025)
Inverter Redistribution through Self-Dual and Self-Anti-Dual Function Transformation
por: Wang, Jingren, et al.
Publicado: (2026)
por: Wang, Jingren, et al.
Publicado: (2026)
Wan-Image: Pushing the Boundaries of Generative Visual Intelligence
por: Mao, Chaojie, et al.
Publicado: (2026)
por: Mao, Chaojie, et al.
Publicado: (2026)
Camera Artist: A Multi-Agent Framework for Cinematic Language Storytelling Video Generation
por: Hu, Haobo, et al.
Publicado: (2026)
por: Hu, Haobo, et al.
Publicado: (2026)
MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation
por: Huang, Mingyang, et al.
Publicado: (2025)
por: Huang, Mingyang, et al.
Publicado: (2025)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
por: Sun, Zhe, et al.
Publicado: (2025)
por: Sun, Zhe, et al.
Publicado: (2025)
Balance Divergence for Knowledge Distillation
por: Qi, Yafei, et al.
Publicado: (2025)
por: Qi, Yafei, et al.
Publicado: (2025)
Wan: Open and Advanced Large-Scale Video Generative Models
por: Wan, Team, et al.
Publicado: (2025)
por: Wan, Team, et al.
Publicado: (2025)
Preconditioned One-Step Generative Modeling for Bayesian Inverse Problems in Function Spaces
por: Cheng, Zilan, et al.
Publicado: (2026)
por: Cheng, Zilan, et al.
Publicado: (2026)
A DeepParticle method for learning and generating aggregation patterns in multi-dimensional Keller-Segel chemotaxis systems
por: Wang, Zhongjian, et al.
Publicado: (2022)
por: Wang, Zhongjian, et al.
Publicado: (2022)
A Novel Stochastic Interacting Particle-Field Algorithm for 3D Parabolic-Parabolic Keller-Segel Chemotaxis System
por: Wang, Zhongjian, et al.
Publicado: (2023)
por: Wang, Zhongjian, et al.
Publicado: (2023)
ShotVerse: Advancing Cinematic Camera Control for Text-Driven Multi-Shot Video Creation
por: Yang, Songlin, et al.
Publicado: (2026)
por: Yang, Songlin, et al.
Publicado: (2026)
Exceptional extensions of local fields and the Carlitz--Wan conjecture
por: Ding, Zhiguo, et al.
Publicado: (2025)
por: Ding, Zhiguo, et al.
Publicado: (2025)
A structure-preserving scheme for computing effective diffusivity and anomalous diffusion phenomena of random flows
por: Zhang, Tan, et al.
Publicado: (2024)
por: Zhang, Tan, et al.
Publicado: (2024)
A Bidirectional DeepParticle Method for Efficiently Solving Low-dimensional Transport Map Problems
por: Zhang, Tan, et al.
Publicado: (2025)
por: Zhang, Tan, et al.
Publicado: (2025)
Ejemplares similares
-
Wan-Animate: Unified Character Animation and Replacement with Holistic Replication
por: Cheng, Gang, et al.
Publicado: (2025) -
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
por: Wang, Zhongjian, et al.
Publicado: (2025) -
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
por: Meng, Dechao, et al.
Publicado: (2025) -
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
por: Tian, Linrui, et al.
Publicado: (2025) -
SyncAnyone: Implicit Disentanglement via Progressive Self-Correction for Lip-Syncing in the wild
por: Zhang, Xindi, et al.
Publicado: (2025)