X-Streamer: Unified Human World Modeling with Audiovisual Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Xie, You, Gu, Tianpei, Li, Zenan, Zhang, Chenxu, Song, Guoxian, Zhao, Xiaochen, Liang, Chao, Jiang, Jianwen, Xu, Hongyi, Luo, Linjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
by: Song, Guoxian, et al.
Published: (2025)
by: Song, Guoxian, et al.
Published: (2025)
X-Actor: Emotional and Expressive Long-Range Portrait Acting from Audio
by: Zhang, Chenxu, et al.
Published: (2025)
by: Zhang, Chenxu, et al.
Published: (2025)
Plan-X: Instruct Video Generation via Semantic Planning
by: Huang, Lun, et al.
Published: (2025)
by: Huang, Lun, et al.
Published: (2025)
X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention
by: Zhao, Xiaochen, et al.
Published: (2025)
by: Zhao, Xiaochen, et al.
Published: (2025)
X-Dancer: Expressive Music to Human Dance Video Generation
by: Chen, Zeyuan, et al.
Published: (2025)
by: Chen, Zeyuan, et al.
Published: (2025)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
by: Xie, You, et al.
Published: (2024)
by: Xie, You, et al.
Published: (2024)
X-Dyna: Expressive Dynamic Human Image Animation
by: Chang, Di, et al.
Published: (2025)
by: Chang, Di, et al.
Published: (2025)
DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis
by: Gu, Yuming, et al.
Published: (2023)
by: Gu, Yuming, et al.
Published: (2023)
High Quality Human Image Animation using Regional Supervision and Motion Blur Condition
by: Xu, Zhongcong, et al.
Published: (2024)
by: Xu, Zhongcong, et al.
Published: (2024)
Lynx: Towards High-Fidelity Personalized Video Generation
by: Sang, Shen, et al.
Published: (2025)
by: Sang, Shen, et al.
Published: (2025)
Video-As-Prompt: Unified Semantic Control for Video Generation
by: Bian, Yuxuan, et al.
Published: (2025)
by: Bian, Yuxuan, et al.
Published: (2025)
Bridging Your Imagination with Audio-Video Generation via a Unified Director
by: Zhang, Jiaxu, et al.
Published: (2025)
by: Zhang, Jiaxu, et al.
Published: (2025)
OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
by: Lin, Gaojie, et al.
Published: (2025)
by: Lin, Gaojie, et al.
Published: (2025)
CU-Mamba: Selective State Space Models with Channel Learning for Image Restoration
by: Deng, Rui, et al.
Published: (2024)
by: Deng, Rui, et al.
Published: (2024)
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
by: Yang, Yiming, et al.
Published: (2025)
by: Yang, Yiming, et al.
Published: (2025)
How Does Mandated Non‐Financial Disclosure Affect Corporate Cash Holdings? Evidence From the CSR Disclosure Mandate in China
by: Adrian Cheung, et al.
Published: (2026)
by: Adrian Cheung, et al.
Published: (2026)
Simulating Human Audiovisual Search Behavior
by: Cho, Hyunsung, et al.
Published: (2026)
by: Cho, Hyunsung, et al.
Published: (2026)
U-Mind: A Unified Framework for Real-Time Multimodal Interaction with Audiovisual Generation
by: Deng, Xiang, et al.
Published: (2026)
by: Deng, Xiang, et al.
Published: (2026)
MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
by: Chang, Di, et al.
Published: (2023)
by: Chang, Di, et al.
Published: (2023)
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
by: Liang, Chao, et al.
Published: (2025)
by: Liang, Chao, et al.
Published: (2025)
Duo Streamers: A Streaming Gesture Recognition Framework
by: Zhu, Boxuan, et al.
Published: (2025)
by: Zhu, Boxuan, et al.
Published: (2025)
InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction
by: Ren, Pengzhen, et al.
Published: (2024)
by: Ren, Pengzhen, et al.
Published: (2024)
Revealing AI Reasoning Increases Trust but Crowds Out Unique Human Knowledge
by: Chen, Zenan, et al.
Published: (2025)
by: Chen, Zenan, et al.
Published: (2025)
From Interaction to Collaboration: Reforming Audiovisual Communication Education in the Era of AIGC
by: Fang Xie1, Yu Rui2*
Published: (2026)
by: Fang Xie1, Yu Rui2*
Published: (2026)
The Reservoir of the Per-emb-2 Streamer
by: Taniguchi, Kotomi, et al.
Published: (2024)
by: Taniguchi, Kotomi, et al.
Published: (2024)
The Effect of NO 2 on the Corrosion Mechanism of P110 Tubing Steel Used for CO 2 Storage
by: Zhen Yang, et al.
Published: (2025)
by: Zhen Yang, et al.
Published: (2025)
Computational Model for Photoionization in Pure SF6 Streamer at 1-15 atm
by: Feng, Zihao, et al.
Published: (2025)
by: Feng, Zihao, et al.
Published: (2025)
Unified Understanding of Environment, Task, and Human for Human-Robot Interaction in Real-World Environments
by: Yano, Yuga, et al.
Published: (2024)
by: Yano, Yuga, et al.
Published: (2024)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
by: Cai, Songjin, et al.
Published: (2026)
by: Cai, Songjin, et al.
Published: (2026)
Dynamics of Streamers and Pseudostreamers and Implications for the Solar Wind
by: Dey, Sahel, et al.
Published: (2025)
by: Dey, Sahel, et al.
Published: (2025)
"The Intangible Victory", Interactive Audiovisual Installation
by: Tsioutas, Konstantinos, et al.
Published: (2026)
by: Tsioutas, Konstantinos, et al.
Published: (2026)
Effect of Electric Field Distribution on Branching Characteristics of Streamers in Transformer Oil
by: Tao Zhu, et al.
Published: (2025)
by: Tao Zhu, et al.
Published: (2025)
Spectral conditions for spanning $k$-trees or $k$-ended-trees of $t$-connected graphs
by: Lin, Jifu, et al.
Published: (2024)
by: Lin, Jifu, et al.
Published: (2024)
Hydrodynamics for asymmetric simple exclusion on a finite segment with Glauber-type source
by: Xu, Lu, et al.
Published: (2024)
by: Xu, Lu, et al.
Published: (2024)
Nonequilibrium fluctuations for the occupation time of the SSEP in $d \geq 2$
by: Xu, Tiecheng, et al.
Published: (2025)
by: Xu, Tiecheng, et al.
Published: (2025)
FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation
by: Zhong, Tianyun, et al.
Published: (2024)
by: Zhong, Tianyun, et al.
Published: (2024)
A Unified Interaction Control Framework for Safe Robotic Ultrasound Scanning with Human-Intention-Aware Compliance
by: Yan, Xiangjie, et al.
Published: (2024)
by: Yan, Xiangjie, et al.
Published: (2024)
Semantic Satellite Communications for Synchronized Audiovisual Reconstruction
by: Liu, Fangyu, et al.
Published: (2026)
by: Liu, Fangyu, et al.
Published: (2026)
When Virtual Meets Authenticity: How Virtual Streamers Shape Consumer Decision‐Making Confidence
by: Ying Xie, et al.
Published: (2025)
by: Ying Xie, et al.
Published: (2025)
Bio‐Inspired Multiple Responsive NIR II Nanophosphors for Reversible and Environment‐Interactive Information Encryption
by: Tianpei He, et al.
Published: (2024)
by: Tianpei He, et al.
Published: (2024)
Similar Items
-
X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
by: Song, Guoxian, et al.
Published: (2025) -
X-Actor: Emotional and Expressive Long-Range Portrait Acting from Audio
by: Zhang, Chenxu, et al.
Published: (2025) -
Plan-X: Instruct Video Generation via Semantic Planning
by: Huang, Lun, et al.
Published: (2025) -
X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention
by: Zhao, Xiaochen, et al.
Published: (2025) -
X-Dancer: Expressive Music to Human Dance Video Generation
by: Chen, Zeyuan, et al.
Published: (2025)