Gespeichert in:
| Hauptverfasser: | Zhao, Chengfeng, Shu, Jiazhi, Zhao, Yubo, Huang, Tianyu, Lu, Jiahao, Gu, Zekai, Ren, Chengwei, Dou, Zhiyang, Shuai, Qing, Liu, Yuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.10632 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GO-Renderer: Generative Object Rendering with 3D-aware Controllable Video Diffusion Models
von: Gu, Zekai, et al.
Veröffentlicht: (2026)
von: Gu, Zekai, et al.
Veröffentlicht: (2026)
HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos
von: Xiong, Weitao, et al.
Veröffentlicht: (2025)
von: Xiong, Weitao, et al.
Veröffentlicht: (2025)
ViMo: Generating Motions from Casual Videos
von: Qiu, Liangdong, et al.
Veröffentlicht: (2024)
von: Qiu, Liangdong, et al.
Veröffentlicht: (2024)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
von: Gu, Zekai, et al.
Veröffentlicht: (2025)
SyncHuman: Synchronizing 2D and 3D Generative Models for Single-view Human Reconstruction
von: Chen, Wenyue, et al.
Veröffentlicht: (2025)
von: Chen, Wenyue, et al.
Veröffentlicht: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
von: Zhao, Yubo, et al.
Veröffentlicht: (2026)
von: Zhao, Yubo, et al.
Veröffentlicht: (2026)
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
CoMo: Compositional Motion Customization for Text-to-Video Generation
von: Xu, Youcan, et al.
Veröffentlicht: (2025)
von: Xu, Youcan, et al.
Veröffentlicht: (2025)
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
UltraViCo: Breaking Extrapolation Limits in Video Diffusion Transformers
von: Zhao, Min, et al.
Veröffentlicht: (2025)
von: Zhao, Min, et al.
Veröffentlicht: (2025)
ScaMo: Exploring the Scaling Law in Autoregressive Motion Generation Model
von: Lu, Shunlin, et al.
Veröffentlicht: (2024)
von: Lu, Shunlin, et al.
Veröffentlicht: (2024)
CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects
von: Pi, Huaijin, et al.
Veröffentlicht: (2025)
von: Pi, Huaijin, et al.
Veröffentlicht: (2025)
ReCoM: Realistic Co-Speech Motion Generation with Recurrent Embedded Transformer
von: Xie, Yong, et al.
Veröffentlicht: (2025)
von: Xie, Yong, et al.
Veröffentlicht: (2025)
UNIC: Neural Garment Deformation Field for Real-time Clothed Character Animation
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
Realistic Human Motion Generation with Cross-Diffusion Models
von: Ren, Zeping, et al.
Veröffentlicht: (2023)
von: Ren, Zeping, et al.
Veröffentlicht: (2023)
TLControl: Trajectory and Language Control for Human Motion Synthesis
von: Wan, Weilin, et al.
Veröffentlicht: (2023)
von: Wan, Weilin, et al.
Veröffentlicht: (2023)
ViMoNet: A Multimodal Vision-Language Framework for Human Behavior Understanding from Motion and Video
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
von: Gupta, Rajan Das, et al.
Veröffentlicht: (2025)
MotionPRO: Exploring the Role of Pressure in Human MoCap and Beyond
von: Ren, Shenghao, et al.
Veröffentlicht: (2025)
von: Ren, Shenghao, et al.
Veröffentlicht: (2025)
Co-Alignment: Rethinking Alignment as Bidirectional Human-AI Cognitive Adaptation
von: Li, Yubo, et al.
Veröffentlicht: (2025)
von: Li, Yubo, et al.
Veröffentlicht: (2025)
Can You Count to Nine? A Human Evaluation Benchmark for Counting Limits in Modern Text-to-Video Models
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
CoViLLM: An Adaptive Human-Robot Collaborative Assembly Framework Using Large Language Models
von: Zhao, Jiabao, et al.
Veröffentlicht: (2026)
von: Zhao, Jiabao, et al.
Veröffentlicht: (2026)
Dynamic Realms: 4D Content Analysis, Recovery and Generation with Geometric, Topological and Physical Priors
von: Dou, Zhiyang
Veröffentlicht: (2024)
von: Dou, Zhiyang
Veröffentlicht: (2024)
ViBE: Co-Optimizing Workload Skew and Hardware Variability for MoE Serving
von: Go, Seokjin, et al.
Veröffentlicht: (2026)
von: Go, Seokjin, et al.
Veröffentlicht: (2026)
AnyAct: Towards Human Reenactment of Character Motion From Video
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
TrackingWorld: World-centric Monocular 3D Tracking of Almost All Pixels
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
LaserHuman: Language-guided Scene-aware Human Motion Generation in Free Environment
von: Cong, Peishan, et al.
Veröffentlicht: (2024)
von: Cong, Peishan, et al.
Veröffentlicht: (2024)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023)
von: Wang, Tan, et al.
Veröffentlicht: (2023)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
von: Sheng, Mingzhi, et al.
Veröffentlicht: (2026)
von: Sheng, Mingzhi, et al.
Veröffentlicht: (2026)
CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning
von: Yang, Jiange, et al.
Veröffentlicht: (2025)
von: Yang, Jiange, et al.
Veröffentlicht: (2025)
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs
von: Wu, Qi, et al.
Veröffentlicht: (2024)
von: Wu, Qi, et al.
Veröffentlicht: (2024)
MOSPA: Human Motion Generation Driven by Spatial Audio
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
DeCo: Decoupled Human-Centered Diffusion Video Editing with Motion Consistency
von: Zhong, Xiaojing, et al.
Veröffentlicht: (2024)
von: Zhong, Xiaojing, et al.
Veröffentlicht: (2024)
Integrating Physician Diagnostic Logic into Large Language Models: Preference Learning from Process Feedback
von: Dou, Chengfeng, et al.
Veröffentlicht: (2024)
von: Dou, Chengfeng, et al.
Veröffentlicht: (2024)
GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure
von: Gu, Leslie, et al.
Veröffentlicht: (2025)
von: Gu, Leslie, et al.
Veröffentlicht: (2025)
EgoReAct: Egocentric Video-Driven 3D Human Reaction Generation
von: Zhang, Libo, et al.
Veröffentlicht: (2025)
von: Zhang, Libo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GO-Renderer: Generative Object Rendering with 3D-aware Controllable Video Diffusion Models
von: Gu, Zekai, et al.
Veröffentlicht: (2026) -
HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos
von: Xiong, Weitao, et al.
Veröffentlicht: (2025) -
ViMo: Generating Motions from Casual Videos
von: Qiu, Liangdong, et al.
Veröffentlicht: (2024) -
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
von: Gu, Zekai, et al.
Veröffentlicht: (2025) -
SyncHuman: Synchronizing 2D and 3D Generative Models for Single-view Human Reconstruction
von: Chen, Wenyue, et al.
Veröffentlicht: (2025)