Salvato in:
| Autori principali: | Gan, Qijun, Ren, Yi, Zhang, Chen, Ye, Zhenhui, Xie, Pan, Yin, Xiang, Yuan, Zehuan, Peng, Bingyue, Zhu, Jianke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2502.04847 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
InfinityHuman: Towards Long-Term Audio-Driven Human
di: Li, Xiaodi, et al.
Pubblicazione: (2025)
di: Li, Xiaodi, et al.
Pubblicazione: (2025)
PianoMotion10M: Dataset and Benchmark for Hand Motion Generation in Piano Performance
di: Gan, Qijun, et al.
Pubblicazione: (2024)
di: Gan, Qijun, et al.
Pubblicazione: (2024)
HLLM: Enhancing Sequential Recommendations via Hierarchical Large Language Models for Item and User Modeling
di: Chen, Junyi, et al.
Pubblicazione: (2024)
di: Chen, Junyi, et al.
Pubblicazione: (2024)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
di: Guo, Ying, et al.
Pubblicazione: (2026)
di: Guo, Ying, et al.
Pubblicazione: (2026)
Fine-Grained Multi-View Hand Reconstruction Using Inverse Rendering
di: Gan, Qijun, et al.
Pubblicazione: (2024)
di: Gan, Qijun, et al.
Pubblicazione: (2024)
XHand: Real-time Expressive Hand Avatar
di: Gan, Qijun, et al.
Pubblicazione: (2024)
di: Gan, Qijun, et al.
Pubblicazione: (2024)
HyperMotionX: The Dataset and Benchmark with DiT-Based Pose-Guided Human Image Animation of Complex Motions
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
di: Tian, Keyu, et al.
Pubblicazione: (2024)
di: Tian, Keyu, et al.
Pubblicazione: (2024)
Generative Refinement Networks for Visual Synthesis
di: Han, Jian, et al.
Pubblicazione: (2026)
di: Han, Jian, et al.
Pubblicazione: (2026)
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
di: Sun, Peize, et al.
Pubblicazione: (2024)
di: Sun, Peize, et al.
Pubblicazione: (2024)
OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
di: Gan, Qijun, et al.
Pubblicazione: (2025)
di: Gan, Qijun, et al.
Pubblicazione: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
Do As I Do: Pose Guided Human Motion Copy
di: Wu, Sifan, et al.
Pubblicazione: (2024)
di: Wu, Sifan, et al.
Pubblicazione: (2024)
HLLM-Creator: Hierarchical LLM-based Personalized Creative Generation
di: Chen, Junyi, et al.
Pubblicazione: (2025)
di: Chen, Junyi, et al.
Pubblicazione: (2025)
VC-LLM: Automated Advertisement Video Creation from Raw Footage using Multi-modal LLMs
di: Qian, Dongjun, et al.
Pubblicazione: (2025)
di: Qian, Dongjun, et al.
Pubblicazione: (2025)
Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer
di: Shao, Ruizhi, et al.
Pubblicazione: (2024)
di: Shao, Ruizhi, et al.
Pubblicazione: (2024)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
FlashVideo: Flowing Fidelity to Detail for Efficient High-Resolution Video Generation
di: Zhang, Shilong, et al.
Pubblicazione: (2025)
di: Zhang, Shilong, et al.
Pubblicazione: (2025)
Validation of Human Pose Estimation and Human Mesh Recovery for Extracting Clinically Relevant Motion Data from Videos
di: Armstrong, Kai, et al.
Pubblicazione: (2025)
di: Armstrong, Kai, et al.
Pubblicazione: (2025)
Language-Guided Transformer Tokenizer for Human Motion Generation
di: Yan, Sheng, et al.
Pubblicazione: (2026)
di: Yan, Sheng, et al.
Pubblicazione: (2026)
Waver: Wave Your Way to Lifelike Video Generation
di: Zhang, Yifu, et al.
Pubblicazione: (2025)
di: Zhang, Yifu, et al.
Pubblicazione: (2025)
HyperDiff: Hypergraph Guided Diffusion Model for 3D Human Pose Estimation
di: Han, Bing, et al.
Pubblicazione: (2025)
di: Han, Bing, et al.
Pubblicazione: (2025)
Rethinking Generative Human Video Coding with Implicit Motion Transformation
di: Chen, Bolin, et al.
Pubblicazione: (2025)
di: Chen, Bolin, et al.
Pubblicazione: (2025)
VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers
di: Zheng, Jun, et al.
Pubblicazione: (2024)
di: Zheng, Jun, et al.
Pubblicazione: (2024)
DiHuR: Diffusion-Guided Generalizable Human Reconstruction
di: Chen, Jinnan, et al.
Pubblicazione: (2024)
di: Chen, Jinnan, et al.
Pubblicazione: (2024)
Mask$^2$DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation
di: Qi, Tianhao, et al.
Pubblicazione: (2025)
di: Qi, Tianhao, et al.
Pubblicazione: (2025)
$\text{Di}^2\text{Pose}$: Discrete Diffusion Model for Occluded 3D Human Pose Estimation
di: Wang, Weiquan, et al.
Pubblicazione: (2024)
di: Wang, Weiquan, et al.
Pubblicazione: (2024)
AnimaX: Animating the Inanimate in 3D with Joint Video-Pose Diffusion Models
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
UniTok: A Unified Tokenizer for Visual Generation and Understanding
di: Ma, Chuofan, et al.
Pubblicazione: (2025)
di: Ma, Chuofan, et al.
Pubblicazione: (2025)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
di: Han, Jian, et al.
Pubblicazione: (2024)
di: Han, Jian, et al.
Pubblicazione: (2024)
HumanScore: Benchmarking Human Motions in Generated Videos
di: Fang, Yusu, et al.
Pubblicazione: (2026)
di: Fang, Yusu, et al.
Pubblicazione: (2026)
Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton
di: Kang, Hongbo, et al.
Pubblicazione: (2024)
di: Kang, Hongbo, et al.
Pubblicazione: (2024)
Target Pose Guided Whole-body Grasping Motion Generation for Digital Humans
di: Shao, Quanquan, et al.
Pubblicazione: (2024)
di: Shao, Quanquan, et al.
Pubblicazione: (2024)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
di: Zhang, Yuang, et al.
Pubblicazione: (2024)
di: Zhang, Yuang, et al.
Pubblicazione: (2024)
FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers
di: He, Xuanhua, et al.
Pubblicazione: (2025)
di: He, Xuanhua, et al.
Pubblicazione: (2025)
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
di: Chen, Pengtao, et al.
Pubblicazione: (2025)
di: Chen, Pengtao, et al.
Pubblicazione: (2025)
PoseGen: In-Context LoRA Finetuning for Pose-Controllable Long Human Video Generation
di: He, Jingxuan, et al.
Pubblicazione: (2025)
di: He, Jingxuan, et al.
Pubblicazione: (2025)
SMooDi: Stylized Motion Diffusion Model
di: Zhong, Lei, et al.
Pubblicazione: (2024)
di: Zhong, Lei, et al.
Pubblicazione: (2024)
Kinematics Modeling Network for Video-based Human Pose Estimation
di: Dang, Yonghao, et al.
Pubblicazione: (2022)
di: Dang, Yonghao, et al.
Pubblicazione: (2022)
UnPose: Uncertainty-Guided Diffusion Priors for Zero-Shot Pose Estimation
di: Jiang, Zhaodong, et al.
Pubblicazione: (2025)
di: Jiang, Zhaodong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
InfinityHuman: Towards Long-Term Audio-Driven Human
di: Li, Xiaodi, et al.
Pubblicazione: (2025) -
PianoMotion10M: Dataset and Benchmark for Hand Motion Generation in Piano Performance
di: Gan, Qijun, et al.
Pubblicazione: (2024) -
HLLM: Enhancing Sequential Recommendations via Hierarchical Large Language Models for Item and User Modeling
di: Chen, Junyi, et al.
Pubblicazione: (2024) -
ALIVE: Animate Your World with Lifelike Audio-Video Generation
di: Guo, Ying, et al.
Pubblicazione: (2026) -
Fine-Grained Multi-View Hand Reconstruction Using Inverse Rendering
di: Gan, Qijun, et al.
Pubblicazione: (2024)