Gespeichert in:
| Hauptverfasser: | Wang, Lizhen, Xia, Zhurong, Hu, Tianshu, Wang, Pengrui, Wei, Pengfei, Zheng, Zerong, Zhou, Ming, Zhang, Yuan, Gao, Mingyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.10568 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
DreamVVT: Mastering Realistic Video Virtual Try-On in the Wild via a Stage-Wise Diffusion Transformer Framework
von: Zuo, Tongchun, et al.
Veröffentlicht: (2025)
von: Zuo, Tongchun, et al.
Veröffentlicht: (2025)
Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer
von: Shao, Ruizhi, et al.
Veröffentlicht: (2024)
von: Shao, Ruizhi, et al.
Veröffentlicht: (2024)
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
von: Liang, Chao, et al.
Veröffentlicht: (2025)
von: Liang, Chao, et al.
Veröffentlicht: (2025)
FlowAct-R1: Towards Interactive Humanoid Video Generation
von: Wang, Lizhen, et al.
Veröffentlicht: (2026)
von: Wang, Lizhen, et al.
Veröffentlicht: (2026)
Animatable and Relightable Gaussians for High-fidelity Human Avatar Modeling
von: Li, Zhe, et al.
Veröffentlicht: (2023)
von: Li, Zhe, et al.
Veröffentlicht: (2023)
Towards Imbalanced Motion: Part-Decoupling Network for Video Portrait Segmentation
von: Yu, Tianshu, et al.
Veröffentlicht: (2023)
von: Yu, Tianshu, et al.
Veröffentlicht: (2023)
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation
von: Jiang, Jianwen, et al.
Veröffentlicht: (2025)
von: Jiang, Jianwen, et al.
Veröffentlicht: (2025)
DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer
von: Guo, Xu, et al.
Veröffentlicht: (2026)
von: Guo, Xu, et al.
Veröffentlicht: (2026)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior
von: Huang, Junjia, et al.
Veröffentlicht: (2026)
von: Huang, Junjia, et al.
Veröffentlicht: (2026)
DreamVideo: High-Fidelity Image-to-Video Generation with Image Retention and Text Guidance
von: Wang, Cong, et al.
Veröffentlicht: (2023)
von: Wang, Cong, et al.
Veröffentlicht: (2023)
RMD: A Simple Baseline for More General Human Motion Generation via Training-free Retrieval-Augmented Motion Diffuse
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
Video Diffusion Transformers are In-Context Learners
von: Fei, Zhengcong, et al.
Veröffentlicht: (2024)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2024)
Structural Analysis of Phosphorus and Arsenic Clusters: A Comparative DFT and MP2 Study
von: Zerong Daniel Wang
Veröffentlicht: (2025)
von: Zerong Daniel Wang
Veröffentlicht: (2025)
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
DreamFuse: Adaptive Image Fusion with Diffusion Transformer
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
Make Your Actor Talk: Generalizable and High-Fidelity Lip Sync with Motion and Appearance Disentanglement
von: Yu, Runyi, et al.
Veröffentlicht: (2024)
von: Yu, Runyi, et al.
Veröffentlicht: (2024)
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
Demonstrating Record Fidelity for the Quantum Fourier Transform
von: Aumann, Philipp, et al.
Veröffentlicht: (2026)
von: Aumann, Philipp, et al.
Veröffentlicht: (2026)
Motion Diffusion Autoencoders: Enabling Attribute Manipulation in Human Motion Demonstrated on Karate Techniques
von: Richardson, Anthony, et al.
Veröffentlicht: (2025)
von: Richardson, Anthony, et al.
Veröffentlicht: (2025)
AI Spillover is Different: Flat and Lean Firms as Engines of AI Diffusion and Productivity Gain
von: Wang, Xiaoning, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoning, et al.
Veröffentlicht: (2025)
Ingredients: Blending Custom Photos with Video Diffusion Transformers
von: Fei, Zhengcong, et al.
Veröffentlicht: (2025)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2025)
DreamFoley: Scalable VLMs for High-Fidelity Video-to-Audio Generation
von: Li, Fu, et al.
Veröffentlicht: (2025)
von: Li, Fu, et al.
Veröffentlicht: (2025)
DreamText: High Fidelity Scene Text Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
Video Motion Transfer with Diffusion Transformers
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
An improved evolutionary structure optimization method considering stress minimization and smooth design
von: Leijia Wang, et al.
Veröffentlicht: (2024)
von: Leijia Wang, et al.
Veröffentlicht: (2024)
Semantics-Aware Human Motion Generation from Audio Instructions
von: Wang, Zi-An, et al.
Veröffentlicht: (2025)
von: Wang, Zi-An, et al.
Veröffentlicht: (2025)
Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations
von: Tang, Chao, et al.
Veröffentlicht: (2026)
von: Tang, Chao, et al.
Veröffentlicht: (2026)
MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation
von: Shi, Shuwei, et al.
Veröffentlicht: (2024)
von: Shi, Shuwei, et al.
Veröffentlicht: (2024)
MeshAvatar: Learning High-quality Triangular Human Avatars from Multi-view Videos
von: Chen, Yushuo, et al.
Veröffentlicht: (2024)
von: Chen, Yushuo, et al.
Veröffentlicht: (2024)
DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
Balancing Privacy and Efficiency: Music Information Retrieval via Additive Homomorphic Encryption
von: Wang, William Zerong, et al.
Veröffentlicht: (2025)
von: Wang, William Zerong, et al.
Veröffentlicht: (2025)
SkyReels-A1: Expressive Portrait Animation in Video Diffusion Transformers
von: Qiu, Di, et al.
Veröffentlicht: (2025)
von: Qiu, Di, et al.
Veröffentlicht: (2025)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025) -
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026) -
DreamVVT: Mastering Realistic Video Virtual Try-On in the Wild via a Stage-Wise Diffusion Transformer Framework
von: Zuo, Tongchun, et al.
Veröffentlicht: (2025) -
Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer
von: Shao, Ruizhi, et al.
Veröffentlicht: (2024) -
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
von: Liang, Chao, et al.
Veröffentlicht: (2025)