Boximator: Generating Rich and Controllable Motions for Video Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jiawei, Zhang, Yuchen, Zou, Jiaxin, Zeng, Yan, Wei, Guoqiang, Yuan, Liping, Li, Hang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding
von: Yuan, Liping, et al.
Veröffentlicht: (2025)
von: Yuan, Liping, et al.
Veröffentlicht: (2025)
Image Conductor: Precision Control for Interactive Video Synthesis
von: Li, Yaowei, et al.
Veröffentlicht: (2024)
von: Li, Yaowei, et al.
Veröffentlicht: (2024)
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023)
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023)
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
von: Yan, Xin, et al.
Veröffentlicht: (2024)
von: Yan, Xin, et al.
Veröffentlicht: (2024)
TokenMotion: Decoupled Motion Control via Token Disentanglement for Human-centric Video Generation
von: Li, Ruineng, et al.
Veröffentlicht: (2025)
von: Li, Ruineng, et al.
Veröffentlicht: (2025)
Quantitative Video World Model Evaluation for Geometric-Consistency
von: Wu, Jiaxin, et al.
Veröffentlicht: (2026)
von: Wu, Jiaxin, et al.
Veröffentlicht: (2026)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
Representation Learning for Compressed Video Action Recognition via Attentive Cross-modal Interaction with Motion Enhancement
von: Li, Bing, et al.
Veröffentlicht: (2022)
von: Li, Bing, et al.
Veröffentlicht: (2022)
Resource-Efficient Motion Control for Video Generation via Dynamic Mask Guidance
von: Feng, Sicong, et al.
Veröffentlicht: (2025)
von: Feng, Sicong, et al.
Veröffentlicht: (2025)
Conditional Video Generation for High-Efficiency Video Compression
von: Yi, Fangqiu, et al.
Veröffentlicht: (2025)
von: Yi, Fangqiu, et al.
Veröffentlicht: (2025)
I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
von: Feng, Wanquan, et al.
Veröffentlicht: (2024)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
GSV3D: Gaussian Splatting-based Geometric Distillation with Stable Video Diffusion for Single-Image 3D Object Generation
von: Tao, Ye, et al.
Veröffentlicht: (2025)
von: Tao, Ye, et al.
Veröffentlicht: (2025)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance
von: Zeng, Ziyun, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2026)
Motion-Aware Caching for Efficient Autoregressive Video Generation
von: Xu, Jing, et al.
Veröffentlicht: (2026)
von: Xu, Jing, et al.
Veröffentlicht: (2026)
Rethinking Video Generation Model for the Embodied World
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
FAVOR-Bench: A Comprehensive Benchmark for Fine-Grained Video Motion Understanding
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
von: Chen, Zhifei, et al.
Veröffentlicht: (2025)
von: Chen, Zhifei, et al.
Veröffentlicht: (2025)
Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
MEPG:Multi-Expert Planning and Generation for Compositionally-Rich Image Generation
von: Zhao, Yuan, et al.
Veröffentlicht: (2025)
von: Zhao, Yuan, et al.
Veröffentlicht: (2025)
Boosting Camera Motion Control for Video Diffusion Transformers
von: Cheong, Soon Yau, et al.
Veröffentlicht: (2024)
von: Cheong, Soon Yau, et al.
Veröffentlicht: (2024)
TrajMamba: An Ego-Motion-Guided Mamba Model for Pedestrian Trajectory Prediction from an Egocentric Perspective
von: Peng, Yusheng, et al.
Veröffentlicht: (2026)
von: Peng, Yusheng, et al.
Veröffentlicht: (2026)
Learn the Force We Can: Enabling Sparse Motion Control in Multi-Object Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2023)
von: Davtyan, Aram, et al.
Veröffentlicht: (2023)
DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers
von: Wang, Lizhen, et al.
Veröffentlicht: (2025)
von: Wang, Lizhen, et al.
Veröffentlicht: (2025)
DEMO: Disentangled Motion Latent Flow Matching for Fine-Grained Controllable Talking Portrait Synthesis
von: Chen, Peiyin, et al.
Veröffentlicht: (2025)
von: Chen, Peiyin, et al.
Veröffentlicht: (2025)
Improved Belief-Attention in Vision Task
von: Zhang, Guoqiang
Veröffentlicht: (2026)
von: Zhang, Guoqiang
Veröffentlicht: (2026)
MambaOVSR: Multiscale Fusion with Global Motion Modeling for Chinese Opera Video Super-Resolution
von: Chang, Hua, et al.
Veröffentlicht: (2025)
von: Chang, Hua, et al.
Veröffentlicht: (2025)
PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation
von: Huang, Yidong, et al.
Veröffentlicht: (2026)
von: Huang, Yidong, et al.
Veröffentlicht: (2026)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
von: Li, Quanhao, et al.
Veröffentlicht: (2025)
von: Li, Quanhao, et al.
Veröffentlicht: (2025)
Video Text Preservation with Synthetic Text-Rich Videos
von: Liu, Ziyang, et al.
Veröffentlicht: (2025)
von: Liu, Ziyang, et al.
Veröffentlicht: (2025)
MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators
von: Zhang, Yaqi, et al.
Veröffentlicht: (2023)
von: Zhang, Yaqi, et al.
Veröffentlicht: (2023)
Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance
von: Lin, Yiqi, et al.
Veröffentlicht: (2026)
von: Lin, Yiqi, et al.
Veröffentlicht: (2026)
Physics-Guided Motion Loss for Video Generation Model
von: Xue, Bowen, et al.
Veröffentlicht: (2025)
von: Xue, Bowen, et al.
Veröffentlicht: (2025)
Ctrl-VI: Controllable Video Synthesis via Variational Inference
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
Efficient Training for Human Video Generation with Entropy-Guided Prioritized Progressive Learning
von: Li, Changlin, et al.
Veröffentlicht: (2025)
von: Li, Changlin, et al.
Veröffentlicht: (2025)
Video-As-Prompt: Unified Semantic Control for Video Generation
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation
von: Wang, Weimin, et al.
Veröffentlicht: (2024)
von: Wang, Weimin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding
von: Yuan, Liping, et al.
Veröffentlicht: (2025) -
Image Conductor: Precision Control for Interactive Video Synthesis
von: Li, Yaowei, et al.
Veröffentlicht: (2024) -
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
von: Shi, Fengyuan, et al.
Veröffentlicht: (2023) -
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025) -
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
von: Yan, Xin, et al.
Veröffentlicht: (2024)