Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yariv, Guy, Kirstain, Yuval, Zohar, Amit, Sheynin, Shelly, Taigman, Yaniv, Adi, Yossi, Benaim, Sagie, Polyak, Adam |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
di: Chefer, Hila, et al.
Pubblicazione: (2025)
di: Chefer, Hila, et al.
Pubblicazione: (2025)
Video Editing via Factorized Diffusion Distillation
di: Singer, Uriel, et al.
Pubblicazione: (2024)
di: Singer, Uriel, et al.
Pubblicazione: (2024)
LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
di: Yariv, Guy, et al.
Pubblicazione: (2024)
di: Yariv, Guy, et al.
Pubblicazione: (2024)
DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion
di: Issachar, Noam, et al.
Pubblicazione: (2025)
di: Issachar, Noam, et al.
Pubblicazione: (2025)
RewardSDS: Aligning Score Distillation via Reward-Weighted Sampling
di: Chachy, Itay, et al.
Pubblicazione: (2025)
di: Chachy, Itay, et al.
Pubblicazione: (2025)
Generating Intermediate Representations for Compositional Text-To-Image Generation
di: Galun, Ran, et al.
Pubblicazione: (2024)
di: Galun, Ran, et al.
Pubblicazione: (2024)
RealMaster: Lifting Rendered Scenes into Photorealistic Video
di: Cohen-Bar, Dana, et al.
Pubblicazione: (2026)
di: Cohen-Bar, Dana, et al.
Pubblicazione: (2026)
PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions
di: Benishu, Omer, et al.
Pubblicazione: (2026)
di: Benishu, Omer, et al.
Pubblicazione: (2026)
SemanticMoments: Training-Free Motion Similarity via Third Moment Features
di: Huberman, Saar, et al.
Pubblicazione: (2026)
di: Huberman, Saar, et al.
Pubblicazione: (2026)
Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillation
di: Shavin, David, et al.
Pubblicazione: (2026)
di: Shavin, David, et al.
Pubblicazione: (2026)
RAD: Retrieval-Augmented Monocular Metric Depth Estimation for Underrepresented Classes
di: Baltaxe, Michael, et al.
Pubblicazione: (2026)
di: Baltaxe, Michael, et al.
Pubblicazione: (2026)
Colored Noise Diffusion Sampling
di: Davidson, Hadar, et al.
Pubblicazione: (2026)
di: Davidson, Hadar, et al.
Pubblicazione: (2026)
Mask2IV: Interaction-Centric Video Generation via Mask Trajectories
di: Li, Gen, et al.
Pubblicazione: (2025)
di: Li, Gen, et al.
Pubblicazione: (2025)
MV-RAG: Retrieval Augmented Multiview Diffusion
di: Dayani, Yosef, et al.
Pubblicazione: (2025)
di: Dayani, Yosef, et al.
Pubblicazione: (2025)
Designing a Conditional Prior Distribution for Flow-Based Generative Models
di: Issachar, Noam, et al.
Pubblicazione: (2025)
di: Issachar, Noam, et al.
Pubblicazione: (2025)
Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score Distillation
di: Fiebelman, Gal, et al.
Pubblicazione: (2025)
di: Fiebelman, Gal, et al.
Pubblicazione: (2025)
ReMoMask: Retrieval-Augmented Masked Motion Generation
di: Li, Zhengdao, et al.
Pubblicazione: (2025)
di: Li, Zhengdao, et al.
Pubblicazione: (2025)
DGD: Dynamic 3D Gaussians Distillation
di: Labe, Isaac, et al.
Pubblicazione: (2024)
di: Labe, Isaac, et al.
Pubblicazione: (2024)
Structurally Disentangled Feature Fields Distillation for 3D Understanding and Editing
di: Levy, Yoel, et al.
Pubblicazione: (2025)
di: Levy, Yoel, et al.
Pubblicazione: (2025)
Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes
di: Krakovsky, Shai, et al.
Pubblicazione: (2025)
di: Krakovsky, Shai, et al.
Pubblicazione: (2025)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
di: Zhang, Xiangyue, et al.
Pubblicazione: (2025)
di: Zhang, Xiangyue, et al.
Pubblicazione: (2025)
Edit as You See: Image-guided Video Editing via Masked Motion Modeling
di: Huang, Zhi-Lin, et al.
Pubblicazione: (2025)
di: Huang, Zhi-Lin, et al.
Pubblicazione: (2025)
Mask$^2$DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation
di: Qi, Tianhao, et al.
Pubblicazione: (2025)
di: Qi, Tianhao, et al.
Pubblicazione: (2025)
MMM: Generative Masked Motion Model
di: Pinyoanuntapong, Ekkasit, et al.
Pubblicazione: (2023)
di: Pinyoanuntapong, Ekkasit, et al.
Pubblicazione: (2023)
MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
di: Pinyoanuntapong, Ekkasit, et al.
Pubblicazione: (2024)
di: Pinyoanuntapong, Ekkasit, et al.
Pubblicazione: (2024)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
di: Li, Zhe, et al.
Pubblicazione: (2025)
di: Li, Zhe, et al.
Pubblicazione: (2025)
SMILE: Infusing Spatial and Motion Semantics in Masked Video Learning
di: Thoker, Fida Mohammad, et al.
Pubblicazione: (2025)
di: Thoker, Fida Mohammad, et al.
Pubblicazione: (2025)
Resource-Efficient Motion Control for Video Generation via Dynamic Mask Guidance
di: Feng, Sicong, et al.
Pubblicazione: (2025)
di: Feng, Sicong, et al.
Pubblicazione: (2025)
Motion Guided Token Compression for Efficient Masked Video Modeling
di: Feng, Yukun, et al.
Pubblicazione: (2024)
di: Feng, Yukun, et al.
Pubblicazione: (2024)
Text-driven Human Motion Generation with Motion Masked Diffusion Model
di: Chen, Xingyu
Pubblicazione: (2024)
di: Chen, Xingyu
Pubblicazione: (2024)
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
di: Luo, Yifu, et al.
Pubblicazione: (2025)
di: Luo, Yifu, et al.
Pubblicazione: (2025)
Spectral Progressive Diffusion for Efficient Image and Video Generation
di: Xiao, Howard, et al.
Pubblicazione: (2026)
di: Xiao, Howard, et al.
Pubblicazione: (2026)
MaskFocus: Focusing Policy Optimization on Critical Steps for Masked Image Generation
di: Zhang, Guohui, et al.
Pubblicazione: (2025)
di: Zhang, Guohui, et al.
Pubblicazione: (2025)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
di: Chen, Wenchao, et al.
Pubblicazione: (2024)
di: Chen, Wenchao, et al.
Pubblicazione: (2024)
Mask-ControlNet: Higher-Quality Image Generation with An Additional Mask Prompt
di: Huang, Zhiqi, et al.
Pubblicazione: (2024)
di: Huang, Zhiqi, et al.
Pubblicazione: (2024)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
di: Wang, Siyuan, et al.
Pubblicazione: (2025)
di: Wang, Siyuan, et al.
Pubblicazione: (2025)
Discriminative Class Tokens for Text-to-Image Diffusion Models
di: Schwartz, Idan, et al.
Pubblicazione: (2023)
di: Schwartz, Idan, et al.
Pubblicazione: (2023)
Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation
di: Chao, Brian, et al.
Pubblicazione: (2026)
di: Chao, Brian, et al.
Pubblicazione: (2026)
KMM: Key Frame Mask Mamba for Extended Motion Generation
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
HU-based Foreground Masking for 3D Medical Masked Image Modeling
di: Lee, Jin, et al.
Pubblicazione: (2025)
di: Lee, Jin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
di: Chefer, Hila, et al.
Pubblicazione: (2025) -
Video Editing via Factorized Diffusion Distillation
di: Singer, Uriel, et al.
Pubblicazione: (2024) -
LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
di: Yariv, Guy, et al.
Pubblicazione: (2024) -
DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion
di: Issachar, Noam, et al.
Pubblicazione: (2025) -
RewardSDS: Aligning Score Distillation via Reward-Weighted Sampling
di: Chachy, Itay, et al.
Pubblicazione: (2025)