VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Chefer, Hila, Singer, Uriel, Zohar, Amit, Kirstain, Yuval, Polyak, Adam, Taigman, Yaniv, Wolf, Lior, Sheynin, Shelly |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Video Editing via Factorized Diffusion Distillation
di: Singer, Uriel, et al.
Pubblicazione: (2024)
di: Singer, Uriel, et al.
Pubblicazione: (2024)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
di: Yariv, Guy, et al.
Pubblicazione: (2025)
di: Yariv, Guy, et al.
Pubblicazione: (2025)
FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation
di: Shaulov, Ariel, et al.
Pubblicazione: (2025)
di: Shaulov, Ariel, et al.
Pubblicazione: (2025)
A Meaningful Perturbation Metric for Evaluating Explainability Methods
di: Cohen, Danielle, et al.
Pubblicazione: (2025)
di: Cohen, Danielle, et al.
Pubblicazione: (2025)
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
di: Bakish, Yarden, et al.
Pubblicazione: (2025)
di: Bakish, Yarden, et al.
Pubblicazione: (2025)
RealMaster: Lifting Rendered Scenes into Photorealistic Video
di: Cohen-Bar, Dana, et al.
Pubblicazione: (2026)
di: Cohen-Bar, Dana, et al.
Pubblicazione: (2026)
Still-Moving: Customized Video Generation without Customized Video Data
di: Chefer, Hila, et al.
Pubblicazione: (2024)
di: Chefer, Hila, et al.
Pubblicazione: (2024)
JointTuner: Appearance-Motion Adaptive Joint Training for Customized Video Generation
di: Chen, Fangda, et al.
Pubblicazione: (2025)
di: Chen, Fangda, et al.
Pubblicazione: (2025)
CourtMotion: Learning Event-Driven Motion Representations from Skeletal Data for Basketball
di: Sela, Omer, et al.
Pubblicazione: (2025)
di: Sela, Omer, et al.
Pubblicazione: (2025)
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching
di: Kwon, Mingi, et al.
Pubblicazione: (2025)
di: Kwon, Mingi, et al.
Pubblicazione: (2025)
Motion by Queries: Identity-Motion Trade-offs in Text-to-Video Generation
di: Atzmon, Yuval, et al.
Pubblicazione: (2024)
di: Atzmon, Yuval, et al.
Pubblicazione: (2024)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
di: Zhai, Yuanhao, et al.
Pubblicazione: (2024)
di: Zhai, Yuanhao, et al.
Pubblicazione: (2024)
Discriminative Class Tokens for Text-to-Image Diffusion Models
di: Schwartz, Idan, et al.
Pubblicazione: (2023)
di: Schwartz, Idan, et al.
Pubblicazione: (2023)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
di: Sheng, Mingzhi, et al.
Pubblicazione: (2026)
di: Sheng, Mingzhi, et al.
Pubblicazione: (2026)
IlluSign: Illustrating Sign Language Videos by Leveraging the Attention Mechanism
di: Bruner, Janna, et al.
Pubblicazione: (2025)
di: Bruner, Janna, et al.
Pubblicazione: (2025)
SphereUFormer: A U-Shaped Transformer for Spherical 360 Perception
di: Benny, Yaniv, et al.
Pubblicazione: (2024)
di: Benny, Yaniv, et al.
Pubblicazione: (2024)
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
di: Kong, Zhe, et al.
Pubblicazione: (2025)
di: Kong, Zhe, et al.
Pubblicazione: (2025)
TokenTrim: Inference-Time Token Pruning for Autoregressive Long Video Generation
di: Shaulov, Ariel, et al.
Pubblicazione: (2026)
di: Shaulov, Ariel, et al.
Pubblicazione: (2026)
VideoSPatS: Video SPatiotemporal Splines for Disentangled Occlusion, Appearance and Motion Modeling and Editing
di: Bello, Juan Luis Gonzalez, et al.
Pubblicazione: (2025)
di: Bello, Juan Luis Gonzalez, et al.
Pubblicazione: (2025)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
di: Liu, Huijie, et al.
Pubblicazione: (2025)
di: Liu, Huijie, et al.
Pubblicazione: (2025)
Compositional Video Generation via Inference-Time Guidance
di: Shaulov, Ariel, et al.
Pubblicazione: (2026)
di: Shaulov, Ariel, et al.
Pubblicazione: (2026)
MACE-Dance: Motion-Appearance Cascaded Experts for Music-Driven Dance Video Generation
di: Yang, Kaixing, et al.
Pubblicazione: (2025)
di: Yang, Kaixing, et al.
Pubblicazione: (2025)
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
di: Zhang, Zhenghao, et al.
Pubblicazione: (2025)
di: Zhang, Zhenghao, et al.
Pubblicazione: (2025)
PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
di: Gao, Mingju, et al.
Pubblicazione: (2026)
di: Gao, Mingju, et al.
Pubblicazione: (2026)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
di: Bi, Xiuli, et al.
Pubblicazione: (2024)
di: Bi, Xiuli, et al.
Pubblicazione: (2024)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
di: Zhao, Shuling, et al.
Pubblicazione: (2024)
di: Zhao, Shuling, et al.
Pubblicazione: (2024)
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
di: Zhou, Hang, et al.
Pubblicazione: (2024)
di: Zhou, Hang, et al.
Pubblicazione: (2024)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
MOT FCG++: Enhanced Representation of Spatio-temporal Motion and Appearance Features
di: Fang, Yanzhao
Pubblicazione: (2024)
di: Fang, Yanzhao
Pubblicazione: (2024)
LocoMotion: Learning Motion-Focused Video-Language Representations
di: Doughty, Hazel, et al.
Pubblicazione: (2024)
di: Doughty, Hazel, et al.
Pubblicazione: (2024)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
di: Zhan, Yu-Wei, et al.
Pubblicazione: (2025)
di: Zhan, Yu-Wei, et al.
Pubblicazione: (2025)
A Self-supervised Motion Representation for Portrait Video Generation
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
SMRABooth: Subject and Motion Representation Alignment for Customized Video Generation
di: Xu, Xuancheng, et al.
Pubblicazione: (2025)
di: Xu, Xuancheng, et al.
Pubblicazione: (2025)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
di: Singer, Assaf, et al.
Pubblicazione: (2025)
di: Singer, Assaf, et al.
Pubblicazione: (2025)
Motion Prompting: Controlling Video Generation with Motion Trajectories
di: Geng, Daniel, et al.
Pubblicazione: (2024)
di: Geng, Daniel, et al.
Pubblicazione: (2024)
UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
di: Bai, Jianhong, et al.
Pubblicazione: (2024)
di: Bai, Jianhong, et al.
Pubblicazione: (2024)
Motion Control for Enhanced Complex Action Video Generation
di: Zhou, Qiang, et al.
Pubblicazione: (2024)
di: Zhou, Qiang, et al.
Pubblicazione: (2024)
Motion Attribution for Video Generation
di: Wu, Xindi, et al.
Pubblicazione: (2026)
di: Wu, Xindi, et al.
Pubblicazione: (2026)
Joint-Motion Mutual Learning for Pose Estimation in Videos
di: Wu, Sifan, et al.
Pubblicazione: (2024)
di: Wu, Sifan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Video Editing via Factorized Diffusion Distillation
di: Singer, Uriel, et al.
Pubblicazione: (2024) -
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
di: Yariv, Guy, et al.
Pubblicazione: (2025) -
FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation
di: Shaulov, Ariel, et al.
Pubblicazione: (2025) -
A Meaningful Perturbation Metric for Evaluating Explainability Methods
di: Cohen, Danielle, et al.
Pubblicazione: (2025) -
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
di: Bakish, Yarden, et al.
Pubblicazione: (2025)