SyncVP: Joint Diffusion for Synchronous Multi-Modal Video Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Pallotta, Enrico, Azar, Sina Mokhtarzadeh, Li, Shuai, Zatsarynna, Olga, Gall, Juergen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sequence-Adaptive Video Prediction in Continuous Streams using Diffusion Noise Optimization
di: Azar, Sina Mokhtarzadeh, et al.
Pubblicazione: (2025)
di: Azar, Sina Mokhtarzadeh, et al.
Pubblicazione: (2025)
EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses
di: Pallotta, Enrico, et al.
Pubblicazione: (2025)
di: Pallotta, Enrico, et al.
Pubblicazione: (2025)
CamC2V: Context-aware Controllable Video Generation
di: Denninger, Luis, et al.
Pubblicazione: (2025)
di: Denninger, Luis, et al.
Pubblicazione: (2025)
Gated Temporal Diffusion for Stochastic Long-Term Dense Anticipation
di: Zatsarynna, Olga, et al.
Pubblicazione: (2024)
di: Zatsarynna, Olga, et al.
Pubblicazione: (2024)
MANTA: Diffusion Mamba for Efficient and Effective Stochastic Long-Term Dense Anticipation
di: Zatsarynna, Olga, et al.
Pubblicazione: (2025)
di: Zatsarynna, Olga, et al.
Pubblicazione: (2025)
STRIVE: Structured Spatiotemporal Exploration for Reinforcement Learning in Video Question Answering
di: Bahrami, Emad, et al.
Pubblicazione: (2026)
di: Bahrami, Emad, et al.
Pubblicazione: (2026)
Towards Generalizing Temporal Action Segmentation to Unseen Views
di: Bahrami, Emad, et al.
Pubblicazione: (2025)
di: Bahrami, Emad, et al.
Pubblicazione: (2025)
Privacy-Preserving Semantic Segmentation from Ultra-Low-Resolution RGB Inputs
di: Huang, Xuying, et al.
Pubblicazione: (2025)
di: Huang, Xuying, et al.
Pubblicazione: (2025)
MixANT: Observation-dependent Memory Propagation for Stochastic Dense Action Anticipation
di: Wasim, Syed Talal, et al.
Pubblicazione: (2025)
di: Wasim, Syed Talal, et al.
Pubblicazione: (2025)
Looking into the Unknown: Exploring Action Discovery for Segmentation of Known and Unknown Actions
di: Spurio, Federico, et al.
Pubblicazione: (2025)
di: Spurio, Federico, et al.
Pubblicazione: (2025)
SyncMV4D: Synchronized Multi-view Joint Diffusion of Appearance and Motion for Hand-Object Interaction Synthesis
di: Dang, Lingwei, et al.
Pubblicazione: (2025)
di: Dang, Lingwei, et al.
Pubblicazione: (2025)
SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning
di: Cheng, Xin, et al.
Pubblicazione: (2026)
di: Cheng, Xin, et al.
Pubblicazione: (2026)
SyncVIS: Synchronized Video Instance Segmentation
di: Zheng, Rongkun, et al.
Pubblicazione: (2024)
di: Zheng, Rongkun, et al.
Pubblicazione: (2024)
Learning a Neural Association Network for Self-supervised Multi-Object Tracking
di: Li, Shuai, et al.
Pubblicazione: (2024)
di: Li, Shuai, et al.
Pubblicazione: (2024)
DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
di: Dong, Yue-Jiang, et al.
Pubblicazione: (2025)
di: Dong, Yue-Jiang, et al.
Pubblicazione: (2025)
Video Panels for Long Video Understanding
di: Doorenbos, Lars, et al.
Pubblicazione: (2025)
di: Doorenbos, Lars, et al.
Pubblicazione: (2025)
OmniSync: Towards Universal Lip Synchronization via Diffusion Transformers
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
SyncTweedies: A General Generative Framework Based on Synchronized Diffusions
di: Kim, Jaihoon, et al.
Pubblicazione: (2024)
di: Kim, Jaihoon, et al.
Pubblicazione: (2024)
SyncTrack4D: Cross-Video Motion Alignment and Video Synchronization for Multi-Video 4D Gaussian Splatting
di: Lee, Yonghan, et al.
Pubblicazione: (2025)
di: Lee, Yonghan, et al.
Pubblicazione: (2025)
SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization
di: Li, Deming, et al.
Pubblicazione: (2026)
di: Li, Deming, et al.
Pubblicazione: (2026)
HighSync: High-Quality Lip Synchronization via Latent Diffusion Models
di: Daghigh, Saeed Firouzi, et al.
Pubblicazione: (2026)
di: Daghigh, Saeed Firouzi, et al.
Pubblicazione: (2026)
PhysioSync: Temporal and Cross-Modal Contrastive Learning Inspired by Physiological Synchronization for EEG-Based Emotion Recognition
di: Cui, Kai, et al.
Pubblicazione: (2025)
di: Cui, Kai, et al.
Pubblicazione: (2025)
CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing
di: Cong, Gaoxiang, et al.
Pubblicazione: (2026)
di: Cong, Gaoxiang, et al.
Pubblicazione: (2026)
SyncSDE: A Probabilistic Framework for Diffusion Synchronization
di: Lee, Hyunjun, et al.
Pubblicazione: (2025)
di: Lee, Hyunjun, et al.
Pubblicazione: (2025)
ADA-Track++: End-to-End Multi-Camera 3D Multi-Object Tracking with Alternating Detection and Association
di: Ding, Shuxiao, et al.
Pubblicazione: (2024)
di: Ding, Shuxiao, et al.
Pubblicazione: (2024)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
di: Li, Chunyu, et al.
Pubblicazione: (2024)
di: Li, Chunyu, et al.
Pubblicazione: (2024)
StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
StableMamba: Distillation-free Scaling of Large SSMs for Images and Videos
di: Suleman, Hamid, et al.
Pubblicazione: (2024)
di: Suleman, Hamid, et al.
Pubblicazione: (2024)
FlowNar: Scalable Streaming Narration for Long-Form Videos
di: Zhong, Zeyun, et al.
Pubblicazione: (2026)
di: Zhong, Zeyun, et al.
Pubblicazione: (2026)
Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis
di: He, Wenkun, et al.
Pubblicazione: (2024)
di: He, Wenkun, et al.
Pubblicazione: (2024)
LC-SLab -- An Object-based Deep Learning Framework for Large-scale Land Cover Classification from Satellite Imagery and Sparse In-situ Labels
di: Leonhardt, Johannes, et al.
Pubblicazione: (2025)
di: Leonhardt, Johannes, et al.
Pubblicazione: (2025)
Identifying Spatio-Temporal Drivers of Extreme Events
di: Eddin, Mohamad Hakam Shams, et al.
Pubblicazione: (2024)
di: Eddin, Mohamad Hakam Shams, et al.
Pubblicazione: (2024)
Enhancing Video-Based Robot Failure Detection Using Task Knowledge
di: Thoduka, Santosh, et al.
Pubblicazione: (2025)
di: Thoduka, Santosh, et al.
Pubblicazione: (2025)
RocSync: Millisecond-Accurate Temporal Synchronization for Heterogeneous Camera Systems
di: Meyer, Jaro, et al.
Pubblicazione: (2025)
di: Meyer, Jaro, et al.
Pubblicazione: (2025)
MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation
di: Li, Liyang, et al.
Pubblicazione: (2026)
di: Li, Liyang, et al.
Pubblicazione: (2026)
Global-Aware Monocular Semantic Scene Completion with State Space Models
di: Li, Shijie, et al.
Pubblicazione: (2025)
di: Li, Shijie, et al.
Pubblicazione: (2025)
UniSync: Towards Generalizable and High-Fidelity Lip Synchronization for Challenging Scenarios
di: Fan, Ruidi, et al.
Pubblicazione: (2026)
di: Fan, Ruidi, et al.
Pubblicazione: (2026)
MV-Match: Multi-View Matching for Domain-Adaptive Identification of Plant Nutrient Deficiencies
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Sequence-Adaptive Video Prediction in Continuous Streams using Diffusion Noise Optimization
di: Azar, Sina Mokhtarzadeh, et al.
Pubblicazione: (2025) -
EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses
di: Pallotta, Enrico, et al.
Pubblicazione: (2025) -
CamC2V: Context-aware Controllable Video Generation
di: Denninger, Luis, et al.
Pubblicazione: (2025) -
Gated Temporal Diffusion for Stochastic Long-Term Dense Anticipation
di: Zatsarynna, Olga, et al.
Pubblicazione: (2024) -
MANTA: Diffusion Mamba for Efficient and Effective Stochastic Long-Term Dense Anticipation
di: Zatsarynna, Olga, et al.
Pubblicazione: (2025)