PLACID: Identity-Preserving Multi-Object Compositing via Video Diffusion with Synthetic Trajectories
Fuente:
arXiv
Salvato in:
| Autori principali: | Tarrés, Gemma Canet, Baradad, Manel, Moreno-Noguer, Francesc, Li, Yumeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Thinking Outside the BBox: Unconstrained Generative Object Compositing
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2024)
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2024)
GLVD: Guided Learned Vertex Descent
di: Rico, Pol Caselles, et al.
Pubblicazione: (2025)
di: Rico, Pol Caselles, et al.
Pubblicazione: (2025)
Separating Knowledge and Perception with Procedural Data
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2025)
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2025)
Cost Savings from Automatic Quality Assessment of Generated Images
di: Giro-i-Nieto, Xavier, et al.
Pubblicazione: (2025)
di: Giro-i-Nieto, Xavier, et al.
Pubblicazione: (2025)
Multitwine: Multi-Object Compositing with Text and Layout Control
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2025)
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2025)
PARASOL: Parametric Style Control for Diffusion Image Synthesis
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2023)
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2023)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
Video Text Preservation with Synthetic Text-Rich Videos
di: Liu, Ziyang, et al.
Pubblicazione: (2025)
di: Liu, Ziyang, et al.
Pubblicazione: (2025)
Concat-ID: Towards Universal Identity-Preserving Video Synthesis
di: Zhong, Yong, et al.
Pubblicazione: (2025)
di: Zhong, Yong, et al.
Pubblicazione: (2025)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
di: YU, Mark, et al.
Pubblicazione: (2025)
di: YU, Mark, et al.
Pubblicazione: (2025)
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
Magic-Me: Identity-Specific Video Customized Diffusion
di: Ma, Ze, et al.
Pubblicazione: (2024)
di: Ma, Ze, et al.
Pubblicazione: (2024)
Synthetic Object Compositions for Scalable and Accurate Learning in Detection, Segmentation, and Grounding
di: Huang, Weikai, et al.
Pubblicazione: (2025)
di: Huang, Weikai, et al.
Pubblicazione: (2025)
TranSPORTmer: A Holistic Approach to Trajectory Understanding in Multi-Agent Sports
di: Capellera, Guillem, et al.
Pubblicazione: (2024)
di: Capellera, Guillem, et al.
Pubblicazione: (2024)
MAVIN: Multi-Action Video Generation with Diffusion Models via Transition Video Infilling
di: Zhang, Bowen, et al.
Pubblicazione: (2024)
di: Zhang, Bowen, et al.
Pubblicazione: (2024)
Diffusion Curriculum: Synthetic-to-Real Data Curriculum via Image-Guided Diffusion
di: Liang, Yijun, et al.
Pubblicazione: (2024)
di: Liang, Yijun, et al.
Pubblicazione: (2024)
InstantID: Zero-shot Identity-Preserving Generation in Seconds
di: Wang, Qixun, et al.
Pubblicazione: (2024)
di: Wang, Qixun, et al.
Pubblicazione: (2024)
Video Diffusion Models Excel at Tracking Similar-Looking Objects Without Supervision
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
STORM: End-to-End Referring Multi-Object Tracking in Videos
di: Lu, Zijia, et al.
Pubblicazione: (2026)
di: Lu, Zijia, et al.
Pubblicazione: (2026)
Diffusion as Reasoning: Enhancing Object Navigation via Diffusion Model Conditioned on LLM-based Object-Room Knowledge
di: Ji, Yiming, et al.
Pubblicazione: (2024)
di: Ji, Yiming, et al.
Pubblicazione: (2024)
Controllable Video Object Insertion via Multiview Priors
di: Qi, Xia, et al.
Pubblicazione: (2026)
di: Qi, Xia, et al.
Pubblicazione: (2026)
Reinforcement-Guided Synthetic Data Generation for Privacy-Sensitive Identity Recognition
di: Jia, Xuemei, et al.
Pubblicazione: (2026)
di: Jia, Xuemei, et al.
Pubblicazione: (2026)
M$^3$-VOS: Multi-Phase, Multi-Transition, and Multi-Scenery Video Object Segmentation
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving
di: Huang, Jiehui, et al.
Pubblicazione: (2024)
di: Huang, Jiehui, et al.
Pubblicazione: (2024)
StableAnimator: High-Quality Identity-Preserving Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
di: Qi, Carl, et al.
Pubblicazione: (2024)
di: Qi, Carl, et al.
Pubblicazione: (2024)
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
di: Kienzle, Daniel, et al.
Pubblicazione: (2025)
di: Kienzle, Daniel, et al.
Pubblicazione: (2025)
Zero-shot Synthetic Video Realism Enhancement via Structure-aware Denoising
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
Video-GPT via Next Clip Diffusion
di: Zhuang, Shaobin, et al.
Pubblicazione: (2025)
di: Zhuang, Shaobin, et al.
Pubblicazione: (2025)
MAUGen: A Unified Diffusion Approach for Multi-Identity Facial Expression and AU Label Generation
di: Li, Xiangdong, et al.
Pubblicazione: (2026)
di: Li, Xiangdong, et al.
Pubblicazione: (2026)
TrackDiffusion: Tracklet-Conditioned Video Generation via Diffusion Models
di: Li, Pengxiang, et al.
Pubblicazione: (2023)
di: Li, Pengxiang, et al.
Pubblicazione: (2023)
When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation
di: Zheng, Dongqi
Pubblicazione: (2026)
di: Zheng, Dongqi
Pubblicazione: (2026)
Beyond the Pixels: VLM-based Evaluation of Identity Preservation in Reference-Guided Synthesis
di: Singhania, Aditi, et al.
Pubblicazione: (2025)
di: Singhania, Aditi, et al.
Pubblicazione: (2025)
Letting Trajectories Spread: Quality-Preserving Control for Diverse Flow Matching
di: Wu, Jingxuan, et al.
Pubblicazione: (2025)
di: Wu, Jingxuan, et al.
Pubblicazione: (2025)
Trajectory-Aware Adaptive Inference in Object Detection Models
di: Papanikolaou, Grigorios, et al.
Pubblicazione: (2026)
di: Papanikolaou, Grigorios, et al.
Pubblicazione: (2026)
DP-MDM: Detail-Preserving MR Reconstruction via Multiple Diffusion Models
di: Geng, Mengxiao, et al.
Pubblicazione: (2024)
di: Geng, Mengxiao, et al.
Pubblicazione: (2024)
InstantGeoAvatar: Effective Geometry and Appearance Modeling of Animatable Avatars from Monocular Video
di: Budria, Alvaro, et al.
Pubblicazione: (2024)
di: Budria, Alvaro, et al.
Pubblicazione: (2024)
Multi-identity Human Image Animation with Structural Video Diffusion
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
Advancing Complex Video Object Segmentation via Progressive Concept Construction
di: Zhang, Zhixiong, et al.
Pubblicazione: (2025)
di: Zhang, Zhixiong, et al.
Pubblicazione: (2025)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
di: Xie, Kevin, et al.
Pubblicazione: (2025)
di: Xie, Kevin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Thinking Outside the BBox: Unconstrained Generative Object Compositing
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2024) -
GLVD: Guided Learned Vertex Descent
di: Rico, Pol Caselles, et al.
Pubblicazione: (2025) -
Separating Knowledge and Perception with Procedural Data
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2025) -
Cost Savings from Automatic Quality Assessment of Generated Images
di: Giro-i-Nieto, Xavier, et al.
Pubblicazione: (2025) -
Multitwine: Multi-Object Compositing with Text and Layout Control
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2025)