Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lai, Yixuan, Wang, He, Zhou, Kun, Shao, Tianjia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Concat-ID: Towards Universal Identity-Preserving Video Synthesis
von: Zhong, Yong, et al.
Veröffentlicht: (2025)
von: Zhong, Yong, et al.
Veröffentlicht: (2025)
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025)
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025)
InstantID: Zero-shot Identity-Preserving Generation in Seconds
von: Wang, Qixun, et al.
Veröffentlicht: (2024)
von: Wang, Qixun, et al.
Veröffentlicht: (2024)
Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024)
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024)
ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving
von: Huang, Jiehui, et al.
Veröffentlicht: (2024)
von: Huang, Jiehui, et al.
Veröffentlicht: (2024)
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2026)
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
Transformers and Slot Encoding for Sample Efficient Physical World Modelling
von: Petri, Francesco, et al.
Veröffentlicht: (2024)
von: Petri, Francesco, et al.
Veröffentlicht: (2024)
PLACID: Identity-Preserving Multi-Object Compositing via Video Diffusion with Synthetic Trajectories
von: Tarrés, Gemma Canet, et al.
Veröffentlicht: (2026)
von: Tarrés, Gemma Canet, et al.
Veröffentlicht: (2026)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
GenesisTex2: Stable, Consistent and High-Quality Text-to-Texture Generation
von: Lu, Jiawei, et al.
Veröffentlicht: (2024)
von: Lu, Jiawei, et al.
Veröffentlicht: (2024)
Slot-VLM: SlowFast Slots for Video-Language Modeling
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
Temporally Consistent Object-Centric Learning by Contrasting Slots
von: Manasyan, Anna, et al.
Veröffentlicht: (2024)
von: Manasyan, Anna, et al.
Veröffentlicht: (2024)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
Magic-Me: Identity-Specific Video Customized Diffusion
von: Ma, Ze, et al.
Veröffentlicht: (2024)
von: Ma, Ze, et al.
Veröffentlicht: (2024)
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
von: Xing, Jiazheng, et al.
Veröffentlicht: (2026)
von: Xing, Jiazheng, et al.
Veröffentlicht: (2026)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
Beyond the Pixels: VLM-based Evaluation of Identity Preservation in Reference-Guided Synthesis
von: Singhania, Aditi, et al.
Veröffentlicht: (2025)
von: Singhania, Aditi, et al.
Veröffentlicht: (2025)
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
ID-Sim: An Identity-Focused Similarity Metric
von: Chae, Julia, et al.
Veröffentlicht: (2026)
von: Chae, Julia, et al.
Veröffentlicht: (2026)
AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
Predicting Video Slot Attention Queries from Random Slot-Feature Pairs
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
TC-SSA: Token Compression via Semantic Slot Aggregation for Gigapixel Pathology Reasoning
von: Chen, Zhuo, et al.
Veröffentlicht: (2026)
von: Chen, Zhuo, et al.
Veröffentlicht: (2026)
Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in Retinal Diagnosis
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
Show and Polish: Reference-Guided Identity Preservation in Face Video Restoration
von: Han, Wenkang, et al.
Veröffentlicht: (2025)
von: Han, Wenkang, et al.
Veröffentlicht: (2025)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
CGSA: Class-Guided Slot-Aware Adaptation for Source-Free Object Detection
von: Dai, Boyang, et al.
Veröffentlicht: (2026)
von: Dai, Boyang, et al.
Veröffentlicht: (2026)
PAS: A Training-Free Stabilizer for Temporal Encoding in Video LLMs
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
von: Zhang, Erhang, et al.
Veröffentlicht: (2025)
von: Zhang, Erhang, et al.
Veröffentlicht: (2025)
SlotPi: Physics-informed Object-centric Reasoning Models
von: Li, Jian, et al.
Veröffentlicht: (2025)
von: Li, Jian, et al.
Veröffentlicht: (2025)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
von: Wu, Mingyang, et al.
Veröffentlicht: (2026)
von: Wu, Mingyang, et al.
Veröffentlicht: (2026)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
DualReal: Adaptive Joint Training for Lossless Identity-Motion Fusion in Video Customization
von: Wang, Wenchuan, et al.
Veröffentlicht: (2025)
von: Wang, Wenchuan, et al.
Veröffentlicht: (2025)
When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation
von: Zheng, Dongqi
Veröffentlicht: (2026)
von: Zheng, Dongqi
Veröffentlicht: (2026)
VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video Reasoning
von: Ding, Yang, et al.
Veröffentlicht: (2025)
von: Ding, Yang, et al.
Veröffentlicht: (2025)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
Characterizing Motion Encoding in Video Diffusion Timesteps
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Concat-ID: Towards Universal Identity-Preserving Video Synthesis
von: Zhong, Yong, et al.
Veröffentlicht: (2025) -
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025) -
InstantID: Zero-shot Identity-Preserving Generation in Seconds
von: Wang, Qixun, et al.
Veröffentlicht: (2024) -
Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024) -
ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving
von: Huang, Jiehui, et al.
Veröffentlicht: (2024)