Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xue, Bowen, Duan, Zheng-Peng, Yan, Qixin, Wang, Wenjing, Liu, Hao, Guo, Chun-Le, Li, Chongyi, Li, Chen, Lyu, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
by: Mi, Yilin, et al.
Published: (2025)
by: Mi, Yilin, et al.
Published: (2025)
Improving Reconstruction of Representation Autoencoder
by: Liu, Siyu, et al.
Published: (2026)
by: Liu, Siyu, et al.
Published: (2026)
Identity as Presence: Towards Appearance and Voice Personalized Joint Audio-Video Generation
by: Chen, Yingjie, et al.
Published: (2026)
by: Chen, Yingjie, et al.
Published: (2026)
Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
by: Yan, Zhiyuan, et al.
Published: (2024)
by: Yan, Zhiyuan, et al.
Published: (2024)
Plug-and-Play Versatile Compressed Video Enhancement
by: Zeng, Huimin, et al.
Published: (2025)
by: Zeng, Huimin, et al.
Published: (2025)
Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation
by: Zhu, Lunjie, et al.
Published: (2026)
by: Zhu, Lunjie, et al.
Published: (2026)
VTinker: Guided Flow Upsampling and Texture Mapping for High-Resolution Video Frame Interpolation
by: Wu, Chenyang, et al.
Published: (2025)
by: Wu, Chenyang, et al.
Published: (2025)
Lighting Every Darkness with 3DGS: Fast Training and Real-Time Rendering for HDR View Synthesis
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
A Diffusion-Based Framework for Occluded Object Movement
by: Duan, Zheng-Peng, et al.
Published: (2025)
by: Duan, Zheng-Peng, et al.
Published: (2025)
Time-Aware One Step Diffusion Network for Real-World Image Super-Resolution
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
by: Chang, Zewei, et al.
Published: (2025)
by: Chang, Zewei, et al.
Published: (2025)
VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
by: Bian, Yuxuan, et al.
Published: (2025)
by: Bian, Yuxuan, et al.
Published: (2025)
Video Generation with Stable Transparency via Shiftable RGB-A Distribution Learner
by: Dong, Haotian, et al.
Published: (2025)
by: Dong, Haotian, et al.
Published: (2025)
ReMA: A Training-Free Plug-and-Play Mixing Augmentation for Video Behavior Recognition
by: Cui, Feng-Qi, et al.
Published: (2026)
by: Cui, Feng-Qi, et al.
Published: (2026)
DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution
by: Duan, Zheng-Peng, et al.
Published: (2025)
by: Duan, Zheng-Peng, et al.
Published: (2025)
FaceMe: Robust Blind Face Restoration with Personal Identification
by: Liu, Siyu, et al.
Published: (2025)
by: Liu, Siyu, et al.
Published: (2025)
From Understanding to Erasing: Towards Complete and Stable Video Object Removal
by: Liu, Dingming, et al.
Published: (2026)
by: Liu, Dingming, et al.
Published: (2026)
Synergistic Multiscale Detail Refinement via Intrinsic Supervision for Underwater Image Enhancement
by: Zhang, Dehuan, et al.
Published: (2023)
by: Zhang, Dehuan, et al.
Published: (2023)
Stochastic Generative Plug-and-Play Priors
by: Park, Chicago Y., et al.
Published: (2026)
by: Park, Chicago Y., et al.
Published: (2026)
Enhancing Logits Distillation with Plug\&Play Kendall's $τ$ Ranking Loss
by: Guan, Yuchen, et al.
Published: (2024)
by: Guan, Yuchen, et al.
Published: (2024)
UltraLED: Learning to See Everything in Ultra-High Dynamic Range Scenes
by: Meng, Yuang, et al.
Published: (2025)
by: Meng, Yuang, et al.
Published: (2025)
PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments
by: Peng, Kebin, et al.
Published: (2024)
by: Peng, Kebin, et al.
Published: (2024)
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
by: Duan, Zheng-Peng, et al.
Published: (2024)
by: Duan, Zheng-Peng, et al.
Published: (2024)
Plug-and-Play Diffusion Distillation
by: Hsiao, Yi-Ting, et al.
Published: (2024)
by: Hsiao, Yi-Ting, et al.
Published: (2024)
Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
PLUS: Plug-and-Play Enhanced Liver Lesion Diagnosis Model on Non-Contrast CT Scans
by: Hao, Jiacheng, et al.
Published: (2025)
by: Hao, Jiacheng, et al.
Published: (2025)
nnSAM: Plug-and-play Segment Anything Model Improves nnUNet Performance
by: Li, Yunxiang, et al.
Published: (2023)
by: Li, Yunxiang, et al.
Published: (2023)
NFCDS: A Plug-and-Play Noise Frequency-Controlled Diffusion Sampling Strategy for Image Restoration
by: Wang, Zhen, et al.
Published: (2026)
by: Wang, Zhen, et al.
Published: (2026)
Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models
by: Huang, Gexin, et al.
Published: (2026)
by: Huang, Gexin, et al.
Published: (2026)
Fancy123: One Image to High-Quality 3D Mesh Generation via Plug-and-Play Deformation
by: Yu, Qiao, et al.
Published: (2024)
by: Yu, Qiao, et al.
Published: (2024)
Plug-and-Play Posterior Sampling for Blind Inverse Problems
by: Li, Anqi, et al.
Published: (2025)
by: Li, Anqi, et al.
Published: (2025)
Plug-and-Play Context Feature Reuse for Efficient Masked Generation
by: Liu, Xuejie, et al.
Published: (2025)
by: Liu, Xuejie, et al.
Published: (2025)
YOSE: You Only Select Essential Tokens for Efficient DiT-based Video Object Removal
by: Wu, Chenyang, et al.
Published: (2026)
by: Wu, Chenyang, et al.
Published: (2026)
PADS: Plug-and-Play 3D Human Pose Analysis via Diffusion Generative Modeling
by: Ji, Haorui, et al.
Published: (2024)
by: Ji, Haorui, et al.
Published: (2024)
Protecting NeRFs' Copyright via Plug-And-Play Watermarking Base Model
by: Song, Qi, et al.
Published: (2024)
by: Song, Qi, et al.
Published: (2024)
PAT-VCM: Plug-and-Play Auxiliary Tokens for Video Coding for Machines
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
Plug, Play, and Fortify: A Low-Cost Module for Robust Multimodal Image Understanding Models
by: Lu, Siqi, et al.
Published: (2026)
by: Lu, Siqi, et al.
Published: (2026)
Kalman-Inspired Feature Propagation for Video Face Super-Resolution
by: Feng, Ruicheng, et al.
Published: (2024)
by: Feng, Ruicheng, et al.
Published: (2024)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
by: Li, Senmao, et al.
Published: (2025)
by: Li, Senmao, et al.
Published: (2025)
RAG-Adapter: A Plug-and-Play RAG-enhanced Framework for Long Video Understanding
by: Tan, Xichen, et al.
Published: (2025)
by: Tan, Xichen, et al.
Published: (2025)
Similar Items
-
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
by: Mi, Yilin, et al.
Published: (2025) -
Improving Reconstruction of Representation Autoencoder
by: Liu, Siyu, et al.
Published: (2026) -
Identity as Presence: Towards Appearance and Voice Personalized Joint Audio-Video Generation
by: Chen, Yingjie, et al.
Published: (2026) -
Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
by: Yan, Zhiyuan, et al.
Published: (2024) -
Plug-and-Play Versatile Compressed Video Enhancement
by: Zeng, Huimin, et al.
Published: (2025)