InstantID: Zero-shot Identity-Preserving Generation in Seconds
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Qixun, Bai, Xu, Wang, Haofan, Qin, Zekui, Chen, Anthony, Li, Huaxia, Tang, Xu, Hu, Yao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
InstantStyle-Plus: Style Transfer with Content-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
InstantIR: Blind Image Restoration with Instant Generative Reference
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2024)
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2024)
Generating Synthetic Data via Augmentations for Improved Facial Resemblance in DreamBooth and InstantID
von: Ulusan, Koray, et al.
Veröffentlicht: (2025)
von: Ulusan, Koray, et al.
Veröffentlicht: (2025)
InstantCharacter: Personalize Any Characters with a Scalable Diffusion Transformer Framework
von: Tao, Jiale, et al.
Veröffentlicht: (2025)
von: Tao, Jiale, et al.
Veröffentlicht: (2025)
InstantFamily: Masked Attention for Zero-shot Multi-ID Image Generation
von: Kim, Chanran, et al.
Veröffentlicht: (2024)
von: Kim, Chanran, et al.
Veröffentlicht: (2024)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
CSGO: Content-Style Composition in Text-to-Image Generation
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
Unified Video-Language Pre-training with Synchronized Audio
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
InstantSplat: Sparse-view Gaussian Splatting in Seconds
von: Fan, Zhiwen, et al.
Veröffentlicht: (2024)
von: Fan, Zhiwen, et al.
Veröffentlicht: (2024)
Target-Driven Distillation: Consistency Distillation with Target Timestep Selection and Decoupled Guidance
von: Wang, Cunzheng, et al.
Veröffentlicht: (2024)
von: Wang, Cunzheng, et al.
Veröffentlicht: (2024)
StableGarment: Garment-Centric Generation via Stable Diffusion
von: Wang, Rui, et al.
Veröffentlicht: (2024)
von: Wang, Rui, et al.
Veröffentlicht: (2024)
Zero-shot Face Editing via ID-Attribute Decoupled Inversion
von: Hou, Yang, et al.
Veröffentlicht: (2025)
von: Hou, Yang, et al.
Veröffentlicht: (2025)
StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
von: Zhou, Zhengguang, et al.
Veröffentlicht: (2024)
von: Zhou, Zhengguang, et al.
Veröffentlicht: (2024)
ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving
von: Huang, Jiehui, et al.
Veröffentlicht: (2024)
von: Huang, Jiehui, et al.
Veröffentlicht: (2024)
FantasyID: Face Knowledge Enhanced ID-Preserving Video Generation
von: Zhang, Yunpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yunpeng, et al.
Veröffentlicht: (2025)
GAIA: Zero-shot Talking Avatar Generation
von: He, Tianyu, et al.
Veröffentlicht: (2023)
von: He, Tianyu, et al.
Veröffentlicht: (2023)
WildActor: Unconstrained Identity-Preserving Video Generation
von: Guo, Qin, et al.
Veröffentlicht: (2026)
von: Guo, Qin, et al.
Veröffentlicht: (2026)
HiFi-Portrait: Zero-shot Identity-preserved Portrait Generation with High-fidelity Multi-face Fusion
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
PointDreamer: Zero-shot 3D Textured Mesh Reconstruction from Colored Point Cloud
von: Yu, Qiao, et al.
Veröffentlicht: (2024)
von: Yu, Qiao, et al.
Veröffentlicht: (2024)
SIGMA: Selective-Interleaved Generation with Multi-Attribute Tokens
von: Zhang, Xiaoyan, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaoyan, et al.
Veröffentlicht: (2026)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
von: Wu, Mingyang, et al.
Veröffentlicht: (2026)
von: Wu, Mingyang, et al.
Veröffentlicht: (2026)
PSVMA+: Exploring Multi-granularity Semantic-visual Adaption for Generalized Zero-shot Learning
von: Liu, Man, et al.
Veröffentlicht: (2024)
von: Liu, Man, et al.
Veröffentlicht: (2024)
LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation
von: Song, Wenhui, et al.
Veröffentlicht: (2025)
von: Song, Wenhui, et al.
Veröffentlicht: (2025)
AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
von: Xiang, Qiang, et al.
Veröffentlicht: (2025)
von: Xiang, Qiang, et al.
Veröffentlicht: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
ID$^3$: Identity-Preserving-yet-Diversified Diffusion Models for Synthetic Face Recognition
von: Li, Shen, et al.
Veröffentlicht: (2024)
von: Li, Shen, et al.
Veröffentlicht: (2024)
SSR-Encoder: Encoding Selective Subject Representation for Subject-Driven Generation
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2023)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2023)
Diff-PC: Identity-preserving and 3D-aware Controllable Diffusion for Zero-shot Portrait Customization
von: Xu, Yifang, et al.
Veröffentlicht: (2026)
von: Xu, Yifang, et al.
Veröffentlicht: (2026)
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
ZeroStereo: Zero-shot Stereo Matching from Single Images
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment
von: Xu, Yuanfeng, et al.
Veröffentlicht: (2024)
von: Xu, Yuanfeng, et al.
Veröffentlicht: (2024)
DreamSalon: A Staged Diffusion Framework for Preserving Identity-Context in Editable Face Generation
von: Lin, Haonan, et al.
Veröffentlicht: (2024)
von: Lin, Haonan, et al.
Veröffentlicht: (2024)
Zero-shot Composed Text-Image Retrieval
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
von: Chen, Weifeng, et al.
Veröffentlicht: (2024)
Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot Generalization
von: Shen, Yang, et al.
Veröffentlicht: (2024)
von: Shen, Yang, et al.
Veröffentlicht: (2024)
Multimodal Sense-Informed Prediction of 3D Human Motions
von: Lou, Zhenyu, et al.
Veröffentlicht: (2024)
von: Lou, Zhenyu, et al.
Veröffentlicht: (2024)
IdentityStory: Taming Your Identity-Preserving Generator for Human-Centric Story Generation
von: Zhou, Donghao, et al.
Veröffentlicht: (2025)
von: Zhou, Donghao, et al.
Veröffentlicht: (2025)
Skeleton and Font Generation Network for Zero-shot Chinese Character Generation
von: Xue, Mobai, et al.
Veröffentlicht: (2025)
von: Xue, Mobai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024) -
InstantStyle-Plus: Style Transfer with Content-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024) -
InstantIR: Blind Image Restoration with Instant Generative Reference
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2024) -
Generating Synthetic Data via Augmentations for Improved Facial Resemblance in DreamBooth and InstantID
von: Ulusan, Koray, et al.
Veröffentlicht: (2025) -
InstantCharacter: Personalize Any Characters with a Scalable Diffusion Transformer Framework
von: Tao, Jiale, et al.
Veröffentlicht: (2025)