A Training-Free Approach for Multi-ID Customization via Attention Adjustment and Spatial Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Jiawei, Jiao, Guanlong, Xu, Jianjin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StreamGVE: Training-Free Video Editing via Few-Step Streaming Video Generation
von: Jiao, Guanlong, et al.
Veröffentlicht: (2026)
von: Jiao, Guanlong, et al.
Veröffentlicht: (2026)
EditID: Training-Free Editable ID Customization for Text-to-Image Generation
von: Li, Guandong, et al.
Veröffentlicht: (2025)
von: Li, Guandong, et al.
Veröffentlicht: (2025)
CusEnhancer: A Zero-Shot Scene and Controllability Enhancement Method for Photo Customization via ResInversion
von: Ren, Maoye, et al.
Veröffentlicht: (2025)
von: Ren, Maoye, et al.
Veröffentlicht: (2025)
MagicView: Multi-View Consistent Identity Customization via Priors-Guided In-Context Learning
von: Li, Hengjia, et al.
Veröffentlicht: (2025)
von: Li, Hengjia, et al.
Veröffentlicht: (2025)
Multi-party Collaborative Attention Control for Image Customization
von: Yang, Han, et al.
Veröffentlicht: (2025)
von: Yang, Han, et al.
Veröffentlicht: (2025)
Tuning-Free Visual Customization via View Iterative Self-Attention Control
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
FreeControl: Efficient, Training-Free Structural Control via One-Step Attention Extraction
von: Lin, Jiang, et al.
Veröffentlicht: (2025)
von: Lin, Jiang, et al.
Veröffentlicht: (2025)
SwitchCraft: Training-Free Multi-Event Video Generation with Attention Controls
von: Xu, Qianxun, et al.
Veröffentlicht: (2026)
von: Xu, Qianxun, et al.
Veröffentlicht: (2026)
OmniPhysGS: 3D Constitutive Gaussians for General Physics-Based Dynamics Generation
von: Lin, Yuchen, et al.
Veröffentlicht: (2025)
von: Lin, Yuchen, et al.
Veröffentlicht: (2025)
MVPortrait: Text-Guided Motion and Emotion Control for Multi-view Vivid Portrait Animation
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
UniEdit-Flow: Unleashing Inversion and Editing in the Era of Flow Models
von: Jiao, Guanlong, et al.
Veröffentlicht: (2025)
von: Jiao, Guanlong, et al.
Veröffentlicht: (2025)
ISAC: Training-Free Instance-to-Semantic Attention Control for Improving Multi-Instance Generation
von: Jo, Sanghyun, et al.
Veröffentlicht: (2025)
von: Jo, Sanghyun, et al.
Veröffentlicht: (2025)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
Training-Free Open-Ended Object Detection and Segmentation via Attention as Prompts
von: Lin, Zhiwei, et al.
Veröffentlicht: (2024)
von: Lin, Zhiwei, et al.
Veröffentlicht: (2024)
RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global Complementation
von: Sun, Zhaoyang, et al.
Veröffentlicht: (2024)
von: Sun, Zhaoyang, et al.
Veröffentlicht: (2024)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
von: Xia, Tian, et al.
Veröffentlicht: (2024)
von: Xia, Tian, et al.
Veröffentlicht: (2024)
HAM: A Training-Free Style Transfer Approach via Heterogeneous Attention Modulation for Diffusion Models
von: He, Yeqi, et al.
Veröffentlicht: (2026)
von: He, Yeqi, et al.
Veröffentlicht: (2026)
TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction
von: Zheng, Zhijie, et al.
Veröffentlicht: (2026)
von: Zheng, Zhijie, et al.
Veröffentlicht: (2026)
Proteus-ID: ID-Consistent and Motion-Coherent Video Customization
von: Zhang, Guiyu, et al.
Veröffentlicht: (2025)
von: Zhang, Guiyu, et al.
Veröffentlicht: (2025)
Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis
von: Han, Woojung, et al.
Veröffentlicht: (2025)
von: Han, Woojung, et al.
Veröffentlicht: (2025)
DisenStudio: Customized Multi-subject Text-to-Video Generation with Disentangled Spatial Control
von: Chen, Hong, et al.
Veröffentlicht: (2024)
von: Chen, Hong, et al.
Veröffentlicht: (2024)
Information Bottleneck Approach to Spatial Attention Learning
von: Lai, Qiuxia, et al.
Veröffentlicht: (2021)
von: Lai, Qiuxia, et al.
Veröffentlicht: (2021)
PuLID: Pure and Lightning ID Customization via Contrastive Alignment
von: Guo, Zinan, et al.
Veröffentlicht: (2024)
von: Guo, Zinan, et al.
Veröffentlicht: (2024)
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
Search2Motion: Training-Free Object-Level Motion Control via Attention-Consensus Search
von: Liu, Sainan, et al.
Veröffentlicht: (2026)
von: Liu, Sainan, et al.
Veröffentlicht: (2026)
CARE: Training-Free Controllable Restoration for Medical Images via Dual-Latent Steering
von: Liu, Xu
Veröffentlicht: (2026)
von: Liu, Xu
Veröffentlicht: (2026)
One-to-More: High-Fidelity Training-Free Anomaly Generation with Attention Control
von: Rao, Haoxiang, et al.
Veröffentlicht: (2026)
von: Rao, Haoxiang, et al.
Veröffentlicht: (2026)
Inv-Adapter: ID Customization Generation via Image Inversion and Lightweight Adapter
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
Training-free Regional Prompting for Diffusion Transformers
von: Chen, Anthony, et al.
Veröffentlicht: (2024)
von: Chen, Anthony, et al.
Veröffentlicht: (2024)
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
MUSAR: Exploring Multi-Subject Customization from Single-Subject Dataset via Attention Routing
von: Guo, Zinan, et al.
Veröffentlicht: (2025)
von: Guo, Zinan, et al.
Veröffentlicht: (2025)
LoRA-Composer: Leveraging Low-Rank Adaptation for Multi-Concept Customization in Training-Free Diffusion Models
von: Yang, Yang, et al.
Veröffentlicht: (2024)
von: Yang, Yang, et al.
Veröffentlicht: (2024)
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
von: Zhang, Zhexin, et al.
Veröffentlicht: (2026)
von: Zhang, Zhexin, et al.
Veröffentlicht: (2026)
Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning
von: Deng, Yang, et al.
Veröffentlicht: (2026)
von: Deng, Yang, et al.
Veröffentlicht: (2026)
MagicID: Hybrid Preference Optimization for ID-Consistent and Dynamic-Preserved Video Customization
von: Li, Hengjia, et al.
Veröffentlicht: (2025)
von: Li, Hengjia, et al.
Veröffentlicht: (2025)
FaceSnap: Enhanced ID-fidelity Network for Tuning-free Portrait Customization
von: Zhai, Benxiang, et al.
Veröffentlicht: (2026)
von: Zhai, Benxiang, et al.
Veröffentlicht: (2026)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
von: Bi, Xiuli, et al.
Veröffentlicht: (2024)
von: Bi, Xiuli, et al.
Veröffentlicht: (2024)
MS-CustomNet: Controllable Multi-Subject Customization with Hierarchical Relational Semantics
von: Cai, Pengxiang, et al.
Veröffentlicht: (2026)
von: Cai, Pengxiang, et al.
Veröffentlicht: (2026)
From Zero to Hero: Training-Free Custom Concept Spawning in World Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2026)
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2026)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
StreamGVE: Training-Free Video Editing via Few-Step Streaming Video Generation
von: Jiao, Guanlong, et al.
Veröffentlicht: (2026) -
EditID: Training-Free Editable ID Customization for Text-to-Image Generation
von: Li, Guandong, et al.
Veröffentlicht: (2025) -
CusEnhancer: A Zero-Shot Scene and Controllability Enhancement Method for Photo Customization via ResInversion
von: Ren, Maoye, et al.
Veröffentlicht: (2025) -
MagicView: Multi-View Consistent Identity Customization via Priors-Guided In-Context Learning
von: Li, Hengjia, et al.
Veröffentlicht: (2025) -
Multi-party Collaborative Attention Control for Image Customization
von: Yang, Han, et al.
Veröffentlicht: (2025)