Omni-ID: Holistic Identity Representation Designed for Generative Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qian, Guocheng, Wang, Kuan-Chieh, Patashnik, Or, Heravi, Negin, Ostashev, Daniil, Tulyakov, Sergey, Cohen-Or, Daniel, Aberman, Kfir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Nested Attention: Semantic-aware Attention Values for Concept Personalization
von: Patashnik, Or, et al.
Veröffentlicht: (2025)
von: Patashnik, Or, et al.
Veröffentlicht: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
Object-level Visual Prompts for Compositional Image Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
Scaling Group Inference for Diverse and High-Quality Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
von: Dahary, Omer, et al.
Veröffentlicht: (2024)
von: Dahary, Omer, et al.
Veröffentlicht: (2024)
Preventing Shortcuts in Adapter Training via Providing the Shortcuts
von: Goyal, Anujraaj Argo, et al.
Veröffentlicht: (2025)
von: Goyal, Anujraaj Argo, et al.
Veröffentlicht: (2025)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
von: Dahary, Omer, et al.
Veröffentlicht: (2025)
von: Dahary, Omer, et al.
Veröffentlicht: (2025)
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
MyVLM: Personalizing VLMs for User-Specific Queries
von: Alaluf, Yuval, et al.
Veröffentlicht: (2024)
von: Alaluf, Yuval, et al.
Veröffentlicht: (2024)
Dynamic Concepts Personalization from Single Videos
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
Continuous Control of Editing Models via Adaptive-Origin Guidance
von: Wolf, Alon, et al.
Veröffentlicht: (2026)
von: Wolf, Alon, et al.
Veröffentlicht: (2026)
Stable Flow: Vital Layers for Training-Free Image Editing
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
LayerComposer: Multi-Human Personalized Generation via Layered Canvas
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
Tuning-free Visual Effect Transfer across Videos
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
von: Mi, Zhenxing, et al.
Veröffentlicht: (2025)
von: Mi, Zhenxing, et al.
Veröffentlicht: (2025)
Visual Personalization Turing Test
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
Interpreting the Weight Space of Customized Diffusion Models
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
AToM: Amortized Text-to-Mesh using 2D Diffusion
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
Multi-subject Open-set Personalization in Video Generation
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
Orthogonal Adaptation for Modular Customization of Diffusion Models
von: Po, Ryan, et al.
Veröffentlicht: (2023)
von: Po, Ryan, et al.
Veröffentlicht: (2023)
Efficient Training with Denoised Neural Weights
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
3D PixBrush: Image-Guided Local Texture Synthesis
von: Decatur, Dale, et al.
Veröffentlicht: (2025)
von: Decatur, Dale, et al.
Veröffentlicht: (2025)
SemanticMoments: Training-Free Motion Similarity via Third Moment Features
von: Huberman, Saar, et al.
Veröffentlicht: (2026)
von: Huberman, Saar, et al.
Veröffentlicht: (2026)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
E$^{2}$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
MotioNet: 3D Human Motion Reconstruction from Monocular Video with Skeleton Consistency
von: Shi, Mingyi, et al.
Veröffentlicht: (2020)
von: Shi, Mingyi, et al.
Veröffentlicht: (2020)
AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention
von: Zhang, Howard, et al.
Veröffentlicht: (2024)
von: Zhang, Howard, et al.
Veröffentlicht: (2024)
HoLa: B-Rep Generation using a Holistic Latent Representation
von: Liu, Yilin, et al.
Veröffentlicht: (2025)
von: Liu, Yilin, et al.
Veröffentlicht: (2025)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
von: Deutch, Gilad, et al.
Veröffentlicht: (2024)
Image Generation from Contextually-Contradictory Prompts
von: Huberman, Saar, et al.
Veröffentlicht: (2025)
von: Huberman, Saar, et al.
Veröffentlicht: (2025)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
Wonderland: Navigating 3D Scenes from a Single Image
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
ArtifactLens: Hundreds of Labels Are Enough for Artifact Detection with VLMs
von: Burgess, James, et al.
Veröffentlicht: (2026)
von: Burgess, James, et al.
Veröffentlicht: (2026)
In-Context Sync-LoRA for Portrait Video Editing
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Nested Attention: Semantic-aware Attention Values for Concept Personalization
von: Patashnik, Or, et al.
Veröffentlicht: (2025) -
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024) -
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025) -
Object-level Visual Prompts for Compositional Image Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025) -
Scaling Group Inference for Diverse and High-Quality Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)