Orthogonal Adaptation for Modular Customization of Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Po, Ryan, Yang, Guandao, Aberman, Kfir, Wetzstein, Gordon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Interpreting the Weight Space of Customized Diffusion Models
di: Dravid, Amil, et al.
Pubblicazione: (2024)
di: Dravid, Amil, et al.
Pubblicazione: (2024)
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
di: Wu, Tong, et al.
Pubblicazione: (2024)
di: Wu, Tong, et al.
Pubblicazione: (2024)
Towards Vision-Language-Garment Models for Web Knowledge Garment Understanding and Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
di: Po, Ryan, et al.
Pubblicazione: (2025)
di: Po, Ryan, et al.
Pubblicazione: (2025)
Continuous Control of Editing Models via Adaptive-Origin Guidance
di: Wolf, Alon, et al.
Pubblicazione: (2026)
di: Wolf, Alon, et al.
Pubblicazione: (2026)
Modular Customization of Diffusion Models via Blockwise-Parameterized Low-Rank Adaptation
di: Zhu, Mingkang, et al.
Pubblicazione: (2025)
di: Zhu, Mingkang, et al.
Pubblicazione: (2025)
Video World Models with Long-term Spatial Memory
di: Wu, Tong, et al.
Pubblicazione: (2025)
di: Wu, Tong, et al.
Pubblicazione: (2025)
Self-Calibrating Gaussian Splatting for Large Field of View Reconstruction
di: Deng, Youming, et al.
Pubblicazione: (2025)
di: Deng, Youming, et al.
Pubblicazione: (2025)
3D PixBrush: Image-Guided Local Texture Synthesis
di: Decatur, Dale, et al.
Pubblicazione: (2025)
di: Decatur, Dale, et al.
Pubblicazione: (2025)
GPT-4V(ision) is a Human-Aligned Evaluator for Text-to-3D Generation
di: Wu, Tong, et al.
Pubblicazione: (2024)
di: Wu, Tong, et al.
Pubblicazione: (2024)
MegaScenes: Scene-Level View Synthesis at Scale
di: Tung, Joseph, et al.
Pubblicazione: (2024)
di: Tung, Joseph, et al.
Pubblicazione: (2024)
MyVLM: Personalizing VLMs for User-Specific Queries
di: Alaluf, Yuval, et al.
Pubblicazione: (2024)
di: Alaluf, Yuval, et al.
Pubblicazione: (2024)
AIpparel: A Multimodal Foundation Model for Digital Garments
di: Nakayama, Kiyohiro, et al.
Pubblicazione: (2024)
di: Nakayama, Kiyohiro, et al.
Pubblicazione: (2024)
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
di: Po, Ryan, et al.
Pubblicazione: (2026)
di: Po, Ryan, et al.
Pubblicazione: (2026)
Long-Context State-Space Video World Models
di: Po, Ryan, et al.
Pubblicazione: (2025)
di: Po, Ryan, et al.
Pubblicazione: (2025)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
di: Dahary, Omer, et al.
Pubblicazione: (2024)
di: Dahary, Omer, et al.
Pubblicazione: (2024)
Spectral Progressive Diffusion for Efficient Image and Video Generation
di: Xiao, Howard, et al.
Pubblicazione: (2026)
di: Xiao, Howard, et al.
Pubblicazione: (2026)
Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation
di: Chao, Brian, et al.
Pubblicazione: (2026)
di: Chao, Brian, et al.
Pubblicazione: (2026)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
di: Qian, Guocheng Gordon, et al.
Pubblicazione: (2025)
di: Qian, Guocheng Gordon, et al.
Pubblicazione: (2025)
Robust Symmetry Detection via Riemannian Langevin Dynamics
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention
di: Zhang, Howard, et al.
Pubblicazione: (2024)
di: Zhang, Howard, et al.
Pubblicazione: (2024)
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models
di: Zhang, Lvmin, et al.
Pubblicazione: (2025)
di: Zhang, Lvmin, et al.
Pubblicazione: (2025)
ImageGem: In-the-wild Generative Image Interaction Dataset for Generative Model Personalization
di: Guo, Yuanhe, et al.
Pubblicazione: (2025)
di: Guo, Yuanhe, et al.
Pubblicazione: (2025)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
di: Simsar, Enis, et al.
Pubblicazione: (2024)
di: Simsar, Enis, et al.
Pubblicazione: (2024)
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
di: Huang, Ian, et al.
Pubblicazione: (2024)
di: Huang, Ian, et al.
Pubblicazione: (2024)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
di: Dalva, Yusuf, et al.
Pubblicazione: (2025)
di: Dalva, Yusuf, et al.
Pubblicazione: (2025)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
di: Dahary, Omer, et al.
Pubblicazione: (2025)
di: Dahary, Omer, et al.
Pubblicazione: (2025)
PhysAvatar: Learning the Physics of Dressed 3D Avatars from Visual Observations
di: Zheng, Yang, et al.
Pubblicazione: (2024)
di: Zheng, Yang, et al.
Pubblicazione: (2024)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
di: Wang, Kuan-Chieh, et al.
Pubblicazione: (2024)
di: Wang, Kuan-Chieh, et al.
Pubblicazione: (2024)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
di: Patashnik, Or, et al.
Pubblicazione: (2025)
di: Patashnik, Or, et al.
Pubblicazione: (2025)
VideoMage: Multi-Subject and Motion Customization of Text-to-Video Diffusion Models
di: Huang, Chi-Pin, et al.
Pubblicazione: (2025)
di: Huang, Chi-Pin, et al.
Pubblicazione: (2025)
Asymmetric Flow Models
di: Chen, Hansheng, et al.
Pubblicazione: (2026)
di: Chen, Hansheng, et al.
Pubblicazione: (2026)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
di: Qian, Guocheng, et al.
Pubblicazione: (2024)
di: Qian, Guocheng, et al.
Pubblicazione: (2024)
LoRA-Composer: Leveraging Low-Rank Adaptation for Multi-Concept Customization in Training-Free Diffusion Models
di: Yang, Yang, et al.
Pubblicazione: (2024)
di: Yang, Yang, et al.
Pubblicazione: (2024)
Flying with Photons: Rendering Novel Views of Propagating Light
di: Malik, Anagh, et al.
Pubblicazione: (2024)
di: Malik, Anagh, et al.
Pubblicazione: (2024)
Infinite Gaze Generation for Videos with Autoregressive Diffusion
di: Kang, Jenna, et al.
Pubblicazione: (2026)
di: Kang, Jenna, et al.
Pubblicazione: (2026)
InfoGaussian: Structure-Aware Dynamic Gaussians through Lightweight Information Shaping
di: Zhang, Yunchao, et al.
Pubblicazione: (2024)
di: Zhang, Yunchao, et al.
Pubblicazione: (2024)
Policy-based Foveated Imaging and Perception
di: Xiao, Howard, et al.
Pubblicazione: (2026)
di: Xiao, Howard, et al.
Pubblicazione: (2026)
CustomVideoX: 3D Reference Attention Driven Dynamic Adaptation for Zero-Shot Customized Video Diffusion Transformers
di: She, D., et al.
Pubblicazione: (2025)
di: She, D., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Interpreting the Weight Space of Customized Diffusion Models
di: Dravid, Amil, et al.
Pubblicazione: (2024) -
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
di: Wu, Tong, et al.
Pubblicazione: (2024) -
Towards Vision-Language-Garment Models for Web Knowledge Garment Understanding and Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2025) -
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
di: Po, Ryan, et al.
Pubblicazione: (2025) -
Continuous Control of Editing Models via Adaptive-Origin Guidance
di: Wolf, Alon, et al.
Pubblicazione: (2026)