From Zero to Hero: Training-Free Custom Concept Spawning in World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Akdemir, Kiymet, Yanardag, Pinar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2024)
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2024)
Audit & Repair: An Agentic Framework for Consistent Story Visualization in Text-to-Image Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025)
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025)
MIST: Mitigating Intersectional Bias with Disentangled Cross-Attention Editing in Text-to-Image Diffusion Models
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025)
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025)
LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
von: Simsar, Enis, et al.
Veröffentlicht: (2024)
von: Simsar, Enis, et al.
Veröffentlicht: (2024)
The Curious Case of End Token: A Zero-Shot Disentangled Image Editing using CLIP
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
AdaState: Self-Evolving Anchors for Streaming Video Generation
von: Dalva, Yusuf, et al.
Veröffentlicht: (2026)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2026)
RAVEL: Rare Concept Generation and Editing via Graph-driven Relational Guidance
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2024)
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2024)
LoRAverse: A Submodular Framework to Retrieve Diverse Adapters for Diffusion Models
von: Sonmezer, Mert, et al.
Veröffentlicht: (2025)
von: Sonmezer, Mert, et al.
Veröffentlicht: (2025)
Explaining in Diffusion: Explaining a Classifier Through Hierarchical Semantics with Text-to-Image Diffusion Models
von: Kazimi, Tahira, et al.
Veröffentlicht: (2024)
von: Kazimi, Tahira, et al.
Veröffentlicht: (2024)
GANTASTIC: GAN-based Transfer of Interpretable Directions for Disentangled Image Editing in Text-to-Image Diffusion Models
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
Dynamic View Synthesis as an Inverse Problem
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2025)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2025)
DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2026)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2026)
Diverse Video Generation with Determinantal Point Process-Guided Policy Optimization
von: Kazimi, Tahira, et al.
Veröffentlicht: (2025)
von: Kazimi, Tahira, et al.
Veröffentlicht: (2025)
CREA: A Collaborative Multi-Agent Framework for Creative Image Editing and Generation
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2025)
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2025)
FluxSpace: Disentangled Semantic Editing in Rectified Flow Transformers
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024)
Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization
von: Dunlop, Connor, et al.
Veröffentlicht: (2025)
von: Dunlop, Connor, et al.
Veröffentlicht: (2025)
ConceptAttention: Diffusion Transformers Learn Highly Interpretable Features
von: Helbling, Alec, et al.
Veröffentlicht: (2025)
von: Helbling, Alec, et al.
Veröffentlicht: (2025)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2025)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2025)
Stylebreeder: Exploring and Democratizing Artistic Styles through Text-to-Image Models
von: Zheng, Matthew, et al.
Veröffentlicht: (2024)
von: Zheng, Matthew, et al.
Veröffentlicht: (2024)
LoRA-Composer: Leveraging Low-Rank Adaptation for Multi-Concept Customization in Training-Free Diffusion Models
von: Yang, Yang, et al.
Veröffentlicht: (2024)
von: Yang, Yang, et al.
Veröffentlicht: (2024)
ConceptPose: Training-Free Zero-Shot Object Pose Estimation using Concept Vectors
von: Kuang, Liming, et al.
Veröffentlicht: (2025)
von: Kuang, Liming, et al.
Veröffentlicht: (2025)
MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2024)
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2024)
Contrastive Test-Time Composition of Multiple LoRA Models for Image Generation
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2024)
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2024)
CusConcept: Customized Visual Concept Decomposition with Diffusion Models
von: Xu, Zhi, et al.
Veröffentlicht: (2024)
von: Xu, Zhi, et al.
Veröffentlicht: (2024)
Aligning Latent Geometry for Spherical Flow Matching in Image Generation
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2026)
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2026)
Zero-to-Hero: Enhancing Zero-Shot Novel View Synthesis via Attention Map Filtering
von: Sobol, Ido, et al.
Veröffentlicht: (2024)
von: Sobol, Ido, et al.
Veröffentlicht: (2024)
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
Non-confusing Generation of Customized Concepts in Diffusion Models
von: Lin, Wang, et al.
Veröffentlicht: (2024)
von: Lin, Wang, et al.
Veröffentlicht: (2024)
Training-Free Multi-Concept Image Editing
von: Foteinopoulou, Niki, et al.
Veröffentlicht: (2026)
von: Foteinopoulou, Niki, et al.
Veröffentlicht: (2026)
CustomCrafter: Customized Video Generation with Preserving Motion and Concept Composition Abilities
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
EditID: Training-Free Editable ID Customization for Text-to-Image Generation
von: Li, Guandong, et al.
Veröffentlicht: (2025)
von: Li, Guandong, et al.
Veröffentlicht: (2025)
ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
Differential Vector Erasure: Unified Training-Free Concept Erasure for Flow Matching Models
von: Zhang, Zhiqi, et al.
Veröffentlicht: (2026)
von: Zhang, Zhiqi, et al.
Veröffentlicht: (2026)
Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models
von: Han, Chaolei, et al.
Veröffentlicht: (2025)
von: Han, Chaolei, et al.
Veröffentlicht: (2025)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
von: Sun, Shitong, et al.
Veröffentlicht: (2023)
von: Sun, Shitong, et al.
Veröffentlicht: (2023)
Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing
von: Su, Tongtong, et al.
Veröffentlicht: (2025)
von: Su, Tongtong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2024) -
Audit & Repair: An Agentic Framework for Consistent Story Visualization in Text-to-Image Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025) -
MIST: Mitigating Intersectional Bias with Disentangled Cross-Attention Editing in Text-to-Image Diffusion Models
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2024) -
Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models
von: Akdemir, Kiymet, et al.
Veröffentlicht: (2025) -
LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)