LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Mingkang, Chen, Xi, Wang, Zhongdao, Zhao, Hengshuang, Jia, Jiaya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Modular Customization of Diffusion Models via Blockwise-Parameterized Low-Rank Adaptation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
Query-Efficient Video Adversarial Attack with Stylized Logo
von: Tang, Duoxun, et al.
Veröffentlicht: (2024)
von: Tang, Duoxun, et al.
Veröffentlicht: (2024)
GridMask Data Augmentation
von: Chen, Pengguang, et al.
Veröffentlicht: (2020)
von: Chen, Pengguang, et al.
Veröffentlicht: (2020)
AdvLogo: Adversarial Patch Attack against Object Detectors based on Diffusion Models
von: Miao, Boming, et al.
Veröffentlicht: (2024)
von: Miao, Boming, et al.
Veröffentlicht: (2024)
SLANT: Spurious Logo ANalysis Toolkit
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
Enhancing LLM Knowledge Learning through Generalization
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
Animated Stickers: Bringing Stickers to Life with Video Diffusion
von: Yan, David, et al.
Veröffentlicht: (2024)
von: Yan, David, et al.
Veröffentlicht: (2024)
FashionLOGO: Prompting Multimodal Large Language Models for Fashion Logo Embeddings
von: Wang, Zhen, et al.
Veröffentlicht: (2023)
von: Wang, Zhen, et al.
Veröffentlicht: (2023)
ViSurf: Visual Supervised-and-Reinforcement Fine-Tuning for Large Vision-and-Language Models
von: Liu, Yuqi, et al.
Veröffentlicht: (2025)
von: Liu, Yuqi, et al.
Veröffentlicht: (2025)
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
von: Li, Sifan, et al.
Veröffentlicht: (2025)
von: Li, Sifan, et al.
Veröffentlicht: (2025)
Brand Visibility in Packaging: A Deep Learning Approach for Logo Detection, Saliency-Map Prediction, and Logo Placement Analysis
von: Hosseini, Alireza, et al.
Veröffentlicht: (2024)
von: Hosseini, Alireza, et al.
Veröffentlicht: (2024)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
AnyDoor: Zero-shot Object-level Image Customization
von: Chen, Xi, et al.
Veröffentlicht: (2023)
von: Chen, Xi, et al.
Veröffentlicht: (2023)
TGDPO: Harnessing Token-Level Reward Guidance for Enhancing Direct Preference Optimization
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
A New Method for Vehicle Logo Recognition Based on Swin Transformer
von: Li, Yang, et al.
Veröffentlicht: (2024)
von: Li, Yang, et al.
Veröffentlicht: (2024)
Logos as a Well-Tempered Pre-train for Sign Language Recognition
von: Ovodov, Ilya, et al.
Veröffentlicht: (2025)
von: Ovodov, Ilya, et al.
Veröffentlicht: (2025)
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
Multi-Label Logo Recognition and Retrieval based on Weighted Fusion of Neural Features
von: Bernabeu, Marisa, et al.
Veröffentlicht: (2022)
von: Bernabeu, Marisa, et al.
Veröffentlicht: (2022)
GroupContrast: Semantic-aware Self-supervised Representation Learning for 3D Understanding
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
Toward Intelligent Scene Augmentation for Context-Aware Object Placement and Sponsor-Logo Integration
von: Saraswat, Unnati, et al.
Veröffentlicht: (2025)
von: Saraswat, Unnati, et al.
Veröffentlicht: (2025)
Text-to-Sticker: Style Tailoring Latent Diffusion Models for Human Expression
von: Sinha, Animesh, et al.
Veröffentlicht: (2023)
von: Sinha, Animesh, et al.
Veröffentlicht: (2023)
Mind the Interference: Retaining Pre-trained Knowledge in Parameter Efficient Continual Learning of Vision-Language Models
von: Tang, Longxiang, et al.
Veröffentlicht: (2024)
von: Tang, Longxiang, et al.
Veröffentlicht: (2024)
GDRO: Group-level Reward Post-training Suitable for Diffusion Models
von: Wang, Yiyang, et al.
Veröffentlicht: (2026)
von: Wang, Yiyang, et al.
Veröffentlicht: (2026)
GSE: Evaluating Sticker Visual Semantic Similarity via a General Sticker Encoder
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
ExposureEngine: Oriented Logo Detection and Sponsor Visibility Analytics in Sports Broadcasts
von: Sarkhoosh, Mehdi Houshmand, et al.
Veröffentlicht: (2025)
von: Sarkhoosh, Mehdi Houshmand, et al.
Veröffentlicht: (2025)
Non-confusing Generation of Customized Concepts in Diffusion Models
von: Lin, Wang, et al.
Veröffentlicht: (2024)
von: Lin, Wang, et al.
Veröffentlicht: (2024)
Infrared Adversarial Car Stickers
von: Zhu, Xiaopei, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaopei, et al.
Veröffentlicht: (2024)
DiffDoctor: Diagnosing Image Diffusion Models Before Treating
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
LayerFlow: A Unified Model for Layer-aware Video Generation
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
LLMGA: Multimodal Large Language Model based Generation Assistant
von: Xia, Bin, et al.
Veröffentlicht: (2023)
von: Xia, Bin, et al.
Veröffentlicht: (2023)
Beyond Inserting: Learning Identity Embedding for Semantic-Fidelity Personalized Diffusion Generation
von: Li, Yang, et al.
Veröffentlicht: (2024)
von: Li, Yang, et al.
Veröffentlicht: (2024)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
FashionComposer: Compositional Fashion Image Generation
von: Ji, Sihui, et al.
Veröffentlicht: (2024)
von: Ji, Sihui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Modular Customization of Diffusion Models via Blockwise-Parameterized Low-Rank Adaptation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025) -
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
von: Kang, Mingyu, et al.
Veröffentlicht: (2026) -
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2023) -
AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024) -
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
von: Liang, Zichen, et al.
Veröffentlicht: (2025)