AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jinghao, Qian, Wen, Luo, Hao, Wang, Fan, Zhao, Feng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
FreeTuner: Any Subject in Any Style with Training-free Diffusion
von: Xu, Youcan, et al.
Veröffentlicht: (2024)
von: Xu, Youcan, et al.
Veröffentlicht: (2024)
DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing
von: Wang, Weitao, et al.
Veröffentlicht: (2025)
von: Wang, Weitao, et al.
Veröffentlicht: (2025)
Prototype Clustered Diffusion Models for Versatile Inverse Problems
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
AVID: Any-Length Video Inpainting with Diffusion Model
von: Zhang, Zhixing, et al.
Veröffentlicht: (2023)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2023)
AnyDressing: Customizable Multi-Garment Virtual Dressing via Latent Diffusion Models
von: Li, Xinghui, et al.
Veröffentlicht: (2024)
von: Li, Xinghui, et al.
Veröffentlicht: (2024)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
FuseAnyPart: Diffusion-Driven Facial Parts Swapping via Multiple Reference Images
von: Yu, Zheng, et al.
Veröffentlicht: (2024)
von: Yu, Zheng, et al.
Veröffentlicht: (2024)
Uncovering the Text Embedding in Text-to-Image Diffusion Models
von: Yu, Hu, et al.
Veröffentlicht: (2024)
von: Yu, Hu, et al.
Veröffentlicht: (2024)
Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image Dehazing
von: Yang, Zizheng, et al.
Veröffentlicht: (2025)
von: Yang, Zizheng, et al.
Veröffentlicht: (2025)
AnyI2V: Animating Any Conditional Image with Motion Control
von: Li, Ziye, et al.
Veröffentlicht: (2025)
von: Li, Ziye, et al.
Veröffentlicht: (2025)
Decomposition Ascribed Synergistic Learning for Unified Image Restoration
von: Zhang, Jinghao, et al.
Veröffentlicht: (2023)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2023)
DualFast: Dual-Speedup Framework for Fast Sampling of Diffusion Models
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
von: Yang, Yang, et al.
Veröffentlicht: (2025)
von: Yang, Yang, et al.
Veröffentlicht: (2025)
EchoGen: Generating Visual Echoes in Any Scene via Feed-Forward Subject-Driven Auto-Regressive Model
von: Dong, Ruixiao, et al.
Veröffentlicht: (2025)
von: Dong, Ruixiao, et al.
Veröffentlicht: (2025)
Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
von: Chan, Kelvin C. K., et al.
Veröffentlicht: (2024)
von: Chan, Kelvin C. K., et al.
Veröffentlicht: (2024)
SSR-Encoder: Encoding Selective Subject Representation for Subject-Driven Generation
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2023)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2023)
SwapAnyone: Consistent and Realistic Video Synthesis for Swapping Any Person into Any Video
von: Zhao, Chengshu, et al.
Veröffentlicht: (2025)
von: Zhao, Chengshu, et al.
Veröffentlicht: (2025)
StoryTailor:A Zero-Shot Pipeline for Action-Rich Multi-Subject Visual Narratives
von: Hu, Jinghao, et al.
Veröffentlicht: (2026)
von: Hu, Jinghao, et al.
Veröffentlicht: (2026)
Any-to-3D Generation via Hybrid Diffusion Supervision
von: Fan, Yijun, et al.
Veröffentlicht: (2024)
von: Fan, Yijun, et al.
Veröffentlicht: (2024)
AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
Depth Anything with Any Prior
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
Drive Any Mesh: 4D Latent Diffusion for Mesh Deformation from Video
von: Shi, Yahao, et al.
Veröffentlicht: (2025)
von: Shi, Yahao, et al.
Veröffentlicht: (2025)
AIComposer: Any Style and Content Image Composition via Feature Integration
von: Li, Haowen, et al.
Veröffentlicht: (2025)
von: Li, Haowen, et al.
Veröffentlicht: (2025)
Compress Any Segment Anything Model (SAM)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
von: Chen, Jinshu, et al.
Veröffentlicht: (2025)
von: Chen, Jinshu, et al.
Veröffentlicht: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
AniDoc: Animation Creation Made Easier
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
Qihoo-T2X: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Any-Task
von: Wang, Jing, et al.
Veröffentlicht: (2024)
von: Wang, Jing, et al.
Veröffentlicht: (2024)
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
SAP: Segment Any 4K Panorama
von: Jiang, Lutao, et al.
Veröffentlicht: (2026)
von: Jiang, Lutao, et al.
Veröffentlicht: (2026)
Stable at Any Speed: Speed-Driven Multi-Object Tracking with Learnable Kalman Filtering
von: Gong, Yan, et al.
Veröffentlicht: (2025)
von: Gong, Yan, et al.
Veröffentlicht: (2025)
AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation
von: Wu, Zijie, et al.
Veröffentlicht: (2026)
von: Wu, Zijie, et al.
Veröffentlicht: (2026)
AnimateAnyMesh: A Feed-Forward 4D Foundation Model for Text-Driven Universal Mesh Animation
von: Wu, Zijie, et al.
Veröffentlicht: (2025)
von: Wu, Zijie, et al.
Veröffentlicht: (2025)
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
AnyTrans: Translate AnyText in the Image with Large Scale Models
von: Qian, Zhipeng, et al.
Veröffentlicht: (2024)
von: Qian, Zhipeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024) -
FreeTuner: Any Subject in Any Style with Training-free Diffusion
von: Xu, Youcan, et al.
Veröffentlicht: (2024) -
DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing
von: Wang, Weitao, et al.
Veröffentlicht: (2025) -
Prototype Clustered Diffusion Models for Versatile Inverse Problems
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024) -
AVID: Any-Length Video Inpainting with Diffusion Model
von: Zhang, Zhixing, et al.
Veröffentlicht: (2023)