Fine-Tuning Stable Diffusion XL for Stylistic Icon Generation: A Comparison of Caption Size
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sultan, Youssef, Ma, Jiangqin, Liao, Yu-Ying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Image-to-Recipe Translation
von: Ma, Jiangqin, et al.
Veröffentlicht: (2024)
von: Ma, Jiangqin, et al.
Veröffentlicht: (2024)
Stylistic Attribute Control in Latent Diffusion Models
von: Reimann, Max, et al.
Veröffentlicht: (2026)
von: Reimann, Max, et al.
Veröffentlicht: (2026)
Efficient Image Restoration through Low-Rank Adaptation and Stable Diffusion XL
von: Zhao, Haiyang
Veröffentlicht: (2024)
von: Zhao, Haiyang
Veröffentlicht: (2024)
Generalized SAM: Efficient Fine-Tuning of SAM for Variable Input Image Sizes
von: Kato, Sota, et al.
Veröffentlicht: (2024)
von: Kato, Sota, et al.
Veröffentlicht: (2024)
Stylecodes: Encoding Stylistic Information For Image Generation
von: Rowles, Ciara
Veröffentlicht: (2024)
von: Rowles, Ciara
Veröffentlicht: (2024)
FeRA: Frequency-Energy Constrained Routing for Effective Diffusion Adaptation Fine-Tuning
von: Yin, Bo, et al.
Veröffentlicht: (2025)
von: Yin, Bo, et al.
Veröffentlicht: (2025)
DeepIcon: A Hierarchical Network for Layer-wise Icon Vectorization
von: Bing, Qi, et al.
Veröffentlicht: (2024)
von: Bing, Qi, et al.
Veröffentlicht: (2024)
Detecting Dataset Abuse in Fine-Tuning Stable Diffusion Models for Text-to-Image Synthesis
von: Wang, Songrui, et al.
Veröffentlicht: (2024)
von: Wang, Songrui, et al.
Veröffentlicht: (2024)
Diffusion-RSCC: Diffusion Probabilistic Model for Change Captioning in Remote Sensing Images
von: Yu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Yu, Xiaofei, et al.
Veröffentlicht: (2024)
EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
DomainStudio: Fine-Tuning Diffusion Models for Domain-Driven Image Generation using Limited Data
von: Zhu, Jingyuan, et al.
Veröffentlicht: (2023)
von: Zhu, Jingyuan, et al.
Veröffentlicht: (2023)
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
Gungnir: Exploiting Stylistic Features in Images for Backdoor Attacks on Diffusion Models
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
Quantitative Comparison of Fine-Tuning Techniques for Pretrained Latent Diffusion Models in the Generation of Unseen SAR Images
von: Debuysère, Solène, et al.
Veröffentlicht: (2025)
von: Debuysère, Solène, et al.
Veröffentlicht: (2025)
Sparse Fine-Tuning of Transformers for Generative Tasks
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
Progressive Knowledge Distillation Of Stable Diffusion XL Using Layer Level Loss
von: Gupta, Yatharth, et al.
Veröffentlicht: (2024)
von: Gupta, Yatharth, et al.
Veröffentlicht: (2024)
CycleCap: Improving VLMs Captioning Performance via Self-Supervised Cycle Consistency Fine-Tuning
von: Krestenitis, Marios, et al.
Veröffentlicht: (2026)
von: Krestenitis, Marios, et al.
Veröffentlicht: (2026)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
Unconditional Priors Matter! Improving Conditional Generation of Fine-Tuned Diffusion Models
von: Phunyaphibarn, Prin, et al.
Veröffentlicht: (2025)
von: Phunyaphibarn, Prin, et al.
Veröffentlicht: (2025)
COCONut-PanCap: Joint Panoptic Segmentation and Grounded Captions for Fine-Grained Understanding and Generation
von: Deng, Xueqing, et al.
Veröffentlicht: (2025)
von: Deng, Xueqing, et al.
Veröffentlicht: (2025)
Peregrine: One-Shot Fine-Tuning for FHE Inference of General Deep CNNs
von: Ling, Huaming, et al.
Veröffentlicht: (2025)
von: Ling, Huaming, et al.
Veröffentlicht: (2025)
Only-Style: Stylistic Consistency in Image Generation without Content Leakage
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2025)
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2025)
MultiModal Fine-tuning with Synthetic Captions
von: Enomoto, Shohei, et al.
Veröffentlicht: (2026)
von: Enomoto, Shohei, et al.
Veröffentlicht: (2026)
Synthetically Trained Icon Proposals for Parsing and Summarizing Infographics
von: Madan, Spandan, et al.
Veröffentlicht: (2018)
von: Madan, Spandan, et al.
Veröffentlicht: (2018)
Fine-Tuning Text-To-Image Diffusion Models for Class-Wise Spurious Feature Generation
von: MaungMaung, AprilPyone, et al.
Veröffentlicht: (2024)
von: MaungMaung, AprilPyone, et al.
Veröffentlicht: (2024)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
von: Zhou, Hongyu, et al.
Veröffentlicht: (2026)
von: Zhou, Hongyu, et al.
Veröffentlicht: (2026)
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
von: Tang, Yunlong, et al.
Veröffentlicht: (2025)
von: Tang, Yunlong, et al.
Veröffentlicht: (2025)
Disentangling Fine-Tuning from Pre-Training in Visual Captioning with Hybrid Markov Logic
von: Shah, Monika, et al.
Veröffentlicht: (2025)
von: Shah, Monika, et al.
Veröffentlicht: (2025)
Bridging the Visual Gap: Fine-Tuning Multimodal Models with Knowledge-Adapted Captions
von: Yanuka, Moran, et al.
Veröffentlicht: (2024)
von: Yanuka, Moran, et al.
Veröffentlicht: (2024)
DECap: Towards Generalized Explicit Caption Editing via Diffusion Mechanism
von: Wang, Zhen, et al.
Veröffentlicht: (2023)
von: Wang, Zhen, et al.
Veröffentlicht: (2023)
Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Scene Graph-guided SegCaptioning Transformer with Fine-grained Alignment for Controllable Video Segmentation and Captioning
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
FingerCap: Fine-grained Finger-level Hand Motion Captioning
von: Shen, Xin, et al.
Veröffentlicht: (2025)
von: Shen, Xin, et al.
Veröffentlicht: (2025)
Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-Tuning
von: Liu, Jian, et al.
Veröffentlicht: (2025)
von: Liu, Jian, et al.
Veröffentlicht: (2025)
Grounding Stylistic Domain Generalization with Quantitative Domain Shift Measures and Synthetic Scene Images
von: Luo, Yiran, et al.
Veröffentlicht: (2024)
von: Luo, Yiran, et al.
Veröffentlicht: (2024)
MeshXL: Neural Coordinate Field for Generative 3D Foundation Models
von: Chen, Sijin, et al.
Veröffentlicht: (2024)
von: Chen, Sijin, et al.
Veröffentlicht: (2024)
VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL
von: Dai, Fengyuan, et al.
Veröffentlicht: (2025)
von: Dai, Fengyuan, et al.
Veröffentlicht: (2025)
Image-Caption Encoding for Improving Zero-Shot Generalization
von: Yu, Eric Yang, et al.
Veröffentlicht: (2024)
von: Yu, Eric Yang, et al.
Veröffentlicht: (2024)
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
Fine Tuning Text-to-Image Diffusion Models for Correcting Anomalous Images
von: Yoo, Hyunwoo
Veröffentlicht: (2024)
von: Yoo, Hyunwoo
Veröffentlicht: (2024)
Ähnliche Einträge
-
Deep Image-to-Recipe Translation
von: Ma, Jiangqin, et al.
Veröffentlicht: (2024) -
Stylistic Attribute Control in Latent Diffusion Models
von: Reimann, Max, et al.
Veröffentlicht: (2026) -
Efficient Image Restoration through Low-Rank Adaptation and Stable Diffusion XL
von: Zhao, Haiyang
Veröffentlicht: (2024) -
Generalized SAM: Efficient Fine-Tuning of SAM for Variable Input Image Sizes
von: Kato, Sota, et al.
Veröffentlicht: (2024) -
Stylecodes: Encoding Stylistic Information For Image Generation
von: Rowles, Ciara
Veröffentlicht: (2024)