Training-Free Sketch-Guided Diffusion with Latent Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Ding, Sandra Zhang, Mao, Jiafeng, Aizawa, Kiyoharu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Guided Image Synthesis via Initial Image Editing in Diffusion Model
di: Mao, Jiafeng, et al.
Pubblicazione: (2023)
di: Mao, Jiafeng, et al.
Pubblicazione: (2023)
The Lottery Ticket Hypothesis in Denoising: Towards Semantic-Driven Initialization
di: Mao, Jiafeng, et al.
Pubblicazione: (2023)
di: Mao, Jiafeng, et al.
Pubblicazione: (2023)
Harnessing PDF Data for Improving Japanese Large Multimodal Models
di: Baek, Jeonghun, et al.
Pubblicazione: (2025)
di: Baek, Jeonghun, et al.
Pubblicazione: (2025)
SketchDeco: Training-Free Latent Composition for Precise Sketch Colourisation
di: Utintu, Chaitat, et al.
Pubblicazione: (2024)
di: Utintu, Chaitat, et al.
Pubblicazione: (2024)
Entity-NeRF: Detecting and Removing Moving Entities in Urban Scenes
di: Otonari, Takashi, et al.
Pubblicazione: (2024)
di: Otonari, Takashi, et al.
Pubblicazione: (2024)
Manga109Dialog: A Large-scale Dialogue Dataset for Comics Speaker Detection
di: Li, Yingxuan, et al.
Pubblicazione: (2023)
di: Li, Yingxuan, et al.
Pubblicazione: (2023)
LORE: Latent Optimization for Precise Semantic Control in Rectified Flow-based Image Editing
di: Ouyang, Liangyang, et al.
Pubblicazione: (2025)
di: Ouyang, Liangyang, et al.
Pubblicazione: (2025)
FoodMLLM-JP: Leveraging Multimodal Large Language Models for Japanese Recipe Generation
di: Imajuku, Yuki, et al.
Pubblicazione: (2024)
di: Imajuku, Yuki, et al.
Pubblicazione: (2024)
MangaUB: A Manga Understanding Benchmark for Large Multimodal Models
di: Ikuta, Hikaru, et al.
Pubblicazione: (2024)
di: Ikuta, Hikaru, et al.
Pubblicazione: (2024)
Saliency Guided Optimization of Diffusion Latents
di: Wang, Xiwen, et al.
Pubblicazione: (2024)
di: Wang, Xiwen, et al.
Pubblicazione: (2024)
PerFace: Metric Learning in Perceptual Facial Similarity for Enhanced Face Anonymization
di: Kumagai, Haruka, et al.
Pubblicazione: (2025)
di: Kumagai, Haruka, et al.
Pubblicazione: (2025)
GL-MCM: Global and Local Maximum Concept Matching for Zero-Shot Out-of-Distribution Detection
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2023)
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2023)
EGLOCE: Training-Free Energy-Guided Latent Optimization for Concept Erasure
di: Ahn, Junyeong, et al.
Pubblicazione: (2026)
di: Ahn, Junyeong, et al.
Pubblicazione: (2026)
Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal Fusion
di: Li, Yingxuan, et al.
Pubblicazione: (2024)
di: Li, Yingxuan, et al.
Pubblicazione: (2024)
FoodLogAthl-218: Constructing a Real-World Food Image Dataset Using Dietary Management Applications
di: Watanabe, Mitsuki, et al.
Pubblicazione: (2025)
di: Watanabe, Mitsuki, et al.
Pubblicazione: (2025)
Difficulty Controlled Diffusion Model for Synthesizing Effective Training Data
di: Wang, Zerun, et al.
Pubblicazione: (2024)
di: Wang, Zerun, et al.
Pubblicazione: (2024)
MangaDiT: Reference-Guided Line Art Colorization with Hierarchical Attention in Diffusion Transformers
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
A Benchmark and Evaluation for Real-World Out-of-Distribution Detection Using Vision-Language Models
di: Noda, Shiho, et al.
Pubblicazione: (2025)
di: Noda, Shiho, et al.
Pubblicazione: (2025)
Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation
di: Horita, Daichi, et al.
Pubblicazione: (2023)
di: Horita, Daichi, et al.
Pubblicazione: (2023)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
di: Xing, Ximing, et al.
Pubblicazione: (2023)
di: Xing, Ximing, et al.
Pubblicazione: (2023)
Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch Generation
di: Yang, Rui, et al.
Pubblicazione: (2025)
di: Yang, Rui, et al.
Pubblicazione: (2025)
JMMMU-Pro: Image-based Japanese Multi-discipline Multimodal Understanding Benchmark via Vibe Benchmark Construction
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2025)
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2025)
Decoupling Training-Free Guided Diffusion by ADMM
di: Zhang, Youyuan, et al.
Pubblicazione: (2024)
di: Zhang, Youyuan, et al.
Pubblicazione: (2024)
GaitProtector: Impersonation-Driven Gait De-Identification via Training-Free Diffusion Latent Optimization
di: Duan, Huiran, et al.
Pubblicazione: (2026)
di: Duan, Huiran, et al.
Pubblicazione: (2026)
Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding
di: Baek, Jeonghun, et al.
Pubblicazione: (2026)
di: Baek, Jeonghun, et al.
Pubblicazione: (2026)
Beyond and Free from Diffusion: Invertible Guided Consistency Training
di: Hsu, Chia-Hong, et al.
Pubblicazione: (2025)
di: Hsu, Chia-Hong, et al.
Pubblicazione: (2025)
Training-Free Disentangled Text-Guided Image Editing via Sparse Latent Constraints
di: Shabrina, Mutiara, et al.
Pubblicazione: (2025)
di: Shabrina, Mutiara, et al.
Pubblicazione: (2025)
Harnessing the Latent Diffusion Model for Training-Free Image Style Transfer
di: Masui, Kento, et al.
Pubblicazione: (2024)
di: Masui, Kento, et al.
Pubblicazione: (2024)
Exploring Palette based Color Guidance in Diffusion Models
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs
di: Tong, Jintao, et al.
Pubblicazione: (2025)
di: Tong, Jintao, et al.
Pubblicazione: (2025)
AEROBLADE: Training-Free Detection of Latent Diffusion Images Using Autoencoder Reconstruction Error
di: Ricker, Jonas, et al.
Pubblicazione: (2024)
di: Ricker, Jonas, et al.
Pubblicazione: (2024)
Annotation-Free Human Sketch Quality Assessment
di: Yang, Lan, et al.
Pubblicazione: (2025)
di: Yang, Lan, et al.
Pubblicazione: (2025)
Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model
di: Mao, Qi, et al.
Pubblicazione: (2025)
di: Mao, Qi, et al.
Pubblicazione: (2025)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
di: Yang, Ruolin, et al.
Pubblicazione: (2025)
di: Yang, Ruolin, et al.
Pubblicazione: (2025)
It's All About Your Sketch: Democratising Sketch Control in Diffusion Models
di: Koley, Subhadeep, et al.
Pubblicazione: (2024)
di: Koley, Subhadeep, et al.
Pubblicazione: (2024)
SwiftSketch: A Diffusion Model for Image-to-Vector Sketch Generation
di: Arar, Ellie, et al.
Pubblicazione: (2025)
di: Arar, Ellie, et al.
Pubblicazione: (2025)
Sketch-Guided Scene Image Generation
di: Zhang, Tianyu, et al.
Pubblicazione: (2024)
di: Zhang, Tianyu, et al.
Pubblicazione: (2024)
Jr. AI Scientist and Its Risk Report: Autonomous Scientific Exploration from a Baseline Paper
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2025)
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2025)
LatentINDIGO: An INN-Guided Latent Diffusion Algorithm for Image Restoration
di: You, Di, et al.
Pubblicazione: (2025)
di: You, Di, et al.
Pubblicazione: (2025)
Alias-Free Latent Diffusion Models: Improving Fractional Shift Equivariance of Diffusion Latent Space
di: Zhou, Yifan, et al.
Pubblicazione: (2025)
di: Zhou, Yifan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Guided Image Synthesis via Initial Image Editing in Diffusion Model
di: Mao, Jiafeng, et al.
Pubblicazione: (2023) -
The Lottery Ticket Hypothesis in Denoising: Towards Semantic-Driven Initialization
di: Mao, Jiafeng, et al.
Pubblicazione: (2023) -
Harnessing PDF Data for Improving Japanese Large Multimodal Models
di: Baek, Jeonghun, et al.
Pubblicazione: (2025) -
SketchDeco: Training-Free Latent Composition for Precise Sketch Colourisation
di: Utintu, Chaitat, et al.
Pubblicazione: (2024) -
Entity-NeRF: Detecting and Removing Moving Entities in Urban Scenes
di: Otonari, Takashi, et al.
Pubblicazione: (2024)