KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Navard, Pouyan, Monsefi, Amin Karimi, Zhou, Mengxi, Chao, Wei-Lun, Yilmaz, Alper, Ramnath, Rajiv |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
Controlla: Learning Controllability via Graph-Constrained Latent Geometry
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026)
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026)
SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation
von: Perera, Shehan, et al.
Veröffentlicht: (2024)
von: Perera, Shehan, et al.
Veröffentlicht: (2024)
LLaVA-LE: Large Language-and-Vision Assistant for Lunar Exploration
von: Inal, Gokce, et al.
Veröffentlicht: (2026)
von: Inal, Gokce, et al.
Veröffentlicht: (2026)
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
von: Meyarian, Abolfazl, et al.
Veröffentlicht: (2026)
von: Meyarian, Abolfazl, et al.
Veröffentlicht: (2026)
Loki: Representation over Architecture for Diffusion-Based Portrait Animation
von: Navard, Pouyan, et al.
Veröffentlicht: (2026)
von: Navard, Pouyan, et al.
Veröffentlicht: (2026)
Masked LoGoNet: Fast and Accurate 3D Image Analysis for Medical Domain
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
A Probabilistic-based Drift Correction Module for Visual Inertial SLAMs
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
TaxaAdapter: Vision Taxonomy Models are Key to Fine-grained Image Generation over the Tree of Life
von: Khurana, Mridul, et al.
Veröffentlicht: (2026)
von: Khurana, Mridul, et al.
Veröffentlicht: (2026)
DocParseNet: Advanced Semantic Segmentation and OCR Embeddings for Efficient Scanned Document Annotation
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
CrashFormer: A Multimodal Architecture to Predict the Risk of Crash
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
Sketch & Paint: Stroke-by-Stroke Evolution of Visual Artworks
von: Prudviraj, Jeripothula, et al.
Veröffentlicht: (2025)
von: Prudviraj, Jeripothula, et al.
Veröffentlicht: (2025)
Gen-AI Police Sketches with Stable Diffusion
von: Fidalgo, Nicholas, et al.
Veröffentlicht: (2025)
von: Fidalgo, Nicholas, et al.
Veröffentlicht: (2025)
CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition
von: Peng, Yusen, et al.
Veröffentlicht: (2025)
von: Peng, Yusen, et al.
Veröffentlicht: (2025)
Rethinking the Good Enough Embedding for Easy Few-Shot Learning
von: Karnes, Michael, et al.
Veröffentlicht: (2026)
von: Karnes, Michael, et al.
Veröffentlicht: (2026)
Toward Aristotelian Medical Representations: Backpropagation-Free Layer-wise Analysis for Interpretable Generalized Metric Learning on MedMNIST
von: Karnes, Michael, et al.
Veröffentlicht: (2026)
von: Karnes, Michael, et al.
Veröffentlicht: (2026)
It's All About Your Sketch: Democratising Sketch Control in Diffusion Models
von: Koley, Subhadeep, et al.
Veröffentlicht: (2024)
von: Koley, Subhadeep, et al.
Veröffentlicht: (2024)
DINTR: Tracking via Diffusion-based Interpolation
von: Nguyen, Pha, et al.
Veröffentlicht: (2024)
von: Nguyen, Pha, et al.
Veröffentlicht: (2024)
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
von: Zhan, Ruohao, et al.
Veröffentlicht: (2025)
von: Zhan, Ruohao, et al.
Veröffentlicht: (2025)
SketchPlan: Diffusion Based Drone Planning From Human Sketches
von: Norelius, Sixten, et al.
Veröffentlicht: (2025)
von: Norelius, Sixten, et al.
Veröffentlicht: (2025)
Lightweight Road Environment Segmentation using Vector Quantization
von: Kwag, Jiyong, et al.
Veröffentlicht: (2025)
von: Kwag, Jiyong, et al.
Veröffentlicht: (2025)
ARIAL: An Agentic Framework for Document VQA with Precise Answer Localization
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
MGA-VQA: Secure and Interpretable Graph-Augmented Visual Question Answering with Memory-Guided Protection Against Unauthorized Knowledge Use
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
AnyControl: Create Your Artwork with Versatile Control on Text-to-Image Generation
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
MotivNet: Evolving Meta-Sapiens into an Emotionally Intelligent Foundation Model
von: Medicharla, Rahul, et al.
Veröffentlicht: (2025)
von: Medicharla, Rahul, et al.
Veröffentlicht: (2025)
UAS Visual Navigation in Large and Unseen Environments via a Meta Agent
von: Han, Yuci, et al.
Veröffentlicht: (2025)
von: Han, Yuci, et al.
Veröffentlicht: (2025)
Disruptive Transformation of Artworks in Master-Disciple Relationships: The Case of Ukiyo-e Artworks
von: Shinichi, Honna, et al.
Veröffentlicht: (2025)
von: Shinichi, Honna, et al.
Veröffentlicht: (2025)
DLaVA: Document Language and Vision Assistant for Answer Localization with Enhanced Interpretability and Trustworthiness
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
Stroke of Surprise: Progressive Semantic Illusions in Vector Sketching
von: Cheng, Huai-Hsun, et al.
Veröffentlicht: (2026)
von: Cheng, Huai-Hsun, et al.
Veröffentlicht: (2026)
SwiftSketch: A Diffusion Model for Image-to-Vector Sketch Generation
von: Arar, Ellie, et al.
Veröffentlicht: (2025)
von: Arar, Ellie, et al.
Veröffentlicht: (2025)
TexControl: Sketch-Based Two-Stage Fashion Image Generation Using Diffusion Model
von: Zhang, Yongming, et al.
Veröffentlicht: (2024)
von: Zhang, Yongming, et al.
Veröffentlicht: (2024)
Detecting AI-generated Artwork
von: Li, Meien, et al.
Veröffentlicht: (2025)
von: Li, Meien, et al.
Veröffentlicht: (2025)
Data Augmentation via Latent Diffusion Models for Detecting Smell-Related Objects in Historical Artworks
von: Sheta, Ahmed, et al.
Veröffentlicht: (2025)
von: Sheta, Ahmed, et al.
Veröffentlicht: (2025)
DSV-LFS: Unifying LLM-Driven Semantic Cues with Visual Features for Robust Few-Shot Segmentation
von: Karimi, Amin, et al.
Veröffentlicht: (2025)
von: Karimi, Amin, et al.
Veröffentlicht: (2025)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
von: Liu, An-Lun, et al.
Veröffentlicht: (2025)
von: Liu, An-Lun, et al.
Veröffentlicht: (2025)
A Fusion Model for Artwork Identification Based on Convolutional Neural Networks and Transformers
von: Wang, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024) -
Controlla: Learning Controllability via Graph-Constrained Latent Geometry
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026) -
SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026) -
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation
von: Perera, Shehan, et al.
Veröffentlicht: (2024) -
LLaVA-LE: Large Language-and-Vision Assistant for Lunar Exploration
von: Inal, Gokce, et al.
Veröffentlicht: (2026)