Illustrious: an Open Advanced Illustration Model
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Sang Hyun, Koh, Jun Young, Lee, Junha, Song, Joy, Kim, Dongha, Moon, Hoyeon, Lee, Hyunju, Song, Min |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Let Triggers Control: Frequency-Aware Dropout for Effective Token Control
by: Koh, Junyoung, et al.
Published: (2026)
by: Koh, Junyoung, et al.
Published: (2026)
CAT: Contrastive Adapter Training for Personalized Image Generation
by: Park, Jae Wan, et al.
Published: (2024)
by: Park, Jae Wan, et al.
Published: (2024)
Improving Text Generation on Images with Synthetic Captions
by: Koh, Jun Young, et al.
Published: (2024)
by: Koh, Jun Young, et al.
Published: (2024)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
by: Lee, Seongyun, et al.
Published: (2024)
by: Lee, Seongyun, et al.
Published: (2024)
Affogato: Learning Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale
by: Lee, Junha, et al.
Published: (2025)
by: Lee, Junha, et al.
Published: (2025)
ABM-LoRA: Activation Boundary Matching for Fast Convergence in Low-Rank Adaptation
by: Lee, Dongha, et al.
Published: (2025)
by: Lee, Dongha, et al.
Published: (2025)
DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning
by: Lee, Junha, et al.
Published: (2026)
by: Lee, Junha, et al.
Published: (2026)
Is user feedback always informative? Retrieval Latent Defending for Semi-Supervised Domain Adaptation without Source Data
by: Song, Junha, et al.
Published: (2024)
by: Song, Junha, et al.
Published: (2024)
SpaCeFormer: Fast Proposal-Free Open-Vocabulary 3D Instance Segmentation
by: Choy, Chris, et al.
Published: (2026)
by: Choy, Chris, et al.
Published: (2026)
Scalp Diagnostic System With Label-Free Segmentation and Training-Free Image Translation
by: Kim, Youngmin, et al.
Published: (2024)
by: Kim, Youngmin, et al.
Published: (2024)
3D Geometric Shape Assembly via Efficient Point Cloud Matching
by: Lee, Nahyuk, et al.
Published: (2024)
by: Lee, Nahyuk, et al.
Published: (2024)
Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation
by: Lee, Junha, et al.
Published: (2025)
by: Lee, Junha, et al.
Published: (2025)
M2SFormer: Multi-Spectral and Multi-Scale Attention with Edge-Aware Difficulty Guidance for Image Forgery Localization
by: Nam, Ju-Hyeon, et al.
Published: (2025)
by: Nam, Ju-Hyeon, et al.
Published: (2025)
SC-Pro: Training-Free Framework for Defending Unsafe Image Synthesis Attack
by: Park, Junha, et al.
Published: (2025)
by: Park, Junha, et al.
Published: (2025)
Online Continuous Generalized Category Discovery
by: Park, Keon-Hee, et al.
Published: (2024)
by: Park, Keon-Hee, et al.
Published: (2024)
Development of Image Collection Method Using YOLO and Siamese Network
by: Shin, Chan Young, et al.
Published: (2024)
by: Shin, Chan Young, et al.
Published: (2024)
Personalized Reward Modeling for Text-to-Image Generation
by: Lee, Jeongeun, et al.
Published: (2025)
by: Lee, Jeongeun, et al.
Published: (2025)
TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts
by: Jeon, Jeimin, et al.
Published: (2026)
by: Jeon, Jeimin, et al.
Published: (2026)
Efficient Few-Shot Neural Architecture Search by Counting the Number of Nonlinear Functions
by: Oh, Youngmin, et al.
Published: (2024)
by: Oh, Youngmin, et al.
Published: (2024)
Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation
by: Chung, Jiwoo, et al.
Published: (2025)
by: Chung, Jiwoo, et al.
Published: (2025)
JUDO: A Juxtaposed Domain-Oriented Multimodal Reasoner for Industrial Anomaly QA
by: Kang, Hyunju, et al.
Published: (2026)
by: Kang, Hyunju, et al.
Published: (2026)
Towards Model-Agnostic Dataset Condensation by Heterogeneous Models
by: Moon, Jun-Yeong, et al.
Published: (2024)
by: Moon, Jun-Yeong, et al.
Published: (2024)
HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting
by: Lee, Jeongeun, et al.
Published: (2025)
by: Lee, Jeongeun, et al.
Published: (2025)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
by: Lee, ByeongCheol, et al.
Published: (2026)
by: Lee, ByeongCheol, et al.
Published: (2026)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
by: Jin, Hyun-Jun, et al.
Published: (2025)
by: Jin, Hyun-Jun, et al.
Published: (2025)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
by: Park, Gilhan, et al.
Published: (2024)
by: Park, Gilhan, et al.
Published: (2024)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
by: Kim, Kinam, et al.
Published: (2025)
by: Kim, Kinam, et al.
Published: (2025)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
by: Lee, SuBeen, et al.
Published: (2025)
by: Lee, SuBeen, et al.
Published: (2025)
Dynamic VLM-Guided Negative Prompting for Diffusion Models
by: Chang, Hoyeon, et al.
Published: (2025)
by: Chang, Hoyeon, et al.
Published: (2025)
Detecting Unknown Objects via Energy-based Separation for Open World Object Detection
by: Heo, Jun-Woo, et al.
Published: (2026)
by: Heo, Jun-Woo, et al.
Published: (2026)
G2P: Gaussian-to-Point Attribute Alignment for Boundary-Aware 3D Semantic Segmentation
by: Song, Hojun, et al.
Published: (2026)
by: Song, Hojun, et al.
Published: (2026)
Recovering Dynamic 3D Sketches from Videos
by: Lee, Jaeah, et al.
Published: (2025)
by: Lee, Jaeah, et al.
Published: (2025)
3Doodle: Compact Abstraction of Objects with 3D Strokes
by: Choi, Changwoon, et al.
Published: (2024)
by: Choi, Changwoon, et al.
Published: (2024)
DETACH : Decomposed Spatio-Temporal Alignment for Exocentric Video and Ambient Sensors with Staged Learning
by: Yoon, Junho, et al.
Published: (2025)
by: Yoon, Junho, et al.
Published: (2025)
Beyond-Labels: Advancing Open-Vocabulary Segmentation With Vision-Language Models
by: Rahman, Muhammad Atta ur, et al.
Published: (2025)
by: Rahman, Muhammad Atta ur, et al.
Published: (2025)
CompSplat: Compression-aware 3D Gaussian Splatting for Real-world Video
by: Song, Hojun, et al.
Published: (2026)
by: Song, Hojun, et al.
Published: (2026)
OnlineBEV: Recurrent Temporal Fusion in Bird's Eye View Representations for Multi-Camera 3D Perception
by: Koh, Junho, et al.
Published: (2025)
by: Koh, Junho, et al.
Published: (2025)
MM-SeR: Multimodal Self-Refinement for Lightweight Image Captioning
by: Song, Junha, et al.
Published: (2025)
by: Song, Junha, et al.
Published: (2025)
Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following
by: Kang, Myeongkyun, et al.
Published: (2026)
by: Kang, Myeongkyun, et al.
Published: (2026)
Generative Unlearning for Any Identity
by: Seo, Juwon, et al.
Published: (2024)
by: Seo, Juwon, et al.
Published: (2024)
Similar Items
-
Let Triggers Control: Frequency-Aware Dropout for Effective Token Control
by: Koh, Junyoung, et al.
Published: (2026) -
CAT: Contrastive Adapter Training for Personalized Image Generation
by: Park, Jae Wan, et al.
Published: (2024) -
Improving Text Generation on Images with Synthetic Captions
by: Koh, Jun Young, et al.
Published: (2024) -
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
by: Lee, Seongyun, et al.
Published: (2024) -
Affogato: Learning Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale
by: Lee, Junha, et al.
Published: (2025)