Comparison Reveals Commonality: Customized Image Generation through Contrastive Inversion
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Minseo, Kwon, Minchan, Lee, Dongyeun, Jeon, Yunho, Kim, Junmo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
by: Kim, Minseo, et al.
Published: (2026)
by: Kim, Minseo, et al.
Published: (2026)
MDS-DETR: DETR with Masked Duplicate Suppressor
by: Lee, Chanho, et al.
Published: (2026)
by: Lee, Chanho, et al.
Published: (2026)
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
by: Kwon, Minchan, et al.
Published: (2026)
by: Kwon, Minchan, et al.
Published: (2026)
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
by: Hur, Jiwan, et al.
Published: (2024)
by: Hur, Jiwan, et al.
Published: (2024)
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
by: Lee, Chanho, et al.
Published: (2023)
by: Lee, Chanho, et al.
Published: (2023)
Inlier-Centric Post-Training Quantization for Object Detection Models
by: Kim, Minsu, et al.
Published: (2026)
by: Kim, Minsu, et al.
Published: (2026)
SFLD: Reducing the content bias for AI-generated Image Detection
by: Gye, Seoyeon, et al.
Published: (2025)
by: Gye, Seoyeon, et al.
Published: (2025)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
by: Lee, Dongyeun, et al.
Published: (2025)
by: Lee, Dongyeun, et al.
Published: (2025)
Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation
by: Kwon, Joohyun, et al.
Published: (2025)
by: Kwon, Joohyun, et al.
Published: (2025)
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
by: Lee, Hansang, et al.
Published: (2017)
by: Lee, Hansang, et al.
Published: (2017)
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
by: Lee, Haeil, et al.
Published: (2024)
by: Lee, Haeil, et al.
Published: (2024)
Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation
by: Kim, Hyunsoo, et al.
Published: (2025)
by: Kim, Hyunsoo, et al.
Published: (2025)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
by: Lee, Gayoung, et al.
Published: (2025)
by: Lee, Gayoung, et al.
Published: (2025)
SSG: Scaled Spatial Guidance for Multi-Scale Visual Autoregressive Generation
by: Shin, Youngwoo, et al.
Published: (2026)
by: Shin, Youngwoo, et al.
Published: (2026)
3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation
by: Kim, Hwidong, et al.
Published: (2026)
by: Kim, Hwidong, et al.
Published: (2026)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
by: Lee, Hansang, et al.
Published: (2022)
by: Lee, Hansang, et al.
Published: (2022)
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
by: Lee, Hansang, et al.
Published: (2022)
by: Lee, Hansang, et al.
Published: (2022)
The Effects of Mixed Sample Data Augmentation are Class Dependent
by: Lee, Haeil, et al.
Published: (2023)
by: Lee, Haeil, et al.
Published: (2023)
Diffusion-driven GAN Inversion for Multi-Modal Face Image Generation
by: Kim, Jihyun, et al.
Published: (2024)
by: Kim, Jihyun, et al.
Published: (2024)
Inv-Adapter: ID Customization Generation via Image Inversion and Lightweight Adapter
by: Xing, Peng, et al.
Published: (2024)
by: Xing, Peng, et al.
Published: (2024)
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
by: Han, Gyojin, et al.
Published: (2026)
by: Han, Gyojin, et al.
Published: (2026)
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
by: Han, Gyojin, et al.
Published: (2026)
by: Han, Gyojin, et al.
Published: (2026)
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts
by: Hwang, Taebaek, et al.
Published: (2025)
by: Hwang, Taebaek, et al.
Published: (2025)
DECOR:Decomposition and Projection of Text Embeddings for Text-to-Image Customization
by: Jang, Geonhui, et al.
Published: (2024)
by: Jang, Geonhui, et al.
Published: (2024)
Beyond Semantics: Disentangling Information Scope in Sparse Autoencoders for CLIP
by: Ro, Yusung, et al.
Published: (2026)
by: Ro, Yusung, et al.
Published: (2026)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
by: Jin, Hyun-Jun, et al.
Published: (2025)
by: Jin, Hyun-Jun, et al.
Published: (2025)
IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation
by: Baek, Sunghyun, et al.
Published: (2026)
by: Baek, Sunghyun, et al.
Published: (2026)
Modeling Stereo-Confidence Out of the End-to-End Stereo-Matching Network via Disparity Plane Sweep
by: Lee, Jae Young, et al.
Published: (2024)
by: Lee, Jae Young, et al.
Published: (2024)
Stereo-Matching Knowledge Distilled Monocular Depth Estimation Filtered by Multiple Disparity Consistency
by: Ka, Woonghyun, et al.
Published: (2024)
by: Ka, Woonghyun, et al.
Published: (2024)
Motion Inversion for Video Customization
by: Wang, Luozhou, et al.
Published: (2024)
by: Wang, Luozhou, et al.
Published: (2024)
Inspecting Explainability of Transformer Models with Additional Statistical Information
by: Nguyen, Hoang C., et al.
Published: (2023)
by: Nguyen, Hoang C., et al.
Published: (2023)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
by: Lee, Dong-Jae, et al.
Published: (2026)
by: Lee, Dong-Jae, et al.
Published: (2026)
CustomContrast: A Multilevel Contrastive Perspective For Subject-Driven Text-to-Image Customization
by: Chen, Nan, et al.
Published: (2024)
by: Chen, Nan, et al.
Published: (2024)
UniSpector: Towards Universal Open-set Defect Recognition via Spectral-Contrastive Visual Prompting
by: Kim, Geonuk, et al.
Published: (2026)
by: Kim, Geonuk, et al.
Published: (2026)
Grid Diffusion Models for Text-to-Video Generation
by: Lee, Taegyeong, et al.
Published: (2024)
by: Lee, Taegyeong, et al.
Published: (2024)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
A Simple Baseline with Single-encoder for Referring Image Segmentation
by: Yu, Seonghoon, et al.
Published: (2024)
by: Yu, Seonghoon, et al.
Published: (2024)
ScoreCL: Augmentation-Adaptive Contrastive Learning via Score-Matching Function
by: Kim, Jin-Young, et al.
Published: (2023)
by: Kim, Jin-Young, et al.
Published: (2023)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
by: Kim, Minchan, et al.
Published: (2024)
by: Kim, Minchan, et al.
Published: (2024)
CANVAS: Commonsense-Aware Navigation System for Intuitive Human-Robot Interaction
by: Choi, Suhwan, et al.
Published: (2024)
by: Choi, Suhwan, et al.
Published: (2024)
Similar Items
-
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
by: Kim, Minseo, et al.
Published: (2026) -
MDS-DETR: DETR with Masked Duplicate Suppressor
by: Lee, Chanho, et al.
Published: (2026) -
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
by: Kwon, Minchan, et al.
Published: (2026) -
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
by: Hur, Jiwan, et al.
Published: (2024) -
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
by: Lee, Chanho, et al.
Published: (2023)