Eye-for-an-eye: Appearance Transfer with Semantic Correspondence in Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Go, Sooyeon, Choi, Kyungmook, Shin, Minjung, Uh, Youngjung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MVCustom: Multi-View Customized Diffusion via Geometric Latent Rendering and Completion
von: Shin, Minjung, et al.
Veröffentlicht: (2025)
von: Shin, Minjung, et al.
Veröffentlicht: (2025)
Semantic Image Synthesis with Unconditional Generator
von: Chae, Jungwoo, et al.
Veröffentlicht: (2024)
von: Chae, Jungwoo, et al.
Veröffentlicht: (2024)
ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation
von: Kim, Shin Seong, et al.
Veröffentlicht: (2025)
von: Kim, Shin Seong, et al.
Veröffentlicht: (2025)
Attribute Based Interpretable Evaluation Metrics for Generative Models
von: Kim, Dongkyun, et al.
Veröffentlicht: (2023)
von: Kim, Dongkyun, et al.
Veröffentlicht: (2023)
Training-free Content Injection using h-space in Diffusion Models
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2023)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2023)
Geometric Disentanglement of Text Embeddings for Subject-Consistent Text-to-Image Generation using A Single Prompt
von: Li, Shangxun, et al.
Veröffentlicht: (2025)
von: Li, Shangxun, et al.
Veröffentlicht: (2025)
HARIVO: Harnessing Text-to-Image Models for Video Generation
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
Syncphony: Synchronized Audio-to-Video Generation with Diffusion Transformers
von: Song, Jibin, et al.
Veröffentlicht: (2025)
von: Song, Jibin, et al.
Veröffentlicht: (2025)
Addressing Negative Transfer in Diffusion Models
von: Go, Hyojun, et al.
Veröffentlicht: (2023)
von: Go, Hyojun, et al.
Veröffentlicht: (2023)
Frequency-Adaptive Sharpness Regularization for Improving 3D Gaussian Splatting Generalization
von: Yun, Youngsik, et al.
Veröffentlicht: (2025)
von: Yun, Youngsik, et al.
Veröffentlicht: (2025)
CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer
von: Nie, Wenbo, et al.
Veröffentlicht: (2026)
von: Nie, Wenbo, et al.
Veröffentlicht: (2026)
Semantic Guidance Tuning for Text-To-Image Diffusion Models
von: Kang, Hyun, et al.
Veröffentlicht: (2023)
von: Kang, Hyun, et al.
Veröffentlicht: (2023)
FlowBlending: Stage-Aware Multi-Model Sampling for Fast and High-Fidelity Video Generation
von: Song, Jibin, et al.
Veröffentlicht: (2025)
von: Song, Jibin, et al.
Veröffentlicht: (2025)
ZERO: Industry-ready Vision Foundation Model with Multi-modal Prompts
von: Choi, Sangbum, et al.
Veröffentlicht: (2025)
von: Choi, Sangbum, et al.
Veröffentlicht: (2025)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
von: Kim, Jisoo, et al.
Veröffentlicht: (2025)
von: Kim, Jisoo, et al.
Veröffentlicht: (2025)
Visual Style Prompting with Swapping Self-Attention
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2024)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2024)
StyleKeeper: Prevent Content Leakage using Negative Visual Query Guidance
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
CogME: A Cognition-Inspired Multi-Dimensional Evaluation Metric for Story Understanding
von: Shin, Minjung, et al.
Veröffentlicht: (2021)
von: Shin, Minjung, et al.
Veröffentlicht: (2021)
TetraSDF: Precise Mesh Extraction with Multi-resolution Tetrahedral Grid
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
von: Deng, Fan, et al.
Veröffentlicht: (2024)
von: Deng, Fan, et al.
Veröffentlicht: (2024)
Denoising Task Routing for Diffusion Models
von: Park, Byeongjun, et al.
Veröffentlicht: (2023)
von: Park, Byeongjun, et al.
Veröffentlicht: (2023)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
SimpleMatch: A Simple and Strong Baseline for Semantic Correspondence
von: Jin, Hailing, et al.
Veröffentlicht: (2026)
von: Jin, Hailing, et al.
Veröffentlicht: (2026)
CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird's-Eye-View Semantic Segmentation
von: Hong, Jeongbin, et al.
Veröffentlicht: (2026)
von: Hong, Jeongbin, et al.
Veröffentlicht: (2026)
When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Detection
von: Kim, Jihyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jihyeon, et al.
Veröffentlicht: (2026)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
TCFG: Tangential Damping Classifier-free Guidance
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
Diffusion Model Patching via Mixture-of-Prompts
von: Ham, Seokil, et al.
Veröffentlicht: (2024)
von: Ham, Seokil, et al.
Veröffentlicht: (2024)
Improving Bird's Eye View Semantic Segmentation by Task Decomposition
von: Zhao, Tianhao, et al.
Veröffentlicht: (2024)
von: Zhao, Tianhao, et al.
Veröffentlicht: (2024)
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
von: Sarkar, Sayan Deb, et al.
Veröffentlicht: (2025)
von: Sarkar, Sayan Deb, et al.
Veröffentlicht: (2025)
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
von: Gong, Linrui, et al.
Veröffentlicht: (2024)
von: Gong, Linrui, et al.
Veröffentlicht: (2024)
Open-Set Domain Adaptation for Semantic Segmentation
von: Choe, Seun-An, et al.
Veröffentlicht: (2024)
von: Choe, Seun-An, et al.
Veröffentlicht: (2024)
Stylistic-STORM (ST-STORM) : Perceiving the Semantic Nature of Appearance
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2024)
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2024)
Towards Continual Expansion of Data Coverage: Automatic Text-guided Edge-case Synthesis
von: Go, Kyeongryeol
Veröffentlicht: (2025)
von: Go, Kyeongryeol
Veröffentlicht: (2025)
Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball Rotation
von: Choi, YoungChan, et al.
Veröffentlicht: (2025)
von: Choi, YoungChan, et al.
Veröffentlicht: (2025)
OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
MambaEye: A Size-Agnostic Visual Encoder with Causal Sequential Processing
von: Choi, Changho, et al.
Veröffentlicht: (2025)
von: Choi, Changho, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MVCustom: Multi-View Customized Diffusion via Geometric Latent Rendering and Completion
von: Shin, Minjung, et al.
Veröffentlicht: (2025) -
Semantic Image Synthesis with Unconditional Generator
von: Chae, Jungwoo, et al.
Veröffentlicht: (2024) -
ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation
von: Kim, Shin Seong, et al.
Veröffentlicht: (2025) -
Attribute Based Interpretable Evaluation Metrics for Generative Models
von: Kim, Dongkyun, et al.
Veröffentlicht: (2023) -
Training-free Content Injection using h-space in Diffusion Models
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2023)