Edge-Aware Image Manipulation via Diffusion Models with a Novel Structure-Preservation Loss
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Minsu, Ryu, Nuri, Ok, Jungseul, Cho, Sunghyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Addressing Text Embedding Leakage in Diffusion-based Image Editing
von: Mun, Sunung, et al.
Veröffentlicht: (2024)
von: Mun, Sunung, et al.
Veröffentlicht: (2024)
EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation
von: Mun, Sunung, et al.
Veröffentlicht: (2026)
von: Mun, Sunung, et al.
Veröffentlicht: (2026)
Elevating 3D Models: High-Quality Texture and Geometry Refinement from a Low-Quality Model
von: Ryu, Nuri, et al.
Veröffentlicht: (2025)
von: Ryu, Nuri, et al.
Veröffentlicht: (2025)
CLIPtone: Unsupervised Learning for Text-based Image Tone Adjustment
von: Lee, Hyeongmin, et al.
Veröffentlicht: (2024)
von: Lee, Hyeongmin, et al.
Veröffentlicht: (2024)
POS-ISP: Pipeline Optimization at the Sequence Level for Task-aware ISP
von: Won, Jiyun, et al.
Veröffentlicht: (2026)
von: Won, Jiyun, et al.
Veröffentlicht: (2026)
Diffusion Model Compression for Image-to-Image Translation
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
ActFusion: a Unified Diffusion Model for Action Segmentation and Anticipation
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
Active Label Correction for Semantic Segmentation with Foundation Models
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
Burst Image Super-Resolution with Base Frame Selection
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
BFS: Back-to-Front Layered Image Synthesis via Knowledge Transfer
von: Kang, Kyoungkook, et al.
Veröffentlicht: (2026)
von: Kang, Kyoungkook, et al.
Veröffentlicht: (2026)
Generic Event Boundary Detection via Denoising Diffusion
von: Hwang, Jaejun, et al.
Veröffentlicht: (2025)
von: Hwang, Jaejun, et al.
Veröffentlicht: (2025)
Active Prompt Learning with Vision-Language Model Priors
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
von: Chung, Chaeyeon, et al.
Veröffentlicht: (2024)
von: Chung, Chaeyeon, et al.
Veröffentlicht: (2024)
How Do Large Vision-Language Models See Text in Image? Unveiling the Distinctive Role of OCR Heads
von: Baek, Ingeol, et al.
Veröffentlicht: (2025)
von: Baek, Ingeol, et al.
Veröffentlicht: (2025)
Video Summarization with Large Language Models
von: Lee, Min Jung, et al.
Veröffentlicht: (2025)
von: Lee, Min Jung, et al.
Veröffentlicht: (2025)
Leveraging Learned Image Prior for 3D Gaussian Compression
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image
von: Jeong, Yoonwoo, et al.
Veröffentlicht: (2023)
von: Jeong, Yoonwoo, et al.
Veröffentlicht: (2023)
Locality-Aware Zero-Shot Human-Object Interaction Detection
von: Kim, Sanghyun, et al.
Veröffentlicht: (2025)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2025)
Generalizable Novel-View Synthesis using a Stereo Camera
von: Lee, Haechan, et al.
Veröffentlicht: (2024)
von: Lee, Haechan, et al.
Veröffentlicht: (2024)
Part-Aware Bottom-Up Group Reasoning for Fine-Grained Social Interaction Detection
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
ChimeraLoRA: Multi-Head LoRA-Guided Synthetic Datasets
von: Kim, Hoyoung, et al.
Veröffentlicht: (2026)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2026)
VIRO: Robust and Efficient Neuro-Symbolic Reasoning with Verification for Referring Expression Comprehension
von: Park, Hyejin, et al.
Veröffentlicht: (2026)
von: Park, Hyejin, et al.
Veröffentlicht: (2026)
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
von: Kim, Seungwook, et al.
Veröffentlicht: (2024)
Dynamic Exposure Burst Image Restoration
von: Kim, Woohyeok, et al.
Veröffentlicht: (2026)
von: Kim, Woohyeok, et al.
Veröffentlicht: (2026)
Degradation-Aware and Structure-Preserving Diffusion for Real-World Image Super-Resolution
von: Ji, Yang, et al.
Veröffentlicht: (2026)
von: Ji, Yang, et al.
Veröffentlicht: (2026)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
von: Kim, Seungwook, et al.
Veröffentlicht: (2026)
FloVD: Optical Flow Meets Video Diffusion Model for Enhanced Camera-Controlled Video Synthesis
von: Jin, Wonjoon, et al.
Veröffentlicht: (2025)
von: Jin, Wonjoon, et al.
Veröffentlicht: (2025)
In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation
von: Kang, Dahyun, et al.
Veröffentlicht: (2024)
von: Kang, Dahyun, et al.
Veröffentlicht: (2024)
Leveraging 3D Geometric Priors in 2D Rotation Symmetry Detection
von: Seo, Ahyun, et al.
Veröffentlicht: (2025)
von: Seo, Ahyun, et al.
Veröffentlicht: (2025)
LayeringDiff: Layered Image Synthesis via Generation, then Disassembly with Generative Knowledge
von: Kang, Kyoungkook, et al.
Veröffentlicht: (2025)
von: Kang, Kyoungkook, et al.
Veröffentlicht: (2025)
Steering Guidance for Personalized Text-to-Image Diffusion Models
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
Harnessing the Power of Training-Free Techniques in Text-to-2D Generation for Text-to-3D Generation via Score Distillation Sampling
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
Geometry-Aware Losses for Structure-Preserving Text-to-Sign Language Generation
von: Wu, Zetian, et al.
Veröffentlicht: (2025)
von: Wu, Zetian, et al.
Veröffentlicht: (2025)
Visual Prompt Discovery via Semantic Exploration
von: Kim, Jaechang, et al.
Veröffentlicht: (2026)
von: Kim, Jaechang, et al.
Veröffentlicht: (2026)
DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models
von: Ryu, Hyogon, et al.
Veröffentlicht: (2025)
von: Ryu, Hyogon, et al.
Veröffentlicht: (2025)
Locality-aware Gaussian Compression for Fast and High-quality Rendering
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
RNA: Video Editing with ROI-based Neural Atlas
von: Lee, Jaekyeong, et al.
Veröffentlicht: (2024)
von: Lee, Jaekyeong, et al.
Veröffentlicht: (2024)
Enhancing Cost Efficiency in Active Learning with Candidate Set Query
von: Gwon, Yeho, et al.
Veröffentlicht: (2025)
von: Gwon, Yeho, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Addressing Text Embedding Leakage in Diffusion-based Image Editing
von: Mun, Sunung, et al.
Veröffentlicht: (2024) -
EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation
von: Mun, Sunung, et al.
Veröffentlicht: (2026) -
Elevating 3D Models: High-Quality Texture and Geometry Refinement from a Low-Quality Model
von: Ryu, Nuri, et al.
Veröffentlicht: (2025) -
CLIPtone: Unsupervised Learning for Text-based Image Tone Adjustment
von: Lee, Hyeongmin, et al.
Veröffentlicht: (2024) -
POS-ISP: Pipeline Optimization at the Sequence Level for Task-aware ISP
von: Won, Jiyun, et al.
Veröffentlicht: (2026)