Unified Diffusion Transformer for High-fidelity Text-Aware Image Restoration
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jin Hyeon, Cho, Paul Hyunbin, Kim, Claire, Min, Jaewon, Lee, Jaeeun, Park, Jihye, Choi, Yeji, Kim, Seungryong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Text-Aware Image Restoration with Diffusion Models
by: Min, Jaewon, et al.
Published: (2025)
by: Min, Jaewon, et al.
Published: (2025)
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
by: Kim, Jin Hyeon, et al.
Published: (2026)
by: Kim, Jin Hyeon, et al.
Published: (2026)
DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models
by: Min, Jaewon, et al.
Published: (2026)
by: Min, Jaewon, et al.
Published: (2026)
MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
Projected Representation Conditioning for High-fidelity Novel View Synthesis
by: Kwak, Min-Seop, et al.
Published: (2026)
by: Kwak, Min-Seop, et al.
Published: (2026)
Improving Cone-Beam CT Image Quality with Knowledge Distillation-Enhanced Diffusion Model in Imbalanced Data Settings
by: Hwang, Joonil, et al.
Published: (2024)
by: Hwang, Joonil, et al.
Published: (2024)
Hybrid Video Diffusion Models with 2D Triplane and 3D Wavelet Representation
by: Kim, Kihong, et al.
Published: (2024)
by: Kim, Kihong, et al.
Published: (2024)
WorldKV: Efficient World Memory with World Retrieval and Compression
by: Yi, Jung, et al.
Published: (2026)
by: Yi, Jung, et al.
Published: (2026)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
by: Ahn, Donghoon, et al.
Published: (2024)
by: Ahn, Donghoon, et al.
Published: (2024)
APPLE: Attribute-Preserving Pseudo-Labeling for Diffusion-Based Face Swapping
by: Kang, Jiwon, et al.
Published: (2026)
by: Kang, Jiwon, et al.
Published: (2026)
UniTT-Stereo: Unified Training of Transformer for Enhanced Stereo Matching
by: Kim, Soomin, et al.
Published: (2024)
by: Kim, Soomin, et al.
Published: (2024)
Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual Correspondence
by: Hong, Sunghwan, et al.
Published: (2024)
by: Hong, Sunghwan, et al.
Published: (2024)
Closed-String Mirror Symmetry for Log Calabi-Yau Surfaces
by: Kim, Hyunbin
Published: (2024)
by: Kim, Hyunbin
Published: (2024)
Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression
by: Yi, Jung, et al.
Published: (2025)
by: Yi, Jung, et al.
Published: (2025)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
by: Kim, Chaehyun, et al.
Published: (2025)
by: Kim, Chaehyun, et al.
Published: (2025)
From Threat to Tool: Leveraging Refusal-Aware Injection Attacks for Safety Alignment
by: Chae, Kyubyung, et al.
Published: (2025)
by: Chae, Kyubyung, et al.
Published: (2025)
A Noise is Worth Diffusion Guidance
by: Ahn, Donghoon, et al.
Published: (2024)
by: Ahn, Donghoon, et al.
Published: (2024)
Repurposing Video Diffusion Transformers for Robust Point Tracking
by: Son, Soowon, et al.
Published: (2025)
by: Son, Soowon, et al.
Published: (2025)
Effective Test-Time Scaling of Discrete Diffusion through Iterative Refinement
by: Lee, Sanghyun, et al.
Published: (2025)
by: Lee, Sanghyun, et al.
Published: (2025)
Diffusion Model for Dense Matching
by: Nam, Jisu, et al.
Published: (2023)
by: Nam, Jisu, et al.
Published: (2023)
Lookahead Unmasking Elicits Accurate Decoding in Diffusion Language Models
by: Lee, Sanghyun, et al.
Published: (2025)
by: Lee, Sanghyun, et al.
Published: (2025)
Where and How to Perturb: On the Design of Perturbation Guidance in Diffusion and Flow Models
by: Ahn, Donghoon, et al.
Published: (2025)
by: Ahn, Donghoon, et al.
Published: (2025)
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
by: Seo, Junyoung, et al.
Published: (2023)
by: Seo, Junyoung, et al.
Published: (2023)
CorGi: Contribution-Guided Block-Wise Interval Caching for Training-Free Acceleration of Diffusion Transformers
by: Son, Yonglak, et al.
Published: (2025)
by: Son, Yonglak, et al.
Published: (2025)
DiffuseSlide: Training-Free High Frame Rate Video Generation Diffusion
by: Hwang, Geunmin, et al.
Published: (2025)
by: Hwang, Geunmin, et al.
Published: (2025)
First Attentions Last: Better Exploiting First Attentions for Efficient Transformer Training
by: Kim, Gyudong, et al.
Published: (2025)
by: Kim, Gyudong, et al.
Published: (2025)
DreamMatcher: Appearance Matching Self-Attention for Semantically-Consistent Text-to-Image Personalization
by: Nam, Jisu, et al.
Published: (2024)
by: Nam, Jisu, et al.
Published: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
by: Park, NaHyeon, et al.
Published: (2024)
by: Park, NaHyeon, et al.
Published: (2024)
Performance Plateaus in Inference-Time Scaling for Text-to-Image Diffusion Without External Models
by: Choi, Changhyun, et al.
Published: (2025)
by: Choi, Changhyun, et al.
Published: (2025)
Geometry-Aware Score Distillation via 3D Consistent Noising and Gradient Consistency Modeling
by: Kwak, Min-Seop, et al.
Published: (2024)
by: Kwak, Min-Seop, et al.
Published: (2024)
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
by: Kim, Nam-Gyu, et al.
Published: (2025)
by: Kim, Nam-Gyu, et al.
Published: (2025)
Diffusion Model Compression for Image-to-Image Translation
by: Kim, Geonung, et al.
Published: (2024)
by: Kim, Geonung, et al.
Published: (2024)
Emergent Temporal Correspondences from Video Diffusion Transformers
by: Nam, Jisu, et al.
Published: (2025)
by: Nam, Jisu, et al.
Published: (2025)
Reciprocal Attention Mixing Transformer for Lightweight Image Restoration
by: Choi, Haram, et al.
Published: (2023)
by: Choi, Haram, et al.
Published: (2023)
MaDis-Stereo: Enhanced Stereo Matching via Distilled Masked Image Modeling
by: Ahn, Jihye, et al.
Published: (2024)
by: Ahn, Jihye, et al.
Published: (2024)
InterRVOS: Interaction-aware Referring Video Object Segmentation
by: Jin, Woojeong, et al.
Published: (2025)
by: Jin, Woojeong, et al.
Published: (2025)
Exploring Temporally-Aware Features for Point Tracking
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
Dynamic Exposure Burst Image Restoration
by: Kim, Woohyeok, et al.
Published: (2026)
by: Kim, Woohyeok, et al.
Published: (2026)
GeomGS: LiDAR-Guided Geometry-Aware Gaussian Splatting for Robot Localization
by: Lee, Jaewon, et al.
Published: (2025)
by: Lee, Jaewon, et al.
Published: (2025)
PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
by: Lee, Seokyeong, et al.
Published: (2025)
by: Lee, Seokyeong, et al.
Published: (2025)
Similar Items
-
Text-Aware Image Restoration with Diffusion Models
by: Min, Jaewon, et al.
Published: (2025) -
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
by: Kim, Jin Hyeon, et al.
Published: (2026) -
DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models
by: Min, Jaewon, et al.
Published: (2026) -
MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation
by: Kim, Seyeon, et al.
Published: (2024) -
Projected Representation Conditioning for High-fidelity Novel View Synthesis
by: Kwak, Min-Seop, et al.
Published: (2026)