ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Yongsheng, Zheng, Haitian, Zhang, Zhifei, Zhang, Jianming, Zhou, Yuqian, Barnes, Connelly, Liu, Yuchen, Xiong, Wei, Lin, Zhe, Luo, Jiebo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
by: Zheng, Haitian, et al.
Published: (2025)
by: Zheng, Haitian, et al.
Published: (2025)
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
by: Zheng, Haitian, et al.
Published: (2022)
by: Zheng, Haitian, et al.
Published: (2022)
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
by: Yu, Yongsheng, et al.
Published: (2025)
by: Yu, Yongsheng, et al.
Published: (2025)
PixelDiT: Pixel Diffusion Transformers for Image Generation
by: Yu, Yongsheng, et al.
Published: (2025)
by: Yu, Yongsheng, et al.
Published: (2025)
Diffusion Transformer-to-Mamba Distillation for High-Resolution Image Generation
by: Yao, Yuan, et al.
Published: (2025)
by: Yao, Yuan, et al.
Published: (2025)
FINECAPTION: Compositional Image Captioning Focusing on Wherever You Want at Any Granularity
by: Hua, Hang, et al.
Published: (2024)
by: Hua, Hang, et al.
Published: (2024)
Tri-Prompting: Video Diffusion with Unified Control over Scene, Subject, and Motion
by: Zhou, Zhenghong, et al.
Published: (2026)
by: Zhou, Zhenghong, et al.
Published: (2026)
Chain-of-Thought Prompting for Demographic Inference with Large Multimodal Models
by: Yu, Yongsheng, et al.
Published: (2024)
by: Yu, Yongsheng, et al.
Published: (2024)
High-Resolution Speech Restoration with Latent Diffusion Model
by: Dhyani, Tushar, et al.
Published: (2024)
by: Dhyani, Tushar, et al.
Published: (2024)
Ultra-High-Resolution Image Synthesis with Pyramid Diffusion Model
by: Yang, Jiajie
Published: (2024)
by: Yang, Jiajie
Published: (2024)
Latent-Reframe: Enabling Camera Control for Video Diffusion Model without Training
by: Zhou, Zhenghong, et al.
Published: (2024)
by: Zhou, Zhenghong, et al.
Published: (2024)
TurboFill: Adapting Few-step Text-to-image Model for Fast Image Inpainting
by: Xie, Liangbin, et al.
Published: (2025)
by: Xie, Liangbin, et al.
Published: (2025)
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
by: You, Haoran, et al.
Published: (2024)
by: You, Haoran, et al.
Published: (2024)
Gradient as Conditions: Rethinking HOG for All-in-one Image Restoration
by: Wu, Jiawei, et al.
Published: (2025)
by: Wu, Jiawei, et al.
Published: (2025)
SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration
by: Zhang, Yuhong, et al.
Published: (2024)
by: Zhang, Yuhong, et al.
Published: (2024)
Image Restoration via Diffusion Models with Dynamic Resolution
by: Zheng, Yang, et al.
Published: (2026)
by: Zheng, Yang, et al.
Published: (2026)
Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models
by: Zhang, Jinjin, et al.
Published: (2025)
by: Zhang, Jinjin, et al.
Published: (2025)
Diffusion Posterior Proximal Sampling for Image Restoration
by: Wu, Hongjie, et al.
Published: (2024)
by: Wu, Hongjie, et al.
Published: (2024)
Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction
by: Cai, Yuanhao, et al.
Published: (2024)
by: Cai, Yuanhao, et al.
Published: (2024)
DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution
by: Duan, Zheng-Peng, et al.
Published: (2025)
by: Duan, Zheng-Peng, et al.
Published: (2025)
DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization
by: Ding, Zihan, et al.
Published: (2024)
by: Ding, Zihan, et al.
Published: (2024)
LLaVA-UHD v2: an MLLM Integrating High-Resolution Semantic Pyramid via Hierarchical Window Transformer
by: Zhang, Yipeng, et al.
Published: (2024)
by: Zhang, Yipeng, et al.
Published: (2024)
KnapFormer: An Online Load Balancer for Efficient Diffusion Transformers Training
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Learning Brain Tumor Representation in 3D High-Resolution MR Images via Interpretable State Space Models
by: Hu, Qingqiao, et al.
Published: (2024)
by: Hu, Qingqiao, et al.
Published: (2024)
AutoDIR: Automatic All-in-One Image Restoration with Latent Diffusion
by: Jiang, Yitong, et al.
Published: (2023)
by: Jiang, Yitong, et al.
Published: (2023)
Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis
by: Sigillo, Luigi, et al.
Published: (2025)
by: Sigillo, Luigi, et al.
Published: (2025)
MAT: Multi-Range Attention Transformer for Efficient Image Super-Resolution
by: Xie, Chengxing, et al.
Published: (2024)
by: Xie, Chengxing, et al.
Published: (2024)
UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration
by: He, Chunming, et al.
Published: (2025)
by: He, Chunming, et al.
Published: (2025)
On Inductive Biases That Enable Generalization of Diffusion Transformers
by: An, Jie, et al.
Published: (2024)
by: An, Jie, et al.
Published: (2024)
LatentINDIGO: An INN-Guided Latent Diffusion Algorithm for Image Restoration
by: You, Di, et al.
Published: (2025)
by: You, Di, et al.
Published: (2025)
Rethinking Global Text Conditioning in Diffusion Transformers
by: Starodubcev, Nikita, et al.
Published: (2026)
by: Starodubcev, Nikita, et al.
Published: (2026)
Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation
by: Sauer, Axel, et al.
Published: (2024)
by: Sauer, Axel, et al.
Published: (2024)
RELD: Regularization by Latent Diffusion Models for Image Restoration
by: Cascarano, Pasquale, et al.
Published: (2025)
by: Cascarano, Pasquale, et al.
Published: (2025)
Are Conditional Latent Diffusion Models Effective for Image Restoration?
by: Yuan, Yunchen, et al.
Published: (2024)
by: Yuan, Yunchen, et al.
Published: (2024)
Single-Step Latent Diffusion for Underwater Image Restoration
by: Wu, Jiayi, et al.
Published: (2025)
by: Wu, Jiayi, et al.
Published: (2025)
MatIR: A Hybrid Mamba-Transformer Image Restoration Model
by: Wen, Juan, et al.
Published: (2025)
by: Wen, Juan, et al.
Published: (2025)
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
by: Xie, Enze, et al.
Published: (2024)
by: Xie, Enze, et al.
Published: (2024)
UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Perceive-IR: Learning to Perceive Degradation Better for All-in-One Image Restoration
by: Zhang, Xu, et al.
Published: (2024)
by: Zhang, Xu, et al.
Published: (2024)
MoE-DiffIR: Task-customized Diffusion Priors for Universal Compressed Image Restoration
by: Ren, Yulin, et al.
Published: (2024)
by: Ren, Yulin, et al.
Published: (2024)
Similar Items
-
PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
by: Zheng, Haitian, et al.
Published: (2025) -
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
by: Zheng, Haitian, et al.
Published: (2022) -
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
by: Yu, Yongsheng, et al.
Published: (2025) -
PixelDiT: Pixel Diffusion Transformers for Image Generation
by: Yu, Yongsheng, et al.
Published: (2025) -
Diffusion Transformer-to-Mamba Distillation for High-Resolution Image Generation
by: Yao, Yuan, et al.
Published: (2025)