Saved in:
| Main Authors: | Wang, Shuai, Li, Zexian, zhang, Qipeng, Song, Tianhui, Li, Xubin, Ge, Tiezheng, Zheng, Bo, Wang, Limin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.21114 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging
by: Song, Tianhui, et al.
Published: (2025)
by: Song, Tianhui, et al.
Published: (2025)
Accelerating Image Generation with Sub-path Linear Approximation Model
by: Xu, Chen, et al.
Published: (2024)
by: Xu, Chen, et al.
Published: (2024)
Flowing Backwards: Improving Normalizing Flows via Reverse Representation Alignment
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
by: Chen, Hongyu, et al.
Published: (2024)
by: Chen, Hongyu, et al.
Published: (2024)
Beyond Point-Wise Matching: Structural Representation Alignment for Accelerating Diffusion Transformers
by: Xu, Shaodong, et al.
Published: (2026)
by: Xu, Shaodong, et al.
Published: (2026)
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing
by: Xu, Shaodong, et al.
Published: (2026)
by: Xu, Shaodong, et al.
Published: (2026)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
by: Wang, Xu, et al.
Published: (2026)
by: Wang, Xu, et al.
Published: (2026)
RHanDS: Refining Malformed Hands for Generated Images with Decoupled Structure and Style Guidance
by: Wang, Chengrui, et al.
Published: (2024)
by: Wang, Chengrui, et al.
Published: (2024)
TBStar-Edit: From Image Editing Pattern Shifting to Consistency Enhancement
by: Fang, Hao, et al.
Published: (2025)
by: Fang, Hao, et al.
Published: (2025)
MixFormerV2: Efficient Fully Transformer Tracking
by: Cui, Yutao, et al.
Published: (2023)
by: Cui, Yutao, et al.
Published: (2023)
DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
by: Lu, Cheng, et al.
Published: (2022)
by: Lu, Cheng, et al.
Published: (2022)
ERA-Solver: Error-Robust Adams Solver for Fast Sampling of Diffusion Probabilistic Models
by: Li, Shengming, et al.
Published: (2023)
by: Li, Shengming, et al.
Published: (2023)
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025)
by: Wang, Hebaixu, et al.
Published: (2025)
AtomoVideo: High Fidelity Image-to-Video Generation
by: Gong, Litong, et al.
Published: (2024)
by: Gong, Litong, et al.
Published: (2024)
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
by: Li, Weijie, et al.
Published: (2024)
by: Li, Weijie, et al.
Published: (2024)
Identity-Preserving Image-to-Video Generation via Reward-Guided Optimization
by: Shen, Liao, et al.
Published: (2025)
by: Shen, Liao, et al.
Published: (2025)
Leveraging Previous Steps: A Training-free Fast Solver for Flow Diffusion
by: Song, Kaiyu, et al.
Published: (2024)
by: Song, Kaiyu, et al.
Published: (2024)
Fast and Stable Diffusion Inverse Solver with History Gradient Update
by: He, Linchao, et al.
Published: (2023)
by: He, Linchao, et al.
Published: (2023)
Distilling Parallel Gradients for Fast ODE Solvers of Diffusion Models
by: Zhu, Beier, et al.
Published: (2025)
by: Zhu, Beier, et al.
Published: (2025)
PixNerd: Pixel Neural Field Diffusion
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
DDT: Decoupled Diffusion Transformer
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
by: Du, Yang, et al.
Published: (2025)
by: Du, Yang, et al.
Published: (2025)
Consistency Model is an Effective Posterior Sample Approximation for Diffusion Inverse Solvers
by: Xu, Tongda, et al.
Published: (2024)
by: Xu, Tongda, et al.
Published: (2024)
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
by: Duan, Zheng-Peng, et al.
Published: (2024)
by: Duan, Zheng-Peng, et al.
Published: (2024)
PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering
by: Gao, Yifan, et al.
Published: (2025)
by: Gao, Yifan, et al.
Published: (2025)
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance
by: Song, Wenxuan, et al.
Published: (2026)
by: Song, Wenxuan, et al.
Published: (2026)
Hierarchical Masked 3D Diffusion Model for Video Outpainting
by: Fan, Fanda, et al.
Published: (2023)
by: Fan, Fanda, et al.
Published: (2023)
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models
by: Shaul, Neta, et al.
Published: (2024)
by: Shaul, Neta, et al.
Published: (2024)
DDA-Thinker: Decoupled Dual-Atomic Reinforcement Learning for Reasoning-Driven Image Editing
by: Yang, Hanqing, et al.
Published: (2026)
by: Yang, Hanqing, et al.
Published: (2026)
MaRS: A Fast Sampler for Mean Reverting Diffusion based on ODE and SDE Solvers
by: Li, Ao, et al.
Published: (2025)
by: Li, Ao, et al.
Published: (2025)
Critique Before Thinking: Mitigating Hallucination through Rationale-Augmented Instruction Tuning
by: Yang, Zexian, et al.
Published: (2025)
by: Yang, Zexian, et al.
Published: (2025)
A Large-Scale Study on Video Action Dataset Condensation
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
T-Stars-Poster: A Framework for Product-Centric Advertising Image Design
by: Chen, Hongyu, et al.
Published: (2025)
by: Chen, Hongyu, et al.
Published: (2025)
SceneBooth: Diffusion-based Framework for Subject-preserved Text-to-Image Generation
by: Chai, Shang, et al.
Published: (2025)
by: Chai, Shang, et al.
Published: (2025)
CF-Font: Content Fusion for Few-shot Font Generation
by: Wang, Chi, et al.
Published: (2023)
by: Wang, Chi, et al.
Published: (2023)
Outlier-Robust Diffusion Solvers for Inverse Problems
by: Zheng, Yang, et al.
Published: (2026)
by: Zheng, Yang, et al.
Published: (2026)
Depth-Guided Robust and Fast Point Cloud Fusion NeRF for Sparse Input Views
by: Guo, Shuai, et al.
Published: (2024)
by: Guo, Shuai, et al.
Published: (2024)
Parallel Diffusion Solver via Residual Dirichlet Policy Optimization
by: Wang, Ruoyu, et al.
Published: (2025)
by: Wang, Ruoyu, et al.
Published: (2025)
Analyzing and Improving Fast Sampling of Text-to-Image Diffusion Models
by: Zhou, Zhenyu, et al.
Published: (2026)
by: Zhou, Zhenyu, et al.
Published: (2026)
Similar Items
-
FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution
by: Wang, Shuai, et al.
Published: (2024) -
DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging
by: Song, Tianhui, et al.
Published: (2025) -
Accelerating Image Generation with Sub-path Linear Approximation Model
by: Xu, Chen, et al.
Published: (2024) -
Flowing Backwards: Improving Normalizing Flows via Reverse Representation Alignment
by: Chen, Yang, et al.
Published: (2025) -
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
by: Chen, Hongyu, et al.
Published: (2024)