Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Lei, Li, Senmao, Yang, Fei, Wang, Jianye, Zhang, Ziheng, Liu, Yuhan, Wang, Yaxing, Yang, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WaDi: Weight Direction-aware Distillation for One-step Image Synthesis
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
One-Way Ticket:Time-Independent Unified Encoder for Distilling Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
Adversarial Concept Distillation for One-Step Diffusion Personalization
von: Yang, Yixiong, et al.
Veröffentlicht: (2025)
von: Yang, Yixiong, et al.
Veröffentlicht: (2025)
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2024)
von: Li, Senmao, et al.
Veröffentlicht: (2024)
Faster Diffusion: Rethinking the Role of the Encoder for Diffusion Model Inference
von: Li, Senmao, et al.
Veröffentlicht: (2023)
von: Li, Senmao, et al.
Veröffentlicht: (2023)
Occlusion-Aware 3D Motion Interpretation for Abnormal Behavior Detection
von: Li, Su, et al.
Veröffentlicht: (2024)
von: Li, Su, et al.
Veröffentlicht: (2024)
Training-free image inversion for one-step diffusion models
von: Wu, Tao, et al.
Veröffentlicht: (2026)
von: Wu, Tao, et al.
Veröffentlicht: (2026)
Free-Lunch Color-Texture Disentanglement for Stylized Image Generation
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
von: Li, Senmao, et al.
Veröffentlicht: (2023)
von: Li, Senmao, et al.
Veröffentlicht: (2023)
Data Extrapolation for Text-to-image Generation on Small Datasets
von: Ye, Senmao, et al.
Veröffentlicht: (2024)
von: Ye, Senmao, et al.
Veröffentlicht: (2024)
Diversity Has Always Been There in Your Visual Autoregressive Models
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
InterLCM: Low-Quality Images as Intermediate States of Latent Consistency Models for Effective Blind Face Restoration
von: Li, Senmao, et al.
Veröffentlicht: (2025)
von: Li, Senmao, et al.
Veröffentlicht: (2025)
Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
Soft Masked Mamba Diffusion Model for CT to MRI Conversion
von: Wang, Zhenbin, et al.
Veröffentlicht: (2024)
von: Wang, Zhenbin, et al.
Veröffentlicht: (2024)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
From Cradle to Cane: A Two-Pass Framework for High-Fidelity Lifespan Face Aging
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
Masked Conditional Diffusion Model for Enhancing Deepfake Detection
von: Chen, Tiewen, et al.
Veröffentlicht: (2024)
von: Chen, Tiewen, et al.
Veröffentlicht: (2024)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On
von: Shao, Dingbao, et al.
Veröffentlicht: (2026)
von: Shao, Dingbao, et al.
Veröffentlicht: (2026)
ALTER: All-in-One Layer Pruning and Temporal Expert Routing for Efficient Diffusion Generation
von: Yang, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Yang, Xiaomeng, et al.
Veröffentlicht: (2025)
Structure Matters: Tackling the Semantic Discrepancy in Diffusion Models for Image Inpainting
von: Liu, Haipeng, et al.
Veröffentlicht: (2024)
von: Liu, Haipeng, et al.
Veröffentlicht: (2024)
MARMOT: Masked Autoencoder for Modeling Transient Imaging
von: Shen, Siyuan, et al.
Veröffentlicht: (2025)
von: Shen, Siyuan, et al.
Veröffentlicht: (2025)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
CoSpace: Benchmarking Continuous Space Perception Ability for Vision-Language Models
von: Zhu, Yiqi, et al.
Veröffentlicht: (2025)
von: Zhu, Yiqi, et al.
Veröffentlicht: (2025)
Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing
von: Hu, Taihang, et al.
Veröffentlicht: (2025)
von: Hu, Taihang, et al.
Veröffentlicht: (2025)
RefAlign: Representation Alignment for Reference-to-Video Generation
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
GaitAdapt: Continual Learning for Evolving Gait Recognition
von: Wang, Jingjie, et al.
Veröffentlicht: (2025)
von: Wang, Jingjie, et al.
Veröffentlicht: (2025)
Adaptive quantization with mixed-precision based on low-cost proxy
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
DDAE++: Enhancing Diffusion Models Towards Unified Generative and Discriminative Learning
von: Xiang, Weilai, et al.
Veröffentlicht: (2025)
von: Xiang, Weilai, et al.
Veröffentlicht: (2025)
Masked Diffusion Vision-Language Models for Temporal Action Localization
von: Wang, Fengshun, et al.
Veröffentlicht: (2026)
von: Wang, Fengshun, et al.
Veröffentlicht: (2026)
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
von: Xu, Xingqian, et al.
Veröffentlicht: (2022)
von: Xu, Xingqian, et al.
Veröffentlicht: (2022)
Towards Dense and Accurate Radar Perception Via Efficient Cross-Modal Diffusion Model
von: Zhang, Ruibin, et al.
Veröffentlicht: (2024)
von: Zhang, Ruibin, et al.
Veröffentlicht: (2024)
CtxMIM: Context-Enhanced Masked Image Modeling for Remote Sensing Image Understanding
von: Zhang, Mingming, et al.
Veröffentlicht: (2023)
von: Zhang, Mingming, et al.
Veröffentlicht: (2023)
Improving Masked Autoencoders by Learning Where to Mask
von: Chen, Haijian, et al.
Veröffentlicht: (2023)
von: Chen, Haijian, et al.
Veröffentlicht: (2023)
Filtering Memorization from Parameter-Space in Diffusion Models
von: Zhe, Yu, et al.
Veröffentlicht: (2026)
von: Zhe, Yu, et al.
Veröffentlicht: (2026)
ReMoMask: Retrieval-Augmented Masked Motion Generation
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
MaTe3D: Mask-guided Text-based 3D-aware Portrait Editing
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
WaDi: Weight Direction-aware Distillation for One-step Image Synthesis
von: Wang, Lei, et al.
Veröffentlicht: (2026) -
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
von: Li, Senmao, et al.
Veröffentlicht: (2025) -
One-Way Ticket:Time-Independent Unified Encoder for Distilling Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2025) -
Adversarial Concept Distillation for One-Step Diffusion Personalization
von: Yang, Yixiong, et al.
Veröffentlicht: (2025) -
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
von: Li, Senmao, et al.
Veröffentlicht: (2024)