Effective and Efficient Masked Image Generation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | You, Zebin, Ou, Jingyang, Zhang, Xiaolu, Hu, Jun, Zhou, Jun, Li, Chongxuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLaDA-o: An Effective and Length-Adaptive Omni Diffusion Model
von: You, Zebin, et al.
Veröffentlicht: (2026)
von: You, Zebin, et al.
Veröffentlicht: (2026)
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
von: You, Zebin, et al.
Veröffentlicht: (2025)
von: You, Zebin, et al.
Veröffentlicht: (2025)
Are Images Indistinguishable to Humans Also Indistinguishable to Classifiers?
von: You, Zebin, et al.
Veröffentlicht: (2024)
von: You, Zebin, et al.
Veröffentlicht: (2024)
DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
von: Lu, Cheng, et al.
Veröffentlicht: (2022)
von: Lu, Cheng, et al.
Veröffentlicht: (2022)
The Blessing of Randomness: SDE Beats ODE in General Diffusion-based Image Editing
von: Nie, Shen, et al.
Veröffentlicht: (2023)
von: Nie, Shen, et al.
Veröffentlicht: (2023)
CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model
von: Wang, Zhengyi, et al.
Veröffentlicht: (2024)
von: Wang, Zhengyi, et al.
Veröffentlicht: (2024)
Large Language Diffusion Models
von: Nie, Shen, et al.
Veröffentlicht: (2025)
von: Nie, Shen, et al.
Veröffentlicht: (2025)
Masked Training for Robust Arrhythmia Detection from Digitalized Multiple Layout ECG Images
von: Zhang, Shanwei, et al.
Veröffentlicht: (2025)
von: Zhang, Shanwei, et al.
Veröffentlicht: (2025)
Scaling Diffusion Transformers Efficiently via $μ$P
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
Masked Generative Transformer Is What You Need for Image Editing
von: Chow, Wei, et al.
Veröffentlicht: (2026)
von: Chow, Wei, et al.
Veröffentlicht: (2026)
Towards Understanding Why Data Augmentation Improves Generalization
von: Li, Jingyang, et al.
Veröffentlicht: (2025)
von: Li, Jingyang, et al.
Veröffentlicht: (2025)
AutoDFP: Automatic Data-Free Pruning via Channel Similarity Reconstruction
von: Li, Siqi, et al.
Veröffentlicht: (2024)
von: Li, Siqi, et al.
Veröffentlicht: (2024)
Multiple Code Hashing for Efficient Image Retrieval
von: Li, Ming-Wei, et al.
Veröffentlicht: (2020)
von: Li, Ming-Wei, et al.
Veröffentlicht: (2020)
PRISM: Privacy-Preserving Improved Stochastic Masking for Federated Generative Models
von: Seo, Kyeongkook, et al.
Veröffentlicht: (2025)
von: Seo, Kyeongkook, et al.
Veröffentlicht: (2025)
Controllable Unlearning for Image-to-Image Generative Models via $\varepsilon$-Constrained Optimization
von: Feng, Xiaohua, et al.
Veröffentlicht: (2024)
von: Feng, Xiaohua, et al.
Veröffentlicht: (2024)
Keypoint Aware Masked Image Modelling
von: Krishna, Madhava, et al.
Veröffentlicht: (2024)
von: Krishna, Madhava, et al.
Veröffentlicht: (2024)
Improving Generative Pre-Training: An In-depth Study of Masked Image Modeling and Denoising Models
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
MaSkel: A Model for Human Whole-body X-rays Generation from Human Masking Images
von: Xi, Yingjie, et al.
Veröffentlicht: (2024)
von: Xi, Yingjie, et al.
Veröffentlicht: (2024)
Hyperbolic Binary Neural Network
von: Chen, Jun, et al.
Veröffentlicht: (2025)
von: Chen, Jun, et al.
Veröffentlicht: (2025)
HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation
von: Kumbong, Hermann, et al.
Veröffentlicht: (2025)
von: Kumbong, Hermann, et al.
Veröffentlicht: (2025)
Effective Decision Boundary Learning for Class Incremental Learning
von: Li, Kunchi, et al.
Veröffentlicht: (2023)
von: Li, Kunchi, et al.
Veröffentlicht: (2023)
Masked Conditioning for Deep Generative Models
von: Mueller, Phillip, et al.
Veröffentlicht: (2025)
von: Mueller, Phillip, et al.
Veröffentlicht: (2025)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
von: Jin, Qixuan, et al.
Veröffentlicht: (2024)
von: Jin, Qixuan, et al.
Veröffentlicht: (2024)
Exploring the Coordination of Frequency and Attention in Masked Image Modeling
von: Gui, Jie, et al.
Veröffentlicht: (2022)
von: Gui, Jie, et al.
Veröffentlicht: (2022)
MaskVD: Region Masking for Efficient Video Object Detection
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2024)
Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation
von: Li, Chao, et al.
Veröffentlicht: (2026)
von: Li, Chao, et al.
Veröffentlicht: (2026)
Memory-Efficient 4-bit Preconditioned Stochastic Optimization
von: Li, Jingyang, et al.
Veröffentlicht: (2024)
von: Li, Jingyang, et al.
Veröffentlicht: (2024)
Revisiting Unknowns: Towards Effective and Efficient Open-Set Active Learning
von: Zong, Chen-Chen, et al.
Veröffentlicht: (2026)
von: Zong, Chen-Chen, et al.
Veröffentlicht: (2026)
BayesDiff: Estimating Pixel-wise Uncertainty in Diffusion via Bayesian Inference
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation
von: Chang, Yu, et al.
Veröffentlicht: (2025)
von: Chang, Yu, et al.
Veröffentlicht: (2025)
MaskBit: Embedding-free Image Generation via Bit Tokens
von: Weber, Mark, et al.
Veröffentlicht: (2024)
von: Weber, Mark, et al.
Veröffentlicht: (2024)
MCM: Multi-layer Concept Map for Efficient Concept Learning from Masked Images
von: Sun, Yuwei, et al.
Veröffentlicht: (2025)
von: Sun, Yuwei, et al.
Veröffentlicht: (2025)
SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
Data Attribution for Text-to-Image Models by Unlearning Synthesized Images
von: Wang, Sheng-Yu, et al.
Veröffentlicht: (2024)
von: Wang, Sheng-Yu, et al.
Veröffentlicht: (2024)
SupMAE: Supervised Masked Autoencoders Are Efficient Vision Learners
von: Liang, Feng, et al.
Veröffentlicht: (2022)
von: Liang, Feng, et al.
Veröffentlicht: (2022)
AM-SAM: Automated Prompting and Mask Calibration for Segment Anything Model
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
Beyond [cls]: Exploring the true potential of Masked Image Modeling representations
von: Przewięźlikowski, Marcin, et al.
Veröffentlicht: (2024)
von: Przewięźlikowski, Marcin, et al.
Veröffentlicht: (2024)
Dynamic Allocation Hypernetwork with Adaptive Model Recalibration for FCL
von: Qi, Xiaoming, et al.
Veröffentlicht: (2025)
von: Qi, Xiaoming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLaDA-o: An Effective and Length-Adaptive Omni Diffusion Model
von: You, Zebin, et al.
Veröffentlicht: (2026) -
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
von: You, Zebin, et al.
Veröffentlicht: (2025) -
Are Images Indistinguishable to Humans Also Indistinguishable to Classifiers?
von: You, Zebin, et al.
Veröffentlicht: (2024) -
DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
von: Lu, Cheng, et al.
Veröffentlicht: (2022) -
The Blessing of Randomness: SDE Beats ODE in General Diffusion-based Image Editing
von: Nie, Shen, et al.
Veröffentlicht: (2023)