Improved Masked Image Generation with Knowledge-Augmented Token Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Guotao, Zhang, Baoquan, Wen, Zhiyuan, Han, Zihao, Ye, Yunming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Head-Aware Key-Value Compression for Efficient Autoregressive Image Generation
von: Liang, Guotao, et al.
Veröffentlicht: (2026)
von: Liang, Guotao, et al.
Veröffentlicht: (2026)
Towards Improved Text-Aligned Codebook Learning: Multi-Hierarchical Codebook-Text Alignment with Long Text
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
AsyncDSB: Schedule-Asynchronous Diffusion Schrödinger Bridge for Image Inpainting
von: Han, Zihao, et al.
Veröffentlicht: (2024)
von: Han, Zihao, et al.
Veröffentlicht: (2024)
SJD-PV: Speculative Jacobi Decoding with Phrase Verification for Autoregressive Image Generation
von: Yu, Zhehao, et al.
Veröffentlicht: (2026)
von: Yu, Zhehao, et al.
Veröffentlicht: (2026)
Codebook Transfer with Part-of-Speech for Vector-Quantized Image Modeling
von: Zhang, Baoquan, et al.
Veröffentlicht: (2024)
von: Zhang, Baoquan, et al.
Veröffentlicht: (2024)
LG-VQ: Language-Guided Codebook Learning
von: Liang, Guotao, et al.
Veröffentlicht: (2024)
von: Liang, Guotao, et al.
Veröffentlicht: (2024)
SJD-VP: Speculative Jacobi Decoding with Verification Prediction for Autoregressive Image Generation
von: Shan, Bingqi, et al.
Veröffentlicht: (2026)
von: Shan, Bingqi, et al.
Veröffentlicht: (2026)
HPCR: Holistic Proxy-based Contrastive Replay for Online Continual Learning
von: Lin, Huiwei, et al.
Veröffentlicht: (2023)
von: Lin, Huiwei, et al.
Veröffentlicht: (2023)
Prototype Optimization with Neural ODE for Few-Shot Learning
von: Zhang, Baoquan, et al.
Veröffentlicht: (2024)
von: Zhang, Baoquan, et al.
Veröffentlicht: (2024)
Improving Flexible Image Tokenizers for Autoregressive Image Generation
von: Fu, Zixuan, et al.
Veröffentlicht: (2026)
von: Fu, Zixuan, et al.
Veröffentlicht: (2026)
MCSDNet: Mesoscale Convective System Detection Network via Multi-scale Spatiotemporal Information
von: Liang, Jiajun, et al.
Veröffentlicht: (2024)
von: Liang, Jiajun, et al.
Veröffentlicht: (2024)
Morphing Tokens Draw Strong Masked Image Models
von: Kim, Taekyung, et al.
Veröffentlicht: (2023)
von: Kim, Taekyung, et al.
Veröffentlicht: (2023)
SeiT++: Masked Token Modeling Improves Storage-efficient Training
von: Lee, Minhyun, et al.
Veröffentlicht: (2023)
von: Lee, Minhyun, et al.
Veröffentlicht: (2023)
S2FT: Parameter-Efficient Fine-Tuning in Sparse Spectrum Domain
von: Zhang, Baoquan, et al.
Veröffentlicht: (2026)
von: Zhang, Baoquan, et al.
Veröffentlicht: (2026)
StyleMark: A Robust Watermarking Method for Art Style Images Against Black-Box Arbitrary Style Transfer
von: Zhang, Yunming, et al.
Veröffentlicht: (2024)
von: Zhang, Yunming, et al.
Veröffentlicht: (2024)
DiffCast: A Unified Framework via Residual Diffusion for Precipitation Nowcasting
von: Yu, Demin, et al.
Veröffentlicht: (2023)
von: Yu, Demin, et al.
Veröffentlicht: (2023)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
von: Wang, Ziming, et al.
Veröffentlicht: (2023)
von: Wang, Ziming, et al.
Veröffentlicht: (2023)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
von: Song, Jiawei, et al.
Veröffentlicht: (2024)
von: Song, Jiawei, et al.
Veröffentlicht: (2024)
DeepSound-V1: Start to Think Step-by-Step in the Audio Generation from Videos
von: Liang, Yunming, et al.
Veröffentlicht: (2025)
von: Liang, Yunming, et al.
Veröffentlicht: (2025)
AVT2-DWF: Improving Deepfake Detection with Audio-Visual Fusion and Dynamic Weighting Strategies
von: Wang, Rui, et al.
Veröffentlicht: (2024)
von: Wang, Rui, et al.
Veröffentlicht: (2024)
Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
ReMoMask: Retrieval-Augmented Masked Motion Generation
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
MaskBit: Embedding-free Image Generation via Bit Tokens
von: Weber, Mark, et al.
Veröffentlicht: (2024)
von: Weber, Mark, et al.
Veröffentlicht: (2024)
GPSToken: Gaussian Parameterized Spatially-adaptive Tokenization for Image Representation and Generation
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2025)
Vision Foundation Models as Generalist Tokenizers for Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
Improving Image De-raining Using Reference-Guided Transformers
von: Ye, Zihao, et al.
Veröffentlicht: (2024)
von: Ye, Zihao, et al.
Veröffentlicht: (2024)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
Improving Autoregressive Image Generation through Coarse-to-Fine Token Prediction
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with Minimal 3D Knowledge
von: Wang, Haoru, et al.
Veröffentlicht: (2025)
von: Wang, Haoru, et al.
Veröffentlicht: (2025)
Autoregressive Image Generation with Masked Bit Modeling
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation
von: Lee, Minhyun, et al.
Veröffentlicht: (2024)
von: Lee, Minhyun, et al.
Veröffentlicht: (2024)
NativeTok: Native Visual Tokenization for Improved Image Generation
von: Wu, Bin, et al.
Veröffentlicht: (2026)
von: Wu, Bin, et al.
Veröffentlicht: (2026)
Mask Image Watermarking
von: Hu, Runyi, et al.
Veröffentlicht: (2025)
von: Hu, Runyi, et al.
Veröffentlicht: (2025)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
Medical Referring Image Segmentation via Next-Token Mask Prediction
von: Chen, Xinyu, et al.
Veröffentlicht: (2025)
von: Chen, Xinyu, et al.
Veröffentlicht: (2025)
MLIP: Medical Language-Image Pre-training with Masked Local Representation Learning
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
Retrieval Augmented Comic Image Generation
von: Shui, Yunhao, et al.
Veröffentlicht: (2025)
von: Shui, Yunhao, et al.
Veröffentlicht: (2025)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Head-Aware Key-Value Compression for Efficient Autoregressive Image Generation
von: Liang, Guotao, et al.
Veröffentlicht: (2026) -
Towards Improved Text-Aligned Codebook Learning: Multi-Hierarchical Codebook-Text Alignment with Long Text
von: Liang, Guotao, et al.
Veröffentlicht: (2025) -
AsyncDSB: Schedule-Asynchronous Diffusion Schrödinger Bridge for Image Inpainting
von: Han, Zihao, et al.
Veröffentlicht: (2024) -
SJD-PV: Speculative Jacobi Decoding with Phrase Verification for Autoregressive Image Generation
von: Yu, Zhehao, et al.
Veröffentlicht: (2026) -
Codebook Transfer with Part-of-Speech for Vector-Quantized Image Modeling
von: Zhang, Baoquan, et al.
Veröffentlicht: (2024)