Wavelet-Driven Masked Image Modeling: A Path to Efficient Visual Representation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiang, Wenzhao, Liu, Chang, Yu, Hongyang, Chen, Xilin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AEMIM: Adversarial Examples Meet Masked Image Modeling
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2024)
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2024)
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2026)
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2026)
Learning from the Right Patches: A Two-Stage Wavelet-Driven Masked Autoencoder for Histopathology Representation Learning
von: Younis, Raneen, et al.
Veröffentlicht: (2025)
von: Younis, Raneen, et al.
Veröffentlicht: (2025)
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
SemanticMIM: Marring Masked Image Modeling with Semantics Compression for General Visual Representation
von: Yuan, Yike, et al.
Veröffentlicht: (2024)
von: Yuan, Yike, et al.
Veröffentlicht: (2024)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
WaveDM: Wavelet-Based Diffusion Models for Image Restoration
von: Huang, Yi, et al.
Veröffentlicht: (2023)
von: Huang, Yi, et al.
Veröffentlicht: (2023)
WaveFormer: A 3D Transformer with Wavelet-Driven Feature Representation for Efficient Medical Image Segmentation
von: Hasan, Md Mahfuz Al, et al.
Veröffentlicht: (2025)
von: Hasan, Md Mahfuz Al, et al.
Veröffentlicht: (2025)
Autoregressive Image Generation with Masked Bit Modeling
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
von: Yu, Qihang, et al.
Veröffentlicht: (2026)
TweezeEdit: Consistent and Efficient Image Editing with Path Regularization
von: Mao, Jianda, et al.
Veröffentlicht: (2025)
von: Mao, Jianda, et al.
Veröffentlicht: (2025)
Learning Mask Invariant Mutual Information for Masked Image Modeling
von: Huang, Tao, et al.
Veröffentlicht: (2025)
von: Huang, Tao, et al.
Veröffentlicht: (2025)
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
von: Sun, Peng, et al.
Veröffentlicht: (2026)
von: Sun, Peng, et al.
Veröffentlicht: (2026)
PiLaMIM: Toward Richer Visual Representations by Integrating Pixel and Latent Masked Image Modeling
von: Lee, Junmyeong, et al.
Veröffentlicht: (2025)
von: Lee, Junmyeong, et al.
Veröffentlicht: (2025)
CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation
von: Xu, Yifeng, et al.
Veröffentlicht: (2024)
von: Xu, Yifeng, et al.
Veröffentlicht: (2024)
Less is More: Decoder-Free Masked Modeling for Efficient Skeleton Representation Learning
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2026)
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2026)
Improving Model Generalization by On-manifold Adversarial Augmentation in the Frequency Domain
von: Liu, Chang, et al.
Veröffentlicht: (2023)
von: Liu, Chang, et al.
Veröffentlicht: (2023)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
von: Wei, Yibing, et al.
Veröffentlicht: (2024)
von: Wei, Yibing, et al.
Veröffentlicht: (2024)
Suppressing Non-Semantic Noise in Masked Image Modeling Representations
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2026)
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2026)
A Survey on Interpretability in Visual Recognition
von: Wan, Qiyang, et al.
Veröffentlicht: (2025)
von: Wan, Qiyang, et al.
Veröffentlicht: (2025)
MOODv2: Masked Image Modeling for Out-of-Distribution Detection
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
Efficient Cell Painting Image Representation Learning via Cross-Well Aligned Masked Siamese Network
von: Huang, Pin-Jui, et al.
Veröffentlicht: (2025)
von: Huang, Pin-Jui, et al.
Veröffentlicht: (2025)
From Prototypes to General Distributions: An Efficient Curriculum for Masked Image Modeling
von: Lin, Jinhong, et al.
Veröffentlicht: (2024)
von: Lin, Jinhong, et al.
Veröffentlicht: (2024)
Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition
von: Zhang, Yifei, et al.
Veröffentlicht: (2025)
von: Zhang, Yifei, et al.
Veröffentlicht: (2025)
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
MLIP: Medical Language-Image Pre-training with Masked Local Representation Learning
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition
von: Lin, Tiancheng, et al.
Veröffentlicht: (2024)
von: Lin, Tiancheng, et al.
Veröffentlicht: (2024)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Wavelet-Domain Masked Image Modeling for Color-Consistent HDR Video Reconstruction
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
Path Choice Matters for Clear Attribution in Path Methods
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
von: Xin, Yi, et al.
Veröffentlicht: (2025)
von: Xin, Yi, et al.
Veröffentlicht: (2025)
T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2024)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2024)
Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation
von: Zhang, Yutong, et al.
Veröffentlicht: (2026)
von: Zhang, Yutong, et al.
Veröffentlicht: (2026)
EfficientMT: Efficient Temporal Adaptation for Motion Transfer in Text-to-Video Diffusion Models
von: Cai, Yufei, et al.
Veröffentlicht: (2025)
von: Cai, Yufei, et al.
Veröffentlicht: (2025)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation
von: Li, Pengzhi, et al.
Veröffentlicht: (2025)
von: Li, Pengzhi, et al.
Veröffentlicht: (2025)
Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
von: Jiang, Jerry, et al.
Veröffentlicht: (2026)
von: Jiang, Jerry, et al.
Veröffentlicht: (2026)
PathDiff: Histopathology Image Synthesis with Unpaired Text and Mask Conditions
von: Bhosale, Mahesh, et al.
Veröffentlicht: (2025)
von: Bhosale, Mahesh, et al.
Veröffentlicht: (2025)
Harnessing Massive Satellite Imagery with Efficient Masked Image Modeling
von: Wang, Fengxiang, et al.
Veröffentlicht: (2024)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AEMIM: Adversarial Examples Meet Masked Image Modeling
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2024) -
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2026) -
Learning from the Right Patches: A Two-Stage Wavelet-Driven Masked Autoencoder for Histopathology Representation Learning
von: Younis, Raneen, et al.
Veröffentlicht: (2025) -
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
von: Wang, Chengkun, et al.
Veröffentlicht: (2024) -
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)