Exploring the Coordination of Frequency and Attention in Masked Image Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Gui, Jie, Chen, Tuo, Dong, Minjing, Liu, Zhengqi, Luo, Hao, Kwok, James Tin-Yau, Tang, Yuan Yan |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revisiting Adversarial Training under Hyperspectral Image
by: Zhang, Weihua, et al.
Published: (2025)
by: Zhang, Weihua, et al.
Published: (2025)
A Survey on Small Sample Imbalance Problem: Metrics, Feature Analysis, and Solutions
by: Zhao, Shuxian, et al.
Published: (2025)
by: Zhao, Shuxian, et al.
Published: (2025)
CFVNet: An End-to-End Cancelable Finger Vein Network for Recognition
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
Fooling the Image Dehazing Models by First Order Gradient
by: Gui, Jie, et al.
Published: (2023)
by: Gui, Jie, et al.
Published: (2023)
Improving Fast Adversarial Training via Self-Knowledge Guidance
by: Jiang, Chengze, et al.
Published: (2024)
by: Jiang, Chengze, et al.
Published: (2024)
ColorVein: Colorful Cancelable Vein Biometrics
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Toward Fine-Grained Facial Control in 3D Talking Head Generation
by: Xie, Shaoyang, et al.
Published: (2026)
by: Xie, Shaoyang, et al.
Published: (2026)
Unrevealed Threats: A Comprehensive Study of the Adversarial Robustness of Underwater Image Enhancement Models
by: Zhai, Siyu, et al.
Published: (2024)
by: Zhai, Siyu, et al.
Published: (2024)
BitC-3DGS: High-Capacity 3D Gaussian Splatting Watermarking via Bit Compression
by: Bi, Yuquan, et al.
Published: (2026)
by: Bi, Yuquan, et al.
Published: (2026)
Underwater Organism Color Enhancement via Color Code Decomposition, Adaptation and Interpolation
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
Backdooring Self-Supervised Contrastive Learning by Noisy Alignment
by: Chen, Tuo, et al.
Published: (2025)
by: Chen, Tuo, et al.
Published: (2025)
Survey of Adversarial Robustness in Multimodal Large Language Models
by: Jiang, Chengze, et al.
Published: (2025)
by: Jiang, Chengze, et al.
Published: (2025)
Diversifying Counterattacks: Orthogonal Exploration for Robust CLIP Inference
by: Jiang, Chengze, et al.
Published: (2025)
by: Jiang, Chengze, et al.
Published: (2025)
Improving Fast Adversarial Training Paradigm: An Example Taxonomy Perspective
by: Gui, Jie, et al.
Published: (2024)
by: Gui, Jie, et al.
Published: (2024)
Efficient Image-to-Image Diffusion Classifier for Adversarial Robustness
by: Mei, Hefei, et al.
Published: (2024)
by: Mei, Hefei, et al.
Published: (2024)
Color Image Set Recognition Based on Quaternionic Grassmannians
by: Wang, Xiang Xiang, et al.
Published: (2025)
by: Wang, Xiang Xiang, et al.
Published: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
by: Zhu, Younan, et al.
Published: (2025)
by: Zhu, Younan, et al.
Published: (2025)
Automated Dominative Subspace Mining for Efficient Neural Architecture Search
by: Chen, Yaofo, et al.
Published: (2022)
by: Chen, Yaofo, et al.
Published: (2022)
Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
Beyond One-Hot Labels: Semantic Mixing for Model Calibration
by: Luo, Haoyang, et al.
Published: (2025)
by: Luo, Haoyang, et al.
Published: (2025)
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
MaskSAM: Towards Auto-prompt SAM with Mask Classification for Volumetric Medical Image Segmentation
by: Xie, Bin, et al.
Published: (2024)
by: Xie, Bin, et al.
Published: (2024)
PA-Attack: Guiding Gray-Box Attacks on LVLM Vision Encoders with Prototypes and Attention
by: Mei, Hefei, et al.
Published: (2026)
by: Mei, Hefei, et al.
Published: (2026)
Frequency Autoregressive Image Generation with Continuous Tokens
by: Yu, Hu, et al.
Published: (2025)
by: Yu, Hu, et al.
Published: (2025)
Efficient Masked Image Compression with Position-Indexed Self-Attention
by: Dai, Chengjie, et al.
Published: (2025)
by: Dai, Chengjie, et al.
Published: (2025)
Timestep-Aware Block Masking for Efficient Diffusion Model Inference
by: He, Haodong, et al.
Published: (2026)
by: He, Haodong, et al.
Published: (2026)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
by: Chen, Wenchao, et al.
Published: (2024)
by: Chen, Wenchao, et al.
Published: (2024)
Efficient Diffusion-Based 3D Human Pose Estimation with Hierarchical Temporal Pruning
by: Bi, Yuquan, et al.
Published: (2025)
by: Bi, Yuquan, et al.
Published: (2025)
Learning Correction Errors via Frequency-Self Attention for Blind Image Super-Resolution
by: Sun, Haochen, et al.
Published: (2024)
by: Sun, Haochen, et al.
Published: (2024)
Feature Clipping for Uncertainty Calibration
by: Tao, Linwei, et al.
Published: (2024)
by: Tao, Linwei, et al.
Published: (2024)
A Semi-supervised Nighttime Dehazing Baseline with Spatial-Frequency Aware and Realistic Brightness Constraint
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
Restore Anything with Masks: Leveraging Mask Image Modeling for Blind All-in-One Image Restoration
by: Qin, Chu-Jie, et al.
Published: (2024)
by: Qin, Chu-Jie, et al.
Published: (2024)
Training-Free Pretrained Model Merging
by: Xu, Zhengqi, et al.
Published: (2024)
by: Xu, Zhengqi, et al.
Published: (2024)
Pure-Pass: Fine-Grained, Adaptive Masking for Dynamic Token-Mixing Routing in Lightweight Image Super-Resolution
by: Wu, Junyu, et al.
Published: (2025)
by: Wu, Junyu, et al.
Published: (2025)
VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language Models
by: Mei, Hefei, et al.
Published: (2025)
by: Mei, Hefei, et al.
Published: (2025)
Coordinate-Based Dual-Constrained Autoregressive Motion Generation
by: Ding, Kang, et al.
Published: (2026)
by: Ding, Kang, et al.
Published: (2026)
Towards Efficient Diffusion-Based Image Editing with Instant Attention Masks
by: Zou, Siyu, et al.
Published: (2024)
by: Zou, Siyu, et al.
Published: (2024)
VSSD: Vision Mamba with Non-Causal State Space Duality
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
Catching the Details: Self-Distilled RoI Predictors for Fine-Grained MLLM Perception
by: Shi, Yuheng, et al.
Published: (2025)
by: Shi, Yuheng, et al.
Published: (2025)
Efficient Rectified Flow for Image Fusion
by: Wang, Zirui, et al.
Published: (2025)
by: Wang, Zirui, et al.
Published: (2025)
Similar Items
-
Revisiting Adversarial Training under Hyperspectral Image
by: Zhang, Weihua, et al.
Published: (2025) -
A Survey on Small Sample Imbalance Problem: Metrics, Feature Analysis, and Solutions
by: Zhao, Shuxian, et al.
Published: (2025) -
CFVNet: An End-to-End Cancelable Finger Vein Network for Recognition
by: Wang, Yifan, et al.
Published: (2024) -
Fooling the Image Dehazing Models by First Order Gradient
by: Gui, Jie, et al.
Published: (2023) -
Improving Fast Adversarial Training via Self-Knowledge Guidance
by: Jiang, Chengze, et al.
Published: (2024)