Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yibing, Gupta, Abhinav, Morgado, Pedro |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
by: Hermosilla, Pedro, et al.
Published: (2025)
by: Hermosilla, Pedro, et al.
Published: (2025)
Multi-Scale Neighborhood Occupancy Masked Autoencoder for Self-Supervised Learning in LiDAR Point Clouds
by: Abdelsamad, Mohamed, et al.
Published: (2025)
by: Abdelsamad, Mohamed, et al.
Published: (2025)
INoD: Injected Noise Discriminator for Self-Supervised Representation Learning in Agricultural Fields
by: Hindel, Julia, et al.
Published: (2023)
by: Hindel, Julia, et al.
Published: (2023)
Masked Modeling for Self-supervised Representation Learning on Vision and Beyond
by: Li, Siyuan, et al.
Published: (2023)
by: Li, Siyuan, et al.
Published: (2023)
Boosting Semi-Supervised Medical Image Segmentation via Masked Image Consistency and Discrepancy Learning
by: Zhou, Pengcheng, et al.
Published: (2025)
by: Zhou, Pengcheng, et al.
Published: (2025)
MaskAnyNet: Rethinking Masked Image Regions as Valuable Information in Supervised Learning
by: Hong, Jingshan, et al.
Published: (2025)
by: Hong, Jingshan, et al.
Published: (2025)
Cluster and Predict Latent Patches for Improved Masked Image Modeling
by: Darcet, Timothée, et al.
Published: (2025)
by: Darcet, Timothée, et al.
Published: (2025)
Mask & Match: Learning to Recognize Handwritten Math with Self-Supervised Attention
by: Mitra, Shree, et al.
Published: (2025)
by: Mitra, Shree, et al.
Published: (2025)
Self-Supervised Learning for Interventional Image Analytics: Towards Robust Device Trackers
by: Islam, Saahil, et al.
Published: (2024)
by: Islam, Saahil, et al.
Published: (2024)
Zero-Shot and Supervised Bird Image Segmentation Using Foundation Models: A Dual-Pipeline Approach with Grounding DINO~1.5, YOLOv11, and SAM~2.1
by: Munagala, Abhinav
Published: (2026)
by: Munagala, Abhinav
Published: (2026)
MAESIL: Masked Autoencoder for Enhanced Self-supervised Medical Image Learning
by: Kim, Kyeonghun, et al.
Published: (2026)
by: Kim, Kyeonghun, et al.
Published: (2026)
BiGR: Harnessing Binary Latent Codes for Image Generation and Improved Visual Representation Capabilities
by: Hao, Shaozhe, et al.
Published: (2024)
by: Hao, Shaozhe, et al.
Published: (2024)
UniCorn: Towards Self-Improving Unified Multimodal Models through Self-Generated Supervision
by: Han, Ruiyan, et al.
Published: (2026)
by: Han, Ruiyan, et al.
Published: (2026)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
by: Lee, Jaeseong, et al.
Published: (2024)
by: Lee, Jaeseong, et al.
Published: (2024)
FSFM: A Generalizable Face Security Foundation Model via Self-Supervised Facial Representation Learning
by: Wang, Gaojian, et al.
Published: (2024)
by: Wang, Gaojian, et al.
Published: (2024)
AdvDINO: Domain-Adversarial Self-Supervised Representation Learning for Spatial Proteomics
by: Su, Stella, et al.
Published: (2025)
by: Su, Stella, et al.
Published: (2025)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
From Pixels to Components: Eigenvector Masking for Visual Representation Learning
by: Bizeul, Alice, et al.
Published: (2025)
by: Bizeul, Alice, et al.
Published: (2025)
DiSSECT: Structuring Transfer-Ready Medical Image Representations through Discrete Self-Supervision
by: Singh, Azad, et al.
Published: (2025)
by: Singh, Azad, et al.
Published: (2025)
Learning Disentangled Representation in Object-Centric Models for Visual Dynamics Prediction via Transformers
by: Gandhi, Sanket, et al.
Published: (2024)
by: Gandhi, Sanket, et al.
Published: (2024)
LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model
by: Mei, Xiaodong, et al.
Published: (2026)
by: Mei, Xiaodong, et al.
Published: (2026)
Learning from Gene Names, Expression Values and Images: Contrastive Masked Text-Image Pretraining for Spatial Transcriptomics Representation Learning
by: Qian, Jiahe, et al.
Published: (2025)
by: Qian, Jiahe, et al.
Published: (2025)
S4-Driver: Scalable Self-Supervised Driving Multimodal Large Language Modelwith Spatio-Temporal Visual Representation
by: Xie, Yichen, et al.
Published: (2025)
by: Xie, Yichen, et al.
Published: (2025)
SCE-MAE: Selective Correspondence Enhancement with Masked Autoencoder for Self-Supervised Landmark Estimation
by: Yin, Kejia, et al.
Published: (2024)
by: Yin, Kejia, et al.
Published: (2024)
Dynamic Entity-Masked Graph Diffusion Model for histopathological image Representation Learning
by: Zhuang, Zhenfeng, et al.
Published: (2024)
by: Zhuang, Zhenfeng, et al.
Published: (2024)
Integrating Sequence and Image Modeling in Irregular Medical Time Series Through Self-Supervised Learning
by: Chen, Liuqing, et al.
Published: (2025)
by: Chen, Liuqing, et al.
Published: (2025)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023)
by: Chin, Zhi-Yi, et al.
Published: (2023)
i-MAE: Are Latent Representations in Masked Autoencoders Linearly Separable?
by: Zhang, Kevin, et al.
Published: (2022)
by: Zhang, Kevin, et al.
Published: (2022)
Self-Supervised Cross-Modal Learning for Image-to-Point Cloud Registration
by: Wang, Xingmei, et al.
Published: (2025)
by: Wang, Xingmei, et al.
Published: (2025)
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025)
by: Lee, Sua, et al.
Published: (2025)
Smiling Women Pitching Down: Auditing Representational and Presentational Gender Biases in Image Generative AI
by: Sun, Luhang, et al.
Published: (2023)
by: Sun, Luhang, et al.
Published: (2023)
Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation
by: Zhu, Chunzheng, et al.
Published: (2026)
by: Zhu, Chunzheng, et al.
Published: (2026)
Masked Latent Transformer with the Random Masking Ratio to Advance the Diagnosis of Dental Fluorosis
by: Wu, Yun, et al.
Published: (2024)
by: Wu, Yun, et al.
Published: (2024)
Self-Supervised Masked Autoencoders with Dense-Unet for Coronary Calcium Removal in limited CT Data
by: Chen, Mo
Published: (2025)
by: Chen, Mo
Published: (2025)
VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs
by: Li, Qiaoru, et al.
Published: (2026)
by: Li, Qiaoru, et al.
Published: (2026)
Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation
by: Si, Qi, et al.
Published: (2025)
by: Si, Qi, et al.
Published: (2025)
Shifting to Machine Supervision: Annotation-Efficient Semi and Self-Supervised Learning for Automatic Medical Image Segmentation and Classification
by: Singh, Pranav, et al.
Published: (2023)
by: Singh, Pranav, et al.
Published: (2023)
Latent Modulated Function for Computational Optimal Continuous Image Representation
by: He, Zongyao, et al.
Published: (2024)
by: He, Zongyao, et al.
Published: (2024)
Prompt-SID: Learning Structural Representation Prompt via Latent Diffusion for Single-Image Denoising
by: Li, Huaqiu, et al.
Published: (2025)
by: Li, Huaqiu, et al.
Published: (2025)
Enhanced Object Tracking by Self-Supervised Auxiliary Depth Estimation Learning
by: Wei, Zhenyu, et al.
Published: (2024)
by: Wei, Zhenyu, et al.
Published: (2024)
Similar Items
-
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
by: Hermosilla, Pedro, et al.
Published: (2025) -
Multi-Scale Neighborhood Occupancy Masked Autoencoder for Self-Supervised Learning in LiDAR Point Clouds
by: Abdelsamad, Mohamed, et al.
Published: (2025) -
INoD: Injected Noise Discriminator for Self-Supervised Representation Learning in Agricultural Fields
by: Hindel, Julia, et al.
Published: (2023) -
Masked Modeling for Self-supervised Representation Learning on Vision and Beyond
by: Li, Siyuan, et al.
Published: (2023) -
Boosting Semi-Supervised Medical Image Segmentation via Masked Image Consistency and Discrepancy Learning
by: Zhou, Pengcheng, et al.
Published: (2025)