Mixed Autoencoder for Self-supervised Visual Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Kai, Liu, Zhili, Hong, Lanqing, Xu, Hang, Li, Zhenguo, Yeung, Dit-Yan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
by: Liu, Zhili, et al.
Published: (2024)
by: Liu, Zhili, et al.
Published: (2024)
TransformMix: Learning Transformation and Mixing Strategies from Data
by: Cheung, Tsz-Him, et al.
Published: (2024)
by: Cheung, Tsz-Him, et al.
Published: (2024)
Implicit Concept Removal of Diffusion Models
by: Liu, Zhili, et al.
Published: (2023)
by: Liu, Zhili, et al.
Published: (2023)
Eyes Closed, Safety On: Protecting Multimodal LLMs via Image-to-Text Transformation
by: Gou, Yunhao, et al.
Published: (2024)
by: Gou, Yunhao, et al.
Published: (2024)
Mixture of Cluster-conditional LoRA Experts for Vision-language Instruction Tuning
by: Gou, Yunhao, et al.
Published: (2023)
by: Gou, Yunhao, et al.
Published: (2023)
MagicDrive: Street View Generation with Diverse 3D Geometry Control
by: Gao, Ruiyuan, et al.
Published: (2023)
by: Gao, Ruiyuan, et al.
Published: (2023)
GeoDiffusion: Text-Prompted Geometric Control for Object Detection Data Generation
by: Chen, Kai, et al.
Published: (2023)
by: Chen, Kai, et al.
Published: (2023)
ECCV 2024 W-CODA: 1st Workshop on Multimodal Perception and Comprehension of Corner Cases in Autonomous Driving
by: Chen, Kai, et al.
Published: (2025)
by: Chen, Kai, et al.
Published: (2025)
Unified Triplet-Level Hallucination Evaluation for Large Vision-Language Models
by: Wu, Junjie, et al.
Published: (2024)
by: Wu, Junjie, et al.
Published: (2024)
Self-supervised Transformation Learning for Equivariant Representations
by: Yu, Jaemyung, et al.
Published: (2025)
by: Yu, Jaemyung, et al.
Published: (2025)
Learning High-resolution Vector Representation from Multi-Camera Images for 3D Object Detection
by: Chen, Zhili, et al.
Published: (2024)
by: Chen, Zhili, et al.
Published: (2024)
SARMAE: Masked Autoencoder for SAR Representation Learning
by: Liu, Danxu, et al.
Published: (2025)
by: Liu, Danxu, et al.
Published: (2025)
Automated Evaluation of Large Vision-Language Models on Self-driving Corner Cases
by: Chen, Kai, et al.
Published: (2024)
by: Chen, Kai, et al.
Published: (2024)
DetDiffusion: Synergizing Generative and Perceptive Models for Enhanced Data Generation and Perception
by: Wang, Yibo, et al.
Published: (2024)
by: Wang, Yibo, et al.
Published: (2024)
TrackDiffusion: Tracklet-Conditioned Video Generation via Diffusion Models
by: Li, Pengxiang, et al.
Published: (2023)
by: Li, Pengxiang, et al.
Published: (2023)
Robust Representation Learning in Masked Autoencoders
by: Shrivastava, Anika, et al.
Published: (2026)
by: Shrivastava, Anika, et al.
Published: (2026)
Self-Supervised Training with Autoencoders for Visual Anomaly Detection
by: Bauer, Alexander, et al.
Published: (2022)
by: Bauer, Alexander, et al.
Published: (2022)
Visualizing the loss landscape of Self-supervised Vision Transformer
by: Lee, Youngwan, et al.
Published: (2024)
by: Lee, Youngwan, et al.
Published: (2024)
Dual Risk Minimization: Towards Next-Level Robustness in Fine-tuning Zero-Shot Models
by: Li, Kaican, et al.
Published: (2024)
by: Li, Kaican, et al.
Published: (2024)
BRIDLE: Generalized Self-supervised Learning with Quantization
by: Nguyen, Hoang M., et al.
Published: (2025)
by: Nguyen, Hoang M., et al.
Published: (2025)
MTS-DMAE: Dual-Masked Autoencoder for Unsupervised Multivariate Time Series Representation Learning
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
Sparse Autoencoders for Interpretable Medical Image Representation Learning
by: Wesp, Philipp, et al.
Published: (2026)
by: Wesp, Philipp, et al.
Published: (2026)
Attention-Guided Masked Autoencoders For Learning Image Representations
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Self-Organizing Visual Prototypes for Non-Parametric Representation Learning
by: Silva, Thalles, et al.
Published: (2025)
by: Silva, Thalles, et al.
Published: (2025)
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning
by: Zhou, Xinning, et al.
Published: (2025)
by: Zhou, Xinning, et al.
Published: (2025)
SRL-SOA: Self-Representation Learning with Sparse 1D-Operational Autoencoder for Hyperspectral Image Band Selection
by: Ahishali, Mete, et al.
Published: (2022)
by: Ahishali, Mete, et al.
Published: (2022)
Diffusion Transformers with Representation Autoencoders
by: Zheng, Boyang, et al.
Published: (2025)
by: Zheng, Boyang, et al.
Published: (2025)
Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling
by: Xie, Tianyu, et al.
Published: (2025)
by: Xie, Tianyu, et al.
Published: (2025)
Reasoning-Aligned Perception Decoupling for Scalable Multi-modal Reasoning
by: Gou, Yunhao, et al.
Published: (2025)
by: Gou, Yunhao, et al.
Published: (2025)
MedFLIP: Medical Vision-and-Language Self-supervised Fast Pre-Training with Masked Autoencoder
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
Semi-Self Representation Learning for Crowdsourced WiFi Trajectories
by: Kuo, Yu-Lin, et al.
Published: (2025)
by: Kuo, Yu-Lin, et al.
Published: (2025)
Beyond Instance Consistency: Investigating View Diversity in Self-supervised Learning
by: Qin, Huaiyuan, et al.
Published: (2025)
by: Qin, Huaiyuan, et al.
Published: (2025)
Augmentation-aware Self-supervised Learning with Conditioned Projector
by: Przewięźlikowski, Marcin, et al.
Published: (2023)
by: Przewięźlikowski, Marcin, et al.
Published: (2023)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
by: Gidaris, Spyros, et al.
Published: (2023)
by: Gidaris, Spyros, et al.
Published: (2023)
Masked Autoencoders for Ultrasound Signals: Robust Representation Learning for Downstream Applications
by: Roßteutscher, Immanuel, et al.
Published: (2025)
by: Roßteutscher, Immanuel, et al.
Published: (2025)
Visual Imitation Learning with Calibrated Contrastive Representation
by: Wang, Yunke, et al.
Published: (2024)
by: Wang, Yunke, et al.
Published: (2024)
Collaboratively Self-supervised Video Representation Learning for Action Recognition
by: Zhang, Jie, et al.
Published: (2024)
by: Zhang, Jie, et al.
Published: (2024)
S-JEA: Stacked Joint Embedding Architectures for Self-Supervised Visual Representation Learning
by: Manová, Alžběta, et al.
Published: (2023)
by: Manová, Alžběta, et al.
Published: (2023)
Extracting Symbolic Sequences from Visual Representations via Self-Supervised Learning
by: Pozos, Victor Sebastian Martinez, et al.
Published: (2025)
by: Pozos, Victor Sebastian Martinez, et al.
Published: (2025)
Self-Supervised Siamese Autoencoders
by: Baier, Friederike, et al.
Published: (2023)
by: Baier, Friederike, et al.
Published: (2023)
Similar Items
-
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
by: Liu, Zhili, et al.
Published: (2024) -
TransformMix: Learning Transformation and Mixing Strategies from Data
by: Cheung, Tsz-Him, et al.
Published: (2024) -
Implicit Concept Removal of Diffusion Models
by: Liu, Zhili, et al.
Published: (2023) -
Eyes Closed, Safety On: Protecting Multimodal LLMs via Image-to-Text Transformation
by: Gou, Yunhao, et al.
Published: (2024) -
Mixture of Cluster-conditional LoRA Experts for Vision-language Instruction Tuning
by: Gou, Yunhao, et al.
Published: (2023)