Decoupling Common and Unique Representations for Multimodal Self-supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yi, Albrecht, Conrad M, Braham, Nassim Ait Ali, Liu, Chenying, Xiong, Zhitong, Zhu, Xiao Xiang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SSL4EO-S12 v1.1: A Multimodal, Multiseasonal Dataset for Pretraining, Updated
by: Blumenstiel, Benedikt, et al.
Published: (2025)
by: Blumenstiel, Benedikt, et al.
Published: (2025)
SpectralEarth-FM: Bringing Hyperspectral Imagery into Multimodal Earth Observation Pretraining
by: Braham, Nassim Ait Ali, et al.
Published: (2026)
by: Braham, Nassim Ait Ali, et al.
Published: (2026)
SpectralEarth: Training Hyperspectral Foundation Models at Scale
by: Braham, Nassim Ait Ali, et al.
Published: (2024)
by: Braham, Nassim Ait Ali, et al.
Published: (2024)
CromSS: Cross-modal pre-training with noisy labels for remote sensing image segmentation
by: Liu, Chenying, et al.
Published: (2024)
by: Liu, Chenying, et al.
Published: (2024)
Hierarchical Semi-Supervised Active Learning for Remote Sensing
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Task Specific Pretraining with Noisy Labels for Remote Sensing Image Segmentation
by: Liu, Chenying, et al.
Published: (2024)
by: Liu, Chenying, et al.
Published: (2024)
Prospects for Mitigating Spectral Variability in Tropical Species Classification Using Self-Supervised Learning
by: Prieur, Colin, et al.
Published: (2025)
by: Prieur, Colin, et al.
Published: (2025)
AIO2: Online Correction of Object Labels for Deep Learning with Incomplete Annotation in Remote Sensing Image Segmentation
by: Liu, Chenying, et al.
Published: (2024)
by: Liu, Chenying, et al.
Published: (2024)
Multi-Label Guided Soft Contrastive Learning for Efficient Earth Observation Pretraining
by: Wang, Yi, et al.
Published: (2024)
by: Wang, Yi, et al.
Published: (2024)
Hyperspectral Vision Transformers for Greenhouse Gas Estimations from Space
by: Avilés, Ruben Gonzalez, et al.
Published: (2025)
by: Avilés, Ruben Gonzalez, et al.
Published: (2025)
One for All: Toward Unified Foundation Models for Earth Vision
by: Xiong, Zhitong, et al.
Published: (2024)
by: Xiong, Zhitong, et al.
Published: (2024)
HyBiomass: Global Hyperspectral Imagery Benchmark Dataset for Evaluating Geospatial Foundation Models in Forest Aboveground Biomass Estimation
by: Banze, Aaron, et al.
Published: (2025)
by: Banze, Aaron, et al.
Published: (2025)
LandSegmenter: Towards a Flexible Foundation Model for Land Use and Land Cover Mapping
by: Liu, Chenying, et al.
Published: (2025)
by: Liu, Chenying, et al.
Published: (2025)
Self-supervised Audiovisual Representation Learning for Remote Sensing Data
by: Heidler, Konrad, et al.
Published: (2021)
by: Heidler, Konrad, et al.
Published: (2021)
Adaptive Gradient Calibration for Single-Positive Multi-Label Learning in Remote Sensing Image Scene Classification
by: Liu, Chenying, et al.
Published: (2025)
by: Liu, Chenying, et al.
Published: (2025)
EarthNets: Empowering AI in Earth Observation
by: Xiong, Zhitong, et al.
Published: (2022)
by: Xiong, Zhitong, et al.
Published: (2022)
ChatEarthNet: A Global-Scale Image-Text Dataset Empowering Vision-Language Geo-Foundation Models
by: Yuan, Zhenghang, et al.
Published: (2024)
by: Yuan, Zhenghang, et al.
Published: (2024)
VG-SSL: Benchmarking Self-supervised Representation Learning Approaches for Visual Geo-localization
by: Xiao, Jiuhong, et al.
Published: (2023)
by: Xiao, Jiuhong, et al.
Published: (2023)
UrbanSARFloods: Sentinel-1 SLC-Based Benchmark Dataset for Urban and Open-Area Flood Mapping
by: Zhao, Jie, et al.
Published: (2024)
by: Zhao, Jie, et al.
Published: (2024)
Representation Space Constrained Learning with Modality Decoupling for Multimodal Object Detection
by: Shao, YiKang, et al.
Published: (2025)
by: Shao, YiKang, et al.
Published: (2025)
Self-Supervised Video Representation Learning in a Heuristic Decoupled Perspective
by: Song, Zeen, et al.
Published: (2024)
by: Song, Zeen, et al.
Published: (2024)
AutoLCZ: Towards Automatized Local Climate Zone Mapping from Rule-Based Remote Sensing
by: Liu, Chenying, et al.
Published: (2024)
by: Liu, Chenying, et al.
Published: (2024)
Towards a Unified Copernicus Foundation Model for Earth Vision
by: Wang, Yi, et al.
Published: (2025)
by: Wang, Yi, et al.
Published: (2025)
PolyGNN: Polyhedron-based Graph Neural Network for 3D Building Reconstruction from Point Clouds
by: Chen, Zhaiyu, et al.
Published: (2023)
by: Chen, Zhaiyu, et al.
Published: (2023)
Slimmable Networks for Contrastive Self-supervised Learning
by: Zhao, Shuai, et al.
Published: (2022)
by: Zhao, Shuai, et al.
Published: (2022)
EO-VAE: Towards A Multi-sensor Tokenizer for Earth Observation Data
by: Lehmann, Nils, et al.
Published: (2026)
by: Lehmann, Nils, et al.
Published: (2026)
Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models
by: Xiao, Junyuan, et al.
Published: (2026)
by: Xiao, Junyuan, et al.
Published: (2026)
Exploring Structural Degradation in Dense Representations for Self-supervised Learning
by: Dai, Siran, et al.
Published: (2025)
by: Dai, Siran, et al.
Published: (2025)
DOFA-CLIP: Multimodal Vision-Language Foundation Models for Earth Observation
by: Xiong, Zhitong, et al.
Published: (2025)
by: Xiong, Zhitong, et al.
Published: (2025)
Self-supervised Learning of Hybrid Part-aware 3D Representations of 2D Gaussians and Superquadrics
by: Gao, Zhirui, et al.
Published: (2024)
by: Gao, Zhirui, et al.
Published: (2024)
Constrained Multiview Representation for Self-supervised Contrastive Learning
by: Dai, Siyuan, et al.
Published: (2024)
by: Dai, Siyuan, et al.
Published: (2024)
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
by: Dai, Siran, et al.
Published: (2025)
by: Dai, Siran, et al.
Published: (2025)
Robust Multimodal Learning via Representation Decoupling
by: Wei, Shicai, et al.
Published: (2024)
by: Wei, Shicai, et al.
Published: (2024)
Neural Plasticity-Inspired Multimodal Foundation Model for Earth Observation
by: Xiong, Zhitong, et al.
Published: (2024)
by: Xiong, Zhitong, et al.
Published: (2024)
Data-Centric Benchmark for Label Noise Estimation and Ranking in Remote Sensing Image Segmentation
by: Nogueira, Keiller, et al.
Published: (2026)
by: Nogueira, Keiller, et al.
Published: (2026)
Collaboratively Self-supervised Video Representation Learning for Action Recognition
by: Zhang, Jie, et al.
Published: (2024)
by: Zhang, Jie, et al.
Published: (2024)
Mixed Autoencoder for Self-supervised Visual Representation Learning
by: Chen, Kai, et al.
Published: (2023)
by: Chen, Kai, et al.
Published: (2023)
SOLAR: Self-supervised Joint Learning for Symmetric Multimodal Retrieval
by: Yang, Wenjie, et al.
Published: (2026)
by: Yang, Wenjie, et al.
Published: (2026)
COMUNI: Decomposing Common and Unique Video Signals for Diffusion-based Video Generation
by: Sun, Mingzhen, et al.
Published: (2024)
by: Sun, Mingzhen, et al.
Published: (2024)
Self-supervised Photographic Image Layout Representation Learning
by: Zhao, Zhaoran, et al.
Published: (2024)
by: Zhao, Zhaoran, et al.
Published: (2024)
Similar Items
-
SSL4EO-S12 v1.1: A Multimodal, Multiseasonal Dataset for Pretraining, Updated
by: Blumenstiel, Benedikt, et al.
Published: (2025) -
SpectralEarth-FM: Bringing Hyperspectral Imagery into Multimodal Earth Observation Pretraining
by: Braham, Nassim Ait Ali, et al.
Published: (2026) -
SpectralEarth: Training Hyperspectral Foundation Models at Scale
by: Braham, Nassim Ait Ali, et al.
Published: (2024) -
CromSS: Cross-modal pre-training with noisy labels for remote sensing image segmentation
by: Liu, Chenying, et al.
Published: (2024) -
Hierarchical Semi-Supervised Active Learning for Remote Sensing
by: Huang, Wei, et al.
Published: (2025)