Self-Supervised Modality-Agnostic Pre-Training of Swin Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Talasila, Abhiroop, Maity, Maitreya, Priyakumar, U. Deva |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection
by: Khurana, Mehar, et al.
Published: (2024)
by: Khurana, Mehar, et al.
Published: (2024)
StrideNET: Swin Transformer for Terrain Recognition with Dynamic Roughness Extraction
by: Shelare, Maitreya, et al.
Published: (2024)
by: Shelare, Maitreya, et al.
Published: (2024)
SparseSwin: Swin Transformer with Sparse Transformer Block
by: Pinasthika, Krisna, et al.
Published: (2023)
by: Pinasthika, Krisna, et al.
Published: (2023)
Brain Hematoma Marker Recognition Using Multitask Learning: SwinTransformer and Swin-Unet
by: Hirata, Kodai, et al.
Published: (2025)
by: Hirata, Kodai, et al.
Published: (2025)
Masked Self-Supervised Pre-Training for Text Recognition Transformers on Large-Scale Datasets
by: Kišš, Martin, et al.
Published: (2025)
by: Kišš, Martin, et al.
Published: (2025)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
by: Cazenavette, George, et al.
Published: (2025)
by: Cazenavette, George, et al.
Published: (2025)
DualSwinFusionSeg: Multimodal Martian Landslide Segmentation via Dual Swin Transformer with Multi-Scale Fusion and UNet++
by: Kabir, Shahriar, et al.
Published: (2026)
by: Kabir, Shahriar, et al.
Published: (2026)
DIAR: Deep Image Alignment and Reconstruction using Swin Transformers
by: Kwiatkowski, Monika, et al.
Published: (2023)
by: Kwiatkowski, Monika, et al.
Published: (2023)
Bridging Diversity and Uncertainty in Active learning with Self-Supervised Pre-Training
by: Doucet, Paul, et al.
Published: (2024)
by: Doucet, Paul, et al.
Published: (2024)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
by: Chatterjee, Abhiroop, et al.
Published: (2025)
by: Chatterjee, Abhiroop, et al.
Published: (2025)
MentalBlackboard: Evaluating Spatial Visualization via Mathematical Transformations
by: Yilmaz, Nilay, et al.
Published: (2026)
by: Yilmaz, Nilay, et al.
Published: (2026)
Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet
by: Niu, Xin, et al.
Published: (2026)
by: Niu, Xin, et al.
Published: (2026)
DiNO-Diffusion. Scaling Medical Diffusion via Self-Supervised Pre-Training
by: Jimenez-Perez, Guillermo, et al.
Published: (2024)
by: Jimenez-Perez, Guillermo, et al.
Published: (2024)
Self-Supervised Vision Transformers for Writer Retrieval
by: Raven, Tim, et al.
Published: (2024)
by: Raven, Tim, et al.
Published: (2024)
Wild Visual Navigation: Fast Traversability Learning via Pre-Trained Models and Online Self-Supervision
by: Mattamala, Matías, et al.
Published: (2024)
by: Mattamala, Matías, et al.
Published: (2024)
IO Transformer: Evaluating SwinV2-Based Reward Models for Computer Vision
by: Meyer, Maxwell, et al.
Published: (2024)
by: Meyer, Maxwell, et al.
Published: (2024)
Self-Supervised Training with Autoencoders for Visual Anomaly Detection
by: Bauer, Alexander, et al.
Published: (2022)
by: Bauer, Alexander, et al.
Published: (2022)
Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-Training
by: Reddy, Arun, et al.
Published: (2023)
by: Reddy, Arun, et al.
Published: (2023)
Axial-UNet: A Neural Weather Model for Precipitation Nowcasting
by: Mamtani, Sumit, et al.
Published: (2025)
by: Mamtani, Sumit, et al.
Published: (2025)
Enhancing Cardiovascular Disease Prediction through Multi-Modal Self-Supervised Learning
by: Girlanda, Francesco, et al.
Published: (2024)
by: Girlanda, Francesco, et al.
Published: (2024)
EA-Swin: An Embedding-Agnostic Swin Transformer for AI-Generated Video Detection
by: Mai, Hung, et al.
Published: (2026)
by: Mai, Hung, et al.
Published: (2026)
Self-Supervised Pre-Training for Deep Image Prior-Based Robust PET Image Denoising
by: Onishi, Yuya, et al.
Published: (2023)
by: Onishi, Yuya, et al.
Published: (2023)
Parallel Swin Transformer-Enhanced 3D MRI-to-CT Synthesis for MRI-Only Radiotherapy Planning
by: Dorjsembe, Zolnamar, et al.
Published: (2026)
by: Dorjsembe, Zolnamar, et al.
Published: (2026)
XTransfer: Modality-Agnostic Few-Shot Model Transfer for Human Sensing at the Edge
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Residual-SwinCA-Net: A Channel-Aware Integrated Residual CNN-Swin Transformer for Malignant Lesion Segmentation in BUSI
by: Naz, Saeeda, et al.
Published: (2025)
by: Naz, Saeeda, et al.
Published: (2025)
Multi-instance Learning as Downstream Task of Self-Supervised Learning-based Pre-trained Model
by: Matsuishi, Koki, et al.
Published: (2025)
by: Matsuishi, Koki, et al.
Published: (2025)
Self-Supervised Learning for Pre-training Capsule Networks: Overcoming Medical Imaging Dataset Challenges
by: El-Shimy, Heba, et al.
Published: (2025)
by: El-Shimy, Heba, et al.
Published: (2025)
Supervised Multi-Modal Fission Learning
by: Mao, Lingchao, et al.
Published: (2024)
by: Mao, Lingchao, et al.
Published: (2024)
Depth-supervised NeRF: Fewer Views and Faster Training for Free
by: Deng, Kangle, et al.
Published: (2021)
by: Deng, Kangle, et al.
Published: (2021)
Enhancing Semi-Supervised Multi-View Graph Convolutional Networks via Supervised Contrastive Learning and Self-Training
by: Xiao, Huaiyuan, et al.
Published: (2025)
by: Xiao, Huaiyuan, et al.
Published: (2025)
Steering Rectified Flow Models in the Vector Field for Controlled Image Generation
by: Patel, Maitreya, et al.
Published: (2024)
by: Patel, Maitreya, et al.
Published: (2024)
Learning to Recorrupt: Noise Distribution Agnostic Self-Supervised Image Denoising
by: Monroy, Brayan, et al.
Published: (2026)
by: Monroy, Brayan, et al.
Published: (2026)
An Empirical Study of Accuracy-Robustness Tradeoff and Training Efficiency in Self-Supervised Learning
by: Ghofrani, Fatemeh, et al.
Published: (2025)
by: Ghofrani, Fatemeh, et al.
Published: (2025)
Graph Memory: A Structured and Interpretable Framework for Modality-Agnostic Embedding-Based Inference
by: Oliveira, Artur A., et al.
Published: (2025)
by: Oliveira, Artur A., et al.
Published: (2025)
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy
by: Roth, Marcel, et al.
Published: (2024)
by: Roth, Marcel, et al.
Published: (2024)
Pre-training Vision Transformers with Formula-driven Supervised Learning
by: Kataoka, Hirokatsu, et al.
Published: (2022)
by: Kataoka, Hirokatsu, et al.
Published: (2022)
ProFeAT: Projected Feature Adversarial Training for Self-Supervised Learning of Robust Representations
by: Addepalli, Sravanti, et al.
Published: (2024)
by: Addepalli, Sravanti, et al.
Published: (2024)
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
by: Patel, Maitreya, et al.
Published: (2026)
by: Patel, Maitreya, et al.
Published: (2026)
Intelligent Anomaly Detection for Lane Rendering Using Transformer with Self-Supervised Pre-Training and Customized Fine-Tuning
by: Dong, Yongqi, et al.
Published: (2023)
by: Dong, Yongqi, et al.
Published: (2023)
ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models
by: Patel, Maitreya, et al.
Published: (2023)
by: Patel, Maitreya, et al.
Published: (2023)
Similar Items
-
Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection
by: Khurana, Mehar, et al.
Published: (2024) -
StrideNET: Swin Transformer for Terrain Recognition with Dynamic Roughness Extraction
by: Shelare, Maitreya, et al.
Published: (2024) -
SparseSwin: Swin Transformer with Sparse Transformer Block
by: Pinasthika, Krisna, et al.
Published: (2023) -
Brain Hematoma Marker Recognition Using Multitask Learning: SwinTransformer and Swin-Unet
by: Hirata, Kodai, et al.
Published: (2025) -
Masked Self-Supervised Pre-Training for Text Recognition Transformers on Large-Scale Datasets
by: Kišš, Martin, et al.
Published: (2025)