Probing the Mid-level Vision Capabilities of Self-Supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xuweiyi, Marks, Markus, Cheng, Zezhou |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
by: Chen, Xuweiyi, et al.
Published: (2026)
by: Chen, Xuweiyi, et al.
Published: (2026)
Empowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
by: Zhou, Wentao, et al.
Published: (2025)
by: Zhou, Wentao, et al.
Published: (2025)
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
by: Wang, Boyang, et al.
Published: (2025)
by: Wang, Boyang, et al.
Published: (2025)
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
by: Chen, Xuweiyi, et al.
Published: (2025)
by: Chen, Xuweiyi, et al.
Published: (2025)
Open Vocabulary Monocular 3D Object Detection
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
by: Chen, Xuweiyi, et al.
Published: (2025)
by: Chen, Xuweiyi, et al.
Published: (2025)
Learning Keypoints for Multi-Agent Behavior Analysis using Self-Supervision
by: Khalil, Daniel, et al.
Published: (2024)
by: Khalil, Daniel, et al.
Published: (2024)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
by: Xia, Tian, et al.
Published: (2024)
by: Xia, Tian, et al.
Published: (2024)
A Closer Look at Benchmarking Self-Supervised Pre-training with Image Classification
by: Marks, Markus, et al.
Published: (2024)
by: Marks, Markus, et al.
Published: (2024)
Adversarial Robustness of Discriminative Self-Supervised Learning in Vision
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
Next-Embedding Prediction Makes Strong Vision Learners
by: Xu, Sihan, et al.
Published: (2025)
by: Xu, Sihan, et al.
Published: (2025)
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
Intervention-Based Self-Supervised Learning: A Causal Probe Paradigm for Remote Photoplethysmography
by: Niu, Zhiyi, et al.
Published: (2026)
by: Niu, Zhiyi, et al.
Published: (2026)
Stealthy Backdoor Attack in Self-Supervised Learning Vision Encoders for Large Vision Language Models
by: Liu, Zhaoyi, et al.
Published: (2025)
by: Liu, Zhaoyi, et al.
Published: (2025)
Self-Supervised Vision Transformer for Enhanced Virtual Clothes Try-On
by: Lu, Lingxiao, et al.
Published: (2024)
by: Lu, Lingxiao, et al.
Published: (2024)
Hierarchical Text-to-Vision Self Supervised Alignment for Improved Histopathology Representation Learning
by: Watawana, Hasindri, et al.
Published: (2024)
by: Watawana, Hasindri, et al.
Published: (2024)
Exploring the Effect of Dataset Diversity in Self-Supervised Learning for Surgical Computer Vision
by: Jaspers, Tim J. M., et al.
Published: (2024)
by: Jaspers, Tim J. M., et al.
Published: (2024)
Pattern Integration and Enhancement Vision Transformer for Self-Supervised Learning in Remote Sensing
by: Lu, Kaixuan, et al.
Published: (2024)
by: Lu, Kaixuan, et al.
Published: (2024)
Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation
by: Yang, Longzhen, et al.
Published: (2025)
by: Yang, Longzhen, et al.
Published: (2025)
SAVeD: Learning to Denoise Low-SNR Video for Improved Downstream Performance
by: Stathatos, Suzanne, et al.
Published: (2025)
by: Stathatos, Suzanne, et al.
Published: (2025)
VideoSSR: Video Self-Supervised Reinforcement Learning
by: He, Zefeng, et al.
Published: (2025)
by: He, Zefeng, et al.
Published: (2025)
Self-Supervised Contrastive Learning for Multi-Label Images
by: Chen, Jiale
Published: (2025)
by: Chen, Jiale
Published: (2025)
Enhancing Representations through Heterogeneous Self-Supervised Learning
by: Li, Zhong-Yu, et al.
Published: (2023)
by: Li, Zhong-Yu, et al.
Published: (2023)
Unleashing the Power of Image-Tabular Self-Supervised Learning via Breaking Cross-Tabular Barriers
by: Fu, Yibing, et al.
Published: (2025)
by: Fu, Yibing, et al.
Published: (2025)
Liveness Detection in Computer Vision: Transformer-based Self-Supervised Learning for Face Anti-Spoofing
by: Keresh, Arman, et al.
Published: (2024)
by: Keresh, Arman, et al.
Published: (2024)
Depth-Wise Representation Development Under Blockwise Self-Supervised Learning for Video Vision Transformers
by: Römer, Jonas, et al.
Published: (2026)
by: Römer, Jonas, et al.
Published: (2026)
Revisiting End-to-End Learning with Slide-level Supervision in Computational Pathology
by: Tang, Wenhao, et al.
Published: (2025)
by: Tang, Wenhao, et al.
Published: (2025)
Multi-Object Hallucination in Vision-Language Models
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
Backdooring Self-Supervised Contrastive Learning by Noisy Alignment
by: Chen, Tuo, et al.
Published: (2025)
by: Chen, Tuo, et al.
Published: (2025)
On the Discriminability of Self-Supervised Representation Learning
by: Song, Zeen, et al.
Published: (2024)
by: Song, Zeen, et al.
Published: (2024)
Information Flow in Self-Supervised Learning
by: Tan, Zhiquan, et al.
Published: (2023)
by: Tan, Zhiquan, et al.
Published: (2023)
Diminishing Returns in Self-Supervised Learning
by: Bridge, Oli, et al.
Published: (2025)
by: Bridge, Oli, et al.
Published: (2025)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
by: Lee, Seungho, et al.
Published: (2024)
by: Lee, Seungho, et al.
Published: (2024)
Entropy Reveals Block Importance in Masked Self-Supervised Vision Transformers
by: Xiang, Peihao, et al.
Published: (2026)
by: Xiang, Peihao, et al.
Published: (2026)
Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology
by: Malik, Hashmat Shadab, et al.
Published: (2025)
by: Malik, Hashmat Shadab, et al.
Published: (2025)
Self-Supervised Learning for Building Robust Pediatric Chest X-ray Classification Models
by: Cheng, Sheng, et al.
Published: (2024)
by: Cheng, Sheng, et al.
Published: (2024)
ViT-5: Vision Transformers for The Mid-2020s
by: Wang, Feng, et al.
Published: (2026)
by: Wang, Feng, et al.
Published: (2026)
Simulated Cortical Magnification Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2025)
by: Yu, Zhengyang, et al.
Published: (2025)
Self-Supervised Representation Learning for Nerve Fiber Distribution Patterns in 3D-PLI
by: Oberstrass, Alexander, et al.
Published: (2024)
by: Oberstrass, Alexander, et al.
Published: (2024)
Similar Items
-
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
by: Chen, Xuweiyi, et al.
Published: (2024) -
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
by: Chen, Xuweiyi, et al.
Published: (2026) -
Empowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
by: Zhou, Wentao, et al.
Published: (2025) -
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
by: Wang, Boyang, et al.
Published: (2025) -
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
by: Chen, Xuweiyi, et al.
Published: (2025)