Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Yang, Cai, Yusen, Hryniewska-Guzik, Weronika, Lin, Qing, Zhang, Mengmi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to See Through a Baby's Eyes: Early Visual Diets Enable Robust Visual Intelligence in Humans and Machines
by: Cai, Yusen, et al.
Published: (2025)
by: Cai, Yusen, et al.
Published: (2025)
X-ray transferable polyrepresentation learning
by: Hryniewska-Guzik, Weronika, et al.
Published: (2025)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2025)
CNN-based explanation ensembling for dataset, representation and explanations evaluation
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
NormEnsembleXAI: Unveiling the Strengths and Weaknesses of XAI Ensemble Techniques
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
From Pretext to Purpose: Batch-Adaptive Self-Supervised Learning
by: Zhang, Jiansong, et al.
Published: (2023)
by: Zhang, Jiansong, et al.
Published: (2023)
Multi-task learning for classification, segmentation, reconstruction, and detection on chest CT scans
by: Hryniewska-Guzik, Weronika, et al.
Published: (2023)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2023)
Progressive Pretext Task Learning for Human Trajectory Prediction
by: Lin, Xiaotong, et al.
Published: (2024)
by: Lin, Xiaotong, et al.
Published: (2024)
A comparative analysis of deep learning models for lung segmentation on X-ray images
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
Pose Prior Learner: Unsupervised Categorical Prior Learning for Pose Estimation
by: Wang, Ziyu, et al.
Published: (2024)
by: Wang, Ziyu, et al.
Published: (2024)
Learning to See the Elephant in the Room: Self-Supervised Context Reasoning in Humans and AI
by: Liu, Xiao, et al.
Published: (2022)
by: Liu, Xiao, et al.
Published: (2022)
Pretext Task Adversarial Learning for Unpaired Low-field to Ultra High-field MRI Synthesis
by: Zhang, Zhenxuan, et al.
Published: (2025)
by: Zhang, Zhenxuan, et al.
Published: (2025)
Make Me Happier: Evoking Emotions Through Image Diffusion Models
by: Lin, Qing, et al.
Published: (2024)
by: Lin, Qing, et al.
Published: (2024)
MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning
by: Nedungadi, Vishal, et al.
Published: (2024)
by: Nedungadi, Vishal, et al.
Published: (2024)
LookWhere? Efficient Visual Recognition by Learning Where to Look and What to See from Self-Supervision
by: Fuller, Anthony, et al.
Published: (2025)
by: Fuller, Anthony, et al.
Published: (2025)
Time2Agri: Temporal Pretext Tasks for Agricultural Monitoring
by: Gupta, Moti Rattan, et al.
Published: (2025)
by: Gupta, Moti Rattan, et al.
Published: (2025)
SSL-Interactions: Pretext Tasks for Interactive Trajectory Prediction
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
by: Tan, Hao, et al.
Published: (2026)
by: Tan, Hao, et al.
Published: (2026)
Benchmarking Robust Self-Supervised Learning Across Diverse Downstream Tasks
by: Kowalczuk, Antoni, et al.
Published: (2024)
by: Kowalczuk, Antoni, et al.
Published: (2024)
PP-SSL : Priority-Perception Self-Supervised Learning for Fine-Grained Recognition
by: Li, ShuaiHeng, et al.
Published: (2024)
by: Li, ShuaiHeng, et al.
Published: (2024)
Information Flow in Self-Supervised Learning
by: Tan, Zhiquan, et al.
Published: (2023)
by: Tan, Zhiquan, et al.
Published: (2023)
AltChart: Enhancing VLM-based Chart Summarization Through Multi-Pretext Tasks
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning
by: Liu, Yuhong, et al.
Published: (2025)
by: Liu, Yuhong, et al.
Published: (2025)
Exploring Transferability of Self-Supervised Learning by Task Conflict Calibration
by: Guo, Huijie, et al.
Published: (2025)
by: Guo, Huijie, et al.
Published: (2025)
SARL: Spatially-Aware Self-Supervised Representation Learning for Visuo-Tactile Perception
by: Khurana, Gurmeher, et al.
Published: (2025)
by: Khurana, Gurmeher, et al.
Published: (2025)
Adversarial Robustness of Discriminative Self-Supervised Learning in Vision
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
Spectral-Spatial Self-Supervised Learning for Few-Shot Hyperspectral Image Classification
by: Chen, Wenchen, et al.
Published: (2025)
by: Chen, Wenchen, et al.
Published: (2025)
Spatial-Aware Self-Supervision for Medical 3D Imaging with Multi-Granularity Observable Tasks
by: Zhang, Yiqin, et al.
Published: (2025)
by: Zhang, Yiqin, et al.
Published: (2025)
Unveiling the Tapestry: the Interplay of Generalization and Forgetting in Continual Learning
by: Shi, Zenglin, et al.
Published: (2022)
by: Shi, Zenglin, et al.
Published: (2022)
Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond
by: Zhang, Jiahang, et al.
Published: (2024)
by: Zhang, Jiahang, et al.
Published: (2024)
VANP: Learning Where to See for Navigation with Self-Supervised Vision-Action Pre-Training
by: Nazeri, Mohammad, et al.
Published: (2024)
by: Nazeri, Mohammad, et al.
Published: (2024)
Image Re-Identification: Where Self-supervision Meets Vision-Language Learning
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Self-Supervised Learning with a Multi-Task Latent Space Objective
by: De Plaen, Pierre-François, et al.
Published: (2026)
by: De Plaen, Pierre-François, et al.
Published: (2026)
Spatial Steerability of GANs via Self-Supervision from Discriminator
by: Wang, Jianyuan, et al.
Published: (2023)
by: Wang, Jianyuan, et al.
Published: (2023)
PRISM: Progressive Reasoning through Iterative Slot Memory for Vision
by: Wang, Ziyu, et al.
Published: (2026)
by: Wang, Ziyu, et al.
Published: (2026)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
by: Khandelwal, Naitik, et al.
Published: (2023)
by: Khandelwal, Naitik, et al.
Published: (2023)
Flow Snapshot Neurons in Action: Deep Neural Networks Generalize to Biological Motion Perception
by: Han, Shuangpeng, et al.
Published: (2024)
by: Han, Shuangpeng, et al.
Published: (2024)
MSSSeg: Learning Multi-Scale Structural Complexity for Self-Supervised Segmentation
by: Li, Haotang, et al.
Published: (2025)
by: Li, Haotang, et al.
Published: (2025)
Pretext Matters: An Empirical Study of SSL Methods in Medical Imaging
by: Ivezić, Vedrana, et al.
Published: (2026)
by: Ivezić, Vedrana, et al.
Published: (2026)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
CLIP-Guided Adaptable Self-Supervised Learning for Human-Centric Visual Tasks
by: Luo, Mingshuang, et al.
Published: (2026)
by: Luo, Mingshuang, et al.
Published: (2026)
Similar Items
-
Learning to See Through a Baby's Eyes: Early Visual Diets Enable Robust Visual Intelligence in Humans and Machines
by: Cai, Yusen, et al.
Published: (2025) -
X-ray transferable polyrepresentation learning
by: Hryniewska-Guzik, Weronika, et al.
Published: (2025) -
CNN-based explanation ensembling for dataset, representation and explanations evaluation
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024) -
NormEnsembleXAI: Unveiling the Strengths and Weaknesses of XAI Ensemble Techniques
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024) -
From Pretext to Purpose: Batch-Adaptive Self-Supervised Learning
by: Zhang, Jiansong, et al.
Published: (2023)