From Pixels to Components: Eigenvector Masking for Visual Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Bizeul, Alice, Sutter, Thomas, Ryser, Alain, Schölkopf, Bernhard, von Kügelgen, Julius, Vogt, Julia E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging the Structure of Medical Data for Improved Representation Learning
by: Agostini, Andrea, et al.
Published: (2025)
by: Agostini, Andrea, et al.
Published: (2025)
Interpretable Diffusion Models with B-cos Networks
by: Bernold, Nicola, et al.
Published: (2025)
by: Bernold, Nicola, et al.
Published: (2025)
Structure is Supervision: Multiview Masked Autoencoders for Radiology
by: Laguna, Sonia, et al.
Published: (2025)
by: Laguna, Sonia, et al.
Published: (2025)
From Logits to Hierarchies: Hierarchical Clustering made Simple
by: Palumbo, Emanuele, et al.
Published: (2024)
by: Palumbo, Emanuele, et al.
Published: (2024)
Self-Supervised Disentanglement by Leveraging Structure in Data Augmentations
by: Eastwood, Cian, et al.
Published: (2023)
by: Eastwood, Cian, et al.
Published: (2023)
Interaction Asymmetry: A General Principle for Learning Composable Abstractions
by: Brady, Jack, et al.
Published: (2024)
by: Brady, Jack, et al.
Published: (2024)
From Pixels to Perception: Interpretable Predictions via Instance-wise Grouped Feature Selection
by: Vandenhirtz, Moritz, et al.
Published: (2025)
by: Vandenhirtz, Moritz, et al.
Published: (2025)
Diffusion-Based Representation Learning
by: Mittal, Sarthak, et al.
Published: (2021)
by: Mittal, Sarthak, et al.
Published: (2021)
PiLaMIM: Toward Richer Visual Representations by Integrating Pixel and Latent Masked Image Modeling
by: Lee, Junmyeong, et al.
Published: (2025)
by: Lee, Junmyeong, et al.
Published: (2025)
Enhancing Radiology Report Generation and Visual Grounding using Reinforcement Learning
by: Gundersen, Benjamin, et al.
Published: (2025)
by: Gundersen, Benjamin, et al.
Published: (2025)
Lightweight Pixel Difference Networks for Efficient Visual Representation Learning
by: Su, Zhuo, et al.
Published: (2024)
by: Su, Zhuo, et al.
Published: (2024)
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
by: Xiang, Wenzhao, et al.
Published: (2026)
by: Xiang, Wenzhao, et al.
Published: (2026)
MammoTracker: Mask-Guided Lesion Tracking in Temporal Mammograms
by: Liu, Xuan, et al.
Published: (2025)
by: Liu, Xuan, et al.
Published: (2025)
From Pixel to Mask: A Survey of Out-of-Distribution Segmentation
by: Zhao, Wenjie, et al.
Published: (2025)
by: Zhao, Wenjie, et al.
Published: (2025)
RadVLM: A Multitask Conversational Vision-Language Model for Radiology
by: Deperrois, Nicolas, et al.
Published: (2025)
by: Deperrois, Nicolas, et al.
Published: (2025)
Pixel Sentence Representation Learning
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
Beyond Independent Frames: Latent Attention Masked Autoencoders for Multi-View Echocardiography
by: Böhi, Simon, et al.
Published: (2026)
by: Böhi, Simon, et al.
Published: (2026)
Boxes2Pixels: Learning Defect Segmentation from Noisy SAM Masks
by: Lendering, Camile, et al.
Published: (2026)
by: Lendering, Camile, et al.
Published: (2026)
From Pixels to Views: Learning Angular-Aware and Physics-Consistent Representations for Light Field Microscopy
by: He, Feng, et al.
Published: (2025)
by: He, Feng, et al.
Published: (2025)
From Waveforms to Pixels: A Survey on Audio-Visual Segmentation
by: Li, Jia, et al.
Published: (2025)
by: Li, Jia, et al.
Published: (2025)
From Web to Pixels: Bringing Agentic Search into Visual Perception
by: Yang, Bokang, et al.
Published: (2026)
by: Yang, Bokang, et al.
Published: (2026)
Structure by Architecture: Structured Representations without Regularization
by: Leeb, Felix, et al.
Published: (2020)
by: Leeb, Felix, et al.
Published: (2020)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Two Is Better Than One: Aligned Representation Pairs for Anomaly Detection
by: Ryser, Alain, et al.
Published: (2024)
by: Ryser, Alain, et al.
Published: (2024)
Explaining Representation Learning with Perceptual Components
by: Yarici, Yavuz, et al.
Published: (2024)
by: Yarici, Yavuz, et al.
Published: (2024)
Identifiable Causal Representation Learning: Unsupervised, Multi-View, and Multi-Environment
by: von Kügelgen, Julius
Published: (2024)
by: von Kügelgen, Julius
Published: (2024)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
by: Wei, Yibing, et al.
Published: (2024)
by: Wei, Yibing, et al.
Published: (2024)
Masked Diffusion Captioning for Visual Feature Learning
by: Feng, Chao, et al.
Published: (2025)
by: Feng, Chao, et al.
Published: (2025)
Wavelet-Driven Masked Image Modeling: A Path to Efficient Visual Representation
by: Xiang, Wenzhao, et al.
Published: (2025)
by: Xiang, Wenzhao, et al.
Published: (2025)
Pixel Motion as Universal Representation for Robot Control
by: Ranasinghe, Kanchana, et al.
Published: (2025)
by: Ranasinghe, Kanchana, et al.
Published: (2025)
Pixel-Perfect Visual Geometry Estimation
by: Xu, Gangwei, et al.
Published: (2026)
by: Xu, Gangwei, et al.
Published: (2026)
Generation is Required for Data-Efficient Perception
by: Brady, Jack, et al.
Published: (2025)
by: Brady, Jack, et al.
Published: (2025)
Learning Pixel-wise Continuous Depth Representation via Clustering for Depth Completion
by: Shenglun, Chen, et al.
Published: (2024)
by: Shenglun, Chen, et al.
Published: (2024)
Verbalized Machine Learning: Revisiting Machine Learning with Language Models
by: Xiao, Tim Z., et al.
Published: (2024)
by: Xiao, Tim Z., et al.
Published: (2024)
SemanticMIM: Marring Masked Image Modeling with Semantics Compression for General Visual Representation
by: Yuan, Yike, et al.
Published: (2024)
by: Yuan, Yike, et al.
Published: (2024)
A Probabilistic Model Behind Self-Supervised Learning
by: Bizeul, Alice, et al.
Published: (2024)
by: Bizeul, Alice, et al.
Published: (2024)
BIMM: Brain Inspired Masked Modeling for Video Representation Learning
by: Wan, Zhifan, et al.
Published: (2024)
by: Wan, Zhifan, et al.
Published: (2024)
T-MAE: Temporal Masked Autoencoders for Point Cloud Representation Learning
by: Wei, Weijie, et al.
Published: (2023)
by: Wei, Weijie, et al.
Published: (2023)
Robust Representation Learning in Masked Autoencoders
by: Shrivastava, Anika, et al.
Published: (2026)
by: Shrivastava, Anika, et al.
Published: (2026)
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation
by: Wei, Wei, et al.
Published: (2025)
by: Wei, Wei, et al.
Published: (2025)
Similar Items
-
Leveraging the Structure of Medical Data for Improved Representation Learning
by: Agostini, Andrea, et al.
Published: (2025) -
Interpretable Diffusion Models with B-cos Networks
by: Bernold, Nicola, et al.
Published: (2025) -
Structure is Supervision: Multiview Masked Autoencoders for Radiology
by: Laguna, Sonia, et al.
Published: (2025) -
From Logits to Hierarchies: Hierarchical Clustering made Simple
by: Palumbo, Emanuele, et al.
Published: (2024) -
Self-Supervised Disentanglement by Leveraging Structure in Data Augmentations
by: Eastwood, Cian, et al.
Published: (2023)