Self-Supervised Spatial Correspondence Across Modalities
Fuente:
arXiv
Saved in:
| Main Authors: | Shrivastava, Ayush, Owens, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Supervised Any-Point Tracking by Contrastive Random Walks
by: Shrivastava, Ayush, et al.
Published: (2024)
by: Shrivastava, Ayush, et al.
Published: (2024)
Point Prompting: Counterfactual Tracking with Video Diffusion Models
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
Fine-grained Defocus Blur Control for Generative Image Models
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection
by: Ghadiya, Ayush, et al.
Published: (2024)
by: Ghadiya, Ayush, et al.
Published: (2024)
Leveraging Motion Information for Better Self-Supervised Video Correspondence Learning
by: Zhou, Zihan, et al.
Published: (2025)
by: Zhou, Zihan, et al.
Published: (2025)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Weakly-Supervised Learning of Dense Functional Correspondences
by: Stojanov, Stefan, et al.
Published: (2025)
by: Stojanov, Stefan, et al.
Published: (2025)
SHIC: Shape-Image Correspondences with no Keypoint Supervision
by: Shtedritski, Aleksandar, et al.
Published: (2024)
by: Shtedritski, Aleksandar, et al.
Published: (2024)
OmniSat: Self-Supervised Modality Fusion for Earth Observation
by: Astruc, Guillaume, et al.
Published: (2024)
by: Astruc, Guillaume, et al.
Published: (2024)
Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis
by: Chefer, Hila, et al.
Published: (2026)
by: Chefer, Hila, et al.
Published: (2026)
Towards Unbiased and Robust Spatio-Temporal Scene Graph Generation and Anticipation
by: Peddi, Rohith, et al.
Published: (2024)
by: Peddi, Rohith, et al.
Published: (2024)
Scaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT Transformers
by: Claessens, Cris, et al.
Published: (2025)
by: Claessens, Cris, et al.
Published: (2025)
Efficient Continuous Video Flow Model for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
S4: Self-Supervised Sensing Across the Spectrum
by: Shenoy, Jayanth, et al.
Published: (2024)
by: Shenoy, Jayanth, et al.
Published: (2024)
Spatial Steerability of GANs via Self-Supervision from Discriminator
by: Wang, Jianyuan, et al.
Published: (2023)
by: Wang, Jianyuan, et al.
Published: (2023)
Generalization of Self-Supervised Vision Transformers for Protein Localization Across Microscopy Domains
by: Isselmann, Ben, et al.
Published: (2026)
by: Isselmann, Ben, et al.
Published: (2026)
Multi-Modal Self-Supervised Semantic Communication
by: Zhao, Hang, et al.
Published: (2025)
by: Zhao, Hang, et al.
Published: (2025)
Jamais Vu: Exposing the Generalization Gap in Supervised Semantic Correspondence
by: Mariotti, Octave, et al.
Published: (2025)
by: Mariotti, Octave, et al.
Published: (2025)
VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning
by: Liao, Tangfei, et al.
Published: (2023)
by: Liao, Tangfei, et al.
Published: (2023)
Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality Signals
by: Fang, Shaoheng, et al.
Published: (2024)
by: Fang, Shaoheng, et al.
Published: (2024)
Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition
by: Halawa, Marah, et al.
Published: (2024)
by: Halawa, Marah, et al.
Published: (2024)
Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media
by: M, Megha Mariam K., et al.
Published: (2026)
by: M, Megha Mariam K., et al.
Published: (2026)
SCE-MAE: Selective Correspondence Enhancement with Masked Autoencoder for Self-Supervised Landmark Estimation
by: Yin, Kejia, et al.
Published: (2024)
by: Yin, Kejia, et al.
Published: (2024)
Self-Supervised Cross-Modal Text-Image Time Series Retrieval in Remote Sensing
by: Hoxha, Genc, et al.
Published: (2025)
by: Hoxha, Genc, et al.
Published: (2025)
UniMRSeg: Unified Modality-Relax Segmentation via Hierarchical Self-Supervised Compensation
by: Zhao, Xiaoqi, et al.
Published: (2025)
by: Zhao, Xiaoqi, et al.
Published: (2025)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
by: Li, Shenglan, et al.
Published: (2025)
by: Li, Shenglan, et al.
Published: (2025)
SIFT-VTON: Geometric Correspondence Supervision on Cross-Attention for Virtual Try-On
by: Takemoto, Kosuke, et al.
Published: (2026)
by: Takemoto, Kosuke, et al.
Published: (2026)
Multi-Modal Monocular Endoscopic Depth and Pose Estimation with Edge-Guided Self-Supervision
by: Ju, Xinwei, et al.
Published: (2026)
by: Ju, Xinwei, et al.
Published: (2026)
Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging
by: Pan, Tan, et al.
Published: (2026)
by: Pan, Tan, et al.
Published: (2026)
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
by: Yang, Shuai, et al.
Published: (2024)
by: Yang, Shuai, et al.
Published: (2024)
UniCorrn: Unified Correspondence Transformer Across 2D and 3D
by: Goswami, Prajnan, et al.
Published: (2026)
by: Goswami, Prajnan, et al.
Published: (2026)
SARL: Spatially-Aware Self-Supervised Representation Learning for Visuo-Tactile Perception
by: Khurana, Gurmeher, et al.
Published: (2025)
by: Khurana, Gurmeher, et al.
Published: (2025)
Spectral-Spatial Self-Supervised Learning for Few-Shot Hyperspectral Image Classification
by: Chen, Wenchen, et al.
Published: (2025)
by: Chen, Wenchen, et al.
Published: (2025)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
Self-Supervised Class-Agnostic Motion Prediction with Spatial and Temporal Consistency Regularizations
by: Wang, Kewei, et al.
Published: (2024)
by: Wang, Kewei, et al.
Published: (2024)
Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning
by: Shen, Yang, et al.
Published: (2026)
by: Shen, Yang, et al.
Published: (2026)
Efficient and High-Fidelity Omni Modality Retrieval
by: Huynh, Chuong, et al.
Published: (2026)
by: Huynh, Chuong, et al.
Published: (2026)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Utilization of Neighbor Information for Image Classification with Different Levels of Supervision
by: Jayatilaka, Gihan, et al.
Published: (2025)
by: Jayatilaka, Gihan, et al.
Published: (2025)
Robust Self-Supervised Cross-Modal Super-Resolution against Real-World Misaligned Observations
by: Dong, Xiaoyu, et al.
Published: (2026)
by: Dong, Xiaoyu, et al.
Published: (2026)
Similar Items
-
Self-Supervised Any-Point Tracking by Contrastive Random Walks
by: Shrivastava, Ayush, et al.
Published: (2024) -
Point Prompting: Counterfactual Tracking with Video Diffusion Models
by: Shrivastava, Ayush, et al.
Published: (2025) -
Fine-grained Defocus Blur Control for Generative Image Models
by: Shrivastava, Ayush, et al.
Published: (2025) -
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection
by: Ghadiya, Ayush, et al.
Published: (2024) -
Leveraging Motion Information for Better Self-Supervised Video Correspondence Learning
by: Zhou, Zihan, et al.
Published: (2025)