When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Yang, Xu, Qianqian, Wen, Peisong, Dai, Siran, Huang, Qingming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring Structural Degradation in Dense Representations for Self-supervised Learning
di: Dai, Siran, et al.
Pubblicazione: (2025)
di: Dai, Siran, et al.
Pubblicazione: (2025)
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
di: Dai, Siran, et al.
Pubblicazione: (2025)
di: Dai, Siran, et al.
Pubblicazione: (2025)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Semantic Concentration for Self-Supervised Dense Representations Learning
di: Wen, Peisong, et al.
Pubblicazione: (2025)
di: Wen, Peisong, et al.
Pubblicazione: (2025)
Not All Pairs are Equal: Hierarchical Learning for Average-Precision-Oriented Video Retrieval
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Bootstrapping Physics-Grounded Video Generation through VLM-Guided Iterative Self-Refinement
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
HiGFA: Hierarchical Guidance for Fine-grained Data Augmentation with Diffusion Models
di: Lu, Zhiguang, et al.
Pubblicazione: (2025)
di: Lu, Zhiguang, et al.
Pubblicazione: (2025)
Semantics Meets Temporal Correspondence: Self-supervised Object-centric Learning in Videos
di: Qian, Rui, et al.
Pubblicazione: (2023)
di: Qian, Rui, et al.
Pubblicazione: (2023)
Regularized Contrastive Partial Multi-view Outlier Detection
di: Wang, Yijia, et al.
Pubblicazione: (2024)
di: Wang, Yijia, et al.
Pubblicazione: (2024)
Learn Faster and Remember More: Balancing Exploration and Exploitation for Continual Test-time Adaptation
di: Yang, Pinci, et al.
Pubblicazione: (2025)
di: Yang, Pinci, et al.
Pubblicazione: (2025)
Top-K Pairwise Ranking: Bridging the Gap Among Ranking-Based Measures for Multi-Label Classification
di: Wang, Zitai, et al.
Pubblicazione: (2024)
di: Wang, Zitai, et al.
Pubblicazione: (2024)
Collaborative Temporal Consistency Learning for Point-supervised Natural Language Video Localization
di: Tao, Zhuo, et al.
Pubblicazione: (2025)
di: Tao, Zhuo, et al.
Pubblicazione: (2025)
AUCSeg: AUC-oriented Pixel-level Long-tail Semantic Segmentation
di: Han, Boyu, et al.
Pubblicazione: (2024)
di: Han, Boyu, et al.
Pubblicazione: (2024)
Decorrelating Structure via Adapters Makes Ensemble Learning Practical for Semi-supervised Learning
di: Wu, Jiaqi, et al.
Pubblicazione: (2024)
di: Wu, Jiaqi, et al.
Pubblicazione: (2024)
Boosting Point-supervised Temporal Action Localization via Text Refinement and Alignment
di: Ma, Yunchuan, et al.
Pubblicazione: (2026)
di: Ma, Yunchuan, et al.
Pubblicazione: (2026)
Collaboratively Self-supervised Video Representation Learning for Action Recognition
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
Leveraging Motion Information for Better Self-Supervised Video Correspondence Learning
di: Zhou, Zihan, et al.
Pubblicazione: (2025)
di: Zhou, Zihan, et al.
Pubblicazione: (2025)
Constrained Multiview Representation for Self-supervised Contrastive Learning
di: Dai, Siyuan, et al.
Pubblicazione: (2024)
di: Dai, Siyuan, et al.
Pubblicazione: (2024)
A Self-supervised Motion Representation for Portrait Video Generation
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
Self-supervised Shape Completion via Involution and Implicit Correspondences
di: Liu, Mengya, et al.
Pubblicazione: (2024)
di: Liu, Mengya, et al.
Pubblicazione: (2024)
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
di: Yang, Shuai, et al.
Pubblicazione: (2024)
di: Yang, Shuai, et al.
Pubblicazione: (2024)
Enhancing Sample Utilization in Noise-Robust Deep Metric Learning With Subgroup-Based Positive-Pair Selection
di: Yu, Zhipeng, et al.
Pubblicazione: (2025)
di: Yu, Zhipeng, et al.
Pubblicazione: (2025)
Rethinking Reward Signals in Video GRPO: When Scores Become Targets
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
Bidirectional Logits Tree: Pursuing Granularity Reconcilement in Fine-Grained Classification
di: Lu, Zhiguang, et al.
Pubblicazione: (2024)
di: Lu, Zhiguang, et al.
Pubblicazione: (2024)
Multi-granularity Correspondence Learning from Long-term Noisy Videos
di: Lin, Yijie, et al.
Pubblicazione: (2024)
di: Lin, Yijie, et al.
Pubblicazione: (2024)
Self-supervised Audiovisual Representation Learning for Remote Sensing Data
di: Heidler, Konrad, et al.
Pubblicazione: (2021)
di: Heidler, Konrad, et al.
Pubblicazione: (2021)
Zero-Shot Video Translation and Editing with Frame Spatial-Temporal Correspondence
di: Yang, Shuai, et al.
Pubblicazione: (2025)
di: Yang, Shuai, et al.
Pubblicazione: (2025)
Past- and Future-Informed KV Cache Policy with Salience Estimation in Autoregressive Video Diffusion
di: Chen, Hanmo, et al.
Pubblicazione: (2026)
di: Chen, Hanmo, et al.
Pubblicazione: (2026)
Uncertainty-aware Long-tailed Weights Model the Utility of Pseudo-labels for Semi-supervised Learning
di: Wu, Jiaqi, et al.
Pubblicazione: (2025)
di: Wu, Jiaqi, et al.
Pubblicazione: (2025)
Mixed Autoencoder for Self-supervised Visual Representation Learning
di: Chen, Kai, et al.
Pubblicazione: (2023)
di: Chen, Kai, et al.
Pubblicazione: (2023)
Foresee-to-Ground: From Predictive Temporal Perception to Evidence-Driven Reasoning for Video Temporal Grounding
di: Zheng, Zelin, et al.
Pubblicazione: (2026)
di: Zheng, Zelin, et al.
Pubblicazione: (2026)
Decoupling Common and Unique Representations for Multimodal Self-supervised Learning
di: Wang, Yi, et al.
Pubblicazione: (2023)
di: Wang, Yi, et al.
Pubblicazione: (2023)
Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
InstaVSR: Taming Diffusion for Efficient and Temporally Consistent Video Super-Resolution
di: Hu, Jintong, et al.
Pubblicazione: (2026)
di: Hu, Jintong, et al.
Pubblicazione: (2026)
CANeRV: Content Adaptive Neural Representation for Video Compression
di: Tang, Lv, et al.
Pubblicazione: (2025)
di: Tang, Lv, et al.
Pubblicazione: (2025)
From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
di: Lo, Ling, et al.
Pubblicazione: (2025)
di: Lo, Ling, et al.
Pubblicazione: (2025)
Skeleton2vec: A Self-supervised Learning Framework with Contextualized Target Representations for Skeleton Sequence
di: Xu, Ruizhuo, et al.
Pubblicazione: (2024)
di: Xu, Ruizhuo, et al.
Pubblicazione: (2024)
Emergent Temporal Correspondences from Video Diffusion Transformers
di: Nam, Jisu, et al.
Pubblicazione: (2025)
di: Nam, Jisu, et al.
Pubblicazione: (2025)
Segment Anything for Video: A Comprehensive Review of Video Object Segmentation and Tracking from Past to Future
di: Xu, Guoping, et al.
Pubblicazione: (2025)
di: Xu, Guoping, et al.
Pubblicazione: (2025)
Self-supervised Learning of Hybrid Part-aware 3D Representations of 2D Gaussians and Superquadrics
di: Gao, Zhirui, et al.
Pubblicazione: (2024)
di: Gao, Zhirui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Exploring Structural Degradation in Dense Representations for Self-supervised Learning
di: Dai, Siran, et al.
Pubblicazione: (2025) -
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
di: Dai, Siran, et al.
Pubblicazione: (2025) -
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
di: Liu, Yang, et al.
Pubblicazione: (2026) -
Semantic Concentration for Self-Supervised Dense Representations Learning
di: Wen, Peisong, et al.
Pubblicazione: (2025) -
Not All Pairs are Equal: Hierarchical Learning for Average-Precision-Oriented Video Retrieval
di: Liu, Yang, et al.
Pubblicazione: (2024)