Collaboratively Self-supervised Video Representation Learning for Action Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Jie, Wan, Zhifan, Hu, Lanqing, Lin, Stephen, Wu, Shuzhe, Shan, Shiguang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hierarchical Compositional Representations for Few-shot Action Recognition
di: Li, Changzhen, et al.
Pubblicazione: (2022)
di: Li, Changzhen, et al.
Pubblicazione: (2022)
BIMM: Brain Inspired Masked Modeling for Video Representation Learning
di: Wan, Zhifan, et al.
Pubblicazione: (2024)
di: Wan, Zhifan, et al.
Pubblicazione: (2024)
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
di: Li, Yong, et al.
Pubblicazione: (2024)
di: Li, Yong, et al.
Pubblicazione: (2024)
Mixed Autoencoder for Self-supervised Visual Representation Learning
di: Chen, Kai, et al.
Pubblicazione: (2023)
di: Chen, Kai, et al.
Pubblicazione: (2023)
Confidence Aware Learning for Reliable Face Anti-spoofing
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
Anonymization Prompt Learning for Facial Privacy-Preserving Text-to-Image Generation
di: Shi, Liang, et al.
Pubblicazione: (2024)
di: Shi, Liang, et al.
Pubblicazione: (2024)
Image to Pseudo-Episode: Boosting Few-Shot Segmentation by Unlabeled Data
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
Steering Vision-Language Pre-trained Models for Incremental Face Presentation Attack Detection
di: Li, Haoze, et al.
Pubblicazione: (2025)
di: Li, Haoze, et al.
Pubblicazione: (2025)
A Self-supervised Motion Representation for Portrait Video Generation
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
Revisiting Face Forgery Detection: From Facial Representation to Forgery Detection
di: Guo, Zonghui, et al.
Pubblicazione: (2024)
di: Guo, Zonghui, et al.
Pubblicazione: (2024)
Generalized Face Liveness Detection via De-fake Face Generator
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos
di: Chen, Yin, et al.
Pubblicazione: (2023)
di: Chen, Yin, et al.
Pubblicazione: (2023)
Idempotent Unsupervised Representation Learning for Skeleton-Based Action Recognition
di: Lin, Lilang, et al.
Pubblicazione: (2024)
di: Lin, Lilang, et al.
Pubblicazione: (2024)
Self-supervised Representation Learning for Cell Event Recognition through Time Arrow Prediction
di: Chen, Cangxiong, et al.
Pubblicazione: (2024)
di: Chen, Cangxiong, et al.
Pubblicazione: (2024)
SBF: An Effective Representation to Augment Skeleton for Video-based Human Action Recognition
di: Peng, Zhuoxuan, et al.
Pubblicazione: (2026)
di: Peng, Zhuoxuan, et al.
Pubblicazione: (2026)
T2VAttack: Adversarial Attack on Text-to-Video Diffusion Models
di: Li, Changzhen, et al.
Pubblicazione: (2025)
di: Li, Changzhen, et al.
Pubblicazione: (2025)
STARS: Self-supervised Tuning for 3D Action Recognition in Skeleton Sequences
di: Mehraban, Soroush, et al.
Pubblicazione: (2024)
di: Mehraban, Soroush, et al.
Pubblicazione: (2024)
Semi-supervised Active Learning for Video Action Detection
di: Singh, Ayush, et al.
Pubblicazione: (2023)
di: Singh, Ayush, et al.
Pubblicazione: (2023)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Self-supervised Audiovisual Representation Learning for Remote Sensing Data
di: Heidler, Konrad, et al.
Pubblicazione: (2021)
di: Heidler, Konrad, et al.
Pubblicazione: (2021)
SUGAR: Learning Skeleton Representation with Visual-Motion Knowledge for Action Recognition
di: Ye, Qilang, et al.
Pubblicazione: (2025)
di: Ye, Qilang, et al.
Pubblicazione: (2025)
Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic Mapping across Heterogeneous Datasets
di: Li, Jia, et al.
Pubblicazione: (2026)
di: Li, Jia, et al.
Pubblicazione: (2026)
EfficientMT: Efficient Temporal Adaptation for Motion Transfer in Text-to-Video Diffusion Models
di: Cai, Yufei, et al.
Pubblicazione: (2025)
di: Cai, Yufei, et al.
Pubblicazione: (2025)
GLip: A Global-Local Integrated Progressive Framework for Robust Visual Speech Recognition
di: Wang, Tianyue, et al.
Pubblicazione: (2025)
di: Wang, Tianyue, et al.
Pubblicazione: (2025)
Semantic or Covariate? A Study on the Intractable Case of Out-of-Distribution Detection
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2024)
di: Wang, Zhongqi, et al.
Pubblicazione: (2024)
FullLoRA: Efficiently Boosting the Robustness of Pretrained Vision Transformers
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
Pre-trained Model Guided Fine-Tuning for Zero-Shot Adversarial Robustness
di: Wang, Sibo, et al.
Pubblicazione: (2024)
di: Wang, Sibo, et al.
Pubblicazione: (2024)
Rethinking the Evaluation of Out-of-Distribution Detection: A Sorites Paradox
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
di: Long, Xingming, et al.
Pubblicazione: (2025)
di: Long, Xingming, et al.
Pubblicazione: (2025)
Assimilation Matters: Model-level Backdoor Detection in Vision-Language Pretrained Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
Feature Hallucination for Self-supervised Action Recognition
di: Wang, Lei, et al.
Pubblicazione: (2025)
di: Wang, Lei, et al.
Pubblicazione: (2025)
Prompt-guided Disentangled Representation for Action Recognition
di: Wu, Tianci, et al.
Pubblicazione: (2025)
di: Wu, Tianci, et al.
Pubblicazione: (2025)
INFACT: A Diagnostic Benchmark for Induced Faithfulness and Factuality Hallucinations in Video-LLMs
di: Yang, Junqi, et al.
Pubblicazione: (2026)
di: Yang, Junqi, et al.
Pubblicazione: (2026)
When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
DO3D: Self-supervised Learning of Decomposed Object-aware 3D Motion and Depth from Monocular Videos
di: Wu, Xiuzhe, et al.
Pubblicazione: (2024)
di: Wu, Xiuzhe, et al.
Pubblicazione: (2024)
TRivia: Self-supervised Fine-tuning of Vision-Language Models for Table Recognition
di: Zhang, Junyuan, et al.
Pubblicazione: (2025)
di: Zhang, Junyuan, et al.
Pubblicazione: (2025)
Balanced Representation Learning for Long-tailed Skeleton-based Action Recognition
di: Liu, Hongda, et al.
Pubblicazione: (2023)
di: Liu, Hongda, et al.
Pubblicazione: (2023)
Constrained Multiview Representation for Self-supervised Contrastive Learning
di: Dai, Siyuan, et al.
Pubblicazione: (2024)
di: Dai, Siyuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hierarchical Compositional Representations for Few-shot Action Recognition
di: Li, Changzhen, et al.
Pubblicazione: (2022) -
BIMM: Brain Inspired Masked Modeling for Video Representation Learning
di: Wan, Zhifan, et al.
Pubblicazione: (2024) -
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
di: Li, Yong, et al.
Pubblicazione: (2024) -
Mixed Autoencoder for Self-supervised Visual Representation Learning
di: Chen, Kai, et al.
Pubblicazione: (2023) -
Confidence Aware Learning for Reliable Face Anti-spoofing
di: Long, Xingming, et al.
Pubblicazione: (2024)