How Effective are Self-Supervised Models for Contact Identification in Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gunawardhana, Malitha, Sadith, Limalka, David, Liel, Harari, Daniel, Khan, Muhammad Haris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
von: Ahamed, Shihab Aaqil, et al.
Veröffentlicht: (2025)
von: Ahamed, Shihab Aaqil, et al.
Veröffentlicht: (2025)
Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos
von: Nagasinghe, Kumaranage Ravindu Yasas, et al.
Veröffentlicht: (2024)
von: Nagasinghe, Kumaranage Ravindu Yasas, et al.
Veröffentlicht: (2024)
Towards Generalizing to Unseen Domains with Few Labels
von: Galappaththige, Chamuditha Jayanga, et al.
Veröffentlicht: (2024)
von: Galappaththige, Chamuditha Jayanga, et al.
Veröffentlicht: (2024)
Action Without Interaction: Probing the Physical Foundations of Video LMMs via Contact-Release Detection
von: Harari, Daniel, et al.
Veröffentlicht: (2025)
von: Harari, Daniel, et al.
Veröffentlicht: (2025)
How good nnU-Net for Segmenting Cardiac MRI: A Comprehensive Evaluation
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2024)
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2024)
Integrating Deep Learning in Cardiology: A Comprehensive Review of Atrial Fibrillation, Left Atrial Scar Segmentation, and the Frontiers of State-of-the-Art Techniques
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2024)
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2024)
Segmenting Bi-Atrial Structures Using ResNext Based Framework
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2025)
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2025)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
von: Khan, Adnan, et al.
Veröffentlicht: (2024)
von: Khan, Adnan, et al.
Veröffentlicht: (2024)
Robust and Label-Efficient Deep Waste Detection
von: Abid, Hassan, et al.
Veröffentlicht: (2025)
von: Abid, Hassan, et al.
Veröffentlicht: (2025)
Self-Supervised Animal Identification for Long Videos
von: Fang, Xuyang, et al.
Veröffentlicht: (2026)
von: Fang, Xuyang, et al.
Veröffentlicht: (2026)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2026)
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2026)
CountZES: Counting via Zero-Shot Exemplar Selection
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2025)
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2025)
Dynamic Position Transformation and Boundary Refinement Network for Left Atrial Segmentation
von: Xu, Fangqiang, et al.
Veröffentlicht: (2024)
von: Xu, Fangqiang, et al.
Veröffentlicht: (2024)
Leveraging Cycle-Consistent Anchor Points for Self-Supervised RGB-D Registration
von: Tourani, Siddharth, et al.
Veröffentlicht: (2025)
von: Tourani, Siddharth, et al.
Veröffentlicht: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2024)
von: Ali, Eman, et al.
Veröffentlicht: (2024)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
von: Tourani, Siddharth, et al.
Veröffentlicht: (2024)
von: Tourani, Siddharth, et al.
Veröffentlicht: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2023)
von: Ali, Eman, et al.
Veröffentlicht: (2023)
Domain-Guided Weight Modulation for Semi-Supervised Domain Generalization
von: Galappaththige, Chamuditha Jayanaga, et al.
Veröffentlicht: (2024)
von: Galappaththige, Chamuditha Jayanaga, et al.
Veröffentlicht: (2024)
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
von: Ali, Eman, et al.
Veröffentlicht: (2025)
von: Ali, Eman, et al.
Veröffentlicht: (2025)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting
von: Paliwal, Avinash, et al.
Veröffentlicht: (2026)
von: Paliwal, Avinash, et al.
Veröffentlicht: (2026)
How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite for Video-LMMs
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
von: Gunduboina, Hariseetharam, et al.
Veröffentlicht: (2025)
von: Gunduboina, Hariseetharam, et al.
Veröffentlicht: (2025)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
von: Ding, Bonan, et al.
Veröffentlicht: (2026)
von: Ding, Bonan, et al.
Veröffentlicht: (2026)
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2024)
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2024)
Gaussian See, Gaussian Do: Semantic 3D Motion Transfer from Multiview Video
von: Bekor, Yarin, et al.
Veröffentlicht: (2025)
von: Bekor, Yarin, et al.
Veröffentlicht: (2025)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
von: Chokuwa, Sharon, et al.
Veröffentlicht: (2024)
von: Chokuwa, Sharon, et al.
Veröffentlicht: (2024)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2024)
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2024)
Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2025)
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2025)
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
von: Maaz, Muhammad, et al.
Veröffentlicht: (2023)
von: Maaz, Muhammad, et al.
Veröffentlicht: (2023)
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
von: Khan, Ufaq, et al.
Veröffentlicht: (2025)
von: Khan, Ufaq, et al.
Veröffentlicht: (2025)
VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
von: Maaz, Muhammad, et al.
Veröffentlicht: (2024)
von: Maaz, Muhammad, et al.
Veröffentlicht: (2024)
Judging from Support-set: A New Way to Utilize Few-Shot Segmentation for Segmentation Refinement Process
von: Moon, Seonghyeon, et al.
Veröffentlicht: (2024)
von: Moon, Seonghyeon, et al.
Veröffentlicht: (2024)
VideoSAVi: Self-Aligned Video Language Models without Human Supervision
von: Kulkarni, Yogesh, et al.
Veröffentlicht: (2024)
von: Kulkarni, Yogesh, et al.
Veröffentlicht: (2024)
OSLoPrompt: Bridging Low-Supervision Challenges and Open-Set Domain Generalization in CLIP
von: C, Mohamad Hassan N, et al.
Veröffentlicht: (2025)
von: C, Mohamad Hassan N, et al.
Veröffentlicht: (2025)
VideoSSR: Video Self-Supervised Reinforcement Learning
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
SelfHVD: Self-Supervised Handheld Video Deblurring
von: Xu, Honglei, et al.
Veröffentlicht: (2025)
von: Xu, Honglei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
von: Ahamed, Shihab Aaqil, et al.
Veröffentlicht: (2025) -
Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos
von: Nagasinghe, Kumaranage Ravindu Yasas, et al.
Veröffentlicht: (2024) -
Towards Generalizing to Unseen Domains with Few Labels
von: Galappaththige, Chamuditha Jayanga, et al.
Veröffentlicht: (2024) -
Action Without Interaction: Probing the Physical Foundations of Video LMMs via Contact-Release Detection
von: Harari, Daniel, et al.
Veröffentlicht: (2025) -
How good nnU-Net for Segmenting Cardiac MRI: A Comprehensive Evaluation
von: Gunawardhana, Malitha, et al.
Veröffentlicht: (2024)