Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Truong, Quang-Trung, Nguyen, Duc Thanh, Hua, Binh-Son, Yeung, Sai-Kit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations
by: Truong, Quang Trung, et al.
Published: (2025)
by: Truong, Quang Trung, et al.
Published: (2025)
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
by: Shum, Ka Chun, et al.
Published: (2023)
by: Shum, Ka Chun, et al.
Published: (2023)
Color Alignment in Diffusion
by: Shum, Ka Chun, et al.
Published: (2025)
by: Shum, Ka Chun, et al.
Published: (2025)
MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning
by: Truong, Quang-Trung, et al.
Published: (2025)
by: Truong, Quang-Trung, et al.
Published: (2025)
Vision-Aware Text Features in Referring Image Segmentation: From Object Understanding to Context Understanding
by: Nguyen-Truong, Hai, et al.
Published: (2024)
by: Nguyen-Truong, Hai, et al.
Published: (2024)
Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues
by: Vu, Tuan-Anh, et al.
Published: (2025)
by: Vu, Tuan-Anh, et al.
Published: (2025)
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
by: Vu, Tuan-Anh, et al.
Published: (2023)
by: Vu, Tuan-Anh, et al.
Published: (2023)
360VOTS: Visual Object Tracking and Segmentation in Omnidirectional Videos
by: Xu, Yinzhe, et al.
Published: (2024)
by: Xu, Yinzhe, et al.
Published: (2024)
SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation
by: Pham, Duc-Hai, et al.
Published: (2024)
by: Pham, Duc-Hai, et al.
Published: (2024)
MonoVQD: Monocular 3D Object Detection with Variational Query Denoising and Self-Distillation
by: Vu, Kiet Dang, et al.
Published: (2025)
by: Vu, Kiet Dang, et al.
Published: (2025)
Advances in 3D Neural Stylization: A Survey
by: Chen, Yingshu, et al.
Published: (2023)
by: Chen, Yingshu, et al.
Published: (2023)
Anatomical Attention Alignment representation for Radiology Report Generation
by: Nguyen, Quang Vinh, et al.
Published: (2025)
by: Nguyen, Quang Vinh, et al.
Published: (2025)
PADM: A Physics-aware Diffusion Model for Attenuation Correction
by: Pham, Trung Kien, et al.
Published: (2025)
by: Pham, Trung Kien, et al.
Published: (2025)
Dream-in-Style: Text-to-3D Generation Using Stylized Score Distillation
by: Kompanowski, Hubert, et al.
Published: (2024)
by: Kompanowski, Hubert, et al.
Published: (2024)
Multi-Perspective Data Augmentation for Few-shot Object Detection
by: Vu, Anh-Khoa Nguyen, et al.
Published: (2025)
by: Vu, Anh-Khoa Nguyen, et al.
Published: (2025)
Frequency Attention for Knowledge Distillation
by: Pham, Cuong, et al.
Published: (2024)
by: Pham, Cuong, et al.
Published: (2024)
Text-to-3D Generation using Jensen-Shannon Score Distillation
by: Do, Khoi, et al.
Published: (2025)
by: Do, Khoi, et al.
Published: (2025)
Betrayed by Attention: A Simple yet Effective Approach for Self-supervised Video Object Segmentation
by: Ding, Shuangrui, et al.
Published: (2023)
by: Ding, Shuangrui, et al.
Published: (2023)
FROMAT: Multiview Material Appearance Transfer via Few-Shot Self-Attention Adaptation
by: Kompanowski, Hubert, et al.
Published: (2025)
by: Kompanowski, Hubert, et al.
Published: (2025)
Efficient and Concise Explanations for Object Detection with Gaussian-Class Activation Mapping Explainer
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation using Reference Image Prompts
by: Tran, Uy Dieu, et al.
Published: (2024)
by: Tran, Uy Dieu, et al.
Published: (2024)
ODExAI: A Comprehensive Object Detection Explainable AI Evaluation
by: Nguyen, Loc Phuc Truong, et al.
Published: (2025)
by: Nguyen, Loc Phuc Truong, et al.
Published: (2025)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
by: Nguyen, Khanh-Binh, et al.
Published: (2025)
by: Nguyen, Khanh-Binh, et al.
Published: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
by: Hoang, Trong-Vu, et al.
Published: (2025)
by: Hoang, Trong-Vu, et al.
Published: (2025)
Retro: Reusing teacher projection head for efficient embedding distillation on Lightweight Models via Self-supervised Learning
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
Point-Unet: A Context-aware Point-based Neural Network for Volumetric Segmentation
by: Ho, Ngoc-Vuong, et al.
Published: (2022)
by: Ho, Ngoc-Vuong, et al.
Published: (2022)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
Mono3DV: Monocular 3D Object Detection with 3D-Aware Bipartite Matching and Variational Query DeNoising
by: Vu, Kiet Dang, et al.
Published: (2026)
by: Vu, Kiet Dang, et al.
Published: (2026)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
by: Nguyen, Huu Tien, et al.
Published: (2025)
by: Nguyen, Huu Tien, et al.
Published: (2025)
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
by: Nguyen, Quang Vinh, et al.
Published: (2024)
by: Nguyen, Quang Vinh, et al.
Published: (2024)
Domain-invariant Mixed-domain Semi-supervised Medical Image Segmentation with Clustered Maximum Mean Discrepancy Alignment
by: Lam, Ba-Thinh, et al.
Published: (2026)
by: Lam, Ba-Thinh, et al.
Published: (2026)
Semantics Meets Temporal Correspondence: Self-supervised Object-centric Learning in Videos
by: Qian, Rui, et al.
Published: (2023)
by: Qian, Rui, et al.
Published: (2023)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
by: Luu, Vinh Quoc, et al.
Published: (2024)
by: Luu, Vinh Quoc, et al.
Published: (2024)
IQBench: How "Smart'' Are Vision-Language Models? A Study with Human IQ Tests
by: Pham, Tan-Hanh, et al.
Published: (2025)
by: Pham, Tan-Hanh, et al.
Published: (2025)
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Schrodinger AI: A Unified Spectral-Dynamical Framework for Classification, Reasoning, and Operator-Based Generalization
by: Nguyen, Truong Son
Published: (2025)
by: Nguyen, Truong Son
Published: (2025)
View-aware Cross-modal Distillation for Multi-view Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
DeVOS: Flow-Guided Deformable Transformer for Video Object Segmentation
by: Fedynyak, Volodymyr, et al.
Published: (2024)
by: Fedynyak, Volodymyr, et al.
Published: (2024)
MambaU-Lite: A Lightweight Model based on Mamba and Integrated Channel-Spatial Attention for Skin Lesion Segmentation
by: Nguyen, Thi-Nhu-Quynh, et al.
Published: (2024)
by: Nguyen, Thi-Nhu-Quynh, et al.
Published: (2024)
Similar Items
-
AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations
by: Truong, Quang Trung, et al.
Published: (2025) -
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
by: Shum, Ka Chun, et al.
Published: (2023) -
Color Alignment in Diffusion
by: Shum, Ka Chun, et al.
Published: (2025) -
MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning
by: Truong, Quang-Trung, et al.
Published: (2025) -
Vision-Aware Text Features in Referring Image Segmentation: From Object Understanding to Context Understanding
by: Nguyen-Truong, Hai, et al.
Published: (2024)