DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
Fuente:
arXiv
Saved in:
| Main Authors: | Tumanyan, Narek, Singer, Assaf, Bagon, Shai, Dekel, Tali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
by: Kaduri, Omri, et al.
Published: (2024)
by: Kaduri, Omri, et al.
Published: (2024)
DIVE: Taming DINO for Subject-Driven Video Editing
by: Huang, Yi, et al.
Published: (2024)
by: Huang, Yi, et al.
Published: (2024)
DINO-Foresight: Looking into the Future with DINO
by: Karypidis, Efstathios, et al.
Published: (2024)
by: Karypidis, Efstathios, et al.
Published: (2024)
DINO-Tok: Adapting DINO for Visual Tokenizers
by: Jia, Mingkai, et al.
Published: (2025)
by: Jia, Mingkai, et al.
Published: (2025)
DRoPS: Dynamic 3D Reconstruction of Pre-Scanned Objects
by: Tumanyan, Narek, et al.
Published: (2026)
by: Tumanyan, Narek, et al.
Published: (2026)
Evaluating Stenosis Detection with Grounding DINO, YOLO, and DINO-DETR
by: Ansari, Muhammad Musab
Published: (2025)
by: Ansari, Muhammad Musab
Published: (2025)
Hearing the Room Through the Shape of the Drum: Modal-Guided Sound Recovery from Multi-Point Surface Vibrations
by: Bagon, Shai, et al.
Published: (2026)
by: Bagon, Shai, et al.
Published: (2026)
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
by: Guo, Hao, et al.
Published: (2024)
by: Guo, Hao, et al.
Published: (2024)
PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training
by: Fu, Weifu, et al.
Published: (2026)
by: Fu, Weifu, et al.
Published: (2026)
DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
by: Gong, Ziren, et al.
Published: (2025)
by: Gong, Ziren, et al.
Published: (2025)
SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3
by: Yang, Sicheng, et al.
Published: (2025)
by: Yang, Sicheng, et al.
Published: (2025)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
by: Liu, Shilong, et al.
Published: (2023)
by: Liu, Shilong, et al.
Published: (2023)
DINO-MX: A Modular & Flexible Framework for Self-Supervised Learning
by: Gokmen, Mahmut Selman, et al.
Published: (2025)
by: Gokmen, Mahmut Selman, et al.
Published: (2025)
Back to the Features: DINO as a Foundation for Video World Models
by: Baldassarre, Federico, et al.
Published: (2025)
by: Baldassarre, Federico, et al.
Published: (2025)
On Partial Prototype Collapse in the DINO Family of Self-Supervised Methods
by: Govindarajan, Hariprasath, et al.
Published: (2024)
by: Govindarajan, Hariprasath, et al.
Published: (2024)
MammoDINO: Anatomically Aware Self-Supervision for Mammographic Images
by: Zhou, Sicheng, et al.
Published: (2025)
by: Zhou, Sicheng, et al.
Published: (2025)
Color-Encoded Illumination for High-Speed Volumetric Scene Reconstruction
by: Novikov, David, et al.
Published: (2026)
by: Novikov, David, et al.
Published: (2026)
AdvDINO: Domain-Adversarial Self-Supervised Representation Learning for Spatial Proteomics
by: Su, Stella, et al.
Published: (2025)
by: Su, Stella, et al.
Published: (2025)
SatDINO: A Deep Dive into Self-Supervised Pretraining for Remote Sensing
by: Straka, Jakub, et al.
Published: (2025)
by: Straka, Jakub, et al.
Published: (2025)
PixelDINO: Semi-Supervised Semantic Segmentation for Detecting Permafrost Disturbances
by: Heidler, Konrad, et al.
Published: (2024)
by: Heidler, Konrad, et al.
Published: (2024)
Fine-Grained DINO Tuning with Dual Supervision for Face Forgery Detection
by: Zhang, Tianxiang, et al.
Published: (2025)
by: Zhang, Tianxiang, et al.
Published: (2025)
Unlocking Generalization in Polyp Segmentation with DINO Self-Attention "keys"
by: Monteiro, Carla, et al.
Published: (2025)
by: Monteiro, Carla, et al.
Published: (2025)
Deploy DINO with Many-to-Many Association
by: Jiang, Haodong, et al.
Published: (2026)
by: Jiang, Haodong, et al.
Published: (2026)
Video-GroundingDINO: Towards Open-Vocabulary Spatio-Temporal Video Grounding
by: Wasim, Syed Talal, et al.
Published: (2023)
by: Wasim, Syed Talal, et al.
Published: (2023)
DINO-AD: Unsupervised Anomaly Detection with Frozen DINO-V3 Features
by: Huo, Jiayu, et al.
Published: (2026)
by: Huo, Jiayu, et al.
Published: (2026)
ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
Control-DINO: Feature Space Conditioning for Controllable Image-to-Video Diffusion
by: Dominici, Edoardo A., et al.
Published: (2026)
by: Dominici, Edoardo A., et al.
Published: (2026)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
by: Zeng, Haoxi, et al.
Published: (2026)
by: Zeng, Haoxi, et al.
Published: (2026)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
D$^3$FlowSLAM: Self-Supervised Dynamic SLAM with Flow Motion Decomposition and DINO Guidance
by: Yu, Xingyuan, et al.
Published: (2022)
by: Yu, Xingyuan, et al.
Published: (2022)
DINO-CoDT: Multi-class Collaborative Detection and Tracking with Vision Foundation Models
by: He, Xunjie, et al.
Published: (2025)
by: He, Xunjie, et al.
Published: (2025)
Text-guided Visual Prompt DINO for Generic Segmentation
by: Guan, Yuchen, et al.
Published: (2025)
by: Guan, Yuchen, et al.
Published: (2025)
Few-Shot Adaptation of Grounding DINO for Agricultural Domain
by: Singh, Rajhans, et al.
Published: (2025)
by: Singh, Rajhans, et al.
Published: (2025)
DINO-YOLO: Self-Supervised Pre-training for Data-Efficient Object Detection in Civil Engineering Applications
by: P, Malaisree, et al.
Published: (2025)
by: P, Malaisree, et al.
Published: (2025)
Oh-A-DINO: Understanding and Enhancing Attribute-Level Information in Self-Supervised Object-Centric Representations
by: Wagner, Stefan Sylvius, et al.
Published: (2025)
by: Wagner, Stefan Sylvius, et al.
Published: (2025)
Learning to See Inside Opaque Liquid Containers using Speckle Vibrometry
by: Kichler, Matan, et al.
Published: (2025)
by: Kichler, Matan, et al.
Published: (2025)
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
by: Jevtić, Aleksandar, et al.
Published: (2025)
by: Jevtić, Aleksandar, et al.
Published: (2025)
Multi-task Image Restoration Guided By Robust DINO Features
by: Lin, Xin, et al.
Published: (2023)
by: Lin, Xin, et al.
Published: (2023)
Simplifying DINO via Coding Rate Regularization
by: Wu, Ziyang, et al.
Published: (2025)
by: Wu, Ziyang, et al.
Published: (2025)
Cross-Modal Knowledge Distillation from Spatial Transcriptomics to Histology
by: Hizmi, Arbel, et al.
Published: (2026)
by: Hizmi, Arbel, et al.
Published: (2026)
Similar Items
-
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
by: Kaduri, Omri, et al.
Published: (2024) -
DIVE: Taming DINO for Subject-Driven Video Editing
by: Huang, Yi, et al.
Published: (2024) -
DINO-Foresight: Looking into the Future with DINO
by: Karypidis, Efstathios, et al.
Published: (2024) -
DINO-Tok: Adapting DINO for Visual Tokenizers
by: Jia, Mingkai, et al.
Published: (2025) -
DRoPS: Dynamic 3D Reconstruction of Pre-Scanned Objects
by: Tumanyan, Narek, et al.
Published: (2026)