BootsTAP: Bootstrapped Training for Tracking-Any-Point
Fuente:
arXiv
Saved in:
| Main Authors: | Doersch, Carl, Luc, Pauline, Yang, Yi, Gokay, Dilara, Koppula, Skanda, Gupta, Ankush, Heyward, Joseph, Rocco, Ignacio, Goroshin, Ross, Carreira, João, Zisserman, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
by: Koppula, Skanda, et al.
Published: (2024)
by: Koppula, Skanda, et al.
Published: (2024)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025)
by: Zholus, Artem, et al.
Published: (2025)
TAPNext++: What's Next for Tracking Any Point (TAP)?
by: Jung, Sebastian, et al.
Published: (2026)
by: Jung, Sebastian, et al.
Published: (2026)
Learning from One Continuous Video Stream
by: Carreira, João, et al.
Published: (2023)
by: Carreira, João, et al.
Published: (2023)
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
by: Hasson, Yana, et al.
Published: (2025)
by: Hasson, Yana, et al.
Published: (2025)
Learning from Streaming Video with Orthogonal Gradients
by: Han, Tengda, et al.
Published: (2025)
by: Han, Tengda, et al.
Published: (2025)
Scaling 4D Representations
by: Carreira, João, et al.
Published: (2024)
by: Carreira, João, et al.
Published: (2024)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
by: Heyward, Joseph, et al.
Published: (2024)
by: Heyward, Joseph, et al.
Published: (2024)
MV-TAP: Tracking Any Point in Multi-View Videos
by: Koo, Jahyeok, et al.
Published: (2025)
by: Koo, Jahyeok, et al.
Published: (2025)
A Simple Recipe for Contrastively Pre-training Video-First Encoders Beyond 16 Frames
by: Papalampidi, Pinelopi, et al.
Published: (2023)
by: Papalampidi, Pinelopi, et al.
Published: (2023)
Moving Off-the-Grid: Scene-Grounded Video Representations
by: van Steenkiste, Sjoerd, et al.
Published: (2024)
by: van Steenkiste, Sjoerd, et al.
Published: (2024)
A Mixed Diet Makes DINO An Omnivorous Vision Encoder
by: Kabra, Rishabh, et al.
Published: (2026)
by: Kabra, Rishabh, et al.
Published: (2026)
Perception Test 2025: Challenge Summary and a Unified VQA Extension
by: Heyward, Joseph, et al.
Published: (2026)
by: Heyward, Joseph, et al.
Published: (2026)
Forecasting Motion in the Wild
by: Thakkar, Neerja, et al.
Published: (2026)
by: Thakkar, Neerja, et al.
Published: (2026)
Unique Lives, Shared World: Learning from Single-Life Videos
by: Han, Tengda, et al.
Published: (2025)
by: Han, Tengda, et al.
Published: (2025)
AnthroTAP: Learning Point Tracking with Real-World Motion
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
HasteBoots: Proving TFHE Programmable Bootstrapping in Seconds
by: Fengrun Liu, et al.
Published: (2026)
by: Fengrun Liu, et al.
Published: (2026)
Efficiently Reconstructing Dynamic Scenes One D4RT at a Time
by: Zhang, Chuhan, et al.
Published: (2025)
by: Zhang, Chuhan, et al.
Published: (2025)
TRecViT: A Recurrent Video Transformer
by: Pătrăucean, Viorica, et al.
Published: (2024)
by: Pătrăucean, Viorica, et al.
Published: (2024)
Direct Motion Models for Assessing Generated Videos
by: Allen, Kelsey, et al.
Published: (2025)
by: Allen, Kelsey, et al.
Published: (2025)
Memory Consolidation Enables Long-Context Video Understanding
by: Balažević, Ivana, et al.
Published: (2024)
by: Balažević, Ivana, et al.
Published: (2024)
Recurrence-based Vanishing Point Detection
by: Bharadwaj, Skanda, et al.
Published: (2024)
by: Bharadwaj, Skanda, et al.
Published: (2024)
Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing
by: van der Vleuten, Noah
Published: (2025)
by: van der Vleuten, Noah
Published: (2025)
ReBoot: Encrypted Training of Deep Neural Networks with CKKS Bootstrapping
by: Pirillo, Alberto, et al.
Published: (2025)
by: Pirillo, Alberto, et al.
Published: (2025)
BootTOD: Bootstrap Task-oriented Dialogue Representations by Aligning Diverse Responses
by: Zeng, Weihao, et al.
Published: (2024)
by: Zeng, Weihao, et al.
Published: (2024)
Twin-Boot: Uncertainty-Aware Optimization via Online Two-Sample Bootstrapping
by: Brito, Carlos Stein
Published: (2025)
by: Brito, Carlos Stein
Published: (2025)
ETAP: Event-based Tracking of Any Point
by: Hamann, Friedhelm, et al.
Published: (2024)
by: Hamann, Friedhelm, et al.
Published: (2024)
TAPTR: Tracking Any Point with Transformers as Detection
by: Li, Hongyang, et al.
Published: (2024)
by: Li, Hongyang, et al.
Published: (2024)
Recurrent Video Masked Autoencoders
by: Zoran, Daniel, et al.
Published: (2025)
by: Zoran, Daniel, et al.
Published: (2025)
BootPIG: Bootstrapping Zero-shot Personalized Image Generation Capabilities in Pretrained Diffusion Models
by: Purushwalkam, Senthil, et al.
Published: (2024)
by: Purushwalkam, Senthil, et al.
Published: (2024)
Making sense of the protests in Turkey (and Brazil): contesting neo-liberal urbanism in ‘Rebel Cities’
by: Bulent Gokay
Published: (2015)
by: Bulent Gokay
Published: (2015)
Self-Supervised Any-Point Tracking by Contrastive Random Walks
by: Shrivastava, Ayush, et al.
Published: (2024)
by: Shrivastava, Ayush, et al.
Published: (2024)
Frozen Forecasting: A Unified Evaluation
by: Walker, Jacob C, et al.
Published: (2025)
by: Walker, Jacob C, et al.
Published: (2025)
Pits and Boots
by: Roy, Michael
Published: (2025)
by: Roy, Michael
Published: (2025)
Méliès Boots
by: Solomon, Matthew
Published: (2022)
by: Solomon, Matthew
Published: (2022)
Doubt's Boots
Published: (2022)
Published: (2022)
Single-Model and Any-Modality for Video Object Tracking
by: Wu, Zongwei, et al.
Published: (2023)
by: Wu, Zongwei, et al.
Published: (2023)
Tracking Any Point Methods for Markerless 3D Tissue Tracking in Endoscopic Stereo Images
by: Reuter, Konrad, et al.
Published: (2025)
by: Reuter, Konrad, et al.
Published: (2025)
Evolution of Panchayati Raj System in India with Special Reference to Telangana: Issues and Challenges
by: Mallesham, Koppula, et al.
Published: (2025)
by: Mallesham, Koppula, et al.
Published: (2025)
Track Any Motions under Any Disturbances
by: Zhang, Zhikai, et al.
Published: (2025)
by: Zhang, Zhikai, et al.
Published: (2025)
Similar Items
-
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
by: Koppula, Skanda, et al.
Published: (2024) -
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025) -
TAPNext++: What's Next for Tracking Any Point (TAP)?
by: Jung, Sebastian, et al.
Published: (2026) -
Learning from One Continuous Video Stream
by: Carreira, João, et al.
Published: (2023) -
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
by: Hasson, Yana, et al.
Published: (2025)