Track Anything Annotate: Video annotation and dataset generation of computer vision models
Fuente:
arXiv
Saved in:
| Main Authors: | Ivanov, Nikita, Klimov, Mark, Glukhikh, Dmitry, Chernysheva, Tatiana, Glukhikh, Igor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Segmentation of temporomandibular joint structures on mri images using neural networks for diagnosis of pathologies
by: Ivanov, Maksim I., et al.
Published: (2025)
by: Ivanov, Maksim I., et al.
Published: (2025)
Annolid: Annotate, Segment, and Track Anything You Need
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
A survey of datasets for computer vision in agriculture
by: Heider, Nico, et al.
Published: (2025)
by: Heider, Nico, et al.
Published: (2025)
Efficient Track Anything
by: Xiong, Yunyang, et al.
Published: (2024)
by: Xiong, Yunyang, et al.
Published: (2024)
Segment Anything for Cell Tracking
by: Chen, Zhu, et al.
Published: (2025)
by: Chen, Zhu, et al.
Published: (2025)
Track Anything Behind Everything: Zero-Shot Amodal Video Object Segmentation
by: Hudson, Finlay G. C., et al.
Published: (2024)
by: Hudson, Finlay G. C., et al.
Published: (2024)
Segment Anything for Video: A Comprehensive Review of Video Object Segmentation and Tracking from Past to Future
by: Xu, Guoping, et al.
Published: (2025)
by: Xu, Guoping, et al.
Published: (2025)
Annotation-Efficient Task Guidance for Medical Segment Anything
by: Ward, Tyler, et al.
Published: (2024)
by: Ward, Tyler, et al.
Published: (2024)
No Free Lunch in Annotation either: An objective evaluation of foundation models for streamlining annotation in animal tracking
by: Mededovic, Emil, et al.
Published: (2025)
by: Mededovic, Emil, et al.
Published: (2025)
A multi-modal vision-language model for generalizable annotation-free pathology localization
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Augmenting Efficient Real-time Surgical Instrument Segmentation in Video with Point Tracking and Segment Anything
by: Wu, Zijian, et al.
Published: (2024)
by: Wu, Zijian, et al.
Published: (2024)
Place Anything into Any Video
by: Liu, Ziling, et al.
Published: (2024)
by: Liu, Ziling, et al.
Published: (2024)
SAM 2++: Tracking Anything at Any Granularity
by: Zhang, Jiaming, et al.
Published: (2025)
by: Zhang, Jiaming, et al.
Published: (2025)
EdgeTAM: On-Device Track Anything Model
by: Zhou, Chong, et al.
Published: (2025)
by: Zhou, Chong, et al.
Published: (2025)
Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos
by: Lin, Weifeng, et al.
Published: (2025)
by: Lin, Weifeng, et al.
Published: (2025)
Get In Video: Add Anything You Want to the Video
by: Zhuang, Shaobin, et al.
Published: (2025)
by: Zhuang, Shaobin, et al.
Published: (2025)
SAM-PD: How Far Can SAM Take Us in Tracking and Segmenting Anything in Videos by Prompt Denoising
by: Zhou, Tao, et al.
Published: (2024)
by: Zhou, Tao, et al.
Published: (2024)
Improving analytical color and texture similarity estimation methods for dataset-agnostic person reidentification
by: Gabdullin, Nikita
Published: (2024)
by: Gabdullin, Nikita
Published: (2024)
MObyGaze: a film dataset of multimodal objectification densely annotated by experts
by: Tores, Julie, et al.
Published: (2025)
by: Tores, Julie, et al.
Published: (2025)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Computer vision training dataset generation for robotic environments using Gaussian splatting
by: Niżeniec, Patryk, et al.
Published: (2025)
by: Niżeniec, Patryk, et al.
Published: (2025)
A strongly annotated passive acoustic dataset for tropical bird monitoring
by: Ruiz, Daniela, et al.
Published: (2026)
by: Ruiz, Daniela, et al.
Published: (2026)
Track Anything Rapter(TAR)
by: Puthanveettil, Tharun V., et al.
Published: (2024)
by: Puthanveettil, Tharun V., et al.
Published: (2024)
SAMJ: Fast Image Annotation on ImageJ/Fiji via Segment Anything Model
by: Garcia-Lopez-de-Haro, Carlos, et al.
Published: (2025)
by: Garcia-Lopez-de-Haro, Carlos, et al.
Published: (2025)
Understanding and evaluating computer vision models through the lens of counterfactuals
by: Shukla, Pushkar
Published: (2025)
by: Shukla, Pushkar
Published: (2025)
AnimateAnything: Consistent and Controllable Animation for Video Generation
by: Lei, Guojun, et al.
Published: (2024)
by: Lei, Guojun, et al.
Published: (2024)
Anything in Any Scene: Photorealistic Video Object Insertion
by: Bai, Chen, et al.
Published: (2024)
by: Bai, Chen, et al.
Published: (2024)
Mitigating annotation shift in cancer classification using single image generative models
by: Arcas, Marta Buetas, et al.
Published: (2024)
by: Arcas, Marta Buetas, et al.
Published: (2024)
Matching Anything by Segmenting Anything
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
CoTracker3: Simpler and Better Point Tracking by Pseudo-Labelling Real Videos
by: Karaev, Nikita, et al.
Published: (2024)
by: Karaev, Nikita, et al.
Published: (2024)
Fine-tuning Segment Anything for Real-Time Tumor Tracking in Cine-MRI
by: Boussot, Valentin, et al.
Published: (2025)
by: Boussot, Valentin, et al.
Published: (2025)
SAM-DA: UAV Tracks Anything at Night with SAM-Powered Domain Adaptation
by: Fu, Changhong, et al.
Published: (2023)
by: Fu, Changhong, et al.
Published: (2023)
Reconstruct Anything Model: a lightweight general model for computational imaging
by: Terris, Matthieu, et al.
Published: (2025)
by: Terris, Matthieu, et al.
Published: (2025)
3AM: 3egment Anything with Geometric Consistency in Videos
by: Sun, Yang-Che, et al.
Published: (2026)
by: Sun, Yang-Che, et al.
Published: (2026)
Biomedical SAM 2: Segment Anything in Biomedical Images and Videos
by: Yan, Zhiling, et al.
Published: (2024)
by: Yan, Zhiling, et al.
Published: (2024)
360Anything: Geometry-Free Lifting of Images and Videos to 360°
by: Wu, Ziyi, et al.
Published: (2026)
by: Wu, Ziyi, et al.
Published: (2026)
Tracking and Segmenting Anything in Any Modality
by: Zhang, Tianlu, et al.
Published: (2025)
by: Zhang, Tianlu, et al.
Published: (2025)
3DGraphLLM: Combining Semantic Graphs and Large Language Models for 3D Scene Understanding
by: Zemskova, Tatiana, et al.
Published: (2024)
by: Zemskova, Tatiana, et al.
Published: (2024)
PLANesT-3D: A new annotated dataset for segmentation of 3D plant point clouds
by: Mertoğlu, Kerem, et al.
Published: (2024)
by: Mertoğlu, Kerem, et al.
Published: (2024)
SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models
by: Makarov, Vladislav, et al.
Published: (2026)
by: Makarov, Vladislav, et al.
Published: (2026)
Similar Items
-
Segmentation of temporomandibular joint structures on mri images using neural networks for diagnosis of pathologies
by: Ivanov, Maksim I., et al.
Published: (2025) -
Annolid: Annotate, Segment, and Track Anything You Need
by: Yang, Chen, et al.
Published: (2024) -
A survey of datasets for computer vision in agriculture
by: Heider, Nico, et al.
Published: (2025) -
Efficient Track Anything
by: Xiong, Yunyang, et al.
Published: (2024) -
Segment Anything for Cell Tracking
by: Chen, Zhu, et al.
Published: (2025)