Adapting SAM 2 for Visual Object Tracking: 1st Place Solution for MMVPR Challenge Multi-Modal Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Cheng-Yen, Huang, Hsiang-Wei, Kim, Pyong-Kun, Kuo, Chien-Kai, Chang, Jui-Wei, Kim, Kwang-Ju, Huang, Chung-I, Hwang, Jenq-Neng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Technical Report for ReID-SAM on SkiTB Visual Tracking Challenge 2025
by: Li, Kunjun, et al.
Published: (2025)
by: Li, Kunjun, et al.
Published: (2025)
SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory
by: Yang, Cheng-Yen, et al.
Published: (2024)
by: Yang, Cheng-Yen, et al.
Published: (2024)
GTA: Global Tracklet Association for Multi-Object Tracking in Sports
by: Sun, Jiacheng, et al.
Published: (2024)
by: Sun, Jiacheng, et al.
Published: (2024)
MambaMOT: State-Space Model as Motion Predictor for Multi-Object Tracking
by: Huang, Hsiang-Wei, et al.
Published: (2024)
by: Huang, Hsiang-Wei, et al.
Published: (2024)
Recent Advances in Embedding Methods for Multi-Object Tracking: A Survey
by: Wang, Gaoang, et al.
Published: (2022)
by: Wang, Gaoang, et al.
Published: (2022)
Detector-in-the-Loop Tracking: Active Memory Rectification for Stable Glottic Opening Localization
by: Wang, Huayu, et al.
Published: (2026)
by: Wang, Huayu, et al.
Published: (2026)
Boosting Online 3D Multi-Object Tracking through Camera-Radar Cross Check
by: Kuan, Sheng-Yao, et al.
Published: (2024)
by: Kuan, Sheng-Yao, et al.
Published: (2024)
A Density-Guided Temporal Attention Transformer for Indiscernible Object Counting in Underwater Video
by: Yang, Cheng-Yen, et al.
Published: (2024)
by: Yang, Cheng-Yen, et al.
Published: (2024)
ToSA: Token Merging with Spatial Awareness
by: Huang, Hsiang-Wei, et al.
Published: (2025)
by: Huang, Hsiang-Wei, et al.
Published: (2025)
Modeling LLM Agent Reviewer Dynamics in Elo-Ranked Review System
by: Huang, Hsiang-Wei, et al.
Published: (2026)
by: Huang, Hsiang-Wei, et al.
Published: (2026)
Exploring Probabilistic Modeling Beyond Domain Generalization for Semantic Segmentation
by: Chen, I-Hsiang, et al.
Published: (2025)
by: Chen, I-Hsiang, et al.
Published: (2025)
Understanding Identity Continuity in Thermal Video through Scene-Level Consistency
by: Sun, Wei-Chieh, et al.
Published: (2026)
by: Sun, Wei-Chieh, et al.
Published: (2026)
PackDiT: Joint Human Motion and Text Generation via Mutual Prompting
by: Jiang, Zhongyu, et al.
Published: (2025)
by: Jiang, Zhongyu, et al.
Published: (2025)
CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models
by: Huang, Hsiang-Wei, et al.
Published: (2026)
by: Huang, Hsiang-Wei, et al.
Published: (2026)
Reasoning Matters for 3D Visual Grounding
by: Huang, Hsiang-Wei, et al.
Published: (2026)
by: Huang, Hsiang-Wei, et al.
Published: (2026)
ToddlerAct: A Toddler Action Recognition Dataset for Gross Motor Development Assessment
by: Huang, Hsiang-Wei, et al.
Published: (2024)
by: Huang, Hsiang-Wei, et al.
Published: (2024)
Warehouse Spatial Question Answering with LLM Agent
by: Huang, Hsiang-Wei, et al.
Published: (2025)
by: Huang, Hsiang-Wei, et al.
Published: (2025)
1st Place Solution for MOSE Track in CVPR 2024 PVUW Workshop: Complex Video Object Segmentation
by: Miao, Deshui, et al.
Published: (2024)
by: Miao, Deshui, et al.
Published: (2024)
Video Object Segmentation via SAM 2: The 4th Solution for LSVOS Challenge VOS Track
by: Pan, Feiyu, et al.
Published: (2024)
by: Pan, Feiyu, et al.
Published: (2024)
1st Place Solution of Multiview Egocentric Hand Tracking Challenge ECCV2024
by: Zou, Minqiang, et al.
Published: (2024)
by: Zou, Minqiang, et al.
Published: (2024)
Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression
by: Li, Kunjun, et al.
Published: (2025)
by: Li, Kunjun, et al.
Published: (2025)
1st Place Solution to the 8th HANDS Workshop Challenge -- ARCTIC Track: 3DGS-based Bimanual Category-agnostic Interaction Reconstruction
by: On, Jeongwan, et al.
Published: (2024)
by: On, Jeongwan, et al.
Published: (2024)
DeconfuseTrack:Dealing with Confusion for Multi-Object Tracking
by: Huang, Cheng, et al.
Published: (2024)
by: Huang, Cheng, et al.
Published: (2024)
Single-image driven 3d viewpoint training data augmentation for effective wine label recognition
by: Huang, Yueh-Cheng, et al.
Published: (2024)
by: Huang, Yueh-Cheng, et al.
Published: (2024)
HopTrack: A Real-time Multi-Object Tracking System for Embedded Devices
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
UNINEXT-Cutie: The 1st Solution for LSVOS Challenge RVOS Track
by: Fang, Hao, et al.
Published: (2024)
by: Fang, Hao, et al.
Published: (2024)
Underwater Camouflaged Object Tracking Meets Vision-Language SAM2
by: Zhang, Chunhui, et al.
Published: (2024)
by: Zhang, Chunhui, et al.
Published: (2024)
Towards Fine-grained Large Object Segmentation 1st Place Solution to 3D AI Challenge 2020 -- Instance Segmentation Track
by: Chen, Zehui, et al.
Published: (2020)
by: Chen, Zehui, et al.
Published: (2020)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
Real-Time Shape Tracking of Facial Landmarks
by: Kim, Hyungjoon, et al.
Published: (2018)
by: Kim, Hyungjoon, et al.
Published: (2018)
The Solution for the CVPR 2023 1st foundation model challenge-Track2
by: Xu, Haonan, et al.
Published: (2024)
by: Xu, Haonan, et al.
Published: (2024)
RegTrack: Simplicity Beneath Complexity in Robust Multi-Modal 3D Multi-Object Tracking
by: Gu, Lipeng, et al.
Published: (2024)
by: Gu, Lipeng, et al.
Published: (2024)
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
by: Yu, Jieming, et al.
Published: (2024)
by: Yu, Jieming, et al.
Published: (2024)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023)
by: Luo, Run, et al.
Published: (2023)
VersaT2I: Improving Text-to-Image Models with Versatile Reward
by: Guo, Jianshu, et al.
Published: (2024)
by: Guo, Jianshu, et al.
Published: (2024)
SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors
by: Fan, Ruijie, et al.
Published: (2025)
by: Fan, Ruijie, et al.
Published: (2025)
VariabilityTrack:Multi-Object Tracking with Variable Speed Object Movement
by: Luo, Run, et al.
Published: (2022)
by: Luo, Run, et al.
Published: (2022)
Adaptive Computational Methods for Robust Object Tracking
by: Ji-Hoon Kwon, et al.
Published: (2024)
by: Ji-Hoon Kwon, et al.
Published: (2024)
Rethinking Memory Design in SAM-Based Visual Object Tracking
by: Alansari, Mohamad, et al.
Published: (2025)
by: Alansari, Mohamad, et al.
Published: (2025)
DAug: Diffusion-based Channel Augmentation for Radiology Image Retrieval and Classification
by: Jin, Ying, et al.
Published: (2024)
by: Jin, Ying, et al.
Published: (2024)
Similar Items
-
Technical Report for ReID-SAM on SkiTB Visual Tracking Challenge 2025
by: Li, Kunjun, et al.
Published: (2025) -
SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory
by: Yang, Cheng-Yen, et al.
Published: (2024) -
GTA: Global Tracklet Association for Multi-Object Tracking in Sports
by: Sun, Jiacheng, et al.
Published: (2024) -
MambaMOT: State-Space Model as Motion Predictor for Multi-Object Tracking
by: Huang, Hsiang-Wei, et al.
Published: (2024) -
Recent Advances in Embedding Methods for Multi-Object Tracking: A Survey
by: Wang, Gaoang, et al.
Published: (2022)