Unveiling the Power of Self-supervision for Multi-view Multi-human Association and Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Wei, Wang, Feifan, Han, Ruize, Qian, Zekun, Wang, Song |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OCTrack: Benchmarking the Open-Corpus Multi-Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration
by: Qian, Zekun, et al.
Published: (2022)
by: Qian, Zekun, et al.
Published: (2022)
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
Point Cloud Self-supervised Learning via 3D to Multi-view Masked Learner
by: Chen, Zhimin, et al.
Published: (2023)
by: Chen, Zhimin, et al.
Published: (2023)
Intra-view and Inter-view Correlation Guided Multi-view Novel Class Discovery
by: Wan, Xinhang, et al.
Published: (2025)
by: Wan, Xinhang, et al.
Published: (2025)
MITracker: Multi-View Integration for Visual Object Tracking
by: Xu, Mengjie, et al.
Published: (2025)
by: Xu, Mengjie, et al.
Published: (2025)
RTAT: A Robust Two-stage Association Tracker for Multi-Object Tracking
by: Guo, Song, et al.
Published: (2024)
by: Guo, Song, et al.
Published: (2024)
Powerful Teachers Matter: Text-Guided Multi-view Knowledge Distillation with Visual Prior Enhancement
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
MMVIAD: Multi-view Multi-task Video Understanding for Industrial Anomaly Detection
by: Zhao, Xiran, et al.
Published: (2026)
by: Zhao, Xiran, et al.
Published: (2026)
Multi-weather Cross-view Geo-localization Using Denoising Diffusion Models
by: Feng, Tongtong, et al.
Published: (2024)
by: Feng, Tongtong, et al.
Published: (2024)
CL-MVSNet: Unsupervised Multi-view Stereo with Dual-level Contrastive Learning
by: Xiong, Kaiqiang, et al.
Published: (2025)
by: Xiong, Kaiqiang, et al.
Published: (2025)
Pink: Unveiling the Power of Referential Comprehension for Multi-modal LLMs
by: Xuan, Shiyu, et al.
Published: (2023)
by: Xuan, Shiyu, et al.
Published: (2023)
GRASPTrack: Geometry-Reasoned Association via Segmentation and Projection for Multi-Object Tracking
by: Han, Xudong, et al.
Published: (2025)
by: Han, Xudong, et al.
Published: (2025)
Balanced Multi-view Clustering
by: Li, Zhenglai, et al.
Published: (2025)
by: Li, Zhenglai, et al.
Published: (2025)
Self-Supervised Multi-Object Tracking with Path Consistency
by: Lu, Zijia, et al.
Published: (2024)
by: Lu, Zijia, et al.
Published: (2024)
FC-Track: Overlap-Aware Post-Association Correction for Online Multi-Object Tracking
by: Ju, Cheng, et al.
Published: (2026)
by: Ju, Cheng, et al.
Published: (2026)
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
Awesome Multi-modal Object Tracking
by: Zhang, Chunhui, et al.
Published: (2024)
by: Zhang, Chunhui, et al.
Published: (2024)
ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking
by: Ge, Jiawei, et al.
Published: (2026)
by: Ge, Jiawei, et al.
Published: (2026)
A Self-supervised Pressure Map human keypoint Detection Approch: Optimizing Generalization and Computational Efficiency Across Datasets
by: Yu, Chengzhang, et al.
Published: (2024)
by: Yu, Chengzhang, et al.
Published: (2024)
Reasoning Path and Latent State Analysis for Multi-view Visual Spatial Reasoning: A Cognitive Science Perspective
by: Xue, Qiyao, et al.
Published: (2025)
by: Xue, Qiyao, et al.
Published: (2025)
FastTrackTr:Towards Fast Multi-Object Tracking with Transformers
by: Liao, Pan, et al.
Published: (2024)
by: Liao, Pan, et al.
Published: (2024)
Unconstrained Multi-view Human Pose Estimation with Algebraic Priors
by: Qin, Xiaolin, et al.
Published: (2026)
by: Qin, Xiaolin, et al.
Published: (2026)
Deformable Image Registration for Self-supervised Cardiac Phase Detection in Multi-View Multi-Disease Cardiac Magnetic Resonance Images
by: Koehler, Sven, et al.
Published: (2025)
by: Koehler, Sven, et al.
Published: (2025)
Learning Progressive Adaptation for Multi-Modal Tracking
by: Wang, He, et al.
Published: (2026)
by: Wang, He, et al.
Published: (2026)
LLMTrack: Semantic Multi-Object Tracking with Multi-modal Large Language Models
by: Liao, Pan, et al.
Published: (2026)
by: Liao, Pan, et al.
Published: (2026)
Semi-supervised Semantic Segmentation for Remote Sensing Images via Multi-scale Uncertainty Consistency and Cross-Teacher-Student Attention
by: Wang, Shanwen, et al.
Published: (2025)
by: Wang, Shanwen, et al.
Published: (2025)
Imputation-free and Alignment-free: Incomplete Multi-view Clustering Driven by Consensus Semantic Learning
by: Dai, Yuzhuo, et al.
Published: (2025)
by: Dai, Yuzhuo, et al.
Published: (2025)
Sparse BEV Fusion with Self-View Consistency for Multi-View Detection and Tracking
by: Toida, Keisuke, et al.
Published: (2025)
by: Toida, Keisuke, et al.
Published: (2025)
Geometry-Guided Reinforcement Learning for Multi-view Consistent 3D Scene Editing
by: Wang, Jiyuan, et al.
Published: (2026)
by: Wang, Jiyuan, et al.
Published: (2026)
Motion Estimation for Multi-Object Tracking using KalmanNet with Semantic-Independent Encoding
by: Song, Jian, et al.
Published: (2025)
by: Song, Jian, et al.
Published: (2025)
Enhanced Parking Perception by Multi-Task Fisheye Cross-view Transformers
by: Musabini, Antonyo, et al.
Published: (2024)
by: Musabini, Antonyo, et al.
Published: (2024)
MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model
by: Tong, Jinguang, et al.
Published: (2026)
by: Tong, Jinguang, et al.
Published: (2026)
DiffPoint: Single and Multi-view Point Cloud Reconstruction with ViT Based Diffusion Model
by: Feng, Yu, et al.
Published: (2024)
by: Feng, Yu, et al.
Published: (2024)
PointCG: Self-supervised Point Cloud Learning via Joint Completion and Generation
by: Liu, Yun, et al.
Published: (2024)
by: Liu, Yun, et al.
Published: (2024)
SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning
by: Guo, Xiaojun, et al.
Published: (2025)
by: Guo, Xiaojun, et al.
Published: (2025)
MuTri: Multi-view Tri-alignment for OCT to OCTA 3D Image Translation
by: Chen, Zhuangzhuang, et al.
Published: (2025)
by: Chen, Zhuangzhuang, et al.
Published: (2025)
Uncertainty-Encoded Multi-Modal Fusion for Robust Object Detection in Autonomous Driving
by: Lou, Yang, et al.
Published: (2023)
by: Lou, Yang, et al.
Published: (2023)
Similar Items
-
OCTrack: Benchmarking the Open-Corpus Multi-Object Tracking
by: Qian, Zekun, et al.
Published: (2024) -
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking
by: Qian, Zekun, et al.
Published: (2024) -
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026) -
From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration
by: Qian, Zekun, et al.
Published: (2022) -
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
by: Qian, Zekun, et al.
Published: (2026)