Autogenic Language Embedding for Coherent Point Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Zikai, Tang, Ying, Luo, Run, Ma, Lintao, Yu, Junqing, Chen, Yi-Ping Phoebe, Yang, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hypergraph-State Collaborative Reasoning for Multi-Object Tracking
by: Song, Zikai, et al.
Published: (2026)
by: Song, Zikai, et al.
Published: (2026)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023)
by: Luo, Run, et al.
Published: (2023)
GateMOT: Q-Gated Attention for Dense Object Tracking
by: Lv, Mingjin, et al.
Published: (2026)
by: Lv, Mingjin, et al.
Published: (2026)
MVP: Winning Solution to SMP Challenge 2025 Video Track
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
CurEvo: Curriculum-Guided Self-Evolution for Video Understanding
by: Zeng, Guiyi, et al.
Published: (2026)
by: Zeng, Guiyi, et al.
Published: (2026)
OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction
by: Ye, Liliang, et al.
Published: (2026)
by: Ye, Liliang, et al.
Published: (2026)
IP-MOT: Instance Prompt Learning for Cross-Domain Multi-Object Tracking
by: Luo, Run, et al.
Published: (2024)
by: Luo, Run, et al.
Published: (2024)
Optimized View and Geometry Distillation from Multi-view Diffuser
by: Zhang, Youjia, et al.
Published: (2023)
by: Zhang, Youjia, et al.
Published: (2023)
SF2T: Self-supervised Fragment Finetuning of Video-LLMs for Fine-Grained Understanding
by: Hu, Yangliu, et al.
Published: (2025)
by: Hu, Yangliu, et al.
Published: (2025)
Ref-GS: Directional Factorization for 2D Gaussian Splatting
by: Zhang, Youjia, et al.
Published: (2024)
by: Zhang, Youjia, et al.
Published: (2024)
Cross-Modality Masked Learning for Survival Prediction in ICI Treated NSCLC Patients
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
TIGER: Text-Instructed 3D Gaussian Retrieval and Coherent Editing
by: Xu, Teng, et al.
Published: (2024)
by: Xu, Teng, et al.
Published: (2024)
VariabilityTrack:Multi-Object Tracking with Variable Speed Object Movement
by: Luo, Run, et al.
Published: (2022)
by: Luo, Run, et al.
Published: (2022)
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Progressive Text-to-Image Diffusion with Soft Latent Direction
by: Ye, YuTeng, et al.
Published: (2023)
by: Ye, YuTeng, et al.
Published: (2023)
PointGS: Point Attention-Aware Sparse View Synthesis with Gaussian Splatting
by: Xiang, Lintao, et al.
Published: (2025)
by: Xiang, Lintao, et al.
Published: (2025)
Learning Instance-Aware Correspondences for Robust Multi-Instance Point Cloud Registration in Cluttered Scenes
by: Yu, Zhiyuan, et al.
Published: (2024)
by: Yu, Zhiyuan, et al.
Published: (2024)
CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
PX2Tooth: Reconstructing the 3D Point Cloud Teeth from a Single Panoramic X-ray
by: Ma, Wen, et al.
Published: (2024)
by: Ma, Wen, et al.
Published: (2024)
TrackSSM: A General Motion Predictor by State-Space Model
by: Hu, Bin, et al.
Published: (2024)
by: Hu, Bin, et al.
Published: (2024)
Densemarks: Learning Canonical Embeddings for Human Heads Images via Point Tracks
by: Pozdeev, Dmitrii, et al.
Published: (2025)
by: Pozdeev, Dmitrii, et al.
Published: (2025)
Solution for Point Tracking Task of ECCV 2nd Perception Test Challenge 2024
by: Zhang, Yuxuan, et al.
Published: (2024)
by: Zhang, Yuxuan, et al.
Published: (2024)
DEEM: Diffusion Models Serve as the Eyes of Large Language Models for Image Perception
by: Luo, Run, et al.
Published: (2024)
by: Luo, Run, et al.
Published: (2024)
Unifying Visual and Vision-Language Tracking via Contrastive Learning
by: Ma, Yinchao, et al.
Published: (2024)
by: Ma, Yinchao, et al.
Published: (2024)
Correlation-Embedded Transformer Tracking: A Single-Branch Framework
by: Xie, Fei, et al.
Published: (2024)
by: Xie, Fei, et al.
Published: (2024)
From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object Tracking
by: Shao, Yuqing, et al.
Published: (2025)
by: Shao, Yuqing, et al.
Published: (2025)
VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation
by: Zhou, Hanyu, et al.
Published: (2025)
by: Zhou, Hanyu, et al.
Published: (2025)
PathoHR: Hierarchical Reasoning for Vision-Language Models in Pathology
by: Huang, Yating, et al.
Published: (2025)
by: Huang, Yating, et al.
Published: (2025)
Hyperbolic Contrastive Learning for Hierarchical 3D Point Cloud Embedding
by: Liu, Yingjie, et al.
Published: (2025)
by: Liu, Yingjie, et al.
Published: (2025)
Exploring Temporally-Aware Features for Point Tracking
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
by: Kim, Inès Hyeonsu, et al.
Published: (2025)
Recent Advances in Embedding Methods for Multi-Object Tracking: A Survey
by: Wang, Gaoang, et al.
Published: (2022)
by: Wang, Gaoang, et al.
Published: (2022)
MATE: Motion-Augmented Temporal Consistency for Event-based Point Tracking
by: Han, Han, et al.
Published: (2024)
by: Han, Han, et al.
Published: (2024)
Multi-View 3D Point Tracking
by: Rajič, Frano, et al.
Published: (2025)
by: Rajič, Frano, et al.
Published: (2025)
Interactive Occlusion Boundary Estimation through Exploitation of Synthetic Data
by: Xu, Lintao, et al.
Published: (2024)
by: Xu, Lintao, et al.
Published: (2024)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025)
by: Zholus, Artem, et al.
Published: (2025)
Dynamic Multimodal Evaluation with Flexible Complexity by Vision-Language Bootstrapping
by: Yang, Yue, et al.
Published: (2024)
by: Yang, Yue, et al.
Published: (2024)
GLAD: Generative Language-Assisted Visual Tracking for Low-Semantic Templates
by: Luo, Xingyu, et al.
Published: (2026)
by: Luo, Xingyu, et al.
Published: (2026)
SMILEtrack: SiMIlarity LEarning for Occlusion-Aware Multiple Object Tracking
by: Wang, Yu-Hsiang, et al.
Published: (2022)
by: Wang, Yu-Hsiang, et al.
Published: (2022)
MaGS: Reconstructing and Simulating Dynamic 3D Objects with Mesh-adsorbed Gaussian Splatting
by: Ma, Shaojie, et al.
Published: (2024)
by: Ma, Shaojie, et al.
Published: (2024)
Similar Items
-
Hypergraph-State Collaborative Reasoning for Multi-Object Tracking
by: Song, Zikai, et al.
Published: (2026) -
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023) -
GateMOT: Q-Gated Attention for Dense Object Tracking
by: Lv, Mingjin, et al.
Published: (2026) -
MVP: Winning Solution to SMP Challenge 2025 Video Track
by: Ye, Liliang, et al.
Published: (2025) -
CurEvo: Curriculum-Guided Self-Evolution for Video Understanding
by: Zeng, Guiyi, et al.
Published: (2026)