Less is More: Token Context-aware Learning for Object Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Chenlong, Zhong, Bineng, Liang, Qihua, Zheng, Yaozong, Li, Guorong, Song, Shuxiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decoupled Spatio-Temporal Consistency Learning for Self-Supervised Tracking
by: Zheng, Yaozong, et al.
Published: (2025)
by: Zheng, Yaozong, et al.
Published: (2025)
Robust Tracking via Mamba-based Context-aware Token Learning
by: Xie, Jinxia, et al.
Published: (2024)
by: Xie, Jinxia, et al.
Published: (2024)
Learning to Track Instance from Single Nature Language Description
by: Zheng, Yaozong, et al.
Published: (2026)
by: Zheng, Yaozong, et al.
Published: (2026)
Towards Universal Modal Tracking with Online Dense Temporal Token Learning
by: Zheng, Yaozong, et al.
Published: (2025)
by: Zheng, Yaozong, et al.
Published: (2025)
MambaLCT: Boosting Tracking via Long-term Context State Space Model
by: Li, Xiaohai, et al.
Published: (2024)
by: Li, Xiaohai, et al.
Published: (2024)
Boosting Self-Supervised Tracking with Contextual Prompts and Noise Learning
by: Zheng, Yaozong, et al.
Published: (2026)
by: Zheng, Yaozong, et al.
Published: (2026)
ODTrack: Online Dense Temporal Token Learning for Visual Tracking
by: Zheng, Yaozong, et al.
Published: (2024)
by: Zheng, Yaozong, et al.
Published: (2024)
Similarity-Guided Layer-Adaptive Vision Transformer for UAV Tracking
by: Xue, Chaocan, et al.
Published: (2025)
by: Xue, Chaocan, et al.
Published: (2025)
Dual-branch Distilled Transformer for Efficient Asymmetric UAV Tracking
by: Yang, Hongtao, et al.
Published: (2026)
by: Yang, Hongtao, et al.
Published: (2026)
An Efficient Token Compression Framework for Visual Object Tracking
by: Wu, Weijing, et al.
Published: (2026)
by: Wu, Weijing, et al.
Published: (2026)
UBATrack: Spatio-Temporal State Space Model for General Multi-Modal Tracking
by: Liang, Qihua, et al.
Published: (2026)
by: Liang, Qihua, et al.
Published: (2026)
Robust RGB-T Tracking via Learnable Visual Fourier Prompt Fine-tuning and Modality Fusion Prompt Generation
by: Yang, Hongtao, et al.
Published: (2025)
by: Yang, Hongtao, et al.
Published: (2025)
Dynamic Updates for Language Adaptation in Visual-Language Tracking
by: Li, Xiaohai, et al.
Published: (2025)
by: Li, Xiaohai, et al.
Published: (2025)
Explicit Visual Prompts for Visual Object Tracking
by: Shi, Liangtao, et al.
Published: (2024)
by: Shi, Liangtao, et al.
Published: (2024)
Explicit Context Reasoning with Supervision for Visual Tracking
by: Zeng, Fansheng, et al.
Published: (2025)
by: Zeng, Fansheng, et al.
Published: (2025)
Adaptive Perception for Unified Visual Multi-modal Object Tracking
by: Hu, Xiantao, et al.
Published: (2025)
by: Hu, Xiantao, et al.
Published: (2025)
Autoregressive Queries for Adaptive Tracking with Spatio-TemporalTransformers
by: Xie, Jinxia, et al.
Published: (2024)
by: Xie, Jinxia, et al.
Published: (2024)
Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking
by: Hu, Xiantao, et al.
Published: (2024)
by: Hu, Xiantao, et al.
Published: (2024)
Observe Less, Understand More: Cost-aware Cross-scale Observation for Remote Sensing Understanding
by: Xie, Zhenghao, et al.
Published: (2026)
by: Xie, Zhenghao, et al.
Published: (2026)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
by: Zhong, Pengzhi, et al.
Published: (2025)
by: Zhong, Pengzhi, et al.
Published: (2025)
ClickTrack: Towards Real-time Interactive Single Object Tracking
by: Wang, Kuiran, et al.
Published: (2024)
by: Wang, Kuiran, et al.
Published: (2024)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
by: Wu, Shaojin, et al.
Published: (2025)
by: Wu, Shaojin, et al.
Published: (2025)
Less is More: Token-Efficient Video-QA via Adaptive Frame-Pruning and Semantic Graph Integration
by: Wang, Shaoguang, et al.
Published: (2025)
by: Wang, Shaoguang, et al.
Published: (2025)
MAFE R-CNN: Selecting More Samples to Learn Category-aware Features for Small Object Detection
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
Context-aware Difference Distilling for Multi-change Captioning
by: Tu, Yunbin, et al.
Published: (2024)
by: Tu, Yunbin, et al.
Published: (2024)
CAMOT: Camera Angle-aware Multi-Object Tracking
by: Limanta, Felix, et al.
Published: (2024)
by: Limanta, Felix, et al.
Published: (2024)
MotionTrack: Learning Motion Predictor for Multiple Object Tracking
by: Xiao, Changcheng, et al.
Published: (2023)
by: Xiao, Changcheng, et al.
Published: (2023)
ContextHOI: Spatial Context Learning for Human-Object Interaction Detection
by: Jia, Mingda, et al.
Published: (2024)
by: Jia, Mingda, et al.
Published: (2024)
Achieving More with Less: Additive Prompt Tuning for Rehearsal-Free Class-Incremental Learning
by: Chen, Haoran, et al.
Published: (2025)
by: Chen, Haoran, et al.
Published: (2025)
Learning More by Seeing Less: Structure First Learning for Efficient, Transferable, and Human-Aligned Vision
by: Li, Tianqin, et al.
Published: (2025)
by: Li, Tianqin, et al.
Published: (2025)
Towards Context-aware Convolutional Network for Image Restoration
by: Hao, Fangwei, et al.
Published: (2024)
by: Hao, Fangwei, et al.
Published: (2024)
Less is More in Semantic Space: Intrinsic Decoupling via Clifford-M for Fundus Image Classification
by: Zheng, Yifeng
Published: (2026)
by: Zheng, Yifeng
Published: (2026)
Target-aware Bidirectional Fusion Transformer for Aerial Object Tracking
by: Sun, Xinglong, et al.
Published: (2025)
by: Sun, Xinglong, et al.
Published: (2025)
Hierarchical Instruction-aware Embodied Visual Tracking
by: Wu, Kui, et al.
Published: (2025)
by: Wu, Kui, et al.
Published: (2025)
MeMix: Writing Less, Remembering More for Streaming 3D Reconstruction
by: Dong, Jiacheng, et al.
Published: (2026)
by: Dong, Jiacheng, et al.
Published: (2026)
Seeing More with Less: Video Capsule Endoscopy with Multi-Task Learning
by: Werner, Julia, et al.
Published: (2025)
by: Werner, Julia, et al.
Published: (2025)
Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective
by: Yue, Zihao, et al.
Published: (2024)
by: Yue, Zihao, et al.
Published: (2024)
Leveraging Visual Tokens for Extended Text Contexts in Multi-Modal Learning
by: Wang, Alex Jinpeng, et al.
Published: (2024)
by: Wang, Alex Jinpeng, et al.
Published: (2024)
LIME: Less Is More for MLLM Evaluation
by: Zhu, King, et al.
Published: (2024)
by: Zhu, King, et al.
Published: (2024)
From Channel Bias to Feature Redundancy: Uncovering the "Less is More" Principle in Few-Shot Learning
by: Zhang, Ji, et al.
Published: (2023)
by: Zhang, Ji, et al.
Published: (2023)
Similar Items
-
Decoupled Spatio-Temporal Consistency Learning for Self-Supervised Tracking
by: Zheng, Yaozong, et al.
Published: (2025) -
Robust Tracking via Mamba-based Context-aware Token Learning
by: Xie, Jinxia, et al.
Published: (2024) -
Learning to Track Instance from Single Nature Language Description
by: Zheng, Yaozong, et al.
Published: (2026) -
Towards Universal Modal Tracking with Online Dense Temporal Token Learning
by: Zheng, Yaozong, et al.
Published: (2025) -
MambaLCT: Boosting Tracking via Long-term Context State Space Model
by: Li, Xiaohai, et al.
Published: (2024)