Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yingquan, Zhang, Pingping, Wang, Dong, Lu, Huchuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
von: Liu, Yangyang, et al.
Veröffentlicht: (2025)
von: Liu, Yangyang, et al.
Veröffentlicht: (2025)
What Makes You Unique? Attribute Prompt Composition for Object Re-Identification
von: Wang, Yingquan, et al.
Veröffentlicht: (2025)
von: Wang, Yingquan, et al.
Veröffentlicht: (2025)
Multi-Scale and Detail-Enhanced Segment Anything Model for Salient Object Detection
von: Gao, Shixuan, et al.
Veröffentlicht: (2024)
von: Gao, Shixuan, et al.
Veröffentlicht: (2024)
Fantastic Animals and Where to Find Them: Segment Any Marine Animal with Dual SAM
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
Unveiling Encoder-Free Vision-Language Models
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
ProFD: Prompt-Guided Feature Disentangling for Occluded Person Re-Identification
von: Cui, Can, et al.
Veröffentlicht: (2024)
von: Cui, Can, et al.
Veröffentlicht: (2024)
Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion
von: Zhu, Yixin, et al.
Veröffentlicht: (2026)
von: Zhu, Yixin, et al.
Veröffentlicht: (2026)
Efficient Token Compression for Vision Transformer with Spatial Information Preserved
von: Mao, Junzhu, et al.
Veröffentlicht: (2025)
von: Mao, Junzhu, et al.
Veröffentlicht: (2025)
Deep Boosting Learning: A Brand-new Cooperative Approach for Image-Text Matching
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
GSSF: Generalized Structural Sparse Function for Deep Cross-modal Metric Learning
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
AToken: A Unified Tokenizer for Vision
von: Lu, Jiasen, et al.
Veröffentlicht: (2025)
von: Lu, Jiasen, et al.
Veröffentlicht: (2025)
Towards Open-Vocabulary Remote Sensing Image Semantic Segmentation
von: Ye, Chengyang, et al.
Veröffentlicht: (2024)
von: Ye, Chengyang, et al.
Veröffentlicht: (2024)
SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
Adaptive Low Light Enhancement via Joint Global-Local Illumination Adjustment
von: Wang, Haodian, et al.
Veröffentlicht: (2025)
von: Wang, Haodian, et al.
Veröffentlicht: (2025)
On-the-Fly Object-aware Representative Point Selection in Point Cloud
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
Robust Duality Learning for Unsupervised Visible-Infrared Person Re-Identification
von: Li, Yongxiang, et al.
Veröffentlicht: (2025)
von: Li, Yongxiang, et al.
Veröffentlicht: (2025)
Hybrid Local-Global Context Learning for Neural Video Compression
von: Zhai, Yongqi, et al.
Veröffentlicht: (2024)
von: Zhai, Yongqi, et al.
Veröffentlicht: (2024)
Test-time adaptation for image compression with distribution regularization
von: Chen, Kecheng, et al.
Veröffentlicht: (2024)
von: Chen, Kecheng, et al.
Veröffentlicht: (2024)
Querying Autonomous Vehicle Point Clouds: Enhanced by 3D Object Counting with CounterNet
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
Parallel Vision Token Scheduling for Fast and Accurate Multimodal LMMs Inference
von: Zhan, Wengyi, et al.
Veröffentlicht: (2025)
von: Zhan, Wengyi, et al.
Veröffentlicht: (2025)
Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training
von: Baraldi, Lorenzo, et al.
Veröffentlicht: (2023)
von: Baraldi, Lorenzo, et al.
Veröffentlicht: (2023)
Serial Low-rank Adaptation of Vision Transformer
von: Zhong, Houqiang, et al.
Veröffentlicht: (2025)
von: Zhong, Houqiang, et al.
Veröffentlicht: (2025)
DT-UFC: Universal Large Model Feature Coding via Peaky-to-Balanced Distribution Transformation
von: Gao, Changsheng, et al.
Veröffentlicht: (2025)
von: Gao, Changsheng, et al.
Veröffentlicht: (2025)
SD-ReID: View-aware Stable Diffusion for Aerial-Ground Person Re-Identification
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
LATex: Leveraging Attribute-based Text Knowledge for Aerial-Ground Person Re-Identification
von: Zhang, Pingping, et al.
Veröffentlicht: (2025)
von: Zhang, Pingping, et al.
Veröffentlicht: (2025)
Local Neighborhood Features for 3D Classification
von: Sheshappanavar, Shivanand Venkanna, et al.
Veröffentlicht: (2022)
von: Sheshappanavar, Shivanand Venkanna, et al.
Veröffentlicht: (2022)
ROGLE: Robust Global-Local Alignment with Automated Region Supervision for Text-Based Person Search
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
DRFormer: A Dual-Regularized Bidirectional Transformer for Person Re-identification
von: Shu, Ying, et al.
Veröffentlicht: (2026)
von: Shu, Ying, et al.
Veröffentlicht: (2026)
Unity is Strength: Unifying Convolutional and Transformeral Features for Better Person Re-Identification
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
Palmprint De-Identification Using Diffusion Model for High-Quality and Diverse Synthesis
von: Yan, Licheng, et al.
Veröffentlicht: (2025)
von: Yan, Licheng, et al.
Veröffentlicht: (2025)
Dual Mutual Learning Network with Global-local Awareness for RGB-D Salient Object Detection
von: Yi, Kang, et al.
Veröffentlicht: (2025)
von: Yi, Kang, et al.
Veröffentlicht: (2025)
When Video Coding Meets Multimodal Large Language Models: A Unified Paradigm for Video Coding
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
SFFNet: Synergistic Feature Fusion Network With Dual-Domain Edge Enhancement for UAV Image Object Detection
von: Zhang, Wenfeng, et al.
Veröffentlicht: (2026)
von: Zhang, Wenfeng, et al.
Veröffentlicht: (2026)
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
CLIP-PCQA: Exploring Subjective-Aligned Vision-Language Modeling for Point Cloud Quality Assessment
von: Liu, Yating, et al.
Veröffentlicht: (2025)
von: Liu, Yating, et al.
Veröffentlicht: (2025)
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection
von: Guo, Junjie, et al.
Veröffentlicht: (2024)
von: Guo, Junjie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification
von: Zhang, Pingping, et al.
Veröffentlicht: (2024) -
IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification
von: Wang, Yuhao, et al.
Veröffentlicht: (2025) -
MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt
von: Wang, Yuhao, et al.
Veröffentlicht: (2024) -
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
von: Liu, Yangyang, et al.
Veröffentlicht: (2025) -
What Makes You Unique? Attribute Prompt Composition for Object Re-Identification
von: Wang, Yingquan, et al.
Veröffentlicht: (2025)