Cross-modulated Attention Transformer for RGBT Tracking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiao, Yun, Zhao, Jiacong, Lu, Andong, Li, Chenglong, Lin, Yin, Yin, Bing, Liu, Cong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation
von: Lu, Andong, et al.
Veröffentlicht: (2024)
von: Lu, Andong, et al.
Veröffentlicht: (2024)
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
von: Lu, Andong, et al.
Veröffentlicht: (2023)
von: Lu, Andong, et al.
Veröffentlicht: (2023)
AFter: Attention-based Fusion Router for RGBT Tracking
von: Lu, Andong, et al.
Veröffentlicht: (2024)
von: Lu, Andong, et al.
Veröffentlicht: (2024)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
von: Sun, Dengdi, et al.
Veröffentlicht: (2024)
von: Sun, Dengdi, et al.
Veröffentlicht: (2024)
RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba
von: Lu, Andong, et al.
Veröffentlicht: (2024)
von: Lu, Andong, et al.
Veröffentlicht: (2024)
Dynamic Disentangled Fusion Network for RGBT Tracking
von: Li, Chenglong, et al.
Veröffentlicht: (2024)
von: Li, Chenglong, et al.
Veröffentlicht: (2024)
Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking
von: Lu, Andong, et al.
Veröffentlicht: (2025)
von: Lu, Andong, et al.
Veröffentlicht: (2025)
Temporal Adaptive RGBT Tracking with Modality Prompt
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
X Modality Assisting RGBT Object Tracking
von: Ding, Zhaisheng, et al.
Veröffentlicht: (2023)
von: Ding, Zhaisheng, et al.
Veröffentlicht: (2023)
Decoupled Cross-Modal Alignment Network for Text-RGBT Person Retrieval and A High-Quality Benchmark
von: Deng, Yifei, et al.
Veröffentlicht: (2025)
von: Deng, Yifei, et al.
Veröffentlicht: (2025)
RAGTrack: Language-aware RGBT Tracking with Retrieval-Augmented Generation
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
1DFormer: a Transformer Architecture Learning 1D Landmark Representations for Facial Landmark Tracking
von: Yin, Shi, et al.
Veröffentlicht: (2023)
von: Yin, Shi, et al.
Veröffentlicht: (2023)
CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
von: Wang, Qishun, et al.
Veröffentlicht: (2025)
von: Wang, Qishun, et al.
Veröffentlicht: (2025)
Alignment-Free RGBT Salient Object Detection: Semantics-guided Asymmetric Correlation Network and A Unified Benchmark
von: Wang, Kunpeng, et al.
Veröffentlicht: (2024)
von: Wang, Kunpeng, et al.
Veröffentlicht: (2024)
Graph-based Semantic Calibration Network for Unaligned UAV RGBT Image Semantic Segmentation and A Large-scale Benchmark
von: Fan, Fangqiang, et al.
Veröffentlicht: (2026)
von: Fan, Fangqiang, et al.
Veröffentlicht: (2026)
Towards General Multimodal Visual Tracking
von: Lu, Andong, et al.
Veröffentlicht: (2025)
von: Lu, Andong, et al.
Veröffentlicht: (2025)
Revisiting RGBT Tracking Benchmarks from the Perspective of Modality Validity: A New Benchmark, Problem, and Solution
von: Tang, Zhangyong, et al.
Veröffentlicht: (2024)
von: Tang, Zhangyong, et al.
Veröffentlicht: (2024)
Towards Robust Optical-SAR Object Detection under Missing Modalities: A Dynamic Quality-Aware Fusion Framework
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2025)
RGB-Sonar Tracking Benchmark and Spatial Cross-Attention Transformer Tracker
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
von: Zhao, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhao, Tianyi, et al.
Veröffentlicht: (2025)
Exploring Part-Informed Visual-Language Learning for Person Re-Identification
von: Lin, Yin, et al.
Veröffentlicht: (2023)
von: Lin, Yin, et al.
Veröffentlicht: (2023)
Nighttime Person Re-Identification via Collaborative Enhancement Network with Multi-domain Learning
von: Lu, Andong, et al.
Veröffentlicht: (2023)
von: Lu, Andong, et al.
Veröffentlicht: (2023)
CattleFace-RGBT: RGB-T Cattle Facial Landmark Benchmark
von: Coffman, Ethan, et al.
Veröffentlicht: (2024)
von: Coffman, Ethan, et al.
Veröffentlicht: (2024)
CRSOT: Cross-Resolution Object Tracking using Unaligned Frame and Event Cameras
von: Zhu, Yabin, et al.
Veröffentlicht: (2024)
von: Zhu, Yabin, et al.
Veröffentlicht: (2024)
COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection
von: Peng, Peiran, et al.
Veröffentlicht: (2025)
von: Peng, Peiran, et al.
Veröffentlicht: (2025)
Bridging the Scale Gap: Balanced Tiny and General Object Detection in Remote Sensing Imagery
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2025)
Spatial Hierarchy and Temporal Attention Guided Cross Masking for Self-supervised Skeleton-based Action Recognition
von: Yin, Xinpeng, et al.
Veröffentlicht: (2024)
von: Yin, Xinpeng, et al.
Veröffentlicht: (2024)
M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision
von: Zhou, Kailai, et al.
Veröffentlicht: (2025)
von: Zhou, Kailai, et al.
Veröffentlicht: (2025)
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework
von: Wan, Dahang, et al.
Veröffentlicht: (2025)
von: Wan, Dahang, et al.
Veröffentlicht: (2025)
Polyline Path Masked Attention for Vision Transformer
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
Rethinking Two-Stage Referring-by-Tracking in Referring Multi-Object Tracking: Make it Strong Again
von: Li, Weize, et al.
Veröffentlicht: (2025)
von: Li, Weize, et al.
Veröffentlicht: (2025)
MTNet: Learning modality-aware representation with transformer for RGBT tracking
von: Hou, Ruichao, et al.
Veröffentlicht: (2025)
von: Hou, Ruichao, et al.
Veröffentlicht: (2025)
Mitigating the Impact of Prominent Position Shift in Drone-based RGBT Object Detection
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking
von: Hu, Guyue, et al.
Veröffentlicht: (2026)
von: Hu, Guyue, et al.
Veröffentlicht: (2026)
ReGLA: Efficient Receptive-Field Modeling with Gated Linear Attention Network
von: Li, Junzhou, et al.
Veröffentlicht: (2026)
von: Li, Junzhou, et al.
Veröffentlicht: (2026)
UrbanFeel: A Comprehensive Benchmark for Temporal and Perceptual Understanding of City Scenes through Human Perspective
von: He, Jun, et al.
Veröffentlicht: (2025)
von: He, Jun, et al.
Veröffentlicht: (2025)
Physics-Constrained Cross-Resolution Enhancement Network for Optics-Guided Thermal UAV Image Super-Resolution
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhao, Zhicheng, et al.
Veröffentlicht: (2026)
Binary-Gaussian: Compact and Progressive Representation for 3D Gaussian Segmentation
von: Yang, An, et al.
Veröffentlicht: (2025)
von: Yang, An, et al.
Veröffentlicht: (2025)
Vision Transformers with Hierarchical Attention
von: Liu, Yun, et al.
Veröffentlicht: (2021)
von: Liu, Yun, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation
von: Lu, Andong, et al.
Veröffentlicht: (2024) -
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
von: Lu, Andong, et al.
Veröffentlicht: (2023) -
AFter: Attention-based Fusion Router for RGBT Tracking
von: Lu, Andong, et al.
Veröffentlicht: (2024) -
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
von: Sun, Dengdi, et al.
Veröffentlicht: (2024) -
RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba
von: Lu, Andong, et al.
Veröffentlicht: (2024)