AFter: Attention-based Fusion Router for RGBT Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Andong, Wang, Wanyu, Li, Chenglong, Tang, Jin, Luo, Bin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking
by: Lu, Andong, et al.
Published: (2025)
by: Lu, Andong, et al.
Published: (2025)
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
by: Lu, Andong, et al.
Published: (2023)
by: Lu, Andong, et al.
Published: (2023)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
by: Sun, Dengdi, et al.
Published: (2024)
by: Sun, Dengdi, et al.
Published: (2024)
Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
Cross-modulated Attention Transformer for RGBT Tracking
by: Xiao, Yun, et al.
Published: (2024)
by: Xiao, Yun, et al.
Published: (2024)
Dynamic Disentangled Fusion Network for RGBT Tracking
by: Li, Chenglong, et al.
Published: (2024)
by: Li, Chenglong, et al.
Published: (2024)
Towards General Multimodal Visual Tracking
by: Lu, Andong, et al.
Published: (2025)
by: Lu, Andong, et al.
Published: (2025)
Nighttime Person Re-Identification via Collaborative Enhancement Network with Multi-domain Learning
by: Lu, Andong, et al.
Published: (2023)
by: Lu, Andong, et al.
Published: (2023)
Alignment-Free RGBT Salient Object Detection: Semantics-guided Asymmetric Correlation Network and A Unified Benchmark
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Graph-based Semantic Calibration Network for Unaligned UAV RGBT Image Semantic Segmentation and A Large-scale Benchmark
by: Fan, Fangqiang, et al.
Published: (2026)
by: Fan, Fangqiang, et al.
Published: (2026)
Decoupled Cross-Modal Alignment Network for Text-RGBT Person Retrieval and A High-Quality Benchmark
by: Deng, Yifei, et al.
Published: (2025)
by: Deng, Yifei, et al.
Published: (2025)
Towards Robust Optical-SAR Object Detection under Missing Modalities: A Dynamic Quality-Aware Fusion Framework
by: Zhao, Zhicheng, et al.
Published: (2025)
by: Zhao, Zhicheng, et al.
Published: (2025)
RAGTrack: Language-aware RGBT Tracking with Retrieval-Augmented Generation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Temporal Adaptive RGBT Tracking with Modality Prompt
by: Wang, Hongyu, et al.
Published: (2024)
by: Wang, Hongyu, et al.
Published: (2024)
CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
by: Wang, Qishun, et al.
Published: (2025)
by: Wang, Qishun, et al.
Published: (2025)
X Modality Assisting RGBT Object Tracking
by: Ding, Zhaisheng, et al.
Published: (2023)
by: Ding, Zhaisheng, et al.
Published: (2023)
Visible-Thermal Multiple Object Tracking: Large-scale Video Dataset and Progressive Fusion Approach
by: Zhu, Yabin, et al.
Published: (2024)
by: Zhu, Yabin, et al.
Published: (2024)
BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-language Models
by: Ye-Bin, Moon, et al.
Published: (2024)
by: Ye-Bin, Moon, et al.
Published: (2024)
Learning Adaptive Fusion Bank for Multi-modal Salient Object Detection
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Revisiting RGBT Tracking Benchmarks from the Perspective of Modality Validity: A New Benchmark, Problem, and Solution
by: Tang, Zhangyong, et al.
Published: (2024)
by: Tang, Zhangyong, et al.
Published: (2024)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
ICPL-ReID: Identity-Conditional Prompt Learning for Multi-Spectral Object Re-Identification
by: Li, Shihao, et al.
Published: (2025)
by: Li, Shihao, et al.
Published: (2025)
Adapting Segment Anything Model to Multi-modal Salient Object Detection with Semantic Feature Fusion Guidance
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Sparse-Dense Mixture of Experts Adapter for Multi-Modal Tracking
by: Zhu, Yabin, et al.
Published: (2026)
by: Zhu, Yabin, et al.
Published: (2026)
DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking
by: Hu, Guyue, et al.
Published: (2026)
by: Hu, Guyue, et al.
Published: (2026)
Mitigating the Impact of Prominent Position Shift in Drone-based RGBT Object Detection
by: Zhang, Yan, et al.
Published: (2025)
by: Zhang, Yan, et al.
Published: (2025)
COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection
by: Peng, Peiran, et al.
Published: (2025)
by: Peng, Peiran, et al.
Published: (2025)
Text-Guided Coarse-to-Fine Fusion Network for Robust Remote Sensing Visual Question Answering
by: Zhao, Zhicheng, et al.
Published: (2024)
by: Zhao, Zhicheng, et al.
Published: (2024)
M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision
by: Zhou, Kailai, et al.
Published: (2025)
by: Zhou, Kailai, et al.
Published: (2025)
Transformer-based RGB-T Tracking with Channel and Spatial Feature Fusion
by: Li, Yunfeng, et al.
Published: (2024)
by: Li, Yunfeng, et al.
Published: (2024)
AttnRouter: Per-Category Attention Routing for Training-Free Image Editing on MMDiT
by: Li, Guandong, et al.
Published: (2026)
by: Li, Guandong, et al.
Published: (2026)
CattleFace-RGBT: RGB-T Cattle Facial Landmark Benchmark
by: Coffman, Ethan, et al.
Published: (2024)
by: Coffman, Ethan, et al.
Published: (2024)
DCG ReID: Disentangling Collaboration and Guidance Fusion Representations for Multi-modal Vehicle Re-Identification
by: Zheng, Aihua, et al.
Published: (2026)
by: Zheng, Aihua, et al.
Published: (2026)
Vehicle-centric Perception via Multimodal Structured Pre-training
by: Wu, Wentao, et al.
Published: (2025)
by: Wu, Wentao, et al.
Published: (2025)
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework
by: Wan, Dahang, et al.
Published: (2025)
by: Wan, Dahang, et al.
Published: (2025)
Mixture of Scale Experts for Alignment-free RGBT Video Object Detection and A Unified Benchmark
by: Wang, Qishun, et al.
Published: (2024)
by: Wang, Qishun, et al.
Published: (2024)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
MTNet: Learning modality-aware representation with transformer for RGBT tracking
by: Hou, Ruichao, et al.
Published: (2025)
by: Hou, Ruichao, et al.
Published: (2025)
Similar Items
-
RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba
by: Lu, Andong, et al.
Published: (2024) -
Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking
by: Lu, Andong, et al.
Published: (2025) -
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
by: Lu, Andong, et al.
Published: (2023) -
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
by: Sun, Dengdi, et al.
Published: (2024) -
Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation
by: Lu, Andong, et al.
Published: (2024)