MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Xing, Yinghui, Yang, Shuo, Wang, Song, Zhang, Shizhou, Liang, Guoqiang, Zhang, Xiuwei, Zhang, Yanning |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
di: Yang, Shuo, et al.
Pubblicazione: (2025)
di: Yang, Shuo, et al.
Pubblicazione: (2025)
Frequency-Guided Spatial Adaptation for Camouflaged Object Detection
di: Zhang, Shizhou, et al.
Pubblicazione: (2024)
di: Zhang, Shizhou, et al.
Pubblicazione: (2024)
Attention Retention for Continual Learning with Vision Transformers
di: Lu, Yue, et al.
Pubblicazione: (2026)
di: Lu, Yue, et al.
Pubblicazione: (2026)
AdaSemiCD: An Adaptive Semi-Supervised Change Detection Method Based on Pseudo-Label Evaluation
di: Lingyan, Ran, et al.
Pubblicazione: (2024)
di: Lingyan, Ran, et al.
Pubblicazione: (2024)
CrossDiff: Exploring Self-Supervised Representation of Pansharpening via Cross-Predictive Diffusion Model
di: Xing, Yinghui, et al.
Pubblicazione: (2024)
di: Xing, Yinghui, et al.
Pubblicazione: (2024)
FreDFT: Frequency Domain Fusion Transformer for Visible-Infrared Object Detection
di: Wu, Wencong, et al.
Pubblicazione: (2025)
di: Wu, Wencong, et al.
Pubblicazione: (2025)
Knowing the Unknown: Interpretable Open-World Object Detection via Concept Decomposition Model
di: Lv, Xueqiang, et al.
Pubblicazione: (2026)
di: Lv, Xueqiang, et al.
Pubblicazione: (2026)
YOLO-IOD: Towards Real Time Incremental Object Detection
di: Zhang, Shizhou, et al.
Pubblicazione: (2025)
di: Zhang, Shizhou, et al.
Pubblicazione: (2025)
Strip-Fusion: Spatiotemporal Fusion for Multispectral Pedestrian Detection
di: Kanu-Asiegbu, Asiegbu Miracle, et al.
Pubblicazione: (2026)
di: Kanu-Asiegbu, Asiegbu Miracle, et al.
Pubblicazione: (2026)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
di: Zhang, Shizhou, et al.
Pubblicazione: (2021)
di: Zhang, Shizhou, et al.
Pubblicazione: (2021)
Cross-Platform Video Person ReID: A New Benchmark Dataset and Adaptation Approach
di: Zhang, Shizhou, et al.
Pubblicazione: (2024)
di: Zhang, Shizhou, et al.
Pubblicazione: (2024)
DMAT: A Dynamic Mask-Aware Transformer for Human De-occlusion
di: Liang, Guoqiang, et al.
Pubblicazione: (2024)
di: Liang, Guoqiang, et al.
Pubblicazione: (2024)
Demystifying Catastrophic Forgetting in Two-Stage Incremental Object Detector
di: Wu, Qirui, et al.
Pubblicazione: (2025)
di: Wu, Qirui, et al.
Pubblicazione: (2025)
AMFD: Distillation via Adaptive Multimodal Fusion for Multispectral Pedestrian Detection
di: Chen, Zizhao, et al.
Pubblicazione: (2024)
di: Chen, Zizhao, et al.
Pubblicazione: (2024)
MM-DETR: An Efficient Multimodal Detection Transformer with Mamba-Driven Dual-Granularity Fusion and Frequency-Aware Modality Adapters
di: Han, Jianhong, et al.
Pubblicazione: (2025)
di: Han, Jianhong, et al.
Pubblicazione: (2025)
Dual-Modal Prompting for Sketch-Based Image Retrieval
di: Gao, Liying, et al.
Pubblicazione: (2024)
di: Gao, Liying, et al.
Pubblicazione: (2024)
Visual Prompt Tuning in Null Space for Continual Learning
di: Lu, Yue, et al.
Pubblicazione: (2024)
di: Lu, Yue, et al.
Pubblicazione: (2024)
Better Matching, Less Forgetting: A Quality-Guided Matcher for Transformer-based Incremental Object Detection
di: Wu, Qirui, et al.
Pubblicazione: (2026)
di: Wu, Qirui, et al.
Pubblicazione: (2026)
MSCoTDet: Language-driven Multi-modal Fusion for Improved Multispectral Pedestrian Detection
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
Multispectral Pedestrian Detection with Sparsely Annotated Label
di: Lee, Chan, et al.
Pubblicazione: (2025)
di: Lee, Chan, et al.
Pubblicazione: (2025)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
di: Zhang, Xue, et al.
Pubblicazione: (2024)
di: Zhang, Xue, et al.
Pubblicazione: (2024)
DuGI-MAE: Improving Infrared Mask Autoencoders via Dual-Domain Guidance
di: Xing, Yinghui, et al.
Pubblicazione: (2025)
di: Xing, Yinghui, et al.
Pubblicazione: (2025)
MRC-DETR: An Adaptive Multi-Residual Coupled Transformer for Bare Board PCB Defect Detection
di: Cao, Jiangzhong, et al.
Pubblicazione: (2025)
di: Cao, Jiangzhong, et al.
Pubblicazione: (2025)
TFDet: Target-Aware Fusion for RGB-T Pedestrian Detection
di: Zhang, Xue, et al.
Pubblicazione: (2023)
di: Zhang, Xue, et al.
Pubblicazione: (2023)
Adaptive Spatial Augmentation for Semi-supervised Semantic Segmentation
di: Ran, Lingyan, et al.
Pubblicazione: (2025)
di: Ran, Lingyan, et al.
Pubblicazione: (2025)
MS-DETR: Efficient DETR Training with Mixed Supervision
di: Zhao, Chuyang, et al.
Pubblicazione: (2024)
di: Zhao, Chuyang, et al.
Pubblicazione: (2024)
CGF-DETR: Cross-Gated Fusion DETR for Enhanced Pneumonia Detection in Chest X-rays
di: Wu, Yefeng, et al.
Pubblicazione: (2025)
di: Wu, Yefeng, et al.
Pubblicazione: (2025)
Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion
di: Zhang, Yinghui, et al.
Pubblicazione: (2025)
di: Zhang, Yinghui, et al.
Pubblicazione: (2025)
Multispectral Detection Transformer with Infrared-Centric Feature Fusion
di: Hwang, Seongmin, et al.
Pubblicazione: (2025)
di: Hwang, Seongmin, et al.
Pubblicazione: (2025)
Revisiting Misalignment in Multispectral Pedestrian Detection: A Language-Driven Approach for Cross-modal Alignment Fusion
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
OVLW-DETR: Open-Vocabulary Light-Weighted Detection Transformer
di: Wang, Yu, et al.
Pubblicazione: (2024)
di: Wang, Yu, et al.
Pubblicazione: (2024)
PedDet: Adaptive Spectral Optimization for Multimodal Pedestrian Detection
di: Zhao, Rui, et al.
Pubblicazione: (2025)
di: Zhao, Rui, et al.
Pubblicazione: (2025)
Multi-level Collaborative Distillation Meets Global Workspace Model: A Unified Framework for OCIL
di: Su, Shibin, et al.
Pubblicazione: (2025)
di: Su, Shibin, et al.
Pubblicazione: (2025)
WCCNet: Wavelet-context Cooperative Network for Efficient Multispectral Pedestrian Detection
di: Wang, Xingjian, et al.
Pubblicazione: (2023)
di: Wang, Xingjian, et al.
Pubblicazione: (2023)
Robust Pedestrian Detection with Uncertain Modality
di: Bie, Qian, et al.
Pubblicazione: (2026)
di: Bie, Qian, et al.
Pubblicazione: (2026)
LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection
di: Chen, Qiang, et al.
Pubblicazione: (2024)
di: Chen, Qiang, et al.
Pubblicazione: (2024)
Semi-Supervised Semantic Segmentation Based on Pseudo-Labels: A Survey
di: Ran, Lingyan, et al.
Pubblicazione: (2024)
di: Ran, Lingyan, et al.
Pubblicazione: (2024)
Causal Mode Multiplexer: A Novel Framework for Unbiased Multispectral Pedestrian Detection
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
di: Kim, Taeheon, et al.
Pubblicazione: (2024)
Salience DETR: Enhancing Detection Transformer with Hierarchical Salience Filtering Refinement
di: Hou, Xiuquan, et al.
Pubblicazione: (2024)
di: Hou, Xiuquan, et al.
Pubblicazione: (2024)
Mr. DETR++: Instructive Multi-Route Training for Detection Transformers with Mixture-of-Experts
di: Zhang, Chang-Bin, et al.
Pubblicazione: (2024)
di: Zhang, Chang-Bin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
di: Yang, Shuo, et al.
Pubblicazione: (2025) -
Frequency-Guided Spatial Adaptation for Camouflaged Object Detection
di: Zhang, Shizhou, et al.
Pubblicazione: (2024) -
Attention Retention for Continual Learning with Vision Transformers
di: Lu, Yue, et al.
Pubblicazione: (2026) -
AdaSemiCD: An Adaptive Semi-Supervised Change Detection Method Based on Pseudo-Label Evaluation
di: Lingyan, Ran, et al.
Pubblicazione: (2024) -
CrossDiff: Exploring Self-Supervised Representation of Pansharpening via Cross-Predictive Diffusion Model
di: Xing, Yinghui, et al.
Pubblicazione: (2024)