Dual-Stream Attention with Multi-Modal Queries for Object Detection in Transportation Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Anwar, Noreen, Bilodeau, Guillaume-Alexandre, Bouachir, Wassim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STF: Spatio-Temporal Fusion Module for Improving Video Object Detection
by: Anwar, Noreen, et al.
Published: (2024)
by: Anwar, Noreen, et al.
Published: (2024)
ReL-SAR: Representation Learning for Skeleton Action Recognition with Convolutional Transformers and BYOL
by: Naimi, Safwen, et al.
Published: (2024)
by: Naimi, Safwen, et al.
Published: (2024)
InceptoFormer: A Multi-Signal Neural Framework for Parkinson's Disease Severity Evaluation from Gait
by: Naimi, Safwen, et al.
Published: (2025)
by: Naimi, Safwen, et al.
Published: (2025)
Detection of Micromobility Vehicles in Urban Traffic Videos
by: Sabri, Khalil, et al.
Published: (2024)
by: Sabri, Khalil, et al.
Published: (2024)
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026)
by: Naimi, Safwen, et al.
Published: (2026)
Learning Data Association for Multi-Object Tracking using Only Coordinates
by: Miah, Mehdi, et al.
Published: (2024)
by: Miah, Mehdi, et al.
Published: (2024)
STDiff: Spatio-temporal Diffusion for Continuous Stochastic Video Prediction
by: Ye, Xi, et al.
Published: (2023)
by: Ye, Xi, et al.
Published: (2023)
How good are deep learning methods for automated road safety analysis using video data? An experimental study
by: Liu, Qingwu, et al.
Published: (2025)
by: Liu, Qingwu, et al.
Published: (2025)
CenterDisks: Real-time instance segmentation with disk covering
by: Litto, Katia Jodogne-Del, et al.
Published: (2024)
by: Litto, Katia Jodogne-Del, et al.
Published: (2024)
PMMA: The Polytechnique Montreal Mobility Aids Dataset
by: Liu, Qingwu, et al.
Published: (2026)
by: Liu, Qingwu, et al.
Published: (2026)
Detection of Autonomous Shuttles in Urban Traffic Images Using Adaptive Residual Context
by: Younes, Mohamed Aziz, et al.
Published: (2026)
by: Younes, Mohamed Aziz, et al.
Published: (2026)
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
by: Jia, Feiyang, et al.
Published: (2025)
by: Jia, Feiyang, et al.
Published: (2025)
Streaming Detection of Queried Event Start
by: Eyzaguirre, Cristobal, et al.
Published: (2024)
by: Eyzaguirre, Cristobal, et al.
Published: (2024)
Improving SAM for Camouflaged Object Detection via Dual Stream Adapters
by: Liu, Jiaming, et al.
Published: (2025)
by: Liu, Jiaming, et al.
Published: (2025)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
DualGazeNet: A Biologically Inspired Dual-Gaze Query Network for Salient Object Detection
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection
by: Chen, Yishuo, et al.
Published: (2024)
by: Chen, Yishuo, et al.
Published: (2024)
DS-Det: Single-Query Paradigm and Attention Disentangled Learning for Flexible Object Detection
by: Cao, Guiping, et al.
Published: (2025)
by: Cao, Guiping, et al.
Published: (2025)
Automated Detection of Mutual Gaze and Joint Attention in Dual-Camera Settings via Dual-Stream Transformers
by: Kosmydel, Jakub, et al.
Published: (2026)
by: Kosmydel, Jakub, et al.
Published: (2026)
Referring Video Object Segmentation with Cross-Modality Proxy Queries
by: Sun, Baoli, et al.
Published: (2025)
by: Sun, Baoli, et al.
Published: (2025)
StreamMOS: Streaming Moving Object Segmentation with Multi-View Perception and Dual-Span Memory
by: Li, Zhiheng, et al.
Published: (2024)
by: Li, Zhiheng, et al.
Published: (2024)
StreamLTS: Query-based Temporal-Spatial LiDAR Fusion for Cooperative Object Detection
by: Yuan, Yunshuang, et al.
Published: (2024)
by: Yuan, Yunshuang, et al.
Published: (2024)
Multi-Modal Guided Multi-Source Domain Adaptation for Object Detection
by: Lee, Sangin, et al.
Published: (2026)
by: Lee, Sangin, et al.
Published: (2026)
M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention
by: Yang, Xiaofan, et al.
Published: (2026)
by: Yang, Xiaofan, et al.
Published: (2026)
Designing Object Detection Models for TinyML: Foundations, Comparative Analysis, Challenges, and Emerging Solutions
by: Zeinaty, Christophe EL, et al.
Published: (2025)
by: Zeinaty, Christophe EL, et al.
Published: (2025)
Rethinking Multi-Modal Object Detection from the Perspective of Mono-Modality Feature Learning
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
Online Episodic Memory Visual Query Localization with Egocentric Streaming Object Memory
by: Manigrasso, Zaira, et al.
Published: (2024)
by: Manigrasso, Zaira, et al.
Published: (2024)
Modality-Decoupled RGB-Thermal Object Detector via Query Fusion
by: Tian, Chao, et al.
Published: (2026)
by: Tian, Chao, et al.
Published: (2026)
BSDP: Brain-inspired Streaming Dual-level Perturbations for Online Open World Object Detection
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
Tracking by Detection and Query: An Efficient End-to-End Framework for Multi-Object Tracking
by: Jia, Shukun, et al.
Published: (2024)
by: Jia, Shukun, et al.
Published: (2024)
DGE-YOLO: Dual-Branch Gathering and Attention for Accurate UAV Object Detection
by: Lv, Kunwei, et al.
Published: (2025)
by: Lv, Kunwei, et al.
Published: (2025)
Modality Prompts for Arbitrary Modality Salient Object Detection
by: Huang, Nianchang, et al.
Published: (2024)
by: Huang, Nianchang, et al.
Published: (2024)
Dual-Stream Spectral Decoupling Distillation for Remote Sensing Object Detection
by: Gao, Xiangyi, et al.
Published: (2025)
by: Gao, Xiangyi, et al.
Published: (2025)
Progressive Multi-Modal Fusion for Robust 3D Object Detection
by: Mohan, Rohit, et al.
Published: (2024)
by: Mohan, Rohit, et al.
Published: (2024)
MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection
by: Wang, Zitian, et al.
Published: (2024)
by: Wang, Zitian, et al.
Published: (2024)
ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop
by: Li, Shuangzhi, et al.
Published: (2026)
by: Li, Shuangzhi, et al.
Published: (2026)
Learning Multi-Modal Prototypes for Cross-Domain Few-Shot Object Detection
by: Wang, Wanqi, et al.
Published: (2026)
by: Wang, Wanqi, et al.
Published: (2026)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
EVT: Efficient View Transformation for Multi-Modal 3D Object Detection
by: Lee, Yongjin, et al.
Published: (2024)
by: Lee, Yongjin, et al.
Published: (2024)
DCMorph: Face Morphing via Dual-Stream Cross-Attention Diffusion
by: Chettaoui, Tahar, et al.
Published: (2026)
by: Chettaoui, Tahar, et al.
Published: (2026)
Similar Items
-
STF: Spatio-Temporal Fusion Module for Improving Video Object Detection
by: Anwar, Noreen, et al.
Published: (2024) -
ReL-SAR: Representation Learning for Skeleton Action Recognition with Convolutional Transformers and BYOL
by: Naimi, Safwen, et al.
Published: (2024) -
InceptoFormer: A Multi-Signal Neural Framework for Parkinson's Disease Severity Evaluation from Gait
by: Naimi, Safwen, et al.
Published: (2025) -
Detection of Micromobility Vehicles in Urban Traffic Videos
by: Sabri, Khalil, et al.
Published: (2024) -
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026)