Tri-Modal Fusion Transformers for UAV-based Object Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Iaboni, Craig, Abichandani, Pramod |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Event-based Spiking Neural Networks for Object Detection: A Review of Datasets, Architectures, Learning Rules, and Implementation
di: Iaboni, Craig, et al.
Pubblicazione: (2024)
di: Iaboni, Craig, et al.
Pubblicazione: (2024)
NU-AIR -- A Neuromorphic Urban Aerial Dataset for Detection and Localization of Pedestrians and Vehicles
di: Iaboni, Craig, et al.
Pubblicazione: (2023)
di: Iaboni, Craig, et al.
Pubblicazione: (2023)
TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR
di: Dharejo, Fayaz Ali, et al.
Pubblicazione: (2026)
di: Dharejo, Fayaz Ali, et al.
Pubblicazione: (2026)
Efficient Feature Fusion for UAV Object Detection
di: Wang, Xudong, et al.
Pubblicazione: (2025)
di: Wang, Xudong, et al.
Pubblicazione: (2025)
Light-Weight Cross-Modal Enhancement Method with Benchmark Construction for UAV-based Open-Vocabulary Object Detection
di: Weng, Zhenhai, et al.
Pubblicazione: (2025)
di: Weng, Zhenhai, et al.
Pubblicazione: (2025)
SwiTrack: Tri-State Switch for Cross-Modal Object Tracking
di: Xu, Boyue, et al.
Pubblicazione: (2025)
di: Xu, Boyue, et al.
Pubblicazione: (2025)
DPFT: Dual Perspective Fusion Transformer for Camera-Radar-based Object Detection
di: Fent, Felix, et al.
Pubblicazione: (2024)
di: Fent, Felix, et al.
Pubblicazione: (2024)
Progressive Multi-Modal Fusion for Robust 3D Object Detection
di: Mohan, Rohit, et al.
Pubblicazione: (2024)
di: Mohan, Rohit, et al.
Pubblicazione: (2024)
Adaptive Image Zoom-in with Bounding Box Transformation for UAV Object Detection
di: Wang, Tao, et al.
Pubblicazione: (2026)
di: Wang, Tao, et al.
Pubblicazione: (2026)
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism
di: Agarwal, Lakshita, et al.
Pubblicazione: (2025)
di: Agarwal, Lakshita, et al.
Pubblicazione: (2025)
HiddenObject: Modality-Agnostic Fusion for Multimodal Hidden Object Detection
di: Song, Harris, et al.
Pubblicazione: (2025)
di: Song, Harris, et al.
Pubblicazione: (2025)
Cross-modal Offset-guided Dynamic Alignment and Fusion for Weakly Aligned UAV Object Detection
di: Zongzhen, Liu, et al.
Pubblicazione: (2025)
di: Zongzhen, Liu, et al.
Pubblicazione: (2025)
GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection
di: Mia, Md Sohag, et al.
Pubblicazione: (2025)
di: Mia, Md Sohag, et al.
Pubblicazione: (2025)
Fusion Meets Diverse Conditions: A High-diversity Benchmark and Baseline for UAV-based Multimodal Object Detection with Condition Cues
di: Chen, Chen, et al.
Pubblicazione: (2025)
di: Chen, Chen, et al.
Pubblicazione: (2025)
RT-DETR++ for UAV Object Detection
di: Shufang, Yuan
Pubblicazione: (2025)
di: Shufang, Yuan
Pubblicazione: (2025)
FastHMR: Accelerating Human Mesh Recovery via Token and Layer Merging with Diffusion Decoding
di: Mehraban, Soroush, et al.
Pubblicazione: (2025)
di: Mehraban, Soroush, et al.
Pubblicazione: (2025)
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
di: Jia, Feiyang, et al.
Pubblicazione: (2025)
di: Jia, Feiyang, et al.
Pubblicazione: (2025)
Domain-invariant Progressive Knowledge Distillation for UAV-based Object Detection
di: Yao, Liang, et al.
Pubblicazione: (2024)
di: Yao, Liang, et al.
Pubblicazione: (2024)
Fusion is Not Enough: Single Modal Attacks on Fusion Models for 3D Object Detection
di: Cheng, Zhiyuan, et al.
Pubblicazione: (2023)
di: Cheng, Zhiyuan, et al.
Pubblicazione: (2023)
CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object Detection
di: Wu, Yuchen, et al.
Pubblicazione: (2026)
di: Wu, Yuchen, et al.
Pubblicazione: (2026)
RoboFusion: Towards Robust Multi-Modal 3D Object Detection via SAM
di: Song, Ziying, et al.
Pubblicazione: (2024)
di: Song, Ziying, et al.
Pubblicazione: (2024)
PoIFusion: Multi-Modal 3D Object Detection via Fusion at Points of Interest
di: Deng, Jiajun, et al.
Pubblicazione: (2024)
di: Deng, Jiajun, et al.
Pubblicazione: (2024)
EVT: Efficient View Transformation for Multi-Modal 3D Object Detection
di: Lee, Yongjin, et al.
Pubblicazione: (2024)
di: Lee, Yongjin, et al.
Pubblicazione: (2024)
VoxelNextFusion: A Simple, Unified and Effective Voxel Fusion Framework for Multi-Modal 3D Object Detection
di: Song, Ziying, et al.
Pubblicazione: (2024)
di: Song, Ziying, et al.
Pubblicazione: (2024)
FreDFT: Frequency Domain Fusion Transformer for Visible-Infrared Object Detection
di: Wu, Wencong, et al.
Pubblicazione: (2025)
di: Wu, Wencong, et al.
Pubblicazione: (2025)
Exploring Modality-Aware Fusion and Decoupled Temporal Propagation for Multi-Modal Object Tracking
di: Wang, Shilei, et al.
Pubblicazione: (2026)
di: Wang, Shilei, et al.
Pubblicazione: (2026)
Modality Prompts for Arbitrary Modality Salient Object Detection
di: Huang, Nianchang, et al.
Pubblicazione: (2024)
di: Huang, Nianchang, et al.
Pubblicazione: (2024)
OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on
di: Xu, Yuhao, et al.
Pubblicazione: (2024)
di: Xu, Yuhao, et al.
Pubblicazione: (2024)
DREB-Net: Dual-stream Restoration Embedding Blur-feature Fusion Network for High-mobility UAV Object Detection
di: Li, Qingpeng, et al.
Pubblicazione: (2024)
di: Li, Qingpeng, et al.
Pubblicazione: (2024)
Semantic-Guided Natural Language and Visual Fusion for Cross-Modal Interaction Based on Tiny Object Detection
di: Huang, Xian-Hong, et al.
Pubblicazione: (2025)
di: Huang, Xian-Hong, et al.
Pubblicazione: (2025)
Dual-Domain Homogeneous Fusion with Cross-Modal Mamba and Progressive Decoder for 3D Object Detection
di: Hu, Xuzhong, et al.
Pubblicazione: (2025)
di: Hu, Xuzhong, et al.
Pubblicazione: (2025)
SFFNet: Synergistic Feature Fusion Network With Dual-Domain Edge Enhancement for UAV Image Object Detection
di: Zhang, Wenfeng, et al.
Pubblicazione: (2026)
di: Zhang, Wenfeng, et al.
Pubblicazione: (2026)
BEVDilation: LiDAR-Centric Multi-Modal Fusion for 3D Object Detection
di: Zhang, Guowen, et al.
Pubblicazione: (2025)
di: Zhang, Guowen, et al.
Pubblicazione: (2025)
JCo-MVTON: Jointly Controllable Multi-Modal Diffusion Transformer for Mask-Free Virtual Try-on
di: Wang, Aowen, et al.
Pubblicazione: (2025)
di: Wang, Aowen, et al.
Pubblicazione: (2025)
MambaRefine-YOLO: A Dual-Modality Small Object Detector for UAV Imagery
di: Cao, Shuyu, et al.
Pubblicazione: (2025)
di: Cao, Shuyu, et al.
Pubblicazione: (2025)
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
di: Xing, Yinghui, et al.
Pubblicazione: (2023)
di: Xing, Yinghui, et al.
Pubblicazione: (2023)
WS-DETR: Robust Water Surface Object Detection through Vision-Radar Fusion with Detection Transformer
di: Yin, Huilin, et al.
Pubblicazione: (2025)
di: Yin, Huilin, et al.
Pubblicazione: (2025)
Cross-Level Sensor Fusion with Object Lists via Transformer for 3D Object Detection
di: Liu, Xiangzhong, et al.
Pubblicazione: (2025)
di: Liu, Xiangzhong, et al.
Pubblicazione: (2025)
Uncertainty-Encoded Multi-Modal Fusion for Robust Object Detection in Autonomous Driving
di: Lou, Yang, et al.
Pubblicazione: (2023)
di: Lou, Yang, et al.
Pubblicazione: (2023)
MambaSOD: Dual Mamba-Driven Cross-Modal Fusion Network for RGB-D Salient Object Detection
di: Zhan, Yue, et al.
Pubblicazione: (2024)
di: Zhan, Yue, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Event-based Spiking Neural Networks for Object Detection: A Review of Datasets, Architectures, Learning Rules, and Implementation
di: Iaboni, Craig, et al.
Pubblicazione: (2024) -
NU-AIR -- A Neuromorphic Urban Aerial Dataset for Detection and Localization of Pedestrians and Vehicles
di: Iaboni, Craig, et al.
Pubblicazione: (2023) -
TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR
di: Dharejo, Fayaz Ali, et al.
Pubblicazione: (2026) -
Efficient Feature Fusion for UAV Object Detection
di: Wang, Xudong, et al.
Pubblicazione: (2025) -
Light-Weight Cross-Modal Enhancement Method with Benchmark Construction for UAV-based Open-Vocabulary Object Detection
di: Weng, Zhenhai, et al.
Pubblicazione: (2025)