AlignFreeNet: Is Cross-Modal Pre-Alignment Necessary? An End-to-End Alignment-Free Lightweight Network for Visible-Infrared Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Dingkun, Zhang, Haote, Gu, Lipeng, Quan, Wuzhou, Wang, Fu Lee, Fan, Honghui, Tang, Jiali, Xie, Haoran, Zhang, Xiaoping, Wei, Mingqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lost in UNet: Improving Infrared Small Target Detection by Underappreciated Local Features
von: Quan, Wuzhou, et al.
Veröffentlicht: (2024)
von: Quan, Wuzhou, et al.
Veröffentlicht: (2024)
Equal is Not Always Fair: A New Perspective on Hyperspectral Representation Non-Uniformity
von: Quan, Wuzhou, et al.
Veröffentlicht: (2025)
von: Quan, Wuzhou, et al.
Veröffentlicht: (2025)
Perceive, Act and Correct: Confidence Is Not Enough for Hyperspectral Classification
von: Yang, Muzhou, et al.
Veröffentlicht: (2025)
von: Yang, Muzhou, et al.
Veröffentlicht: (2025)
DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment
von: Issam, Abderrahmane, et al.
Veröffentlicht: (2025)
von: Issam, Abderrahmane, et al.
Veröffentlicht: (2025)
RegTrack: Simplicity Beneath Complexity in Robust Multi-Modal 3D Multi-Object Tracking
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
CrossTracker: Robust Multi-modal 3D Multi-Object Tracking via Cross Correction
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
Beyond Hungarian: Match-Free Supervision for End-to-End Object Detection
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
Unsupervised Visible-Infrared ReID via Pseudo-label Correction and Modality-level Alignment
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
An End-to-End Model for Photo-Sharing Multi-modal Dialogue Generation
von: Guo, Peiming, et al.
Veröffentlicht: (2024)
von: Guo, Peiming, et al.
Veröffentlicht: (2024)
Align-DETR: Enhancing End-to-end Object Detection with Aligned Loss
von: Cai, Zhi, et al.
Veröffentlicht: (2023)
von: Cai, Zhi, et al.
Veröffentlicht: (2023)
Lean Learning Beyond Clouds: Efficient Discrepancy-Conditioned Optical-SAR Fusion for Semantic Segmentation
von: Meng, Chenxing, et al.
Veröffentlicht: (2026)
von: Meng, Chenxing, et al.
Veröffentlicht: (2026)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
YOLO26: An Analysis of NMS-Free End to End Framework for Real-Time Object Detection
von: Chakrabarty, Sudip
Veröffentlicht: (2026)
von: Chakrabarty, Sudip
Veröffentlicht: (2026)
AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving
von: Wu, Yanhao, et al.
Veröffentlicht: (2026)
von: Wu, Yanhao, et al.
Veröffentlicht: (2026)
Rethinking the Spatio-Temporal Alignment of End-to-End 3D Perception
von: Li, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Li, Xiaoyu, et al.
Veröffentlicht: (2025)
Is Free Self-Alignment Possible?
von: Adila, Dyah, et al.
Veröffentlicht: (2024)
von: Adila, Dyah, et al.
Veröffentlicht: (2024)
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
Hyperbolic Cycle Alignment for Infrared-Visible Image Fusion
von: Li, Timing, et al.
Veröffentlicht: (2025)
von: Li, Timing, et al.
Veröffentlicht: (2025)
Towards End-to-End Alignment of User Satisfaction via Questionnaire in Video Recommendation
von: Li, Na, et al.
Veröffentlicht: (2026)
von: Li, Na, et al.
Veröffentlicht: (2026)
Long-Form End-to-End Speech Translation via Latent Alignment Segmentation
von: Polák, Peter, et al.
Veröffentlicht: (2023)
von: Polák, Peter, et al.
Veröffentlicht: (2023)
MAIN: Mutual Alignment Is Necessary for instruction tuning
von: Yang, Fanyi, et al.
Veröffentlicht: (2025)
von: Yang, Fanyi, et al.
Veröffentlicht: (2025)
DualComp: End-to-End Learning of a Unified Dual-Modality Lossless Compressor
von: Zhao, Yan, et al.
Veröffentlicht: (2025)
von: Zhao, Yan, et al.
Veröffentlicht: (2025)
VIFNet: An End-to-end Visible-Infrared Fusion Network for Image Dehazing
von: Yu, Meng, et al.
Veröffentlicht: (2024)
von: Yu, Meng, et al.
Veröffentlicht: (2024)
Bridging the Gap: Multi-Level Cross-Modality Joint Alignment for Visible-Infrared Person Re-Identification
von: Liang, Tengfei, et al.
Veröffentlicht: (2023)
von: Liang, Tengfei, et al.
Veröffentlicht: (2023)
UHR-DETR: Efficient End-to-End Small Object Detection for Ultra-High-Resolution Remote Sensing Imagery
von: Li, Jingfang, et al.
Veröffentlicht: (2026)
von: Li, Jingfang, et al.
Veröffentlicht: (2026)
Adaptive End-to-End Transceiver Design for NextG Pilot-Free and CP-Free Wireless Systems
von: Cheng, Jiaming, et al.
Veröffentlicht: (2025)
von: Cheng, Jiaming, et al.
Veröffentlicht: (2025)
Safety-Aligned 3D Object Detection: Single-Vehicle, Cooperative, and End-to-End Perspectives
von: Liao, Brian Hsuan-Cheng, et al.
Veröffentlicht: (2026)
von: Liao, Brian Hsuan-Cheng, et al.
Veröffentlicht: (2026)
FD2-Net: Frequency-Driven Feature Decomposition Network for Infrared-Visible Object Detection
von: Li, Ke, et al.
Veröffentlicht: (2024)
von: Li, Ke, et al.
Veröffentlicht: (2024)
SeMe: Training-Free Language Model Merging via Semantic Alignment
von: Gu, Jian, et al.
Veröffentlicht: (2025)
von: Gu, Jian, et al.
Veröffentlicht: (2025)
Diverse Semantics-Guided Feature Alignment and Decoupling for Visible-Infrared Person Re-Identification
von: Dong, Neng, et al.
Veröffentlicht: (2025)
von: Dong, Neng, et al.
Veröffentlicht: (2025)
PIT: A Dynamic Personalized Item Tokenizer for End-to-End Generative Recommendation
von: Wang, Huanjie, et al.
Veröffentlicht: (2026)
von: Wang, Huanjie, et al.
Veröffentlicht: (2026)
An End-to-End, Segmentation-Free, Arabic Handwritten Recognition Model on KHATT
von: Aabed, Sondos, et al.
Veröffentlicht: (2024)
von: Aabed, Sondos, et al.
Veröffentlicht: (2024)
DiffVLA++: Bridging Cognitive Reasoning and End-to-End Driving through Metric-Guided Alignment
von: Gao, Yu, et al.
Veröffentlicht: (2025)
von: Gao, Yu, et al.
Veröffentlicht: (2025)
Autoregressive End-to-End Planning with Time-Invariant Spatial Alignment and Multi-Objective Policy Refinement
von: Zhao, Jianbo, et al.
Veröffentlicht: (2025)
von: Zhao, Jianbo, et al.
Veröffentlicht: (2025)
Anatomy of the Modality Gap: Dissecting the Internal States of End-to-End Speech LLMs
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2026)
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2026)
Aligning the True Semantics: Constrained Decoupling and Distribution Sampling for Cross-Modal Alignment
von: Ma, Xiang, et al.
Veröffentlicht: (2026)
von: Ma, Xiang, et al.
Veröffentlicht: (2026)
PaCo-FR: Patch-Pixel Aligned End-to-End Codebook Learning for Facial Representation Pre-training
von: Xie, Yin, et al.
Veröffentlicht: (2025)
von: Xie, Yin, et al.
Veröffentlicht: (2025)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
FLY-TTS: Fast, Lightweight and High-Quality End-to-End Text-to-Speech Synthesis
von: Guo, Yinlin, et al.
Veröffentlicht: (2024)
von: Guo, Yinlin, et al.
Veröffentlicht: (2024)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lost in UNet: Improving Infrared Small Target Detection by Underappreciated Local Features
von: Quan, Wuzhou, et al.
Veröffentlicht: (2024) -
Equal is Not Always Fair: A New Perspective on Hyperspectral Representation Non-Uniformity
von: Quan, Wuzhou, et al.
Veröffentlicht: (2025) -
Perceive, Act and Correct: Confidence Is Not Enough for Hyperspectral Classification
von: Yang, Muzhou, et al.
Veröffentlicht: (2025) -
DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment
von: Issam, Abderrahmane, et al.
Veröffentlicht: (2025) -
RegTrack: Simplicity Beneath Complexity in Robust Multi-Modal 3D Multi-Object Tracking
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)