SeaDATE: Remedy Dual-Attention Transformer with Semantic Alignment via Contrast Learning for Multimodal Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Shuhan, Li, Yunsong, Xie, Weiying, Zhang, Jiaqing, Tian, Jiayuan, Yang, Danian, Lei, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SwiMDiff: Scene-wide Matching Contrastive Learning with Diffusion Constraint for Remote Sensing Image
by: Tian, Jiayuan, et al.
Published: (2024)
by: Tian, Jiayuan, et al.
Published: (2024)
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
FoRA: Low-Rank Adaptation Model beyond Multimodal Siamese Network
by: Xie, Weiying, et al.
Published: (2024)
by: Xie, Weiying, et al.
Published: (2024)
Distribution-aware Interactive Attention Network and Large-scale Cloud Recognition Benchmark on FY-4A Satellite Image
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Multi-scale direction-aware SAR object detection network via global information fusion
by: Cao, Mingxiang, et al.
Published: (2023)
by: Cao, Mingxiang, et al.
Published: (2023)
E2E-MFD: Towards End-to-End Synchronous Multimodal Fusion Detection
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Hyperspectral Anomaly Detection with Self-Supervised Anomaly Prior
by: Liu, Yidan, et al.
Published: (2024)
by: Liu, Yidan, et al.
Published: (2024)
Exploring Hyperspectral Anomaly Detection with Human Vision: A Small Target Aware Detector
by: Ma, Jitao, et al.
Published: (2024)
by: Ma, Jitao, et al.
Published: (2024)
M$^3$amba: CLIP-driven Mamba Model for Multi-modal Remote Sensing Classification
by: Cao, Mingxiang, et al.
Published: (2025)
by: Cao, Mingxiang, et al.
Published: (2025)
Domain Adaptation for Large-Vocabulary Object Detectors
by: Jiang, Kai, et al.
Published: (2024)
by: Jiang, Kai, et al.
Published: (2024)
Hyperspectral Mamba for Hyperspectral Object Tracking
by: Gao, Long, et al.
Published: (2025)
by: Gao, Long, et al.
Published: (2025)
BSDM: Background Suppression Diffusion Model for Hyperspectral Anomaly Detection
by: Ma, Jitao, et al.
Published: (2023)
by: Ma, Jitao, et al.
Published: (2023)
Physics Inspired Criterion for Pruning-Quantization Joint Learning
by: Xie, Weiying, et al.
Published: (2023)
by: Xie, Weiying, et al.
Published: (2023)
Hyperspectral Adapter for Object Tracking based on Hyperspectral Video
by: Gao, Long, et al.
Published: (2025)
by: Gao, Long, et al.
Published: (2025)
Spanning Training Progress: Temporal Dual-Depth Scoring (TDDS) for Enhanced Dataset Pruning
by: Zhang, Xin, et al.
Published: (2023)
by: Zhang, Xin, et al.
Published: (2023)
ELEY, Geoff y NIELD, Keith El futuro de la clase en la historia ¿Qué queda de lo social?, Valencia, PUV, 2010, 244 pp. - ISBN 978-84-370-7823-6
by: Danián López
Published: (2011)
by: Danián López
Published: (2011)
OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment
by: Liu, Qi, et al.
Published: (2025)
by: Liu, Qi, et al.
Published: (2025)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
DiffCLIP: Few-shot Language-driven Multimodal Classifier
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Reducing Spurious Correlation for Federated Domain Generalization
by: Ma, Shuran, et al.
Published: (2024)
by: Ma, Shuran, et al.
Published: (2024)
RS-DGC: Exploring Neighborhood Statistics for Dynamic Gradient Compression on Remote Sensing Image Interpretation
by: Xie, Weiying, et al.
Published: (2023)
by: Xie, Weiying, et al.
Published: (2023)
LASFNet: A Lightweight Attention-Guided Self-Modulation Feature Fusion Network for Multimodal Object Detection
by: Hao, Lei, et al.
Published: (2025)
by: Hao, Lei, et al.
Published: (2025)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
by: Chang, Kai-Po, et al.
Published: (2025)
by: Chang, Kai-Po, et al.
Published: (2025)
A Review of Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment
by: Qiang, Sunyuan, et al.
Published: (2024)
by: Qiang, Sunyuan, et al.
Published: (2024)
ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
FedFQ: Federated Learning with Fine-Grained Quantization
by: Li, Haowei, et al.
Published: (2024)
by: Li, Haowei, et al.
Published: (2024)
DA-BEV: Unsupervised Domain Adaptation for Bird's Eye View Perception
by: Jiang, Kai, et al.
Published: (2024)
by: Jiang, Kai, et al.
Published: (2024)
Universal Domain Adaptive Object Detection via Dual Probabilistic Alignment
by: Zheng, Yuanfan, et al.
Published: (2024)
by: Zheng, Yuanfan, et al.
Published: (2024)
Improving Multimodal Contrastive Learning of Sentence Embeddings with Object-Phrase Alignment
by: Zhao, Kaiyan, et al.
Published: (2025)
by: Zhao, Kaiyan, et al.
Published: (2025)
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
by: Xu, Kepeng, et al.
Published: (2024)
by: Xu, Kepeng, et al.
Published: (2024)
Representation Alignment Contrastive Regularization for Multi-Object Tracking
by: Liu, Zhonglin, et al.
Published: (2024)
by: Liu, Zhonglin, et al.
Published: (2024)
Automated Glaucoma Report Generation via Dual-Attention Semantic Parallel-LSTM and Multimodal Clinical Data Integration
by: Huang, Cheng, et al.
Published: (2025)
by: Huang, Cheng, et al.
Published: (2025)
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
by: Wang, Yuxiao, et al.
Published: (2024)
by: Wang, Yuxiao, et al.
Published: (2024)
An Enhanced Dual Transformer Contrastive Network for Multimodal Sentiment Analysis
by: Dao, Phuong Q., et al.
Published: (2025)
by: Dao, Phuong Q., et al.
Published: (2025)
THE DATE OF PUBLICATION OF ANTONS VERZEICHNISS DER CONCHYLIEN
by: Cernohorsky, Walter Oliver.
Published: (1978)
by: Cernohorsky, Walter Oliver.
Published: (1978)
THE EURO AREA ENLARGEMENT: THE TARGET DATE PROBLEM
by: Arūnas Dulkys
Published: (2009)
by: Arūnas Dulkys
Published: (2009)
Generative AI ‐Supported Student Video Creation in Communication Education: A Mixed‐Methods Study of Learning Motivation, Career Confidence and Creativity
by: Yunsong Wang, et al.
Published: (2026)
by: Yunsong Wang, et al.
Published: (2026)
Grapevine Disease Prediction Using Climate Variables from Multi-Sensor Remote Sensing Imagery via a Transformer Model
by: Zhao, Weiying, et al.
Published: (2024)
by: Zhao, Weiying, et al.
Published: (2024)
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
by: Li, Shanghao, et al.
Published: (2025)
by: Li, Shanghao, et al.
Published: (2025)
Similar Items
-
SwiMDiff: Scene-wide Matching Contrastive Learning with Diffusion Constraint for Remote Sensing Image
by: Tian, Jiayuan, et al.
Published: (2024) -
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification
by: Zhang, Jiaqing, et al.
Published: (2024) -
FoRA: Low-Rank Adaptation Model beyond Multimodal Siamese Network
by: Xie, Weiying, et al.
Published: (2024) -
Distribution-aware Interactive Attention Network and Large-scale Cloud Recognition Benchmark on FY-4A Satellite Image
by: Zhang, Jiaqing, et al.
Published: (2024) -
Multi-scale direction-aware SAR object detection network via global information fusion
by: Cao, Mingxiang, et al.
Published: (2023)