Semantic-Aware Ship Detection with Vision-Language Integration
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jiahao, Pan, Jiancheng, Sun, Yuze, Huang, Xiaomeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhance Then Search: An Augmentation-Search Strategy with Foundation Models for Cross-Domain Few-Shot Object Detection
by: Pan, Jiancheng, et al.
Published: (2025)
by: Pan, Jiancheng, et al.
Published: (2025)
EarthSynth: Generating Informative Earth Observation with Diffusion Models
by: Pan, Jiancheng, et al.
Published: (2025)
by: Pan, Jiancheng, et al.
Published: (2025)
Asymmetric Visual Semantic Embedding Framework for Efficient Vision-Language Alignment
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community
by: Pan, Jiancheng, et al.
Published: (2024)
by: Pan, Jiancheng, et al.
Published: (2024)
Forgetting-Resistant and Lesion-Aware Source-Free Domain Adaptive Fundus Image Analysis with Vision-Language Model
by: Huai, Zheang, et al.
Published: (2026)
by: Huai, Zheang, et al.
Published: (2026)
SAViL-Det: Semantic-Aware Vision-Language Model for Multi-Script Text Detection
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
SMPISD-MTPNet: Scene Semantic Prior-Assisted Infrared Ship Detection Using Multi-Task Perception Networks
by: Hu, Chen, et al.
Published: (2024)
by: Hu, Chen, et al.
Published: (2024)
Effective and Efficient Adversarial Detection for Vision-Language Models via A Single Vector
by: Huang, Youcheng, et al.
Published: (2024)
by: Huang, Youcheng, et al.
Published: (2024)
Beyond Semantics: Rediscovering Spatial Awareness in Vision-Language Models
by: Qi, Jianing, et al.
Published: (2025)
by: Qi, Jianing, et al.
Published: (2025)
GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations
by: Li, Zesheng, et al.
Published: (2026)
by: Li, Zesheng, et al.
Published: (2026)
Unleashing Vision-Language Semantics for Deepfake Video Detection
by: Zhu, Jiawen, et al.
Published: (2026)
by: Zhu, Jiawen, et al.
Published: (2026)
Automatic Detection of Dark Ship-to-Ship Transfers using Deep Learning and Satellite Imagery
by: Ballinger, Ollie
Published: (2024)
by: Ballinger, Ollie
Published: (2024)
Toward Semantic-Agnostic and Shape-Aware Vision-Language Segmentation Models
by: Seutin, Corentin, et al.
Published: (2026)
by: Seutin, Corentin, et al.
Published: (2026)
O2Former:Direction-Aware and Multi-Scale Query Enhancement for SAR Ship Instance Segmentation
by: Gao, F., et al.
Published: (2025)
by: Gao, F., et al.
Published: (2025)
Aligning Information Capacity Between Vision and Language via Dense-to-Sparse Feature Distillation for Image-Text Matching
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models
by: Peng, Xiaomeng, et al.
Published: (2026)
by: Peng, Xiaomeng, et al.
Published: (2026)
Exploring Interactive Semantic Alignment for Efficient HOI Detection with Vision-language Model
by: Dong, Jihao, et al.
Published: (2024)
by: Dong, Jihao, et al.
Published: (2024)
SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification
by: Chen, Furui, et al.
Published: (2026)
by: Chen, Furui, et al.
Published: (2026)
Denoising-Enhanced YOLO for Robust SAR Ship Detection
by: Zhao, Xiaojing, et al.
Published: (2026)
by: Zhao, Xiaojing, et al.
Published: (2026)
R-Sparse R-CNN: SAR Ship Detection Based on Background-Aware Sparse Learnable Proposals
by: Kamirul, Kamirul, et al.
Published: (2025)
by: Kamirul, Kamirul, et al.
Published: (2025)
UAU-Net: Uncertainty-aware Representation Learning and Evidential Classification for Facial Action Unit Detection
by: Li, Yuze, et al.
Published: (2026)
by: Li, Yuze, et al.
Published: (2026)
Adaptive Augmentation-Aware Latent Learning for Robust LiDAR Semantic Segmentation
by: Li, Wangkai, et al.
Published: (2026)
by: Li, Wangkai, et al.
Published: (2026)
Distribution-Aware Calibration for Object Detection with Noisy Bounding Boxes
by: Zhou, Donghao, et al.
Published: (2023)
by: Zhou, Donghao, et al.
Published: (2023)
Enabling Generalized Zero-shot Learning Towards Unseen Domains by Intrinsic Learning from Redundant LLM Semantics
by: Yue, Jiaqi, et al.
Published: (2024)
by: Yue, Jiaqi, et al.
Published: (2024)
PriorCLIP: Visual Prior Guided Vision-Language Model for Remote Sensing Image-Text Retrieval
by: Pan, Jiancheng, et al.
Published: (2024)
by: Pan, Jiancheng, et al.
Published: (2024)
Rewrite Caption Semantics: Bridging Semantic Gaps for Language-Supervised Semantic Segmentation
by: Xing, Yun, et al.
Published: (2023)
by: Xing, Yun, et al.
Published: (2023)
Lightweight SAR Ship Detection via Contrastive Distillation
by: Devasundaram, Surendar, et al.
Published: (2026)
by: Devasundaram, Surendar, et al.
Published: (2026)
S&D Messenger: Exchanging Semantic and Domain Knowledge for Generic Semi-Supervised Medical Image Segmentation
by: Zhang, Qixiang, et al.
Published: (2024)
by: Zhang, Qixiang, et al.
Published: (2024)
Vision-Language Models as Differentiable Semantic and Spatial Rewards for Text-to-3D Generation
by: Bai, Weimin, et al.
Published: (2025)
by: Bai, Weimin, et al.
Published: (2025)
Convolutional Feature Enhancement and Attention Fusion BiFPN for Ship Detection in SAR Images
by: Meng, Liangjie, et al.
Published: (2025)
by: Meng, Liangjie, et al.
Published: (2025)
DuSSS: Dual Semantic Similarity-Supervised Vision-Language Model for Semi-Supervised Medical Image Segmentation
by: Pan, Qingtao, et al.
Published: (2024)
by: Pan, Qingtao, et al.
Published: (2024)
ViTCoP: Accelerating Large Vision-Language Models via Visual and Textual Semantic Collaborative Pruning
by: Luo, Wen, et al.
Published: (2026)
by: Luo, Wen, et al.
Published: (2026)
An Empirical Study of Parameter Efficient Fine-tuning on Vision-Language Pre-train Model
by: Tian, Yuxin, et al.
Published: (2024)
by: Tian, Yuxin, et al.
Published: (2024)
Adversarial Patch Attack for Ship Detection via Localized Augmentation
by: Liu, Chun, et al.
Published: (2025)
by: Liu, Chun, et al.
Published: (2025)
Video Anomaly Detection with Semantics-Aware Information Bottleneck
by: Li, Juntong, et al.
Published: (2025)
by: Li, Juntong, et al.
Published: (2025)
Vision Transformers with Natural Language Semantics
by: Kim, Young Kyung, et al.
Published: (2024)
by: Kim, Young Kyung, et al.
Published: (2024)
Let Language Constrain Geometry: Vision-Language Models as Semantic and Spatial Critics for 3D Generation
by: Bai, Weimin, et al.
Published: (2025)
by: Bai, Weimin, et al.
Published: (2025)
UniVRSE: Unified Vision-conditioned Response Semantic Entropy for Hallucination Detection in Medical Vision-Language Models
by: Liao, Zehui, et al.
Published: (2025)
by: Liao, Zehui, et al.
Published: (2025)
GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation
by: Yang, Jiahao, et al.
Published: (2026)
by: Yang, Jiahao, et al.
Published: (2026)
Probabilistic Prototype Calibration of Vision-Language Models for Generalized Few-shot Semantic Segmentation
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Similar Items
-
Enhance Then Search: An Augmentation-Search Strategy with Foundation Models for Cross-Domain Few-Shot Object Detection
by: Pan, Jiancheng, et al.
Published: (2025) -
EarthSynth: Generating Informative Earth Observation with Diffusion Models
by: Pan, Jiancheng, et al.
Published: (2025) -
Asymmetric Visual Semantic Embedding Framework for Efficient Vision-Language Alignment
by: Liu, Yang, et al.
Published: (2025) -
Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community
by: Pan, Jiancheng, et al.
Published: (2024) -
Forgetting-Resistant and Lesion-Aware Source-Free Domain Adaptive Fundus Image Analysis with Vision-Language Model
by: Huai, Zheang, et al.
Published: (2026)