OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Xiaoyu, Dong, Jun, Cheng, Jintao, Fan, Rui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
You Sense Only Once Beneath: Ultra-Light Real-Time Underwater Object Detection
by: Dong, Jun, et al.
Published: (2025)
by: Dong, Jun, et al.
Published: (2025)
GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
MFDS-Net: Multi-Scale Feature Depth-Supervised Network for Remote Sensing Change Detection with Global Semantic and Detail Information
by: Huang, Zhenyang, et al.
Published: (2024)
by: Huang, Zhenyang, et al.
Published: (2024)
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
by: Li, Ke, et al.
Published: (2025)
by: Li, Ke, et al.
Published: (2025)
A Pseudo Global Fusion Paradigm-Based Cross-View Network for LiDAR-Based Place Recognition
by: Cheng, Jintao, et al.
Published: (2025)
by: Cheng, Jintao, et al.
Published: (2025)
LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation
by: Wang, Jun, et al.
Published: (2026)
by: Wang, Jun, et al.
Published: (2026)
Efficient Adaptation For Remote Sensing Visual Grounding
by: Moughnieh, Hasan, et al.
Published: (2025)
by: Moughnieh, Hasan, et al.
Published: (2025)
A Large-Scale Referring Remote Sensing Image Segmentation Dataset and Benchmark
by: Yang, Zhigang, et al.
Published: (2025)
by: Yang, Zhigang, et al.
Published: (2025)
YOLO-Ant: A Lightweight Detector via Depthwise Separable Convolutional and Large Kernel Design for Antenna Interference Source Detection
by: Tang, Xiaoyu, et al.
Published: (2024)
by: Tang, Xiaoyu, et al.
Published: (2024)
Think and Answer ME: Benchmarking and Exploring Multi-Entity Reasoning Grounding in Remote Sensing
by: Lyu, Shuchang, et al.
Published: (2026)
by: Lyu, Shuchang, et al.
Published: (2026)
Cross-Scale MAE: A Tale of Multi-Scale Exploitation in Remote Sensing
by: Tang, Maofeng, et al.
Published: (2024)
by: Tang, Maofeng, et al.
Published: (2024)
OptiCorNet: Optimizing Sequence-Based Context Correlation for Visual Place Recognition
by: Li, Zhenyu, et al.
Published: (2025)
by: Li, Zhenyu, et al.
Published: (2025)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
by: Zhang, Peirong, et al.
Published: (2025)
by: Zhang, Peirong, et al.
Published: (2025)
SOMA-1M: A Large-Scale SAR-Optical Multi-resolution Alignment Dataset for Multi-Task Remote Sensing
by: Wu, Peihao, et al.
Published: (2026)
by: Wu, Peihao, et al.
Published: (2026)
SATGround: A Spatially-Aware Approach for Visual Grounding in Remote Sensing
by: Toker, Aysim, et al.
Published: (2025)
by: Toker, Aysim, et al.
Published: (2025)
Source-Free Domain Adaptive Semantic Segmentation of Remote Sensing Images with Diffusion-Guided Label Enrichment
by: Liu, Wenjie, et al.
Published: (2025)
by: Liu, Wenjie, et al.
Published: (2025)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs
by: Zhou, Ziheng, et al.
Published: (2026)
by: Zhou, Ziheng, et al.
Published: (2026)
Scaling Remote Sensing Foundation Models: Data Domain Tradeoffs at the Peta-Scale
by: Wickrema, Charith, et al.
Published: (2025)
by: Wickrema, Charith, et al.
Published: (2025)
CrossEarth-SAR: A SAR-Centric and Billion-Scale Geospatial Foundation Model for Domain Generalizable Semantic Segmentation
by: Ye, Ziqi, et al.
Published: (2026)
by: Ye, Ziqi, et al.
Published: (2026)
SAR-TEXT: A Large-Scale SAR Image-Text Dataset Built with SAR-Narrator and A Progressive Learning Strategy for Downstream Tasks
by: He, Yiguo, et al.
Published: (2025)
by: He, Yiguo, et al.
Published: (2025)
MambaFlow: A Novel and Flow-guided State Space Model for Scene Flow Estimation
by: Luo, Jiehao, et al.
Published: (2025)
by: Luo, Jiehao, et al.
Published: (2025)
OpenEarthSensing: Large-Scale Fine-Grained Benchmark for Open-World Remote Sensing
by: Xiang, Xiang, et al.
Published: (2025)
by: Xiang, Xiang, et al.
Published: (2025)
CV-MOS: A Cross-View Model for Motion Segmentation
by: Tang, Xiaoyu, et al.
Published: (2024)
by: Tang, Xiaoyu, et al.
Published: (2024)
GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing
by: Shabbir, Akashah, et al.
Published: (2025)
by: Shabbir, Akashah, et al.
Published: (2025)
Real-Time AIoT for AAV Antenna Interference Detection via Edge-Cloud Collaboration
by: Dong, Jun, et al.
Published: (2024)
by: Dong, Jun, et al.
Published: (2024)
Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration
by: Li, Zhili, et al.
Published: (2026)
by: Li, Zhili, et al.
Published: (2026)
CrossEarth-Gate: Fisher-Guided Adaptive Tuning Engine for Efficient Adaptation of Cross-Domain Remote Sensing Semantic Segmentation
by: Cao, Shilei, et al.
Published: (2025)
by: Cao, Shilei, et al.
Published: (2025)
SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models
by: Zhong, Chen, et al.
Published: (2026)
by: Zhong, Chen, et al.
Published: (2026)
A Controlled Benchmark of Visual State-Space Backbones with Domain-Shift and Boundary Analysis for Remote-Sensing Segmentation
by: Wasalathilaka, Nichula, et al.
Published: (2026)
by: Wasalathilaka, Nichula, et al.
Published: (2026)
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time
by: Cheng, Jintao, et al.
Published: (2025)
by: Cheng, Jintao, et al.
Published: (2025)
ProVG: Progressive Visual Grounding via Language Decoupling for Remote Sensing Imagery
by: Li, Ke, et al.
Published: (2026)
by: Li, Ke, et al.
Published: (2026)
RSGround-R1: Rethinking Remote Sensing Visual Grounding through Spatial Reasoning
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
DehazeMamba: SAR-guided Optical Remote Sensing Image Dehazing with Adaptive State Space Model
by: Zhao, Zhicheng, et al.
Published: (2025)
by: Zhao, Zhicheng, et al.
Published: (2025)
HSANET: A Hybrid Self-Cross Attention Network For Remote Sensing Change Detection
by: Han, Chengxi, et al.
Published: (2025)
by: Han, Chengxi, et al.
Published: (2025)
DDLNet: Boosting Remote Sensing Change Detection with Dual-Domain Learning
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
MineNetCD: A Benchmark for Global Mining Change Detection on Remote Sensing Imagery
by: Yu, Weikang, et al.
Published: (2024)
by: Yu, Weikang, et al.
Published: (2024)
EarthMarker: A Visual Prompting Multi-modal Large Language Model for Remote Sensing
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
CHOICE: Benchmarking the Remote Sensing Capabilities of Large Vision-Language Models
by: An, Xiao, et al.
Published: (2024)
by: An, Xiao, et al.
Published: (2024)
Multisource Collaborative Domain Generalization for Cross-Scene Remote Sensing Image Classification
by: Han, Zhu, et al.
Published: (2024)
by: Han, Zhu, et al.
Published: (2024)
Similar Items
-
You Sense Only Once Beneath: Ultra-Light Real-Time Underwater Object Detection
by: Dong, Jun, et al.
Published: (2025) -
GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding
by: Zhou, Yue, et al.
Published: (2024) -
MFDS-Net: Multi-Scale Feature Depth-Supervised Network for Remote Sensing Change Detection with Global Semantic and Detail Information
by: Huang, Zhenyang, et al.
Published: (2024) -
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
by: Li, Ke, et al.
Published: (2025) -
A Pseudo Global Fusion Paradigm-Based Cross-View Network for LiDAR-Based Place Recognition
by: Cheng, Jintao, et al.
Published: (2025)