Large Language Models for Captioning and Retrieving Remote Sensing Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Silva, João Daniel, Magalhães, João, Tuia, Devis, Martins, Bruno |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Efficient and Effective Encoder Model for Vision and Language Tasks in the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2025)
von: Silva, João Daniel, et al.
Veröffentlicht: (2025)
Multilingual Vision-Language Pre-training for the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
Knowledge-aware Text-Image Retrieval for Remote Sensing Images
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
Multilingual Training-Free Remote Sensing Image Captioning
von: Rebelo, Carlos, et al.
Veröffentlicht: (2025)
von: Rebelo, Carlos, et al.
Veröffentlicht: (2025)
SMARTIES: Spectrum-Aware Multi-Sensor Auto-Encoder for Remote Sensing Images
von: Sumbul, Gencer, et al.
Veröffentlicht: (2025)
von: Sumbul, Gencer, et al.
Veröffentlicht: (2025)
Knowledge-aware Visual Question Generation for Remote Sensing Images
von: Li, Siran, et al.
Veröffentlicht: (2026)
von: Li, Siran, et al.
Veröffentlicht: (2026)
Questions beyond Pixels: Integrating Commonsense Knowledge in Visual Question Generation for Remote Sensing
von: Li, Siran, et al.
Veröffentlicht: (2026)
von: Li, Siran, et al.
Veröffentlicht: (2026)
Retrieval of Surface Solar Radiation through Implicit Albedo Recovery from Temporal Context
von: Frischholz, Yael, et al.
Veröffentlicht: (2025)
von: Frischholz, Yael, et al.
Veröffentlicht: (2025)
Cross-Modal Learning of Housing Quality in Amsterdam
von: Levering, Alex, et al.
Veröffentlicht: (2024)
von: Levering, Alex, et al.
Veröffentlicht: (2024)
POLO -- Point-based, multi-class animal detection
von: May, Giacomo, et al.
Veröffentlicht: (2024)
von: May, Giacomo, et al.
Veröffentlicht: (2024)
BTCChat: Advancing Remote Sensing Bi-temporal Change Captioning with Multimodal Large Language Model
von: Li, Yujie, et al.
Veröffentlicht: (2025)
von: Li, Yujie, et al.
Veröffentlicht: (2025)
Text-Only Training for Image Captioning with Retrieval Augmentation and Modality Gap Correction
von: Fonseca, Rui, et al.
Veröffentlicht: (2025)
von: Fonseca, Rui, et al.
Veröffentlicht: (2025)
EcoWikiRS: Learning Ecological Representation of Satellite Images from Weak Supervision with Species Observations and Wikipedia
von: Zermatten, Valerie, et al.
Veröffentlicht: (2025)
von: Zermatten, Valerie, et al.
Veröffentlicht: (2025)
Evaluating Remote Sensing Image Captions Beyond Metric Biases
von: Chen, Ziyun, et al.
Veröffentlicht: (2026)
von: Chen, Ziyun, et al.
Veröffentlicht: (2026)
RSCaMa: Remote Sensing Image Change Captioning with State Space Model
von: Liu, Chenyang, et al.
Veröffentlicht: (2024)
von: Liu, Chenyang, et al.
Veröffentlicht: (2024)
A Benchmark for Multi-Lingual Vision-Language Learning in Remote Sensing Image Captioning
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
High-resolution Population Maps Derived from Sentinel-1 and Sentinel-2
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
Enhancing Perception of Key Changes in Remote Sensing Image Change Captioning
von: Yang, Cong, et al.
Veröffentlicht: (2024)
von: Yang, Cong, et al.
Veröffentlicht: (2024)
Composed Image Retrieval for Remote Sensing
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
von: Psomas, Bill, et al.
Veröffentlicht: (2024)
Diffusion-RSCC: Diffusion Probabilistic Model for Change Captioning in Remote Sensing Images
von: Yu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Yu, Xiaofei, et al.
Veröffentlicht: (2024)
Multi-Scale Grouped Prototypes for Interpretable Semantic Segmentation
von: Porta, Hugo, et al.
Veröffentlicht: (2024)
von: Porta, Hugo, et al.
Veröffentlicht: (2024)
CanadaFireSat: Toward high-resolution wildfire forecasting with multiple modalities
von: Porta, Hugo, et al.
Veröffentlicht: (2025)
von: Porta, Hugo, et al.
Veröffentlicht: (2025)
GeoExplorer: Active Geo-localization with Curiosity-Driven Exploration
von: Mi, Li, et al.
Veröffentlicht: (2025)
von: Mi, Li, et al.
Veröffentlicht: (2025)
A Lightweight Sparse Focus Transformer for Remote Sensing Image Change Captioning
von: Sun, Dongwei, et al.
Veröffentlicht: (2024)
von: Sun, Dongwei, et al.
Veröffentlicht: (2024)
HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning
von: Wang, Man, et al.
Veröffentlicht: (2026)
von: Wang, Man, et al.
Veröffentlicht: (2026)
A Novel Lightweight Transformer with Edge-Aware Fusion for Remote Sensing Image Captioning
von: Das, Swadhin, et al.
Veröffentlicht: (2025)
von: Das, Swadhin, et al.
Veröffentlicht: (2025)
Semantic-Spatial Feature Fusion with Dynamic Graph Refinement for Remote Sensing Image Captioning
von: Liu, Maofu, et al.
Veröffentlicht: (2025)
von: Liu, Maofu, et al.
Veröffentlicht: (2025)
JSSFF: A Joint Structural-Semantic Fusion Framework for Remote Sensing Image Captioning
von: Das, Swadhin, et al.
Veröffentlicht: (2026)
von: Das, Swadhin, et al.
Veröffentlicht: (2026)
TRUST: Leveraging Text Robustness for Unsupervised Domain Adaptation
von: Litrico, Mattia, et al.
Veröffentlicht: (2025)
von: Litrico, Mattia, et al.
Veröffentlicht: (2025)
SLVideo: A Sign Language Video Moment Retrieval Framework
von: Martins, Gonçalo Vinagre, et al.
Veröffentlicht: (2024)
von: Martins, Gonçalo Vinagre, et al.
Veröffentlicht: (2024)
MV-CC: Mask Enhanced Video Model for Remote Sensing Change Caption
von: Liu, Ruixun, et al.
Veröffentlicht: (2024)
von: Liu, Ruixun, et al.
Veröffentlicht: (2024)
Show and Guide: Instructional-Plan Grounded Vision and Language Model
von: Glória-Silva, Diogo, et al.
Veröffentlicht: (2024)
von: Glória-Silva, Diogo, et al.
Veröffentlicht: (2024)
Multi-Spectral Remote Sensing Image Retrieval Using Geospatial Foundation Models
von: Blumenstiel, Benedikt, et al.
Veröffentlicht: (2024)
von: Blumenstiel, Benedikt, et al.
Veröffentlicht: (2024)
RSCC: A Large-Scale Remote Sensing Change Caption Dataset for Disaster Events
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2025)
Robust Remote Sensing Image-Text Retrieval with Noisy Correspondence
von: Song, Qiya, et al.
Veröffentlicht: (2026)
von: Song, Qiya, et al.
Veröffentlicht: (2026)
Semantic-CC: Boosting Remote Sensing Image Change Captioning via Foundational Knowledge and Semantic Guidance
von: Zhu, Yongshuo, et al.
Veröffentlicht: (2024)
von: Zhu, Yongshuo, et al.
Veröffentlicht: (2024)
Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2026)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2026)
Checkmate: interpretable and explainable RSVQA is the endgame
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2025)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2025)
From Classification to Segmentation with Explainable AI: A Study on Crack Detection and Growth Monitoring
von: Forest, Florent, et al.
Veröffentlicht: (2023)
von: Forest, Florent, et al.
Veröffentlicht: (2023)
RS-MoE: A Vision-Language Model with Mixture of Experts for Remote Sensing Image Captioning and Visual Question Answering
von: Lin, Hui, et al.
Veröffentlicht: (2024)
von: Lin, Hui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
An Efficient and Effective Encoder Model for Vision and Language Tasks in the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2025) -
Multilingual Vision-Language Pre-training for the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2024) -
Knowledge-aware Text-Image Retrieval for Remote Sensing Images
von: Mi, Li, et al.
Veröffentlicht: (2024) -
Multilingual Training-Free Remote Sensing Image Captioning
von: Rebelo, Carlos, et al.
Veröffentlicht: (2025) -
SMARTIES: Spectrum-Aware Multi-Sensor Auto-Encoder for Remote Sensing Images
von: Sumbul, Gencer, et al.
Veröffentlicht: (2025)