Checkmate: interpretable and explainable RSVQA is the endgame
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tosato, Lucrezia, Chappuis, Christel Tartini, Montariol, Syrielle, Weissgerber, Flora, Lobry, Sylvain, Tuia, Devis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can SAR improve RSVQA performance?
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
SAR Strikes Back: A New Hope for RSVQA
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2025)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2025)
Visual Question Answering on Multiple Remote Sensing Image Modalities
von: Boussaid, Hichem, et al.
Veröffentlicht: (2025)
von: Boussaid, Hichem, et al.
Veröffentlicht: (2025)
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
ConVQG: Contrastive Visual Question Generation with Multimodal Guidance
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
ConGeo: Robust Cross-view Geo-localization across Ground View Variations
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2026)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2026)
Cross-Modal Learning of Housing Quality in Amsterdam
von: Levering, Alex, et al.
Veröffentlicht: (2024)
von: Levering, Alex, et al.
Veröffentlicht: (2024)
MaskSDM with Shapley values to improve flexibility, robustness, and explainability in species distribution modeling
von: Zbinden, Robin, et al.
Veröffentlicht: (2025)
von: Zbinden, Robin, et al.
Veröffentlicht: (2025)
SMARTIES: Spectrum-Aware Multi-Sensor Auto-Encoder for Remote Sensing Images
von: Sumbul, Gencer, et al.
Veröffentlicht: (2025)
von: Sumbul, Gencer, et al.
Veröffentlicht: (2025)
POLO -- Point-based, multi-class animal detection
von: May, Giacomo, et al.
Veröffentlicht: (2024)
von: May, Giacomo, et al.
Veröffentlicht: (2024)
Retrieval-Based Interleaved Visual Chain-of-Thought in Real-World Driving Scenarios
von: Corbière, Charles, et al.
Veröffentlicht: (2025)
von: Corbière, Charles, et al.
Veröffentlicht: (2025)
Exploiting temporal information to detect conversational groups in videos and predict the next speaker
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
Retrieval of Surface Solar Radiation through Implicit Albedo Recovery from Temporal Context
von: Frischholz, Yael, et al.
Veröffentlicht: (2025)
von: Frischholz, Yael, et al.
Veröffentlicht: (2025)
CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
An Efficient and Effective Encoder Model for Vision and Language Tasks in the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2025)
von: Silva, João Daniel, et al.
Veröffentlicht: (2025)
Multilingual Vision-Language Pre-training for the Remote Sensing Domain
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
Knowledge-aware Visual Question Generation for Remote Sensing Images
von: Li, Siran, et al.
Veröffentlicht: (2026)
von: Li, Siran, et al.
Veröffentlicht: (2026)
Questions beyond Pixels: Integrating Commonsense Knowledge in Visual Question Generation for Remote Sensing
von: Li, Siran, et al.
Veröffentlicht: (2026)
von: Li, Siran, et al.
Veröffentlicht: (2026)
Large Language Models for Captioning and Retrieving Remote Sensing Images
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
von: Silva, João Daniel, et al.
Veröffentlicht: (2024)
High-resolution Population Maps Derived from Sentinel-1 and Sentinel-2
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
Knowledge-aware Text-Image Retrieval for Remote Sensing Images
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
Multi-Scale Grouped Prototypes for Interpretable Semantic Segmentation
von: Porta, Hugo, et al.
Veröffentlicht: (2024)
von: Porta, Hugo, et al.
Veröffentlicht: (2024)
CanadaFireSat: Toward high-resolution wildfire forecasting with multiple modalities
von: Porta, Hugo, et al.
Veröffentlicht: (2025)
von: Porta, Hugo, et al.
Veröffentlicht: (2025)
GeoExplorer: Active Geo-localization with Curiosity-Driven Exploration
von: Mi, Li, et al.
Veröffentlicht: (2025)
von: Mi, Li, et al.
Veröffentlicht: (2025)
EcoWikiRS: Learning Ecological Representation of Satellite Images from Weak Supervision with Species Observations and Wikipedia
von: Zermatten, Valerie, et al.
Veröffentlicht: (2025)
von: Zermatten, Valerie, et al.
Veröffentlicht: (2025)
TRUST: Leveraging Text Robustness for Unsupervised Domain Adaptation
von: Litrico, Mattia, et al.
Veröffentlicht: (2025)
von: Litrico, Mattia, et al.
Veröffentlicht: (2025)
From Classification to Segmentation with Explainable AI: A Study on Crack Detection and Growth Monitoring
von: Forest, Florent, et al.
Veröffentlicht: (2023)
von: Forest, Florent, et al.
Veröffentlicht: (2023)
Atomizer: Generalizing to new modalities by breaking satellite images down to a set of scalars
von: de Turckheim, Hugo Riffaud, et al.
Veröffentlicht: (2025)
von: de Turckheim, Hugo Riffaud, et al.
Veröffentlicht: (2025)
What to align in multimodal contrastive learning?
von: Dufumier, Benoit, et al.
Veröffentlicht: (2024)
von: Dufumier, Benoit, et al.
Veröffentlicht: (2024)
Analytical Logit Scaling for High-Resolution Sea Ice Topology Retrieval from Weakly Labeled SAR Imagery
von: Elwaradi, Reda, et al.
Veröffentlicht: (2026)
von: Elwaradi, Reda, et al.
Veröffentlicht: (2026)
The Coralscapes Dataset: Semantic Scene Understanding in Coral Reefs
von: Sauder, Jonathan, et al.
Veröffentlicht: (2025)
von: Sauder, Jonathan, et al.
Veröffentlicht: (2025)
VinaBench: Benchmark for Faithful and Consistent Visual Narratives
von: Gao, Silin, et al.
Veröffentlicht: (2025)
von: Gao, Silin, et al.
Veröffentlicht: (2025)
GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data
von: Ferrod, Roger, et al.
Veröffentlicht: (2026)
von: Ferrod, Roger, et al.
Veröffentlicht: (2026)
IC-EO: Interpretable Code-based assistant for Earth Observation
von: Lahouel, Lamia, et al.
Veröffentlicht: (2026)
von: Lahouel, Lamia, et al.
Veröffentlicht: (2026)
RAMEN: Resolution-Adjustable Multimodal Encoder for Earth Observation
von: Houdré, Nicolas, et al.
Veröffentlicht: (2025)
von: Houdré, Nicolas, et al.
Veröffentlicht: (2025)
MammAlps: A multi-view video behavior monitoring dataset of wild mammals in the Swiss Alps
von: Gabeff, Valentin, et al.
Veröffentlicht: (2025)
von: Gabeff, Valentin, et al.
Veröffentlicht: (2025)
Better, Not Just More: Data-Centric Machine Learning for Earth Observation
von: Roscher, Ribana, et al.
Veröffentlicht: (2023)
von: Roscher, Ribana, et al.
Veröffentlicht: (2023)
An approximation-based approach versus an AI one for the study of CT images of abdominal aorta aneurysms
von: Rinelli, Lucrezia, et al.
Veröffentlicht: (2024)
von: Rinelli, Lucrezia, et al.
Veröffentlicht: (2024)
Pathologist-like explainable AI for interpretable Gleason grading in prostate cancer
von: Mittmann, Gesa, et al.
Veröffentlicht: (2024)
von: Mittmann, Gesa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can SAR improve RSVQA performance?
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024) -
SAR Strikes Back: A New Hope for RSVQA
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2025) -
Visual Question Answering on Multiple Remote Sensing Image Modalities
von: Boussaid, Hichem, et al.
Veröffentlicht: (2025) -
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024) -
ConVQG: Contrastive Visual Question Generation with Multimodal Guidance
von: Mi, Li, et al.
Veröffentlicht: (2024)