Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
Fuente:
arXiv
Salvato in:
| Autori principali: | Tosato, Lucrezia, Boussaid, Hichem, Weissgerber, Flora, Kurtz, Camille, Wendling, Laurent, Lobry, Sylvain |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Visual Question Answering on Multiple Remote Sensing Image Modalities
di: Boussaid, Hichem, et al.
Pubblicazione: (2025)
di: Boussaid, Hichem, et al.
Pubblicazione: (2025)
Can SAR improve RSVQA performance?
di: Tosato, Lucrezia, et al.
Pubblicazione: (2024)
di: Tosato, Lucrezia, et al.
Pubblicazione: (2024)
SAR Strikes Back: A New Hope for RSVQA
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025)
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025)
Checkmate: interpretable and explainable RSVQA is the endgame
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025)
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025)
Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning
di: Tosato, Lucrezia, et al.
Pubblicazione: (2026)
di: Tosato, Lucrezia, et al.
Pubblicazione: (2026)
RAMEN: Resolution-Adjustable Multimodal Encoder for Earth Observation
di: Houdré, Nicolas, et al.
Pubblicazione: (2025)
di: Houdré, Nicolas, et al.
Pubblicazione: (2025)
Large Vision-Language Models for Remote Sensing Visual Question Answering
di: Siripong, Surasakdi, et al.
Pubblicazione: (2024)
di: Siripong, Surasakdi, et al.
Pubblicazione: (2024)
Center-guided Classifier for Semantic Segmentation of Remote Sensing Images
di: Zhang, Wei, et al.
Pubblicazione: (2025)
di: Zhang, Wei, et al.
Pubblicazione: (2025)
Copy-Move Forgery Detection and Question Answering for Remote Sensing Image
di: Zhang, Ze, et al.
Pubblicazione: (2024)
di: Zhang, Ze, et al.
Pubblicazione: (2024)
RSAdapter: Adapting Multimodal Models for Remote Sensing Visual Question Answering
di: Wang, Yuduo, et al.
Pubblicazione: (2023)
di: Wang, Yuduo, et al.
Pubblicazione: (2023)
Knowledge-aware Visual Question Generation for Remote Sensing Images
di: Li, Siran, et al.
Pubblicazione: (2026)
di: Li, Siran, et al.
Pubblicazione: (2026)
Text-Guided Coarse-to-Fine Fusion Network for Robust Remote Sensing Visual Question Answering
di: Zhao, Zhicheng, et al.
Pubblicazione: (2024)
di: Zhao, Zhicheng, et al.
Pubblicazione: (2024)
Threshold Attention Network for Semantic Segmentation of Remote Sensing Images
di: Long, Wei, et al.
Pubblicazione: (2025)
di: Long, Wei, et al.
Pubblicazione: (2025)
Exploiting temporal information to detect conversational groups in videos and predict the next speaker
di: Tosato, Lucrezia, et al.
Pubblicazione: (2024)
di: Tosato, Lucrezia, et al.
Pubblicazione: (2024)
AMMUNet: Multi-Scale Attention Map Merging for Remote Sensing Image Segmentation
di: Yang, Yang, et al.
Pubblicazione: (2024)
di: Yang, Yang, et al.
Pubblicazione: (2024)
SegEarth-R2: Towards Comprehensive Language-guided Segmentation for Remote Sensing Images
di: Xin, Zepeng, et al.
Pubblicazione: (2025)
di: Xin, Zepeng, et al.
Pubblicazione: (2025)
RS-MoE: A Vision-Language Model with Mixture of Experts for Remote Sensing Image Captioning and Visual Question Answering
di: Lin, Hui, et al.
Pubblicazione: (2024)
di: Lin, Hui, et al.
Pubblicazione: (2024)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
di: Meng, Tian, et al.
Pubblicazione: (2024)
di: Meng, Tian, et al.
Pubblicazione: (2024)
Questions beyond Pixels: Integrating Commonsense Knowledge in Visual Question Generation for Remote Sensing
di: Li, Siran, et al.
Pubblicazione: (2026)
di: Li, Siran, et al.
Pubblicazione: (2026)
A Semantically-Aware Relevance Measure for Content-Based Medical Image Retrieval Evaluation
di: Wei, Xiaoyang, et al.
Pubblicazione: (2025)
di: Wei, Xiaoyang, et al.
Pubblicazione: (2025)
Multimodal Integration of Human-Like Attention in Visual Question Answering
di: Sood, Ekta, et al.
Pubblicazione: (2021)
di: Sood, Ekta, et al.
Pubblicazione: (2021)
RRSIS: Referring Remote Sensing Image Segmentation
di: Yuan, Zhenghang, et al.
Pubblicazione: (2023)
di: Yuan, Zhenghang, et al.
Pubblicazione: (2023)
RSVLM-QA: A Benchmark Dataset for Remote Sensing Vision Language Model-based Question Answering
di: Zi, Xing, et al.
Pubblicazione: (2025)
di: Zi, Xing, et al.
Pubblicazione: (2025)
Show Me What and Where has Changed? Question Answering and Grounding for Remote Sensing Change Detection
di: Li, Ke, et al.
Pubblicazione: (2024)
di: Li, Ke, et al.
Pubblicazione: (2024)
Dynamic Dictionary Learning for Remote Sensing Image Segmentation
di: Zou, Xuechao, et al.
Pubblicazione: (2025)
di: Zou, Xuechao, et al.
Pubblicazione: (2025)
RS3Mamba: Visual State Space Model for Remote Sensing Images Semantic Segmentation
di: Ma, Xianping, et al.
Pubblicazione: (2024)
di: Ma, Xianping, et al.
Pubblicazione: (2024)
QIRL: Boosting Visual Question Answering via Optimized Question-Image Relation Learning
di: Xu, Quanxing, et al.
Pubblicazione: (2025)
di: Xu, Quanxing, et al.
Pubblicazione: (2025)
Multi-Page Document Visual Question Answering using Self-Attention Scoring Mechanism
di: Kang, Lei, et al.
Pubblicazione: (2024)
di: Kang, Lei, et al.
Pubblicazione: (2024)
ConFoThinking: Consolidated Focused Attention Driven Thinking for Visual Question Answering
di: Wu, Zhaodong, et al.
Pubblicazione: (2026)
di: Wu, Zhaodong, et al.
Pubblicazione: (2026)
Selectively Answering Visual Questions
di: Eisenschlos, Julian Martin, et al.
Pubblicazione: (2024)
di: Eisenschlos, Julian Martin, et al.
Pubblicazione: (2024)
NSegment : Label-specific Deformations for Remote Sensing Image Segmentation
di: Kim, Yechan, et al.
Pubblicazione: (2025)
di: Kim, Yechan, et al.
Pubblicazione: (2025)
Multi-view Remote Sensing Image Segmentation With SAM priors
di: Qi, Zipeng, et al.
Pubblicazione: (2024)
di: Qi, Zipeng, et al.
Pubblicazione: (2024)
Kolmogorov-Arnold Network for Remote Sensing Image Semantic Segmentation
di: Ma, Xianping, et al.
Pubblicazione: (2025)
di: Ma, Xianping, et al.
Pubblicazione: (2025)
Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images
di: Li, Kaiyu, et al.
Pubblicazione: (2025)
di: Li, Kaiyu, et al.
Pubblicazione: (2025)
Visually Interpretable Subtask Reasoning for Visual Question Answering
di: Cheng, Yu, et al.
Pubblicazione: (2025)
di: Cheng, Yu, et al.
Pubblicazione: (2025)
Targeted Visual Prompting for Medical Visual Question Answering
di: Tascon-Morales, Sergio, et al.
Pubblicazione: (2024)
di: Tascon-Morales, Sergio, et al.
Pubblicazione: (2024)
Visual Robustness Benchmark for Visual Question Answering (VQA)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
di: Ahir, Param, et al.
Pubblicazione: (2023)
di: Ahir, Param, et al.
Pubblicazione: (2023)
Question-Aware Gaussian Experts for Audio-Visual Question Answering
di: Kim, Hongyeob, et al.
Pubblicazione: (2025)
di: Kim, Hongyeob, et al.
Pubblicazione: (2025)
A Dual-Attention Learning Network with Word and Sentence Embedding for Medical Visual Question Answering
di: Huang, Xiaofei, et al.
Pubblicazione: (2022)
di: Huang, Xiaofei, et al.
Pubblicazione: (2022)
Documenti analoghi
-
Visual Question Answering on Multiple Remote Sensing Image Modalities
di: Boussaid, Hichem, et al.
Pubblicazione: (2025) -
Can SAR improve RSVQA performance?
di: Tosato, Lucrezia, et al.
Pubblicazione: (2024) -
SAR Strikes Back: A New Hope for RSVQA
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025) -
Checkmate: interpretable and explainable RSVQA is the endgame
di: Tosato, Lucrezia, et al.
Pubblicazione: (2025) -
Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning
di: Tosato, Lucrezia, et al.
Pubblicazione: (2026)