VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Zhiming, Wang, Di, Guo, Haonan, Zhang, Jing, Du, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Building-road Collaborative Extraction from Remotely Sensed Images via Cross-Interaction
von: Guo, Haonan, et al.
Veröffentlicht: (2023)
von: Guo, Haonan, et al.
Veröffentlicht: (2023)
Expediting Building Footprint Extraction from High-resolution Remote Sensing Images via progressive lenient supervision
von: Guo, Haonan, et al.
Veröffentlicht: (2023)
von: Guo, Haonan, et al.
Veröffentlicht: (2023)
SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models
von: Zhong, Chen, et al.
Veröffentlicht: (2026)
von: Zhong, Chen, et al.
Veröffentlicht: (2026)
HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation
von: Wang, Jun, et al.
Veröffentlicht: (2026)
von: Wang, Jun, et al.
Veröffentlicht: (2026)
SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding
von: Luo, Junwei, et al.
Veröffentlicht: (2024)
von: Luo, Junwei, et al.
Veröffentlicht: (2024)
ViLaCD-R1: A Vision-Language Framework for Semantic Change Detection in Remote Sensing
von: Ma, Xingwei, et al.
Veröffentlicht: (2025)
von: Ma, Xingwei, et al.
Veröffentlicht: (2025)
CompareBench: A Benchmark for Visual Comparison Reasoning in Vision-Language Models
von: Cai, Jie, et al.
Veröffentlicht: (2025)
von: Cai, Jie, et al.
Veröffentlicht: (2025)
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
von: Yuan, Botai, et al.
Veröffentlicht: (2025)
von: Yuan, Botai, et al.
Veröffentlicht: (2025)
PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG
von: Han, Jiaju, et al.
Veröffentlicht: (2026)
von: Han, Jiaju, et al.
Veröffentlicht: (2026)
μ-Bench: A Vision-Language Benchmark for Microscopy Understanding
von: Lozano, Alejandro, et al.
Veröffentlicht: (2024)
von: Lozano, Alejandro, et al.
Veröffentlicht: (2024)
FUSE-RSVLM: Feature Fusion Vision-Language Model for Remote Sensing
von: Dang, Yunkai, et al.
Veröffentlicht: (2025)
von: Dang, Yunkai, et al.
Veröffentlicht: (2025)
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
von: Bao, Han, et al.
Veröffentlicht: (2024)
von: Bao, Han, et al.
Veröffentlicht: (2024)
GeoEyes: On-Demand Visual Focusing for Evidence-Grounded Understanding of Ultra-High-Resolution Remote Sensing Imagery
von: Wang, Fengxiang, et al.
Veröffentlicht: (2026)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2026)
VLBiasBench: A Comprehensive Benchmark for Evaluating Bias in Large Vision-Language Model
von: Wang, Sibo, et al.
Veröffentlicht: (2024)
von: Wang, Sibo, et al.
Veröffentlicht: (2024)
ERGeoBench:A Comprehensive Benchmark for Embodied Reasoning and Geo-localization in Multimodal Large Language Models
von: Xue, Kaiwen, et al.
Veröffentlicht: (2026)
von: Xue, Kaiwen, et al.
Veröffentlicht: (2026)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing
von: Zhang, Zilun, et al.
Veröffentlicht: (2025)
von: Zhang, Zilun, et al.
Veröffentlicht: (2025)
Vision-Language Models in Remote Sensing: Current Progress and Future Trends
von: Li, Xiang, et al.
Veröffentlicht: (2023)
von: Li, Xiang, et al.
Veröffentlicht: (2023)
Vision-Language Model Purified Semi-Supervised Semantic Segmentation for Remote Sensing Images
von: Wang, Shanwen, et al.
Veröffentlicht: (2026)
von: Wang, Shanwen, et al.
Veröffentlicht: (2026)
DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
FedRSClip: Federated Learning for Remote Sensing Scene Classification Using Vision-Language Models
von: Lin, Hui, et al.
Veröffentlicht: (2025)
von: Lin, Hui, et al.
Veröffentlicht: (2025)
Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
Benchmarking and Mitigating Sycophancy in Medical Vision Language Models
von: Xu, Juangui, et al.
Veröffentlicht: (2025)
von: Xu, Juangui, et al.
Veröffentlicht: (2025)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces
von: Zhao, Baining, et al.
Veröffentlicht: (2025)
von: Zhao, Baining, et al.
Veröffentlicht: (2025)
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
SpaceSense-Bench: A Large-Scale Multi-Modal Benchmark for Spacecraft Perception and Pose Estimation
von: Wu, Aodi, et al.
Veröffentlicht: (2026)
von: Wu, Aodi, et al.
Veröffentlicht: (2026)
VLM-RobustBench: A Comprehensive Benchmark for Robustness of Vision-Language Models
von: Saxena, Rohit, et al.
Veröffentlicht: (2026)
von: Saxena, Rohit, et al.
Veröffentlicht: (2026)
MMDocBench: Benchmarking Large Vision-Language Models for Fine-Grained Visual Document Understanding
von: Zhu, Fengbin, et al.
Veröffentlicht: (2024)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2024)
MDK12-Bench: A Multi-Discipline Benchmark for Evaluating Reasoning in Multimodal Large Language Models
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
Beyond Classification Accuracy: Neural-MedBench and the Need for Deeper Reasoning Benchmarks
von: Jing, Miao, et al.
Veröffentlicht: (2025)
von: Jing, Miao, et al.
Veröffentlicht: (2025)
RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing
von: Zhang, Zilun, et al.
Veröffentlicht: (2023)
von: Zhang, Zilun, et al.
Veröffentlicht: (2023)
The Illusion of Clinical Reasoning: A Benchmark Reveals the Pervasive Gap in Vision-Language Models for Clinical Competency
von: Wang, Dingyu, et al.
Veröffentlicht: (2025)
von: Wang, Dingyu, et al.
Veröffentlicht: (2025)
RS-MoE: A Vision-Language Model with Mixture of Experts for Remote Sensing Image Captioning and Visual Question Answering
von: Lin, Hui, et al.
Veröffentlicht: (2024)
von: Lin, Hui, et al.
Veröffentlicht: (2024)
MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models
von: Yan, Bei, et al.
Veröffentlicht: (2024)
von: Yan, Bei, et al.
Veröffentlicht: (2024)
SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals
von: Lin, Zihang, et al.
Veröffentlicht: (2026)
von: Lin, Zihang, et al.
Veröffentlicht: (2026)
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
von: Wang, Zhecan, et al.
Veröffentlicht: (2024)
von: Wang, Zhecan, et al.
Veröffentlicht: (2024)
CDH-Bench: A Commonsense-Driven Hallucination Benchmark for Evaluating Visual Fidelity in Vision-Language Models
von: Chen, Kesheng, et al.
Veröffentlicht: (2026)
von: Chen, Kesheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Building-road Collaborative Extraction from Remotely Sensed Images via Cross-Interaction
von: Guo, Haonan, et al.
Veröffentlicht: (2023) -
Expediting Building Footprint Extraction from High-resolution Remote Sensing Images via progressive lenient supervision
von: Guo, Haonan, et al.
Veröffentlicht: (2023) -
SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models
von: Zhong, Chen, et al.
Veröffentlicht: (2026) -
HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026) -
LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation
von: Wang, Jun, et al.
Veröffentlicht: (2026)