Landsat30-AU: A Vision-Language Dataset for Australian Landsat Imagery
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Sai, Li, Zhuang, Taylor, John A |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Landsat-Bench: Datasets and Benchmarks for Landsat Foundation Models
von: Corley, Isaac, et al.
Veröffentlicht: (2025)
von: Corley, Isaac, et al.
Veröffentlicht: (2025)
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
von: Ma, Qiwei, et al.
Veröffentlicht: (2025)
von: Ma, Qiwei, et al.
Veröffentlicht: (2025)
Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data
von: Mansour, Islam, et al.
Veröffentlicht: (2026)
von: Mansour, Islam, et al.
Veröffentlicht: (2026)
SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
von: Wu, Xue, et al.
Veröffentlicht: (2026)
von: Wu, Xue, et al.
Veröffentlicht: (2026)
FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph
von: Liu, Zhangding, et al.
Veröffentlicht: (2025)
von: Liu, Zhangding, et al.
Veröffentlicht: (2025)
Imagery as Inquiry: Exploring A Multimodal Dataset for Conversational Recommendation
von: Yoon, Se-eun, et al.
Veröffentlicht: (2024)
von: Yoon, Se-eun, et al.
Veröffentlicht: (2024)
MCANet: A Multi-Scale Class-Specific Attention Network for Multi-Label Post-Hurricane Damage Assessment using UAV Imagery
von: Liu, Zhangding, et al.
Veröffentlicht: (2025)
von: Liu, Zhangding, et al.
Veröffentlicht: (2025)
Lightweight Multimodal Adaptation of Vision Language Models for Species Recognition and Habitat Context Interpretation in Drone Thermal Imagery
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
CapRecover: A Cross-Modality Feature Inversion Attack Framework on Vision Language Models
von: Xiu, Kedong, et al.
Veröffentlicht: (2025)
von: Xiu, Kedong, et al.
Veröffentlicht: (2025)
Effective Damage Data Generation by Fusing Imagery with Human Knowledge Using Vision-Language Models
von: Wei, Jie, et al.
Veröffentlicht: (2025)
von: Wei, Jie, et al.
Veröffentlicht: (2025)
AllClear: A Comprehensive Dataset and Benchmark for Cloud Removal in Satellite Imagery
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
MEET: A Million-Scale Dataset for Fine-Grained Geospatial Scene Classification with Zoom-Free Remote Sensing Imagery
von: Li, Yansheng, et al.
Veröffentlicht: (2025)
von: Li, Yansheng, et al.
Veröffentlicht: (2025)
Empirical Recipes for Efficient and Compact Vision-Language Models
von: Huang, Jiabo, et al.
Veröffentlicht: (2026)
von: Huang, Jiabo, et al.
Veröffentlicht: (2026)
A Non-Reference Diffusion-Based Restoration Framework for Landsat 7 ETM+ SLC-off Imagery in Antarctica
von: Tang, Leyue, et al.
Veröffentlicht: (2026)
von: Tang, Leyue, et al.
Veröffentlicht: (2026)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
VME: A Satellite Imagery Dataset and Benchmark for Detecting Vehicles in the Middle East and Beyond
von: Al-Emadi, Noora, et al.
Veröffentlicht: (2025)
von: Al-Emadi, Noora, et al.
Veröffentlicht: (2025)
PRISM: A Multi-View Multi-Capability Retail Video Dataset for Embodied Vision-Language Models
von: Rouhi, Amirreza, et al.
Veröffentlicht: (2026)
von: Rouhi, Amirreza, et al.
Veröffentlicht: (2026)
Compositional Attribute Imbalance in Vision Datasets
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
Multimodal Distribution Matching for Vision-Language Dataset Distillation
von: Jeong, Jongoh, et al.
Veröffentlicht: (2026)
von: Jeong, Jongoh, et al.
Veröffentlicht: (2026)
Merlin: A Computed Tomography Vision-Language Foundation Model and Dataset
von: Blankemeier, Louis, et al.
Veröffentlicht: (2024)
von: Blankemeier, Louis, et al.
Veröffentlicht: (2024)
MM-Skin: Enhancing Dermatology Vision-Language Model with an Image-Text Dataset Derived from Textbooks
von: Zeng, Wenqi, et al.
Veröffentlicht: (2025)
von: Zeng, Wenqi, et al.
Veröffentlicht: (2025)
STAR: A First-Ever Dataset and A Large-Scale Benchmark for Scene Graph Generation in Large-Size Satellite Imagery
von: Li, Yansheng, et al.
Veröffentlicht: (2024)
von: Li, Yansheng, et al.
Veröffentlicht: (2024)
Sanitizing Manufacturing Dataset Labels Using Vision-Language Models
von: Mahjourian, Nazanin, et al.
Veröffentlicht: (2025)
von: Mahjourian, Nazanin, et al.
Veröffentlicht: (2025)
FLAME 3 Dataset: Unleashing the Power of Radiometric Thermal UAV Imagery for Wildfire Management
von: Hopkins, Bryce, et al.
Veröffentlicht: (2024)
von: Hopkins, Bryce, et al.
Veröffentlicht: (2024)
Segment Anything for Satellite Imagery: A Strong Baseline and a Regional Dataset for Automatic Field Delineation
von: Scribano, Carmelo, et al.
Veröffentlicht: (2025)
von: Scribano, Carmelo, et al.
Veröffentlicht: (2025)
KV-Efficient VLA: A Method to Speed up Vision Language Models with RNN-Gated Chunked KV Cache
von: Xu, Wanshun, et al.
Veröffentlicht: (2025)
von: Xu, Wanshun, et al.
Veröffentlicht: (2025)
ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
On the Limits of Token Reduction for Efficient Unified Vision Language Training
von: Chen, Siyi, et al.
Veröffentlicht: (2026)
von: Chen, Siyi, et al.
Veröffentlicht: (2026)
PATIMT-Bench: A Multi-Scenario Benchmark for Position-Aware Text Image Machine Translation in Large Vision-Language Models
von: Zhuang, Wanru, et al.
Veröffentlicht: (2025)
von: Zhuang, Wanru, et al.
Veröffentlicht: (2025)
Seeing Symbols, Missing Cultures: Probing Vision-Language Models' Reasoning on Fire Imagery and Cultural Meaning
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
A Benchmark Dataset for Spatially Aligned Road Damage Assessment in Small Uncrewed Aerial Systems Disaster Imagery
von: Manzini, Thomas, et al.
Veröffentlicht: (2025)
von: Manzini, Thomas, et al.
Veröffentlicht: (2025)
ClimateIQA: A New Dataset and Benchmark to Advance Vision-Language Models in Meteorology Anomalies Analysis
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
Are Large Vision Language Models Good Game Players?
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
Uncovering Intrinsic Capabilities: A Paradigm for Data Curation in Vision-Language Models
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
MAUGen: A Unified Diffusion Approach for Multi-Identity Facial Expression and AU Label Generation
von: Li, Xiangdong, et al.
Veröffentlicht: (2026)
von: Li, Xiangdong, et al.
Veröffentlicht: (2026)
Improving Medical Diagnostics with Vision-Language Models: Convex Hull-Based Uncertainty Analysis
von: Catak, Ferhat Ozgur, et al.
Veröffentlicht: (2024)
von: Catak, Ferhat Ozgur, et al.
Veröffentlicht: (2024)
Paved or unpaved? A Deep Learning derived Road Surface Global Dataset from Mapillary Street-View Imagery
von: Randhawa, Sukanya, et al.
Veröffentlicht: (2024)
von: Randhawa, Sukanya, et al.
Veröffentlicht: (2024)
Gastric-X: A Multimodal Multi-Phase Benchmark Dataset for Advancing Vision-Language Models in Gastric Cancer Analysis
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
Position Prediction Self-Supervised Learning for Multimodal Satellite Imagery Semantic Segmentation
von: Waithaka, John, et al.
Veröffentlicht: (2025)
von: Waithaka, John, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Landsat-Bench: Datasets and Benchmarks for Landsat Foundation Models
von: Corley, Isaac, et al.
Veröffentlicht: (2025) -
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
von: Ma, Qiwei, et al.
Veröffentlicht: (2025) -
Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data
von: Mansour, Islam, et al.
Veröffentlicht: (2026) -
SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
von: Wu, Xue, et al.
Veröffentlicht: (2026) -
FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph
von: Liu, Zhangding, et al.
Veröffentlicht: (2025)