Regression in EO: Are VLMs Up to the Challenge?
Fuente:
arXiv
Guardado en:
| Autores principales: | Xue, Xizhe, Zhu, Xiao Xiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Unified Vision Language Models for Forest Ecological Analysis in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2025)
por: Xue, Xizhe, et al.
Publicado: (2025)
REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2024)
por: Xue, Xizhe, et al.
Publicado: (2024)
Faces of the Mind: Unveiling Mental Health States Through Facial Expressions in 11,427 Adolescents
por: Xu, Xiao, et al.
Publicado: (2024)
por: Xu, Xiao, et al.
Publicado: (2024)
Heatmap Guided Query Transformers for Robust Astrocyte Detection across Immunostains and Resolutions
por: Zhang, Xizhe, et al.
Publicado: (2025)
por: Zhang, Xizhe, et al.
Publicado: (2025)
EO-VAE: Towards A Multi-sensor Tokenizer for Earth Observation Data
por: Lehmann, Nils, et al.
Publicado: (2026)
por: Lehmann, Nils, et al.
Publicado: (2026)
CARE: Confidence-Aware Regression Estimation of building density fine-tuning EO Foundation Models
por: Dionelis, Nikolaos, et al.
Publicado: (2025)
por: Dionelis, Nikolaos, et al.
Publicado: (2025)
Level Up Your Tutorials: VLMs for Game Tutorials Quality Assessment
por: Cambrin, Daniele Rege, et al.
Publicado: (2024)
por: Cambrin, Daniele Rege, et al.
Publicado: (2024)
Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
por: Yang, Yuchen, et al.
Publicado: (2026)
por: Yang, Yuchen, et al.
Publicado: (2026)
How to Embed Matters: Evaluation of EO Embedding Design Choices
por: Gilch, Luis, et al.
Publicado: (2026)
por: Gilch, Luis, et al.
Publicado: (2026)
3D-RCNet: Learning from Transformer to Build a 3D Relational ConvNet for Hyperspectral Image Classification
por: Jing, Haizhao, et al.
Publicado: (2024)
por: Jing, Haizhao, et al.
Publicado: (2024)
SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks
por: Song, Zijian, et al.
Publicado: (2025)
por: Song, Zijian, et al.
Publicado: (2025)
Bridging Sensor Gaps via Attention Gated Tuning for Hyperspectral Image Classification
por: Xue, Xizhe, et al.
Publicado: (2023)
por: Xue, Xizhe, et al.
Publicado: (2023)
UVLM: Benchmarking Video Language Model for Underwater World Understanding
por: Xue, Xizhe, et al.
Publicado: (2025)
por: Xue, Xizhe, et al.
Publicado: (2025)
Unlocking Dense Metric Depth Estimation in VLMs
por: Yu, Hanxun, et al.
Publicado: (2026)
por: Yu, Hanxun, et al.
Publicado: (2026)
Scaling Laws for Geospatial Foundation Models: A case study on PhilEO Bench
por: Dionelis, Nikolaos, et al.
Publicado: (2025)
por: Dionelis, Nikolaos, et al.
Publicado: (2025)
Open-Vocabulary Object Detection in UAV Imagery: A Review and Future Perspectives
por: Zhou, Yang, et al.
Publicado: (2025)
por: Zhou, Yang, et al.
Publicado: (2025)
S-EO: A Large-Scale Dataset for Geometry-Aware Shadow Detection in Remote Sensing Applications
por: Masquil, Elías, et al.
Publicado: (2025)
por: Masquil, Elías, et al.
Publicado: (2025)
Sparse Spectral LoRA: Routed Experts for Medical VLMs
por: Manzari, Omid Nejati, et al.
Publicado: (2026)
por: Manzari, Omid Nejati, et al.
Publicado: (2026)
Ground-V: Teaching VLMs to Ground Complex Instructions in Pixels
por: Zong, Yongshuo, et al.
Publicado: (2025)
por: Zong, Yongshuo, et al.
Publicado: (2025)
FlowEO: Generative Unsupervised Domain Adaptation for Earth Observation
por: Bellier, Georges Le, et al.
Publicado: (2025)
por: Bellier, Georges Le, et al.
Publicado: (2025)
IC-EO: Interpretable Code-based assistant for Earth Observation
por: Lahouel, Lamia, et al.
Publicado: (2026)
por: Lahouel, Lamia, et al.
Publicado: (2026)
PhilEO Bench: Evaluating Geo-Spatial Foundation Models
por: Fibaek, Casper, et al.
Publicado: (2024)
por: Fibaek, Casper, et al.
Publicado: (2024)
SSL4EO-S12 v1.1: A Multimodal, Multiseasonal Dataset for Pretraining, Updated
por: Blumenstiel, Benedikt, et al.
Publicado: (2025)
por: Blumenstiel, Benedikt, et al.
Publicado: (2025)
Prithvi-EO-2.0: A Versatile Multi-Temporal Foundation Model for Earth Observation Applications
por: Szwarcman, Daniela, et al.
Publicado: (2024)
por: Szwarcman, Daniela, et al.
Publicado: (2024)
From My View to Yours: Ego-to-Exo Transfer in VLMs for Understanding Activities of Daily Living
por: Reilly, Dominick, et al.
Publicado: (2025)
por: Reilly, Dominick, et al.
Publicado: (2025)
Deep Pre-Alignment for VLMs
por: Yu, Tianyu, et al.
Publicado: (2026)
por: Yu, Tianyu, et al.
Publicado: (2026)
Leveraging band diversity for feature selection in EO data
por: Hussain, Sadia, et al.
Publicado: (2025)
por: Hussain, Sadia, et al.
Publicado: (2025)
EO-VLM: VLM-Guided Energy Overload Attacks on Vision Models
por: Seo, Minjae, et al.
Publicado: (2025)
por: Seo, Minjae, et al.
Publicado: (2025)
Improving EO Foundation Models with Confidence Assessment for enhanced Semantic segmentation
por: Dionelis, Nikolaos, et al.
Publicado: (2024)
por: Dionelis, Nikolaos, et al.
Publicado: (2024)
Mitigating Accuracy-Robustness Trade-off via Balanced Multi-Teacher Adversarial Distillation
por: Zhao, Shiji, et al.
Publicado: (2023)
por: Zhao, Shiji, et al.
Publicado: (2023)
Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs
por: Ma, Wen, et al.
Publicado: (2026)
por: Ma, Wen, et al.
Publicado: (2026)
Are VLMs Really Blind
por: Singh, Ayush, et al.
Publicado: (2024)
por: Singh, Ayush, et al.
Publicado: (2024)
RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
por: Wei, Guoting, et al.
Publicado: (2024)
por: Wei, Guoting, et al.
Publicado: (2024)
Preserving Localized Patch Semantics in VLMs
por: Esmaeilkhani, Parsa, et al.
Publicado: (2026)
por: Esmaeilkhani, Parsa, et al.
Publicado: (2026)
Towards Class-wise Fair Adversarial Training via Anti-Bias Soft Label Distillation
por: Zhao, Shiji, et al.
Publicado: (2025)
por: Zhao, Shiji, et al.
Publicado: (2025)
SAIP-Net: Enhancing Remote Sensing Image Segmentation via Spectral Adaptive Information Propagation
por: Wang, Zhongtao, et al.
Publicado: (2025)
por: Wang, Zhongtao, et al.
Publicado: (2025)
XSPA: Crafting Imperceptible X-Shaped Sparse Adversarial Perturbations for Transferable Attacks on VLMs
por: Hu, Chengyin, et al.
Publicado: (2026)
por: Hu, Chengyin, et al.
Publicado: (2026)
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
por: Xu, Mingjie, et al.
Publicado: (2025)
por: Xu, Mingjie, et al.
Publicado: (2025)
PaliGemma 2: A Family of Versatile VLMs for Transfer
por: Steiner, Andreas, et al.
Publicado: (2024)
por: Steiner, Andreas, et al.
Publicado: (2024)
CLGRPO: Reasoning Ability Enhancement for Small VLMs
por: Wang, Fanyi, et al.
Publicado: (2025)
por: Wang, Fanyi, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Unified Vision Language Models for Forest Ecological Analysis in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2025) -
REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2024) -
Faces of the Mind: Unveiling Mental Health States Through Facial Expressions in 11,427 Adolescents
por: Xu, Xiao, et al.
Publicado: (2024) -
Heatmap Guided Query Transformers for Robust Astrocyte Detection across Immunostains and Resolutions
por: Zhang, Xizhe, et al.
Publicado: (2025) -
EO-VAE: Towards A Multi-sensor Tokenizer for Earth Observation Data
por: Lehmann, Nils, et al.
Publicado: (2026)