GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data
Fuente:
arXiv
Saved in:
| Main Authors: | Ferrod, Roger, Lecene, Maël, Sapkota, Krishna, Leifman, George, Silverman, Vered, Beryozkin, Genady, Lobry, Sylvain |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RS-OVC: Open-Vocabulary Counting for Remote-Sensing Data
by: Shor, Tamir, et al.
Published: (2026)
by: Shor, Tamir, et al.
Published: (2026)
RSRCC: A Remote Sensing Regional Change Comprehension Benchmark Constructed via Retrieval-Augmented Best-of-N Ranking
by: Kazoom, Roie, et al.
Published: (2026)
by: Kazoom, Roie, et al.
Published: (2026)
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization
by: Barzilai, Aviad, et al.
Published: (2025)
by: Barzilai, Aviad, et al.
Published: (2025)
On-the-Fly OVD Adaptation with FLAME: Few-shot Localization via Active Marginal-Samples Exploration
by: Refael, Yehonathan, et al.
Published: (2025)
by: Refael, Yehonathan, et al.
Published: (2025)
Zero-Shot Multi-Spectral Learning: Reimagining a Generalist Multimodal Gemini 2.5 Model for Remote Sensing Applications
by: Mallya, Ganesh, et al.
Published: (2025)
by: Mallya, Ganesh, et al.
Published: (2025)
Towards a multimodal framework for remote sensing image change retrieval and captioning
by: Ferrod, Roger, et al.
Published: (2024)
by: Ferrod, Roger, et al.
Published: (2024)
Revisiting Cross-Modal Knowledge Distillation: A Disentanglement Approach for RGBD Semantic Segmentation
by: Ferrod, Roger, et al.
Published: (2025)
by: Ferrod, Roger, et al.
Published: (2025)
Can SAR improve RSVQA performance?
by: Tosato, Lucrezia, et al.
Published: (2024)
by: Tosato, Lucrezia, et al.
Published: (2024)
SAR Strikes Back: A New Hope for RSVQA
by: Tosato, Lucrezia, et al.
Published: (2025)
by: Tosato, Lucrezia, et al.
Published: (2025)
ReGround: Improving Textual and Spatial Grounding at No Cost
by: Lee, Phillip Y., et al.
Published: (2024)
by: Lee, Phillip Y., et al.
Published: (2024)
A Survey of Video Datasets for Grounded Event Understanding
by: Sanders, Kate, et al.
Published: (2024)
by: Sanders, Kate, et al.
Published: (2024)
RieMind: Geometry-Grounded Spatial Agent for Scene Understanding
by: Ropero, Fernando, et al.
Published: (2026)
by: Ropero, Fernando, et al.
Published: (2026)
GroundingAnomaly: Spatially-Grounded Diffusion for Few-Shot Anomaly Synthesis
by: Liu, Yishen, et al.
Published: (2026)
by: Liu, Yishen, et al.
Published: (2026)
Atomizer: Generalizing to new modalities by breaking satellite images down to a set of scalars
by: de Turckheim, Hugo Riffaud, et al.
Published: (2025)
by: de Turckheim, Hugo Riffaud, et al.
Published: (2025)
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding
by: Clark, Christopher, et al.
Published: (2026)
by: Clark, Christopher, et al.
Published: (2026)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
by: Liu, Shilong, et al.
Published: (2023)
by: Liu, Shilong, et al.
Published: (2023)
Anchored Diffusion for Video Face Reenactment
by: Kligvasser, Idan, et al.
Published: (2024)
by: Kligvasser, Idan, et al.
Published: (2024)
Semi-supervised Quality Evaluation of Colonoscopy Procedures
by: Kligvasser, Idan, et al.
Published: (2023)
by: Kligvasser, Idan, et al.
Published: (2023)
TartanGround: A Large-Scale Dataset for Ground Robot Perception and Navigation
by: Patel, Manthan, et al.
Published: (2025)
by: Patel, Manthan, et al.
Published: (2025)
SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models
by: Cheng, An-Chieh, et al.
Published: (2024)
by: Cheng, An-Chieh, et al.
Published: (2024)
CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography
by: Prakash, Eva, et al.
Published: (2026)
by: Prakash, Eva, et al.
Published: (2026)
Hierarchical Contextual Grounding LVLM: Enhancing Fine-Grained Visual-Language Understanding with Robust Grounding
by: Guo, Leilei, et al.
Published: (2025)
by: Guo, Leilei, et al.
Published: (2025)
Cambrian-P: Pose-Grounded Video Understanding
by: Yang, Jihan, et al.
Published: (2026)
by: Yang, Jihan, et al.
Published: (2026)
Ground4D: Spatially-Grounded Feedforward 4D Reconstruction for Unstructured Off-Road Scenes
by: Wang, Shuo, et al.
Published: (2026)
by: Wang, Shuo, et al.
Published: (2026)
Visual Question Answering on Multiple Remote Sensing Image Modalities
by: Boussaid, Hichem, et al.
Published: (2025)
by: Boussaid, Hichem, et al.
Published: (2025)
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
by: Tosato, Lucrezia, et al.
Published: (2024)
by: Tosato, Lucrezia, et al.
Published: (2024)
Think with Grounding: Curriculum Reinforced Reasoning with Video Grounding for Long Video Understanding
by: Chen, Houlun, et al.
Published: (2026)
by: Chen, Houlun, et al.
Published: (2026)
Beyond Accuracy: Evaluating Visual Grounding In Multimodal Medical Reasoning
by: Zafar, Anas, et al.
Published: (2026)
by: Zafar, Anas, et al.
Published: (2026)
Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
ViCaS: A Dataset for Combining Holistic and Pixel-level Video Understanding using Captions with Grounded Segmentation
by: Athar, Ali, et al.
Published: (2024)
by: Athar, Ali, et al.
Published: (2024)
Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
Harnessing Object Grounding for Time-Sensitive Video Understanding
by: Wu, Tz-Ying, et al.
Published: (2025)
by: Wu, Tz-Ying, et al.
Published: (2025)
Semantically Grounded QFormer for Efficient Vision Language Understanding
by: Choraria, Moulik, et al.
Published: (2023)
by: Choraria, Moulik, et al.
Published: (2023)
Checkmate: interpretable and explainable RSVQA is the endgame
by: Tosato, Lucrezia, et al.
Published: (2025)
by: Tosato, Lucrezia, et al.
Published: (2025)
Medical Image Spatial Grounding with Semantic Sampling
by: Yu, Andrew Seohwan, et al.
Published: (2026)
by: Yu, Andrew Seohwan, et al.
Published: (2026)
Structured Video-Language Modeling with Temporal Grouping and Spatial Grounding
by: Xiong, Yuanhao, et al.
Published: (2023)
by: Xiong, Yuanhao, et al.
Published: (2023)
Grounded 3D-Aware Spatial Vision-Language Modeling
by: Cheng, An-Chieh, et al.
Published: (2026)
by: Cheng, An-Chieh, et al.
Published: (2026)
T2SGrid: Temporal-to-Spatial Gridification for Video Temporal Grounding
by: Guo, Chaohong, et al.
Published: (2026)
by: Guo, Chaohong, et al.
Published: (2026)
Artiverse: A Diverse and Physically Grounded Dataset for Articulated Objects
by: Iliash, Denys, et al.
Published: (2026)
by: Iliash, Denys, et al.
Published: (2026)
Segment Anyword: Mask Prompt Inversion for Open-Set Grounded Segmentation
by: Liu, Zhihua, et al.
Published: (2025)
by: Liu, Zhihua, et al.
Published: (2025)
Similar Items
-
RS-OVC: Open-Vocabulary Counting for Remote-Sensing Data
by: Shor, Tamir, et al.
Published: (2026) -
RSRCC: A Remote Sensing Regional Change Comprehension Benchmark Constructed via Retrieval-Augmented Best-of-N Ranking
by: Kazoom, Roie, et al.
Published: (2026) -
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization
by: Barzilai, Aviad, et al.
Published: (2025) -
On-the-Fly OVD Adaptation with FLAME: Few-shot Localization via Active Marginal-Samples Exploration
by: Refael, Yehonathan, et al.
Published: (2025) -
Zero-Shot Multi-Spectral Learning: Reimagining a Generalist Multimodal Gemini 2.5 Model for Remote Sensing Applications
by: Mallya, Ganesh, et al.
Published: (2025)