MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Kristoffersen, Oskar, Sánchez, Alba Reinders, Hannemose, Morten Rieger, Dahl, Anders Bjorholm, Papadopoulos, Dim P. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations
by: Guo, Changlu, et al.
Published: (2026)
by: Guo, Changlu, et al.
Published: (2026)
SA-UNetv2: Rethinking Spatial Attention U-Net for Retinal Vessel Segmentation
by: Guo, Changlu, et al.
Published: (2025)
by: Guo, Changlu, et al.
Published: (2025)
Diffusion Based Ambiguous Image Segmentation
by: Christensen, Jakob Lønborg, et al.
Published: (2025)
by: Christensen, Jakob Lønborg, et al.
Published: (2025)
Towards Agnostic and Holistic Universal Image Segmentation with Bit Diffusion
by: Christensen, Jakob Lønborg, et al.
Published: (2026)
by: Christensen, Jakob Lønborg, et al.
Published: (2026)
Two Views Are Better than One: Monocular 3D Pose Estimation with Multiview Consistency
by: Ingwersen, Christian Keilstrup, et al.
Published: (2023)
by: Ingwersen, Christian Keilstrup, et al.
Published: (2023)
Rethinking Uncertainty Quantification and Entanglement in Image Segmentation
by: Christensen, Jakob Lønborg, et al.
Published: (2026)
by: Christensen, Jakob Lønborg, et al.
Published: (2026)
Towards High-Quality Image Segmentation: Improving Topology Accuracy by Penalizing Neighbor Pixels
by: Valverde, Juan Miguel, et al.
Published: (2026)
by: Valverde, Juan Miguel, et al.
Published: (2026)
Med-Art: Diffusion Transformer for 2D Medical Text-to-Image Generation
by: Guo, Changlu, et al.
Published: (2025)
by: Guo, Changlu, et al.
Published: (2025)
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training
by: Lu, Kaixuan, et al.
Published: (2025)
by: Lu, Kaixuan, et al.
Published: (2025)
AutoQ-VIS: Improving Unsupervised Video Instance Segmentation via Automatic Quality Assessment
by: Lu, Kaixuan, et al.
Published: (2025)
by: Lu, Kaixuan, et al.
Published: (2025)
HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement
by: Schouten, Marco, et al.
Published: (2026)
by: Schouten, Marco, et al.
Published: (2026)
TopoMortar: A dataset to evaluate image segmentation methods focused on topology accuracy
by: Valverde, Juan Miguel, et al.
Published: (2025)
by: Valverde, Juan Miguel, et al.
Published: (2025)
Fast Sphericity and Roundness approximation in 2D and 3D using Local Thickness
by: Pieta, Pawel Tomasz, et al.
Published: (2025)
by: Pieta, Pawel Tomasz, et al.
Published: (2025)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
by: Delatolas, Thanos, et al.
Published: (2025)
by: Delatolas, Thanos, et al.
Published: (2025)
Visual Autoregressive Models Beat Diffusion Models on Inference Time Scaling
by: Riise, Erik, et al.
Published: (2025)
by: Riise, Erik, et al.
Published: (2025)
Fast and Compact Graph Cuts for the Boykov-Kolmogorov Algorithm
by: Mikkelstrup, Christian Møller, et al.
Published: (2026)
by: Mikkelstrup, Christian Møller, et al.
Published: (2026)
Feature-Centered First Order Structure Tensor Scale-Space in 2D and 3D
by: Pieta, Pawel Tomasz, et al.
Published: (2024)
by: Pieta, Pawel Tomasz, et al.
Published: (2024)
A General Purpose Spectral Foundational Model for Both Proximal and Remote Sensing Spectral Imaging
by: Laprade, William Michael, et al.
Published: (2025)
by: Laprade, William Michael, et al.
Published: (2025)
Visual Context-Aware Person Fall Detection
by: Nagaj, Aleksander, et al.
Published: (2024)
by: Nagaj, Aleksander, et al.
Published: (2024)
Weak Cube R-CNN: Weakly Supervised 3D Detection using only 2D Bounding Boxes
by: Hansen, Andreas Lau, et al.
Published: (2025)
by: Hansen, Andreas Lau, et al.
Published: (2025)
Efficient Test-Time Scaling for Small Vision-Language Models
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
VoDaSuRe: A Large-Scale Dataset Revealing Domain Shift in Volumetric Super-Resolution
by: Høeg, August Leander, et al.
Published: (2026)
by: Høeg, August Leander, et al.
Published: (2026)
POEM: Precise Object-level Editing via MLLM control
by: Schouten, Marco, et al.
Published: (2025)
by: Schouten, Marco, et al.
Published: (2025)
Latent Directions: A Simple Pathway to Bias Mitigation in Generative AI
by: Olmos, Carolina Lopez, et al.
Published: (2024)
by: Olmos, Carolina Lopez, et al.
Published: (2024)
ReLumix: Extending Image Relighting to Video via Video Diffusion Models
by: Wang, Lezhong, et al.
Published: (2025)
by: Wang, Lezhong, et al.
Published: (2025)
Multi-modal data generation with a deep metric variational autoencoder
by: Sundgaard, Josefine Vilsbøll, et al.
Published: (2022)
by: Sundgaard, Josefine Vilsbøll, et al.
Published: (2022)
BugNIST -- a Large Volumetric Dataset for Object Detection under Domain Shift
by: Jensen, Patrick Møller, et al.
Published: (2023)
by: Jensen, Patrick Møller, et al.
Published: (2023)
pix2pockets: Shot Suggestions in 8-Ball Pool from a Single Image in the Wild
by: Schiøtt, Jonas Myhre, et al.
Published: (2025)
by: Schiøtt, Jonas Myhre, et al.
Published: (2025)
MozzaVID: Mozzarella Volumetric Image Dataset
by: Pieta, Pawel Tomasz, et al.
Published: (2024)
by: Pieta, Pawel Tomasz, et al.
Published: (2024)
Disconnect to Connect: A Data Augmentation Method for Improving Topology Accuracy in Image Segmentation
by: Valverde, Juan Miguel, et al.
Published: (2025)
by: Valverde, Juan Miguel, et al.
Published: (2025)
SMGeo: Cross-View Object Geo-Localization with Grid-Level Mixture-of-Experts
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
Learning to Build Shapes by Extrusion
by: Christiansen, Thor Vestergaard, et al.
Published: (2026)
by: Christiansen, Thor Vestergaard, et al.
Published: (2026)
CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark
by: Wang, Wei, et al.
Published: (2026)
by: Wang, Wei, et al.
Published: (2026)
The TYC Dataset for Understanding Instance-Level Semantics and Motions of Cells in Microstructures
by: Reich, Christoph, et al.
Published: (2023)
by: Reich, Christoph, et al.
Published: (2023)
Learning Multi-View Spatial Reasoning from Cross-View Relations
by: Jeong, Suchae, et al.
Published: (2026)
by: Jeong, Suchae, et al.
Published: (2026)
InstanceV: Instance-Level Video Generation
by: Chen, Yuheng, et al.
Published: (2025)
by: Chen, Yuheng, et al.
Published: (2025)
Multi-Level Embedding and Alignment Network with Consistency and Invariance Learning for Cross-View Geo-Localization
by: Chen, Zhongwei, et al.
Published: (2024)
by: Chen, Zhongwei, et al.
Published: (2024)
MRGeo: Robust Cross-View Geo-Localization of Corrupted Images via Spatial and Channel Feature Enhancement
by: Wu, Le, et al.
Published: (2026)
by: Wu, Le, et al.
Published: (2026)
Materialist: Physically Based Editing Using Single-Image Inverse Rendering
by: Wang, Lezhong, et al.
Published: (2025)
by: Wang, Lezhong, et al.
Published: (2025)
VAGeo: View-specific Attention for Cross-View Object Geo-Localization
by: Li, Zhongyang, et al.
Published: (2025)
by: Li, Zhongyang, et al.
Published: (2025)
Similar Items
-
MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations
by: Guo, Changlu, et al.
Published: (2026) -
SA-UNetv2: Rethinking Spatial Attention U-Net for Retinal Vessel Segmentation
by: Guo, Changlu, et al.
Published: (2025) -
Diffusion Based Ambiguous Image Segmentation
by: Christensen, Jakob Lønborg, et al.
Published: (2025) -
Towards Agnostic and Holistic Universal Image Segmentation with Bit Diffusion
by: Christensen, Jakob Lønborg, et al.
Published: (2026) -
Two Views Are Better than One: Monocular 3D Pose Estimation with Multiview Consistency
by: Ingwersen, Christian Keilstrup, et al.
Published: (2023)