WHU-STree: A Multi-modal Benchmark Dataset for Street Tree Inventory
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Ruifei, Chen, Zhe, Fan, Wen, Long, Chen, Xiao, Huijuan, Zeng, Yelu, Dong, Zhen, Yang, Bisheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
by: Long, Chen, et al.
Published: (2026)
by: Long, Chen, et al.
Published: (2026)
WHU-Synthetic: A Synthetic Perception Dataset for 3-D Multitask Model Research
by: Zhou, Jiahao, et al.
Published: (2024)
by: Zhou, Jiahao, et al.
Published: (2024)
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
by: Liu, Chong, et al.
Published: (2026)
by: Liu, Chong, et al.
Published: (2026)
Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion
by: Xu, Hang, et al.
Published: (2024)
by: Xu, Hang, et al.
Published: (2024)
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
by: Chen, Jiabin, et al.
Published: (2025)
by: Chen, Jiabin, et al.
Published: (2025)
OSMLoc: Single Image-Based Visual Localization in OpenStreetMap with Fused Geometric and Semantic Guidance
by: Liao, Youqi, et al.
Published: (2024)
by: Liao, Youqi, et al.
Published: (2024)
TOL: Textual Localization with OpenStreetMap
by: Liao, Youqi, et al.
Published: (2026)
by: Liao, Youqi, et al.
Published: (2026)
LifelongPR: Lifelong point cloud place recognition based on sample replay and prompt learning
by: Zou, Xianghong, et al.
Published: (2025)
by: Zou, Xianghong, et al.
Published: (2025)
DPG-CD: Depth-Prior-Guided Cross-Modal Joint 2D-3D Change Detection
by: Zhang, Luqi, et al.
Published: (2026)
by: Zhang, Luqi, et al.
Published: (2026)
WHU-PCPR: A cross-platform heterogeneous point cloud dataset for place recognition in complex urban scenes
by: Zou, Xianghong, et al.
Published: (2026)
by: Zou, Xianghong, et al.
Published: (2026)
Unleashing the Capabilities of Large Vision-Language Models for Intelligent Perception of Roadside Infrastructure
by: Fu, Luxuan, et al.
Published: (2026)
by: Fu, Luxuan, et al.
Published: (2026)
DeepAAT: Deep Automated Aerial Triangulation for Fast UAV-based Mapping
by: Chen, Zequan, et al.
Published: (2024)
by: Chen, Zequan, et al.
Published: (2024)
ME-CPT: Multi-Task Enhanced Cross-Temporal Point Transformer for Urban 3D Change Detection
by: Zhang, Luqi, et al.
Published: (2025)
by: Zhang, Luqi, et al.
Published: (2025)
Aerial-ground Cross-modal Localization: Dataset, Ground-truth, and Benchmark
by: Yang, Yandi, et al.
Published: (2025)
by: Yang, Yandi, et al.
Published: (2025)
Fine-grained Action Analysis: A Multi-modality and Multi-task Dataset of Figure Skating
by: Liu, Sheng-Lan, et al.
Published: (2023)
by: Liu, Sheng-Lan, et al.
Published: (2023)
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting
by: Peng, Yuning, et al.
Published: (2024)
by: Peng, Yuning, et al.
Published: (2024)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
by: Li, Zeju, et al.
Published: (2024)
by: Li, Zeju, et al.
Published: (2024)
JL1-CD: A New Benchmark for Remote Sensing Change Detection and a Robust Multi-Teacher Knowledge Distillation Framework
by: Liu, Ziyuan, et al.
Published: (2025)
by: Liu, Ziyuan, et al.
Published: (2025)
SMART-Ship: A Comprehensive Synchronized Multi-modal Aligned Remote Sensing Targets Dataset and Benchmark for Berthed Ships Analysis
by: Fan, Chen-Chen, et al.
Published: (2025)
by: Fan, Chen-Chen, et al.
Published: (2025)
Personalized Cell Segmentation: Benchmark and Framework for Reference-Guided Cell Type Segmentation
by: Wang, Bisheng, et al.
Published: (2026)
by: Wang, Bisheng, et al.
Published: (2026)
SaliencyI2PLoc: saliency-guided image-point cloud localization using contrastive learning
by: Li, Yuhao, et al.
Published: (2024)
by: Li, Yuhao, et al.
Published: (2024)
CRAG-MM: Multi-modal Multi-turn Comprehensive RAG Benchmark
by: Wang, Jiaqi, et al.
Published: (2025)
by: Wang, Jiaqi, et al.
Published: (2025)
StreetTree: A Large-Scale Global Benchmark for Fine-Grained Tree Species Classification
by: Li, Jiapeng, et al.
Published: (2026)
by: Li, Jiapeng, et al.
Published: (2026)
Oracle Bone Inscriptions Multi-modal Dataset
by: Li, Bang, et al.
Published: (2024)
by: Li, Bang, et al.
Published: (2024)
M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought
by: Chen, Qiguang, et al.
Published: (2024)
by: Chen, Qiguang, et al.
Published: (2024)
SS3DM: Benchmarking Street-View Surface Reconstruction with a Synthetic 3D Mesh Dataset
by: Hu, Yubin, et al.
Published: (2024)
by: Hu, Yubin, et al.
Published: (2024)
VistaDream: Sampling multiview consistent images for single-view scene reconstruction
by: Wang, Haiping, et al.
Published: (2024)
by: Wang, Haiping, et al.
Published: (2024)
TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types
by: Chen, Jiankang, et al.
Published: (2025)
by: Chen, Jiankang, et al.
Published: (2025)
Building Floor Number Estimation from Crowdsourced Street-Level Images: Munich Dataset and Baseline Method
by: Sun, Yao, et al.
Published: (2025)
by: Sun, Yao, et al.
Published: (2025)
VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models
by: Xu, Weiye, et al.
Published: (2025)
by: Xu, Weiye, et al.
Published: (2025)
StyledStreets: Multi-style Street Simulator with Spatial and Temporal Consistency
by: Chen, Yuyin, et al.
Published: (2025)
by: Chen, Yuyin, et al.
Published: (2025)
Reliable-loc: Robust sequential LiDAR global localization in large-scale street scenes based on verifiable cues
by: Zou, Xianghong, et al.
Published: (2024)
by: Zou, Xianghong, et al.
Published: (2024)
MMM-RS: A Multi-modal, Multi-GSD, Multi-scene Remote Sensing Dataset and Benchmark for Text-to-Image Generation
by: Luo, Jialin, et al.
Published: (2024)
by: Luo, Jialin, et al.
Published: (2024)
Accurate and Efficient Urban Street Tree Inventory with Deep Learning on Mobile Phone Imagery
by: Khan, Asim, et al.
Published: (2024)
by: Khan, Asim, et al.
Published: (2024)
Unified Multi-modal Diagnostic Framework with Reconstruction Pre-training and Heterogeneity-combat Tuning
by: Zhang, Yupei, et al.
Published: (2024)
by: Zhang, Yupei, et al.
Published: (2024)
Mobile-Seed: Joint Semantic Segmentation and Boundary Detection for Mobile Robots
by: Liao, Youqi, et al.
Published: (2023)
by: Liao, Youqi, et al.
Published: (2023)
LiMT: A Multi-task Liver Image Benchmark Dataset
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
Multi-modal Data Spectrum: Multi-modal Datasets are Multi-dimensional
by: Madaan, Divyam, et al.
Published: (2025)
by: Madaan, Divyam, et al.
Published: (2025)
CytoCrowd: A Multi-Annotator Benchmark Dataset for Cytology Image Analysis
by: Si, Yonghao, et al.
Published: (2026)
by: Si, Yonghao, et al.
Published: (2026)
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
by: Wu, Siwei, et al.
Published: (2024)
by: Wu, Siwei, et al.
Published: (2024)
Similar Items
-
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
by: Long, Chen, et al.
Published: (2026) -
WHU-Synthetic: A Synthetic Perception Dataset for 3-D Multitask Model Research
by: Zhou, Jiahao, et al.
Published: (2024) -
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
by: Liu, Chong, et al.
Published: (2026) -
Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion
by: Xu, Hang, et al.
Published: (2024) -
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
by: Chen, Jiabin, et al.
Published: (2025)