Learning Street View Representations with Spatiotemporal Contrast
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yong, Huang, Yingjing, Mai, Gengchen, Zhang, Fan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cross-View Geolocalization and Disaster Mapping with Street-View and VHR Satellite Imagery: A Case Study of Hurricane IAN
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Examining the Commitments and Difficulties Inherent in Multimodal Foundation Models for Street View Imagery
di: Yang, Zhenyuan, et al.
Pubblicazione: (2024)
di: Yang, Zhenyuan, et al.
Pubblicazione: (2024)
GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations
di: Liu, Zeping, et al.
Pubblicazione: (2025)
di: Liu, Zeping, et al.
Pubblicazione: (2025)
Unsupervised Urban Land Use Mapping with Street View Contrastive Clustering and a Geographical Prior
di: Che, Lin, et al.
Pubblicazione: (2025)
di: Che, Lin, et al.
Pubblicazione: (2025)
StreetTree: A Large-Scale Global Benchmark for Fine-Grained Tree Species Classification
di: Li, Jiapeng, et al.
Pubblicazione: (2026)
di: Li, Jiapeng, et al.
Pubblicazione: (2026)
A Contrastive Learning Framework Empowered by Attention-based Feature Adaptation for Street-View Image Classification
di: You, Qi, et al.
Pubblicazione: (2026)
di: You, Qi, et al.
Pubblicazione: (2026)
Combining Deep Learning and Street View Imagery to Map Smallholder Crop Types
di: Soler, Jordi Laguarta, et al.
Pubblicazione: (2023)
di: Soler, Jordi Laguarta, et al.
Pubblicazione: (2023)
CECT-Mamba: a Hierarchical Contrast-enhanced-aware Model for Pancreatic Tumor Subtyping from Multi-phase CECT
di: Gong, Zhifang, et al.
Pubblicazione: (2025)
di: Gong, Zhifang, et al.
Pubblicazione: (2025)
Text2Seg: Remote Sensing Image Semantic Segmentation via Text-Guided Visual Foundation Models
di: Zhang, Jielu, et al.
Pubblicazione: (2023)
di: Zhang, Jielu, et al.
Pubblicazione: (2023)
Joint Imaging-ROI Representation Learning via Cross-View Contrastive Alignment for Brain Disorder Classification
di: Liang, Wei, et al.
Pubblicazione: (2026)
di: Liang, Wei, et al.
Pubblicazione: (2026)
TorchSpatial: A Location Encoding Framework and Benchmark for Spatial Representation Learning
di: Wu, Nemin, et al.
Pubblicazione: (2024)
di: Wu, Nemin, et al.
Pubblicazione: (2024)
MagicDrive: Street View Generation with Diverse 3D Geometry Control
di: Gao, Ruiyuan, et al.
Pubblicazione: (2023)
di: Gao, Ruiyuan, et al.
Pubblicazione: (2023)
Bird Eye-View to Street-View: A Survey
di: Bajbaa, Khawlah, et al.
Pubblicazione: (2024)
di: Bajbaa, Khawlah, et al.
Pubblicazione: (2024)
Is Contrastive Distillation Enough for Learning Comprehensive 3D Representations?
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
di: Zhou, Zhongliang, et al.
Pubblicazione: (2024)
di: Zhou, Zhongliang, et al.
Pubblicazione: (2024)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
di: Gao, Ruiyuan, et al.
Pubblicazione: (2024)
di: Gao, Ruiyuan, et al.
Pubblicazione: (2024)
OpenStreetView-5M: The Many Roads to Global Visual Geolocation
di: Astruc, Guillaume, et al.
Pubblicazione: (2024)
di: Astruc, Guillaume, et al.
Pubblicazione: (2024)
An Integrated Causal Inference Framework for Traffic Safety Modeling with Semantic Street-View Visual Features
di: Sun, Lishan, et al.
Pubblicazione: (2026)
di: Sun, Lishan, et al.
Pubblicazione: (2026)
Cluster Contrast for Unsupervised Visual Representation Learning
di: Giakoumoglou, Nikolaos, et al.
Pubblicazione: (2025)
di: Giakoumoglou, Nikolaos, et al.
Pubblicazione: (2025)
BuildingView: Constructing Urban Building Exteriors Databases with Street View Imagery and Multimodal Large Language Mode
di: Li, Zongrong, et al.
Pubblicazione: (2024)
di: Li, Zongrong, et al.
Pubblicazione: (2024)
Paved or unpaved? A Deep Learning derived Road Surface Global Dataset from Mapillary Street-View Imagery
di: Randhawa, Sukanya, et al.
Pubblicazione: (2024)
di: Randhawa, Sukanya, et al.
Pubblicazione: (2024)
m2sv: A Scalable Benchmark for Map-to-Street-View Spatial Reasoning
di: Shin, Yosub, et al.
Pubblicazione: (2026)
di: Shin, Yosub, et al.
Pubblicazione: (2026)
Learning Content-Aware Multi-Modal Joint Input Pruning via Bird's-Eye-View Representation
di: Li, Yuxin, et al.
Pubblicazione: (2024)
di: Li, Yuxin, et al.
Pubblicazione: (2024)
Learning Depth from Past Selves: Self-Evolution Contrast for Robust Depth Estimation
di: Cao, Jing, et al.
Pubblicazione: (2025)
di: Cao, Jing, et al.
Pubblicazione: (2025)
Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery
di: Yao, Siyuan, et al.
Pubblicazione: (2026)
di: Yao, Siyuan, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
Adaptive Disentangled Representation Learning for Incomplete Multi-View Multi-Label Classification
di: Li, Quanjiang, et al.
Pubblicazione: (2026)
di: Li, Quanjiang, et al.
Pubblicazione: (2026)
DreamDrive: Generative 4D Scene Modeling from Street View Images
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
From Street to Orbit: Training-Free Cross-View Retrieval via Location Semantics and LLM Guidance
di: Min, Jeongho, et al.
Pubblicazione: (2025)
di: Min, Jeongho, et al.
Pubblicazione: (2025)
From Bird's-Eye to Street View: Crafting Diverse and Condition-Aligned Images with Latent Diffusion Model
di: Xu, Xiaojie, et al.
Pubblicazione: (2024)
di: Xu, Xiaojie, et al.
Pubblicazione: (2024)
Satellite-to-Street: Synthesizing Post-Disaster Views from Satellite Imagery via Generative Vision Models
di: Yang, Yifan, et al.
Pubblicazione: (2026)
di: Yang, Yifan, et al.
Pubblicazione: (2026)
CHOSEN: Contrastive Hypothesis Selection for Multi-View Depth Refinement
di: Qiu, Di, et al.
Pubblicazione: (2024)
di: Qiu, Di, et al.
Pubblicazione: (2024)
VELMA: Verbalization Embodiment of LLM Agents for Vision and Language Navigation in Street View
di: Schumann, Raphael, et al.
Pubblicazione: (2023)
di: Schumann, Raphael, et al.
Pubblicazione: (2023)
Revolutionizing Precise Low Back Pain Diagnosis via Contrastive Learning
di: Le, Thanh Binh, et al.
Pubblicazione: (2025)
di: Le, Thanh Binh, et al.
Pubblicazione: (2025)
MR-CLIP: Efficient Metadata-Guided Learning of MRI Contrast Representations
di: Avci, Mehmet Yigit, et al.
Pubblicazione: (2025)
di: Avci, Mehmet Yigit, et al.
Pubblicazione: (2025)
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
di: Luo, Mingshuang, et al.
Pubblicazione: (2026)
di: Luo, Mingshuang, et al.
Pubblicazione: (2026)
OG-Gaussian: Occupancy Based Street Gaussians for Autonomous Driving
di: Shen, Yedong, et al.
Pubblicazione: (2025)
di: Shen, Yedong, et al.
Pubblicazione: (2025)
MCRL4OR: Multimodal Contrastive Representation Learning for Off-Road Environmental Perception
di: Yang, Yi, et al.
Pubblicazione: (2025)
di: Yang, Yi, et al.
Pubblicazione: (2025)
Uncertainty Quantification via Hölder Divergence for Multi-View Representation Learning
di: Zhang, Yan, et al.
Pubblicazione: (2024)
di: Zhang, Yan, et al.
Pubblicazione: (2024)
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
di: Feng, Yu, et al.
Pubblicazione: (2024)
di: Feng, Yu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cross-View Geolocalization and Disaster Mapping with Street-View and VHR Satellite Imagery: A Case Study of Hurricane IAN
di: Li, Hao, et al.
Pubblicazione: (2024) -
Examining the Commitments and Difficulties Inherent in Multimodal Foundation Models for Street View Imagery
di: Yang, Zhenyuan, et al.
Pubblicazione: (2024) -
GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations
di: Liu, Zeping, et al.
Pubblicazione: (2025) -
Unsupervised Urban Land Use Mapping with Street View Contrastive Clustering and a Geographical Prior
di: Che, Lin, et al.
Pubblicazione: (2025) -
StreetTree: A Large-Scale Global Benchmark for Fine-Grained Tree Species Classification
di: Li, Jiapeng, et al.
Pubblicazione: (2026)