Spatially-Weighted CLIP for Street-View Geo-localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Ting, Li, Fengjiao, Chen, Chunsong, Huang, Haoling, Chen, Yiping, Wu, Meiliu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic4Safety: Causal Insights from Zero-shot Street View Imagery Segmentation for Urban Road Safety
von: Chen, Huan, et al.
Veröffentlicht: (2025)
von: Chen, Huan, et al.
Veröffentlicht: (2025)
Cross-View Geo-Localization with Street-View and VHR Satellite Imagery in Decentrality Settings
von: Xia, Panwang, et al.
Veröffentlicht: (2024)
von: Xia, Panwang, et al.
Veröffentlicht: (2024)
Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic Segmentation
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
Text2Street: Controllable Text-to-image Generation for Street Views
von: Su, Jinming, et al.
Veröffentlicht: (2024)
von: Su, Jinming, et al.
Veröffentlicht: (2024)
Leveraging Depth and Language for Open-Vocabulary Domain-Generalized Semantic Segmentation
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
GeoReasoner: Geo-localization with Reasoning in Street Views using a Large Vision-Language Model
von: Li, Ling, et al.
Veröffentlicht: (2024)
von: Li, Ling, et al.
Veröffentlicht: (2024)
DamageArbiter: A CLIP-Enhanced Multimodal Arbitration Framework for Hurricane Damage Assessment from Street-View Imagery
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
StyledStreets: Multi-style Street Simulator with Spatial and Temporal Consistency
von: Chen, Yuyin, et al.
Veröffentlicht: (2025)
von: Chen, Yuyin, et al.
Veröffentlicht: (2025)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
von: Li, Weijia, et al.
Veröffentlicht: (2024)
von: Li, Weijia, et al.
Veröffentlicht: (2024)
CityPulse: Fine-Grained Assessment of Urban Change with Street View Time Series
von: Huang, Tianyuan, et al.
Veröffentlicht: (2024)
von: Huang, Tianyuan, et al.
Veröffentlicht: (2024)
Learning Street View Representations with Spatiotemporal Contrast
von: Li, Yong, et al.
Veröffentlicht: (2025)
von: Li, Yong, et al.
Veröffentlicht: (2025)
StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
von: Yan, Yunzhi, et al.
Veröffentlicht: (2024)
von: Yan, Yunzhi, et al.
Veröffentlicht: (2024)
Time2General: Learning Spatiotemporal Invariant Representations for Domain-Generalization Video Semantic Segmentation
von: Chen, Siyu, et al.
Veröffentlicht: (2026)
von: Chen, Siyu, et al.
Veröffentlicht: (2026)
SinGeo: Unlock Single Model's Potential for Robust Cross-View Geo-Localization
von: Chen, Yang, et al.
Veröffentlicht: (2026)
von: Chen, Yang, et al.
Veröffentlicht: (2026)
Seeing through Satellite Images at Street Views
von: Qian, Ming, et al.
Veröffentlicht: (2025)
von: Qian, Ming, et al.
Veröffentlicht: (2025)
Artifacts of Idiosyncracy in Global Street View Data
von: Alpherts, Tim, et al.
Veröffentlicht: (2025)
von: Alpherts, Tim, et al.
Veröffentlicht: (2025)
ViSE: A Systematic Approach to Vision-Only Street-View Extrapolation
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2025)
SVIA: A Street View Image Anonymization Framework for Self-Driving Applications
von: Liu, Dongyu, et al.
Veröffentlicht: (2025)
von: Liu, Dongyu, et al.
Veröffentlicht: (2025)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
von: Song, Dan, et al.
Veröffentlicht: (2023)
von: Song, Dan, et al.
Veröffentlicht: (2023)
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images
von: Lee, Han-Hung, et al.
Veröffentlicht: (2024)
von: Lee, Han-Hung, et al.
Veröffentlicht: (2024)
MOGeo: Beyond One-to-One Cross-View Object Geo-localization
von: Lv, Bo, et al.
Veröffentlicht: (2026)
von: Lv, Bo, et al.
Veröffentlicht: (2026)
CLIP-GS: CLIP-Informed Gaussian Splatting for View-Consistent 3D Indoor Semantic Understanding
von: Liao, Guibiao, et al.
Veröffentlicht: (2024)
von: Liao, Guibiao, et al.
Veröffentlicht: (2024)
CLIP3D-AD: Extending CLIP for 3D Few-Shot Anomaly Detection with Multi-View Images Generation
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
Multi-View People Detection in Large Scenes via Supervised View-Wise Contribution Weighting
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
von: Lin, Yuxuan, et al.
Veröffentlicht: (2025)
von: Lin, Yuxuan, et al.
Veröffentlicht: (2025)
Scene4U: Hierarchical Layered 3D Scene Reconstruction from Single Panoramic Image for Your Immerse Exploration
von: Huang, Zilong, et al.
Veröffentlicht: (2025)
von: Huang, Zilong, et al.
Veröffentlicht: (2025)
Street-View Image Generation from a Bird's-Eye View Layout
von: Swerdlow, Alexander, et al.
Veröffentlicht: (2023)
von: Swerdlow, Alexander, et al.
Veröffentlicht: (2023)
Where am I? Cross-View Geo-localization with Natural Language Descriptions
von: Ye, Junyan, et al.
Veröffentlicht: (2024)
von: Ye, Junyan, et al.
Veröffentlicht: (2024)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
SyntheOcc: Synthesize Geometric-Controlled Street View Images through 3D Semantic MPIs
von: Li, Leheng, et al.
Veröffentlicht: (2024)
von: Li, Leheng, et al.
Veröffentlicht: (2024)
SGS-3D: High-Fidelity 3D Instance Segmentation via Reliable Semantic Mask Splitting and Growing
von: Wang, Chaolei, et al.
Veröffentlicht: (2025)
von: Wang, Chaolei, et al.
Veröffentlicht: (2025)
Breaking the Limits of Open-Weight CLIP: An Optimization Framework for Self-supervised Fine-tuning of CLIP
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
m2sv: A Scalable Benchmark for Map-to-Street-View Spatial Reasoning
von: Shin, Yosub, et al.
Veröffentlicht: (2026)
von: Shin, Yosub, et al.
Veröffentlicht: (2026)
Robust Light-Weight Facial Affective Behavior Recognition with CLIP
von: Lin, Li, et al.
Veröffentlicht: (2024)
von: Lin, Li, et al.
Veröffentlicht: (2024)
Refining CLIP's Spatial Awareness: A Visual-Centric Perspective
von: Qiu, Congpei, et al.
Veröffentlicht: (2025)
von: Qiu, Congpei, et al.
Veröffentlicht: (2025)
SGD: Street View Synthesis with Gaussian Splatting and Diffusion Prior
von: Yu, Zhongrui, et al.
Veröffentlicht: (2024)
von: Yu, Zhongrui, et al.
Veröffentlicht: (2024)
Semantic-Aware Label Placement for Augmented Reality in Street View
von: Jia, Jianqing, et al.
Veröffentlicht: (2019)
von: Jia, Jianqing, et al.
Veröffentlicht: (2019)
Eyes on the Streets: Leveraging Street-Level Imaging to Model Urban Crime Dynamics
von: Qi, Zhixuan, et al.
Veröffentlicht: (2024)
von: Qi, Zhixuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Semantic4Safety: Causal Insights from Zero-shot Street View Imagery Segmentation for Urban Road Safety
von: Chen, Huan, et al.
Veröffentlicht: (2025) -
Cross-View Geo-Localization with Street-View and VHR Satellite Imagery in Decentrality Settings
von: Xia, Panwang, et al.
Veröffentlicht: (2024) -
Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic Segmentation
von: Chen, Siyu, et al.
Veröffentlicht: (2025) -
Text2Street: Controllable Text-to-image Generation for Street Views
von: Su, Jinming, et al.
Veröffentlicht: (2024) -
Leveraging Depth and Language for Open-Vocabulary Domain-Generalized Semantic Segmentation
von: Chen, Siyu, et al.
Veröffentlicht: (2025)