GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Zixuan, Zhang, Jing, Wang, Di, Zhou, Zidie, Liu, Wenbin, Guo, Haonan, Wang, En, Du, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Residual Diffusion Bridge Model for Image Restoration
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
von: Wang, Wei, et al.
Veröffentlicht: (2026)
von: Wang, Wei, et al.
Veröffentlicht: (2026)
Universal Pansharpening Foundation Model
von: Wang, Hebaixu, et al.
Veröffentlicht: (2026)
von: Wang, Hebaixu, et al.
Veröffentlicht: (2026)
ViewBridge:Revisiting Cross-View Localization from Image Matching
von: Xia, Panwang, et al.
Veröffentlicht: (2025)
von: Xia, Panwang, et al.
Veröffentlicht: (2025)
UniABG: Unified Adversarial View Bridging and Graph Correspondence for Unsupervised Cross-View Geo-Localization
von: Chen, Cuiqun, et al.
Veröffentlicht: (2025)
von: Chen, Cuiqun, et al.
Veröffentlicht: (2025)
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
TranX-Adapter: Bridging Artifacts and Semantics within MLLMs for Robust AI-generated Image Detection
von: Wang, Wenbin, et al.
Veröffentlicht: (2026)
von: Wang, Wenbin, et al.
Veröffentlicht: (2026)
Robust Drone-View Geo-Localization via Content-Viewpoint Disentanglement
von: Li, Ke, et al.
Veröffentlicht: (2025)
von: Li, Ke, et al.
Veröffentlicht: (2025)
Scale-Aware UAV-to-Satellite Cross-View Geo-Localization: A Semantic Geometric Approach
von: Ye, Yibin, et al.
Veröffentlicht: (2026)
von: Ye, Yibin, et al.
Veröffentlicht: (2026)
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining
von: Wang, Di, et al.
Veröffentlicht: (2024)
von: Wang, Di, et al.
Veröffentlicht: (2024)
Panoptic Captioning: An Equivalence Bridge for Image and Text
von: Lin, Kun-Yu, et al.
Veröffentlicht: (2025)
von: Lin, Kun-Yu, et al.
Veröffentlicht: (2025)
Cross-View Image Set Geo-Localization
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
GeoZero: Incentivizing Reasoning from Scratch on Geospatial Scenes
von: Wang, Di, et al.
Veröffentlicht: (2025)
von: Wang, Di, et al.
Veröffentlicht: (2025)
Efficient Image-to-Image Schrödinger Bridge for CT Field of View Extension
von: Li, Zhenhao, et al.
Veröffentlicht: (2025)
von: Li, Zhenhao, et al.
Veröffentlicht: (2025)
UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes
von: Ni, Shuo, et al.
Veröffentlicht: (2025)
von: Ni, Shuo, et al.
Veröffentlicht: (2025)
Enhancing Cross-View Geo-Localization Generalization via Global-Local Consistency and Geometric Equivariance
von: Wang, Xiaowei, et al.
Veröffentlicht: (2025)
von: Wang, Xiaowei, et al.
Veröffentlicht: (2025)
Bridging Semantic Logic Gaps: A Cognition Inspired Multimodal Boundary Preserving Network for Image Manipulation Localization
von: Li, Songlin, et al.
Veröffentlicht: (2025)
von: Li, Songlin, et al.
Veröffentlicht: (2025)
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
von: Luo, Zhiming, et al.
Veröffentlicht: (2026)
von: Luo, Zhiming, et al.
Veröffentlicht: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
Inconsistency-aware Multimodal Schrödinger Bridge for Deepfake Localization
von: Xiong, Jiayu, et al.
Veröffentlicht: (2026)
von: Xiong, Jiayu, et al.
Veröffentlicht: (2026)
TrAME: Trajectory-Anchored Multi-View Editing for Text-Guided 3D Gaussian Splatting Manipulation
von: Luo, Chaofan, et al.
Veröffentlicht: (2024)
von: Luo, Chaofan, et al.
Veröffentlicht: (2024)
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
von: He, Haibin, et al.
Veröffentlicht: (2026)
von: He, Haibin, et al.
Veröffentlicht: (2026)
Semantic Anchoring for Robust Personalization in Text-to-Image Diffusion Models
von: Yang, Seoyun, et al.
Veröffentlicht: (2025)
von: Yang, Seoyun, et al.
Veröffentlicht: (2025)
UniV2D: Bridging Visual Restoration and Semantic Perception for Underwater Salient Object Detection
von: Chang, Laibin, et al.
Veröffentlicht: (2026)
von: Chang, Laibin, et al.
Veröffentlicht: (2026)
TiMo: Spatiotemporal Foundation Model for Satellite Image Time Series
von: Qin, Xiaolei, et al.
Veröffentlicht: (2025)
von: Qin, Xiaolei, et al.
Veröffentlicht: (2025)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
Text-Video Retrieval with Global-Local Semantic Consistent Learning
von: Zhang, Haonan, et al.
Veröffentlicht: (2024)
von: Zhang, Haonan, et al.
Veröffentlicht: (2024)
DualGeo: A Dual-View Framework for Worldwide Image Geo-localization
von: Cui, Junchao, et al.
Veröffentlicht: (2026)
von: Cui, Junchao, et al.
Veröffentlicht: (2026)
Bridging the Pose-Semantic Gap: A Cascade Framework for Text-Based Person Anomaly Search
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
MRGeo: Robust Cross-View Geo-Localization of Corrupted Images via Spatial and Channel Feature Enhancement
von: Wu, Le, et al.
Veröffentlicht: (2026)
von: Wu, Le, et al.
Veröffentlicht: (2026)
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
Bridged Semantic Alignment for Zero-shot 3D Medical Image Diagnosis
von: Lai, Haoran, et al.
Veröffentlicht: (2025)
von: Lai, Haoran, et al.
Veröffentlicht: (2025)
Bridging the Micro--Macro Gap: Frequency-Aware Semantic Alignment for Image Manipulation Localization
von: Liang, Xiaojie, et al.
Veröffentlicht: (2026)
von: Liang, Xiaojie, et al.
Veröffentlicht: (2026)
Bridging the Gap: Aligning Text-to-Image Diffusion Models with Specific Feedback
von: Niu, Xuexiang, et al.
Veröffentlicht: (2024)
von: Niu, Xuexiang, et al.
Veröffentlicht: (2024)
InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
von: Ma, Sanggyun, et al.
Veröffentlicht: (2025)
von: Ma, Sanggyun, et al.
Veröffentlicht: (2025)
SAMA: Factorized Semantic Anchoring and Motion Alignment for Instruction-Guided Video Editing
von: Zhang, Xinyao, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyao, et al.
Veröffentlicht: (2026)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
von: Yang, Feng, et al.
Veröffentlicht: (2025)
von: Yang, Feng, et al.
Veröffentlicht: (2025)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Residual Diffusion Bridge Model for Image Restoration
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025) -
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
von: Wang, Wei, et al.
Veröffentlicht: (2026) -
Universal Pansharpening Foundation Model
von: Wang, Hebaixu, et al.
Veröffentlicht: (2026) -
ViewBridge:Revisiting Cross-View Localization from Image Matching
von: Xia, Panwang, et al.
Veröffentlicht: (2025) -
UniABG: Unified Adversarial View Bridging and Graph Correspondence for Unsupervised Cross-View Geo-Localization
von: Chen, Cuiqun, et al.
Veröffentlicht: (2025)