DiffVL: Diffusion-Based Visual Localization on 2D Maps via BEV-Conditioned GPS Denoising
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Li, Sun, Hongyang, Liu, Liu, Li, Yunhao, Cai, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion
by: Sun, Zhigang, et al.
Published: (2025)
by: Sun, Zhigang, et al.
Published: (2025)
DiffStega: Towards Universal Training-Free Coverless Image Steganography with Diffusion Models
by: Yang, Yiwei, et al.
Published: (2024)
by: Yang, Yiwei, et al.
Published: (2024)
DiffDenoise: Self-Supervised Medical Image Denoising with Conditional Diffusion Models
by: Demir, Basar, et al.
Published: (2025)
by: Demir, Basar, et al.
Published: (2025)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024)
by: Zhao, Xiao, et al.
Published: (2024)
6D-Diff: A Keypoint Diffusion Framework for 6D Object Pose Estimation
by: Xu, Li, et al.
Published: (2023)
by: Xu, Li, et al.
Published: (2023)
FoundDiff: Foundational Diffusion Model for Generalizable Low-Dose CT Denoising
by: Chen, Zhihao, et al.
Published: (2025)
by: Chen, Zhihao, et al.
Published: (2025)
AnaMoDiff: 2D Analogical Motion Diffusion via Disentangled Denoising
by: Tanveer, Maham, et al.
Published: (2024)
by: Tanveer, Maham, et al.
Published: (2024)
Youtu-VL: Unleashing Visual Potential via Unified Vision-Language Supervision
by: Wei, Zhixiang, et al.
Published: (2026)
by: Wei, Zhixiang, et al.
Published: (2026)
ChatBEV: A Visual Language Model that Understands BEV Maps
by: Xu, Qingyao, et al.
Published: (2025)
by: Xu, Qingyao, et al.
Published: (2025)
DiffMap: Enhancing Map Segmentation with Map Prior Using Diffusion Model
by: Jia, Peijin, et al.
Published: (2024)
by: Jia, Peijin, et al.
Published: (2024)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
BEV$^2$PR: BEV-Enhanced Visual Place Recognition with Structural Cues
by: Ge, Fudong, et al.
Published: (2024)
by: Ge, Fudong, et al.
Published: (2024)
ClickDiff: Click to Induce Semantic Contact Map for Controllable Grasp Generation with Diffusion Models
by: Li, Peiming, et al.
Published: (2024)
by: Li, Peiming, et al.
Published: (2024)
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
by: Ye, Xin, et al.
Published: (2025)
by: Ye, Xin, et al.
Published: (2025)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
Fleming-VL: Towards Universal Medical Visual Reasoning with Multimodal LLMs
by: Shu, Yan, et al.
Published: (2025)
by: Shu, Yan, et al.
Published: (2025)
GeoDiffMM: Geometry-Guided Conditional Diffusion for Motion Magnification
by: Liu, Xuedeng, et al.
Published: (2025)
by: Liu, Xuedeng, et al.
Published: (2025)
BREATH-VL: Vision-Language-Guided 6-DoF Bronchoscopy Localization via Semantic-Geometric Fusion
by: Tian, Qingyao, et al.
Published: (2026)
by: Tian, Qingyao, et al.
Published: (2026)
AdaptDiff: Cross-Modality Domain Adaptation via Weak Conditional Semantic Diffusion for Retinal Vessel Segmentation
by: Hu, Dewei, et al.
Published: (2024)
by: Hu, Dewei, et al.
Published: (2024)
Test-time Correction: An Online 3D Detection System via Visual Prompting
by: Zhang, Hanxue, et al.
Published: (2024)
by: Zhang, Hanxue, et al.
Published: (2024)
DEMOS: Dynamic Environment Motion Synthesis in 3D Scenes via Local Spherical-BEV Perception
by: Gong, Jingyu, et al.
Published: (2024)
by: Gong, Jingyu, et al.
Published: (2024)
Disentangled Diffusion-Based 3D Human Pose Estimation with Hierarchical Spatial and Temporal Denoiser
by: Cai, Qingyuan, et al.
Published: (2024)
by: Cai, Qingyuan, et al.
Published: (2024)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
by: Song, Yiren, et al.
Published: (2024)
by: Song, Yiren, et al.
Published: (2024)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving
by: Brandstaetter, Felix, et al.
Published: (2025)
by: Brandstaetter, Felix, et al.
Published: (2025)
ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object Detection
by: Chen, Jiwei, et al.
Published: (2024)
by: Chen, Jiwei, et al.
Published: (2024)
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
by: Yang, Zetong, et al.
Published: (2023)
by: Yang, Zetong, et al.
Published: (2023)
Diff-Aid: Inference-time Adaptive Interaction Denoising for Rectified Text-to-Image Generation
by: Li, Binglei, et al.
Published: (2026)
by: Li, Binglei, et al.
Published: (2026)
End-to-End Driving with Online Trajectory Evaluation via BEV World Model
by: Li, Yingyan, et al.
Published: (2025)
by: Li, Yingyan, et al.
Published: (2025)
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic Reasoning
by: Liu, Weimin, et al.
Published: (2026)
by: Liu, Weimin, et al.
Published: (2026)
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
by: Gao, Huan-ang, et al.
Published: (2024)
by: Gao, Huan-ang, et al.
Published: (2024)
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
by: Chen, Zigeng, et al.
Published: (2024)
by: Chen, Zigeng, et al.
Published: (2024)
TS-CGNet: Temporal-Spatial Fusion Meets Centerline-Guided Diffusion for BEV Mapping
by: Hong, Xinying, et al.
Published: (2025)
by: Hong, Xinying, et al.
Published: (2025)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Multi-scale 2D Temporal Map Diffusion Models for Natural Language Video Localization
by: Zhang, Chongzhi, et al.
Published: (2024)
by: Zhang, Chongzhi, et al.
Published: (2024)
VipDiff: Towards Coherent and Diverse Video Inpainting via Training-free Denoising Diffusion Models
by: Xie, Chaohao, et al.
Published: (2025)
by: Xie, Chaohao, et al.
Published: (2025)
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
LowDiff: Efficient Diffusion Sampling with Low-Resolution Condition
by: Xu, Jiuyi, et al.
Published: (2025)
by: Xu, Jiuyi, et al.
Published: (2025)
Similar Items
-
DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion
by: Sun, Zhigang, et al.
Published: (2025) -
DiffStega: Towards Universal Training-Free Coverless Image Steganography with Diffusion Models
by: Yang, Yiwei, et al.
Published: (2024) -
DiffDenoise: Self-Supervised Medical Image Denoising with Conditional Diffusion Models
by: Demir, Basar, et al.
Published: (2025) -
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024) -
6D-Diff: A Keypoint Diffusion Framework for 6D Object Pose Estimation
by: Xu, Li, et al.
Published: (2023)