TransLocNet: Cross-Modal Attention for Aerial-Ground Vehicle Localization with Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Phu, Conover, Damon, Bera, Aniket |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlashSLAM: Accelerated RGB-D SLAM for Real-Time 3D Scene Reconstruction with Gaussian Splatting
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Scalable Multi-Robot Informative Path Planning for Target Mapping via Deep Reinforcement Learning
by: Vashisth, Apoorva, et al.
Published: (2024)
by: Vashisth, Apoorva, et al.
Published: (2024)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)
by: Daiya, Divyanshu, et al.
Published: (2024)
MVGaussian: High-Fidelity text-to-3D Content Generation with Multi-View Guidance and Surface Densification
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
To Move or Not to Move: Constraint-based Planning Enables Zero-Shot Generalization for Interactive Navigation
by: Vashisth, Apoorva, et al.
Published: (2026)
by: Vashisth, Apoorva, et al.
Published: (2026)
AGL-NET: Aerial-Ground Cross-Modal Global Localization with Varying Scales
by: Guan, Tianrui, et al.
Published: (2024)
by: Guan, Tianrui, et al.
Published: (2024)
HessianForge: Scalable LiDAR reconstruction with Physics-Informed Neural Representation and Smoothness Energy Constraints
by: Viswanath, Hrishikesh, et al.
Published: (2025)
by: Viswanath, Hrishikesh, et al.
Published: (2025)
SpatiaLoc: Leveraging Multi-Level Spatial Enhanced Descriptors for Cross-Modal Localization
by: Shang, Tianyi, et al.
Published: (2026)
by: Shang, Tianyi, et al.
Published: (2026)
AstroLoc: Robust Space to Ground Image Localizer
by: Berton, Gabriele, et al.
Published: (2025)
by: Berton, Gabriele, et al.
Published: (2025)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
Unified Multi-Modal Interactive & Reactive 3D Motion Generation via Rectified Flow
by: Gupta, Prerit, et al.
Published: (2025)
by: Gupta, Prerit, et al.
Published: (2025)
Adversarial Agent Behavior Learning in Autonomous Driving Using Deep Reinforcement Learning
by: Srinivasan, Arjun, et al.
Published: (2025)
by: Srinivasan, Arjun, et al.
Published: (2025)
HOTFormerLoc: Hierarchical Octree Transformer for Versatile Lidar Place Recognition Across Ground and Aerial Views
by: Griffiths, Ethan, et al.
Published: (2025)
by: Griffiths, Ethan, et al.
Published: (2025)
FullTransNet: Full Transformer with Local-Global Attention for Video Summarization
by: Lan, Libin, et al.
Published: (2025)
by: Lan, Libin, et al.
Published: (2025)
GroundLoc: Efficient Large-Scale Outdoor LiDAR-Only Localization
by: Steinke, Nicolai, et al.
Published: (2025)
by: Steinke, Nicolai, et al.
Published: (2025)
Loc$^2$: Interpretable Cross-View Localization via Depth-Lifted Local Feature Matching
by: Xia, Zimin, et al.
Published: (2025)
by: Xia, Zimin, et al.
Published: (2025)
DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
LoD-Loc: Aerial Visual Localization using LoD 3D Map with Neural Wireframe Alignment
by: Zhu, Juelin, et al.
Published: (2024)
by: Zhu, Juelin, et al.
Published: (2024)
MapLocNet: Coarse-to-Fine Feature Registration for Visual Re-Localization in Navigation Maps
by: Wu, Hang, et al.
Published: (2024)
by: Wu, Hang, et al.
Published: (2024)
RAFA-Net: Region Attention Network For Food Items And Agricultural Stress Recognition
by: Bera, Asish, et al.
Published: (2024)
by: Bera, Asish, et al.
Published: (2024)
360Loc: A Dataset and Benchmark for Omnidirectional Visual Localization with Cross-device Queries
by: Huang, Huajian, et al.
Published: (2023)
by: Huang, Huajian, et al.
Published: (2023)
SegLocNet: Multimodal Localization Network for Autonomous Driving via Bird's-Eye-View Segmentation
by: Zhou, Zijie, et al.
Published: (2025)
by: Zhou, Zijie, et al.
Published: (2025)
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
by: Bhattacharya, Uttaran, et al.
Published: (2024)
by: Bhattacharya, Uttaran, et al.
Published: (2024)
LightLoc: Learning Outdoor LiDAR Localization at Light Speed
by: Li, Wen, et al.
Published: (2025)
by: Li, Wen, et al.
Published: (2025)
AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis
by: Vuong, Khiem, et al.
Published: (2025)
by: Vuong, Khiem, et al.
Published: (2025)
A Tri-Modal Dataset and a Baseline System for Tracking Unmanned Aerial Vehicles
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
A Unified Attention U-Net Framework for Cross-Modality Tumor Segmentation in MRI and CT
by: Rai, Nishan, et al.
Published: (2026)
by: Rai, Nishan, et al.
Published: (2026)
CACE-Net: Co-guidance Attention and Contrastive Enhancement for Effective Audio-Visual Event Localization
by: He, Xiang, et al.
Published: (2024)
by: He, Xiang, et al.
Published: (2024)
LoD-Loc v2: Aerial Visual Localization over Low Level-of-Detail City Models using Explicit Silhouette Alignment
by: Zhu, Juelin, et al.
Published: (2025)
by: Zhu, Juelin, et al.
Published: (2025)
LunarLoc: Segment-Based Global Localization on the Moon
by: Thomas, Annika, et al.
Published: (2025)
by: Thomas, Annika, et al.
Published: (2025)
UnLoc: Leveraging Depth Uncertainties for Floorplan Localization
by: Wüest, Matthias, et al.
Published: (2025)
by: Wüest, Matthias, et al.
Published: (2025)
DiaLoc: An Iterative Approach to Embodied Dialog Localization
by: Zhang, Chao, et al.
Published: (2024)
by: Zhang, Chao, et al.
Published: (2024)
VFM-Loc: Zero-Shot Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies
by: Lu, Jun, et al.
Published: (2026)
by: Lu, Jun, et al.
Published: (2026)
Dual-Level Cross-Modal Contrastive Clustering
by: Zhang, Haixin, et al.
Published: (2024)
by: Zhang, Haixin, et al.
Published: (2024)
Local-to-Global Cross-Modal Attention-Aware Fusion for HSI-X Semantic Segmentation
by: Zhang, Xuming, et al.
Published: (2024)
by: Zhang, Xuming, et al.
Published: (2024)
Layout-to-Image Generation with Localized Descriptions using ControlNet with Cross-Attention Control
by: Lukovnikov, Denis, et al.
Published: (2024)
by: Lukovnikov, Denis, et al.
Published: (2024)
UniLoc: Towards Universal Place Recognition Using Any Single Modality
by: Xia, Yan, et al.
Published: (2024)
by: Xia, Yan, et al.
Published: (2024)
Dual-Task Learning for Dead Tree Detection and Segmentation with Hybrid Self-Attention U-Nets in Aerial Imagery
by: Rahman, Anis Ur, et al.
Published: (2025)
by: Rahman, Anis Ur, et al.
Published: (2025)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
by: Mullen Jr, James F., et al.
Published: (2022)
by: Mullen Jr, James F., et al.
Published: (2022)
Similar Items
-
FlashSLAM: Accelerated RGB-D SLAM for Real-Time 3D Scene Reconstruction with Gaussian Splatting
by: Pham, Phu, et al.
Published: (2024) -
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
by: Pham, Phu, et al.
Published: (2024) -
Scalable Multi-Robot Informative Path Planning for Target Mapping via Deep Reinforcement Learning
by: Vashisth, Apoorva, et al.
Published: (2024) -
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024) -
MVGaussian: High-Fidelity text-to-3D Content Generation with Multi-View Guidance and Surface Densification
by: Pham, Phu, et al.
Published: (2024)