Statewide Visual Geolocalization in the Wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fervers, Florian, Bullinger, Sebastian, Bodensteiner, Christoph, Arens, Michael, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Neural Radiance Fields for Multi-Date Satellite Data
von: Wagner, Valentin, et al.
Veröffentlicht: (2025)
von: Wagner, Valentin, et al.
Veröffentlicht: (2025)
SatGeo-NeRF: Geometrically Regularized NeRF for Satellite Imagery
von: Wagner, Valentin, et al.
Veröffentlicht: (2026)
von: Wagner, Valentin, et al.
Veröffentlicht: (2026)
Strike the Balance: On-the-Fly Uncertainty based User Interactions for Long-Term Video Object Segmentation
von: Vujasinović, Stéphane, et al.
Veröffentlicht: (2024)
von: Vujasinović, Stéphane, et al.
Veröffentlicht: (2024)
ACCSAMS: Automatic Conversion of Exam Documents to Accessible Learning Material for Blind and Visually Impaired
von: Wilkening, David, et al.
Veröffentlicht: (2024)
von: Wilkening, David, et al.
Veröffentlicht: (2024)
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
von: Vogel, Alexander, et al.
Veröffentlicht: (2025)
von: Vogel, Alexander, et al.
Veröffentlicht: (2025)
SFDLA: Source-Free Document Layout Analysis
von: Tewes, Sebastian, et al.
Veröffentlicht: (2025)
von: Tewes, Sebastian, et al.
Veröffentlicht: (2025)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
Scene-agnostic Pose Regression for Visual Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
GRASPing Anatomy to Improve Pathology Segmentation
von: Li, Keyi, et al.
Veröffentlicht: (2025)
von: Li, Keyi, et al.
Veröffentlicht: (2025)
GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
Is Visual in-Context Learning for Compositional Medical Tasks within Reach?
von: Reiß, Simon, et al.
Veröffentlicht: (2025)
von: Reiß, Simon, et al.
Veröffentlicht: (2025)
Network transferability of adversarial patches in real-time object detection
von: Bayer, Jens, et al.
Veröffentlicht: (2024)
von: Bayer, Jens, et al.
Veröffentlicht: (2024)
A Causally Grounded Taxonomy for Image Degradation Robustness Evaluation
von: Becker, Stefan, et al.
Veröffentlicht: (2026)
von: Becker, Stefan, et al.
Veröffentlicht: (2026)
Self-Aware Object Detection via Degradation Manifolds
von: Becker, Stefan, et al.
Veröffentlicht: (2026)
von: Becker, Stefan, et al.
Veröffentlicht: (2026)
Eigenpatches -- Adversarial Patches from Principal Components
von: Bayer, Jens, et al.
Veröffentlicht: (2023)
von: Bayer, Jens, et al.
Veröffentlicht: (2023)
Deformable Mamba for Wide Field of View Segmentation
von: Hu, Jie, et al.
Veröffentlicht: (2024)
von: Hu, Jie, et al.
Veröffentlicht: (2024)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
von: Schmitt, Jonas, et al.
Veröffentlicht: (2024)
von: Schmitt, Jonas, et al.
Veröffentlicht: (2024)
Skill-Conditioned Visual Geolocation for Vision-Language Models
von: Yang, Chenjie, et al.
Veröffentlicht: (2026)
von: Yang, Chenjie, et al.
Veröffentlicht: (2026)
Towards Activated Muscle Group Estimation in the Wild
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
Vision-language Models for Driver Monitoring Systems: A Driver Activity Description Dataset
von: Lerch, David J., et al.
Veröffentlicht: (2026)
von: Lerch, David J., et al.
Veröffentlicht: (2026)
Visual Text Generation in the Wild
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
HierLoc: Hyperbolic Entity Embeddings for Hierarchical Visual Geolocation
von: Gadi, Hari Krishna, et al.
Veröffentlicht: (2026)
von: Gadi, Hari Krishna, et al.
Veröffentlicht: (2026)
Interpretable Perception and Reasoning for Audiovisual Geolocation
von: Su, Yiyang, et al.
Veröffentlicht: (2026)
von: Su, Yiyang, et al.
Veröffentlicht: (2026)
AltChart: Enhancing VLM-based Chart Summarization Through Multi-Pretext Tasks
von: Moured, Omar, et al.
Veröffentlicht: (2024)
von: Moured, Omar, et al.
Veröffentlicht: (2024)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
ChartFormer: A Large Vision Language Model for Converting Chart Images into Tactile Accessible SVGs
von: Moured, Omar, et al.
Veröffentlicht: (2024)
von: Moured, Omar, et al.
Veröffentlicht: (2024)
Good Enough: Is it Worth Improving your Label Quality?
von: Jaus, Alexander, et al.
Veröffentlicht: (2025)
von: Jaus, Alexander, et al.
Veröffentlicht: (2025)
Towards Visual Query Segmentation in the Wild
von: Fan, Bing, et al.
Veröffentlicht: (2026)
von: Fan, Bing, et al.
Veröffentlicht: (2026)
Traversing the Subspace of Adversarial Patches
von: Bayer, Jens, et al.
Veröffentlicht: (2024)
von: Bayer, Jens, et al.
Veröffentlicht: (2024)
Higher-Order Adversarial Patches for Real-Time Object Detectors
von: Bayer, Jens, et al.
Veröffentlicht: (2026)
von: Bayer, Jens, et al.
Veröffentlicht: (2026)
Utilizing dataset affinity prediction in object detection to assess training data
von: Becker, Stefan, et al.
Veröffentlicht: (2023)
von: Becker, Stefan, et al.
Veröffentlicht: (2023)
Data Diet: Can Trimming PET/CT Datasets Enhance Lesion Segmentation?
von: Jaus, Alexander, et al.
Veröffentlicht: (2024)
von: Jaus, Alexander, et al.
Veröffentlicht: (2024)
Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection
von: Lerch, David J., et al.
Veröffentlicht: (2026)
von: Lerch, David J., et al.
Veröffentlicht: (2026)
Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
von: Grainge, Oliver, et al.
Veröffentlicht: (2025)
von: Grainge, Oliver, et al.
Veröffentlicht: (2025)
Thinking in 360°: Humanoid Visual Search in the Wild
von: Yu, Heyang, et al.
Veröffentlicht: (2025)
von: Yu, Heyang, et al.
Veröffentlicht: (2025)
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
OpenStreetView-5M: The Many Roads to Global Visual Geolocation
von: Astruc, Guillaume, et al.
Veröffentlicht: (2024)
von: Astruc, Guillaume, et al.
Veröffentlicht: (2024)
GaGA: Towards Interactive Global Geolocation Assistant
von: Dou, Zhiyang, et al.
Veröffentlicht: (2024)
von: Dou, Zhiyang, et al.
Veröffentlicht: (2024)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Semantic Neural Radiance Fields for Multi-Date Satellite Data
von: Wagner, Valentin, et al.
Veröffentlicht: (2025) -
SatGeo-NeRF: Geometrically Regularized NeRF for Satellite Imagery
von: Wagner, Valentin, et al.
Veröffentlicht: (2026) -
Strike the Balance: On-the-Fly Uncertainty based User Interactions for Long-Term Video Object Segmentation
von: Vujasinović, Stéphane, et al.
Veröffentlicht: (2024) -
ACCSAMS: Automatic Conversion of Exam Documents to Accessible Learning Material for Blind and Visually Impaired
von: Wilkening, David, et al.
Veröffentlicht: (2024) -
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
von: Vogel, Alexander, et al.
Veröffentlicht: (2025)