"Does the cafe entrance look accessible? Where is the door?" Towards Geospatial AI Agents for Visual Inquiries
Fuente:
arXiv
Saved in:
| Main Authors: | Froehlich, Jon E., Hwang, Jared, Wang, Zeyu, O'Meara, John S., Su, Xia, Huang, William, Zhang, Yang, Fiannaca, Alex, Nelson, Philip, Kane, Shaun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RampNet: A Two-Stage Pipeline for Bootstrapping Curb Ramp Detection in Streetscape Images from Open Government Metadata
by: O'Meara, John S., et al.
Published: (2025)
by: O'Meara, John S., et al.
Published: (2025)
StreetReaderAI: Making Street View Accessible Using Context-Aware Multimodal AI
by: Froehlich, Jon E., et al.
Published: (2025)
by: Froehlich, Jon E., et al.
Published: (2025)
"Where Can I Park?" Understanding Human Perspectives and Scalably Detecting Disability Parking from Aerial Imagery
by: Hwang, Jared, et al.
Published: (2025)
by: Hwang, Jared, et al.
Published: (2025)
A Contextual Inquiry of People with Vision Impairments in Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2024)
by: Li, Franklin Mingzhe, et al.
Published: (2024)
Accessibility Scout: Personalized Accessibility Scans of Built Environments
by: Huang, William, et al.
Published: (2025)
by: Huang, William, et al.
Published: (2025)
SonifyAR: Context-Aware Sound Generation in Augmented Reality
by: Su, Xia, et al.
Published: (2024)
by: Su, Xia, et al.
Published: (2024)
Revolutionising Role-Playing Games with ChatGPT
by: Stampfl, Rita, et al.
Published: (2024)
by: Stampfl, Rita, et al.
Published: (2024)
DepthScape: Authoring 2.5D Designs via Depth Estimation, Semantic Understanding, and Geometry Extraction
by: Su, Xia, et al.
Published: (2025)
by: Su, Xia, et al.
Published: (2025)
Roland Barthes at the Collège de France
by: O'Meara, Lucy
Published: (2017)
by: O'Meara, Lucy
Published: (2017)
The Surprising Non-Triviality of Sharing a Sandwich with Two Other People
by: O'Meara, John
Published: (2024)
by: O'Meara, John
Published: (2024)
A Lie Theoretic Framework for Controlling Open Quantum Systems
by: O'Meara, Corey
Published: (2025)
by: O'Meara, Corey
Published: (2025)
Traffic‐habits and local descent groups: Broadening the scope of ethnographic research into land rights after land claims
by: Dayne O'Meara
Published: (2025)
by: Dayne O'Meara
Published: (2025)
Where are they looking in the operating room?
by: Chen, Keqi, et al.
Published: (2026)
by: Chen, Keqi, et al.
Published: (2026)
Engaging with Children's Artwork in Mixed Visual-Ability Families
by: Chheda-Kothary, Arnavi, et al.
Published: (2024)
by: Chheda-Kothary, Arnavi, et al.
Published: (2024)
FlyMeThrough: Human-AI Collaborative 3D Indoor Mapping with Commodity Drones
by: Su, Xia, et al.
Published: (2025)
by: Su, Xia, et al.
Published: (2025)
A Recipe for Success? Exploring Strategies for Improving Non-Visual Access to Cooking Instructions
by: Li, Franklin Mingzhe, et al.
Published: (2024)
by: Li, Franklin Mingzhe, et al.
Published: (2024)
RASSAR: Room Accessibility and Safety Scanning in Augmented Reality
by: Su, Xia, et al.
Published: (2024)
by: Su, Xia, et al.
Published: (2024)
Making Urban Art Accessible: Current Art Access Techniques, Design Considerations, and the Role of AI
by: Jiang, Lucy, et al.
Published: (2024)
by: Jiang, Lucy, et al.
Published: (2024)
Gensors: Authoring Personalized Visual Sensors with Multimodal Foundation Models and Reasoning
by: Liu, Michael Xieyang, et al.
Published: (2025)
by: Liu, Michael Xieyang, et al.
Published: (2025)
CapNav: Benchmarking Vision Language Models on Capability-conditioned Indoor Navigation
by: Su, Xia, et al.
Published: (2026)
by: Su, Xia, et al.
Published: (2026)
Focusing Where Vision Matters: Selective Training for Large Vision Language Models via Visual Information Gain
by: Lee, Seulbi, et al.
Published: (2026)
by: Lee, Seulbi, et al.
Published: (2026)
Disparities at the entrance door: gender gaps in elementary school
by: Lina Martínez
Published: (2018)
by: Lina Martínez
Published: (2018)
cubic: CUDA-accelerated 3D Bioimage Computing
by: Kalinin, Alexandr A., et al.
Published: (2025)
by: Kalinin, Alexandr A., et al.
Published: (2025)
CrossVIT-augmented Geospatial-Intelligence Visualization System for Tracking Economic Development Dynamics
by: Bai, Yanbing, et al.
Published: (2024)
by: Bai, Yanbing, et al.
Published: (2024)
Route Packing: Geospatially-Accurate Visualization of Route Networks
by: Zhao, Jieqiong, et al.
Published: (2019)
by: Zhao, Jieqiong, et al.
Published: (2019)
Evaluating Effectiveness of Interactivity in Contour-based Geospatial Visualizations
by: Nayeem, Abdullah-Al-Raihan, et al.
Published: (2024)
by: Nayeem, Abdullah-Al-Raihan, et al.
Published: (2024)
Unseen City Canvases: Exploring Blind and Low Vision People's Perspectives on Urban and Public Art Accessibility
by: Jiang, Lucy, et al.
Published: (2026)
by: Jiang, Lucy, et al.
Published: (2026)
Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance
by: Kaur, Amandeep, et al.
Published: (2026)
by: Kaur, Amandeep, et al.
Published: (2026)
Portrait of a Graduate and the National School Library Standards
by: Ehler-Hansen, Deborah, et al.
Published: (2019)
by: Ehler-Hansen, Deborah, et al.
Published: (2019)
Ant communities respond to a large‐scale disturbance along an elevational gradient in Puerto Rico, U.S.A.
by: Shannon O'Meara, et al.
Published: (2024)
by: Shannon O'Meara, et al.
Published: (2024)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
by: Zhang, Peirong, et al.
Published: (2025)
by: Zhang, Peirong, et al.
Published: (2025)
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
by: Mena, Omar, et al.
Published: (2025)
by: Mena, Omar, et al.
Published: (2025)
Where Does Vision Meet Language? Understanding and Refining Visual Fusion in MLLMs via Contrastive Attention
by: Song, Shezheng, et al.
Published: (2026)
by: Song, Shezheng, et al.
Published: (2026)
ArtInsight: Enabling AI-Powered Artwork Engagement for Mixed Visual-Ability Families
by: Chheda-Kothary, Arnavi, et al.
Published: (2025)
by: Chheda-Kothary, Arnavi, et al.
Published: (2025)
egenioussBench: A New Dataset for Geospatial Visual Localisation
by: Fanta-Jende, Phillipp, et al.
Published: (2026)
by: Fanta-Jende, Phillipp, et al.
Published: (2026)
HexTiles and Semantic Icons for MAUP-Aware Multivariate Geospatial Visualizations
by: Kawakami, Yuya, et al.
Published: (2024)
by: Kawakami, Yuya, et al.
Published: (2024)
Where Does Authorship Signal Emerge in Encoder-Based Language Models?
by: Kulumba, Francis, et al.
Published: (2026)
by: Kulumba, Francis, et al.
Published: (2026)
PyTorchFire: A GPU-Accelerated Wildfire Simulator with Differentiable Cellular Automata
by: Xia, Zeyu, et al.
Published: (2025)
by: Xia, Zeyu, et al.
Published: (2025)
Compass vs Railway Tracks: Unpacking User Mental Models for Communicating Long-Horizon Work to Humans vs. AI
by: Petridis, Savvas, et al.
Published: (2026)
by: Petridis, Savvas, et al.
Published: (2026)
Human-centered Geospatial Data Science
by: Kang, Yuhao
Published: (2025)
by: Kang, Yuhao
Published: (2025)
Similar Items
-
RampNet: A Two-Stage Pipeline for Bootstrapping Curb Ramp Detection in Streetscape Images from Open Government Metadata
by: O'Meara, John S., et al.
Published: (2025) -
StreetReaderAI: Making Street View Accessible Using Context-Aware Multimodal AI
by: Froehlich, Jon E., et al.
Published: (2025) -
"Where Can I Park?" Understanding Human Perspectives and Scalably Detecting Disability Parking from Aerial Imagery
by: Hwang, Jared, et al.
Published: (2025) -
A Contextual Inquiry of People with Vision Impairments in Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2024) -
Accessibility Scout: Personalized Accessibility Scans of Built Environments
by: Huang, William, et al.
Published: (2025)