Can Foundation Models Reliably Identify Spatial Hazards? A Case Study on Curb Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Sheng, Diwei, Hamilton-Fletcher, Giles, Beheshti, Mahya, Feng, Chen, Rizzo, John-Ross |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating OCR Performance for Assistive Technology: Effects of Walking Speed, Camera Placement, and Camera Type
by: Feng, Junchi, et al.
Published: (2026)
by: Feng, Junchi, et al.
Published: (2026)
Scene Change Detection with Vision-Language Representation Learning
by: Sheng, Diwei, et al.
Published: (2026)
by: Sheng, Diwei, et al.
Published: (2026)
NYC-Indoor-VPR: A Long-Term Indoor Visual Place Recognition Dataset with Semi-Automatic Annotation
by: Sheng, Diwei, et al.
Published: (2024)
by: Sheng, Diwei, et al.
Published: (2024)
Robust Computer-Vision based Construction Site Detection for Assistive-Technology Applications
by: Feng, Junchi, et al.
Published: (2025)
by: Feng, Junchi, et al.
Published: (2025)
Multi-faceted Sensory Substitution for Curb Alerting: A Pilot Investigation in Persons with Blindness and Low Vision
by: Ruan, Ligao, et al.
Published: (2024)
by: Ruan, Ligao, et al.
Published: (2024)
Does Embodiment Matter to Biomechanics and Function? A Comparative Analysis of Head-Mounted and Hand-Held Assistive Devices for Individuals with Blindness and Low Vision
by: Seth, Gaurav, et al.
Published: (2025)
by: Seth, Gaurav, et al.
Published: (2025)
CurbNet: Curb Detection Framework Based on LiDAR Point Cloud Segmentation
by: Zhao, Guoyang, et al.
Published: (2024)
by: Zhao, Guoyang, et al.
Published: (2024)
Exploring the Use of VLMs for Navigation Assistance for People with Blindness and Low Vision
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
Iris-SAM: Iris Segmentation Using a Foundation Model
by: Farmanifard, Parisa, et al.
Published: (2024)
by: Farmanifard, Parisa, et al.
Published: (2024)
Haptics-based, higher-order Sensory Substitution designed for Object Negotiation in Blindness and Low Vision: Virtual Whiskers
by: Feng, Junchi, et al.
Published: (2024)
by: Feng, Junchi, et al.
Published: (2024)
Adapting Segment Anything Model for Power Transmission Corridor Hazard Segmentation
by: Chen, Hang, et al.
Published: (2025)
by: Chen, Hang, et al.
Published: (2025)
Segment Anything Model Can Not Segment Anything: Assessing AI Foundation Model's Generalizability in Permafrost Mapping
by: Li, Wenwen, et al.
Published: (2024)
by: Li, Wenwen, et al.
Published: (2024)
A Multimodal Assistive System for Product Localization and Retrieval for People who are Blind or have Low Vision
by: Ruan, Ligao, et al.
Published: (2026)
by: Ruan, Ligao, et al.
Published: (2026)
Annotation-Free Curb Detection Leveraging Altitude Difference Image
by: Ma, Fulong, et al.
Published: (2024)
by: Ma, Fulong, et al.
Published: (2024)
FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
AGIR: Assessing 3D Gait Impairment with Reasoning based on LLMs
by: Wang, Diwei, et al.
Published: (2025)
by: Wang, Diwei, et al.
Published: (2025)
Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models
by: Cai, Yuzhu, et al.
Published: (2024)
by: Cai, Yuzhu, et al.
Published: (2024)
A Multi-Modal Foundation Model to Assist People with Blindness and Low Vision in Environmental Interaction
by: Hao, Yu, et al.
Published: (2023)
by: Hao, Yu, et al.
Published: (2023)
Landslide Hazard Mapping with Geospatial Foundation Models: Geographical Generalizability, Data Scarcity, and Band Adaptability
by: Li, Wenwen, et al.
Published: (2025)
by: Li, Wenwen, et al.
Published: (2025)
Zero-shot Hazard Identification in Autonomous Driving: A Case Study on the COOOL Benchmark
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
Reliability in Semantic Segmentation: Can We Use Synthetic Data?
by: Loiseau, Thibaut, et al.
Published: (2023)
by: Loiseau, Thibaut, et al.
Published: (2023)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
by: Cao, Haozhi, et al.
Published: (2024)
by: Cao, Haozhi, et al.
Published: (2024)
High-Performance Few-Shot Segmentation with Foundation Models: An Empirical Study
by: Chang, Shijie, et al.
Published: (2024)
by: Chang, Shijie, et al.
Published: (2024)
Distillation Improves Visual Place Recognition for Low Quality Images
by: Yang, Anbang, et al.
Published: (2023)
by: Yang, Anbang, et al.
Published: (2023)
Layer-Specific Fine-Tuning for Improved Negation Handling in Medical Vision-Language Models
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
MAGIC: Map-Guided Few-Shot Audio-Visual Acoustics Modeling
by: Huang, Diwei, et al.
Published: (2024)
by: Huang, Diwei, et al.
Published: (2024)
Are Vision Foundation Models Foundational for Electron Microscopy Image Segmentation?
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
Robustness Analysis on Foundational Segmentation Models
by: Schiappa, Madeline Chantry, et al.
Published: (2023)
by: Schiappa, Madeline Chantry, et al.
Published: (2023)
Can Foundation Models Predict Fitness for Duty?
by: Tapia, Juan E., et al.
Published: (2025)
by: Tapia, Juan E., et al.
Published: (2025)
SpatialBench: Is Your Spatial Foundation Model an All-Round Player?
by: Peng, Haosong, et al.
Published: (2026)
by: Peng, Haosong, et al.
Published: (2026)
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
by: Mirjalili, Vahid, et al.
Published: (2025)
by: Mirjalili, Vahid, et al.
Published: (2025)
GEM: Boost Simple Network for Glass Surface Segmentation via Vision Foundation Models
by: Hao, Jing, et al.
Published: (2023)
by: Hao, Jing, et al.
Published: (2023)
Annotation-Free Detection of Drivable Areas and Curbs Leveraging LiDAR Point Cloud Maps
by: Ma, Fulong, et al.
Published: (2026)
by: Ma, Fulong, et al.
Published: (2026)
CanViT: Toward Active-Vision Foundation Models
by: Berreby, Yohaï-Eliel, et al.
Published: (2026)
by: Berreby, Yohaï-Eliel, et al.
Published: (2026)
Visual Place Cell Encoding: A Computational Model for Spatial Representation and Cognitive Mapping
by: Hamilton, Chance J., et al.
Published: (2025)
by: Hamilton, Chance J., et al.
Published: (2025)
Enhancing Gait Video Analysis in Neurodegenerative Diseases by Knowledge Augmentation in Vision Language Model
by: Wang, Diwei, et al.
Published: (2024)
by: Wang, Diwei, et al.
Published: (2024)
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering
by: Zhang, Chaoning, et al.
Published: (2023)
by: Zhang, Chaoning, et al.
Published: (2023)
Foundation AI Model for Medical Image Segmentation
by: Bao, Rina, et al.
Published: (2024)
by: Bao, Rina, et al.
Published: (2024)
Enhancing the Reliability of Segment Anything Model for Auto-Prompting Medical Image Segmentation with Uncertainty Rectification
by: Zhang, Yichi, et al.
Published: (2023)
by: Zhang, Yichi, et al.
Published: (2023)
Can LLMs Assist Computer Education? an Empirical Case Study of DeepSeek
by: Xiao, Dongfu, et al.
Published: (2025)
by: Xiao, Dongfu, et al.
Published: (2025)
Similar Items
-
Evaluating OCR Performance for Assistive Technology: Effects of Walking Speed, Camera Placement, and Camera Type
by: Feng, Junchi, et al.
Published: (2026) -
Scene Change Detection with Vision-Language Representation Learning
by: Sheng, Diwei, et al.
Published: (2026) -
NYC-Indoor-VPR: A Long-Term Indoor Visual Place Recognition Dataset with Semi-Automatic Annotation
by: Sheng, Diwei, et al.
Published: (2024) -
Robust Computer-Vision based Construction Site Detection for Assistive-Technology Applications
by: Feng, Junchi, et al.
Published: (2025) -
Multi-faceted Sensory Substitution for Curb Alerting: A Pilot Investigation in Persons with Blindness and Low Vision
by: Ruan, Ligao, et al.
Published: (2024)