ELEV-VISION-SAM: Integrated Vision Language and Foundation Model for Automated Estimation of Building Lowest Floor Elevation
Fuente:
arXiv
Saved in:
| Main Authors: | Ho, Yu-Hsuan, Li, Longxiang, Mostafavi, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ELEV-VISION: Automated Lowest Floor Elevation Estimation from Segmenting Street View Images
by: Ho, Yu-Hsuan, et al.
Published: (2023)
by: Ho, Yu-Hsuan, et al.
Published: (2023)
Flood-DamageSense: Multimodal Mamba with Multitask Learning for Building Flood Damage Assessment using SAR Remote Sensing Imagery
by: Ho, Yu-Hsuan, et al.
Published: (2025)
by: Ho, Yu-Hsuan, et al.
Published: (2025)
Recov-Vision: Linking Street View Imagery and Vision-Language Models for Post-Disaster Recovery
by: Xiao, Yiming, et al.
Published: (2025)
by: Xiao, Yiming, et al.
Published: (2025)
Automated Wildfire Damage Assessment from Multi view Ground level Imagery Via Vision Language Models
by: Esparza, Miguel, et al.
Published: (2025)
by: Esparza, Miguel, et al.
Published: (2025)
MSD: A Benchmark Dataset for Floor Plan Generation of Building Complexes
by: van Engelenburg, Casper, et al.
Published: (2024)
by: van Engelenburg, Casper, et al.
Published: (2024)
DamageCAT: A Deep Learning Transformer Framework for Typology-Based Post-Disaster Building Damage Categorization
by: Xiao, Yiming, et al.
Published: (2025)
by: Xiao, Yiming, et al.
Published: (2025)
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering
by: Zhang, Chaoning, et al.
Published: (2023)
by: Zhang, Chaoning, et al.
Published: (2023)
ED-SAM: An Efficient Diffusion Sampling Approach to Domain Generalization in Vision-Language Foundation Models
by: Truong, Thanh-Dat, et al.
Published: (2024)
by: Truong, Thanh-Dat, et al.
Published: (2024)
FloorSAM: SAM-Guided Floorplan Reconstruction with Semantic-Geometric Fusion
by: Ye, Han, et al.
Published: (2025)
by: Ye, Han, et al.
Published: (2025)
Robust SAM: On the Adversarial Robustness of Vision Foundation Models
by: Long, Jiahuan, et al.
Published: (2025)
by: Long, Jiahuan, et al.
Published: (2025)
Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment
by: Kyem, Blessing Agyei, et al.
Published: (2026)
by: Kyem, Blessing Agyei, et al.
Published: (2026)
Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?
by: Qu, Tianyuan, et al.
Published: (2025)
by: Qu, Tianyuan, et al.
Published: (2025)
Ordinal Scale Traffic Congestion Classification with Multi-Modal Vision-Language and Motion Analysis
by: Lin, Yu-Hsuan
Published: (2025)
by: Lin, Yu-Hsuan
Published: (2025)
AD-SAM: Fine-Tuning the Segment Anything Vision Foundation Model for Autonomous Driving Perception
by: Camarena, Mario, et al.
Published: (2025)
by: Camarena, Mario, et al.
Published: (2025)
Parameter-Efficient Fine-Tuning of Vision Foundation Model for Forest Floor Segmentation from UAV Imagery
by: Wasil, Mohammad, et al.
Published: (2025)
by: Wasil, Mohammad, et al.
Published: (2025)
Building Floor Number Estimation from Crowdsourced Street-Level Images: Munich Dataset and Baseline Method
by: Sun, Yao, et al.
Published: (2025)
by: Sun, Yao, et al.
Published: (2025)
RepSAM: Bridging Foundation Models to Robotic Vision via Representation-Guided Adaptation
by: Chu, Wenhui
Published: (2026)
by: Chu, Wenhui
Published: (2026)
Implicit Modeling for Transferability Estimation of Vision Foundation Models
by: Zheng, Yaoyan, et al.
Published: (2025)
by: Zheng, Yaoyan, et al.
Published: (2025)
VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models
by: Ju, Jeongho, et al.
Published: (2024)
by: Ju, Jeongho, et al.
Published: (2024)
SAM-CLIP: Merging Vision Foundation Models towards Semantic and Spatial Understanding
by: Wang, Haoxiang, et al.
Published: (2023)
by: Wang, Haoxiang, et al.
Published: (2023)
Mind the Interference: Retaining Pre-trained Knowledge in Parameter Efficient Continual Learning of Vision-Language Models
by: Tang, Longxiang, et al.
Published: (2024)
by: Tang, Longxiang, et al.
Published: (2024)
TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration
by: Guo, Yiwei, et al.
Published: (2024)
by: Guo, Yiwei, et al.
Published: (2024)
Iris-SAM: Iris Segmentation Using a Foundation Model
by: Farmanifard, Parisa, et al.
Published: (2024)
by: Farmanifard, Parisa, et al.
Published: (2024)
Vision and Language Reference Prompt into SAM for Few-shot Segmentation
by: Sakurai, Kosuke, et al.
Published: (2025)
by: Sakurai, Kosuke, et al.
Published: (2025)
Foundations and Models in Modern Computer Vision: Key Building Blocks in Landmark Architectures
by: Bourceanu, Radu-Andrei, et al.
Published: (2025)
by: Bourceanu, Radu-Andrei, et al.
Published: (2025)
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model
by: Zhang, Yuxuan, et al.
Published: (2024)
by: Zhang, Yuxuan, et al.
Published: (2024)
DetailSemNet: Elevating Signature Verification through Detail-Semantic Integration
by: Shih, Meng-Cheng, et al.
Published: (2025)
by: Shih, Meng-Cheng, et al.
Published: (2025)
VideoSAM: A Large Vision Foundation Model for High-Speed Video Segmentation
by: Maduabuchi, Chika, et al.
Published: (2024)
by: Maduabuchi, Chika, et al.
Published: (2024)
Zero-Shot Refinement of Buildings' Segmentation Models using SAM
by: Mayladan, Ali, et al.
Published: (2023)
by: Mayladan, Ali, et al.
Published: (2023)
OVER-NAV: Elevating Iterative Vision-and-Language Navigation with Open-Vocabulary Detection and StructurEd Representation
by: Zhao, Ganlong, et al.
Published: (2024)
by: Zhao, Ganlong, et al.
Published: (2024)
Depth Any Panoramas: A Foundation Model for Panoramic Depth Estimation
by: Lin, Xin, et al.
Published: (2025)
by: Lin, Xin, et al.
Published: (2025)
Data Adaptive Traceback for Vision-Language Foundation Models in Image Classification
by: Peng, Wenshuo, et al.
Published: (2024)
by: Peng, Wenshuo, et al.
Published: (2024)
WPS-SAM: Towards Weakly-Supervised Part Segmentation with Foundation Models
by: Wu, Xinjian, et al.
Published: (2024)
by: Wu, Xinjian, et al.
Published: (2024)
SolarSAM: Building-scale Photovoltaic Potential Assessment Based on Segment Anything Model (SAM) and Remote Sensing for Emerging City
by: Wang, Guohao
Published: (2024)
by: Wang, Guohao
Published: (2024)
Vision-Language Semantic Aggregation Leveraging Foundation Model for Generalizable Medical Image Segmentation
by: Yu, Wenjun, et al.
Published: (2025)
by: Yu, Wenjun, et al.
Published: (2025)
Vision Also You Need: Navigating Out-of-Distribution Detection with Multimodal Large Language Model
by: Xu, Haoran, et al.
Published: (2026)
by: Xu, Haoran, et al.
Published: (2026)
Look Before Acting: Enhancing Vision Foundation Representations for Vision-Language-Action Models
by: Luo, Yulin, et al.
Published: (2026)
by: Luo, Yulin, et al.
Published: (2026)
Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
EVLF-FM: Explainable Vision Language Foundation Model for Medicine
by: Bai, Yang, et al.
Published: (2025)
by: Bai, Yang, et al.
Published: (2025)
Solar PV Installation Potential Assessment on Building Facades Based on Vision and Language Foundation Models
by: Liu, Ruyu, et al.
Published: (2025)
by: Liu, Ruyu, et al.
Published: (2025)
Similar Items
-
ELEV-VISION: Automated Lowest Floor Elevation Estimation from Segmenting Street View Images
by: Ho, Yu-Hsuan, et al.
Published: (2023) -
Flood-DamageSense: Multimodal Mamba with Multitask Learning for Building Flood Damage Assessment using SAR Remote Sensing Imagery
by: Ho, Yu-Hsuan, et al.
Published: (2025) -
Recov-Vision: Linking Street View Imagery and Vision-Language Models for Post-Disaster Recovery
by: Xiao, Yiming, et al.
Published: (2025) -
Automated Wildfire Damage Assessment from Multi view Ground level Imagery Via Vision Language Models
by: Esparza, Miguel, et al.
Published: (2025) -
MSD: A Benchmark Dataset for Floor Plan Generation of Building Complexes
by: van Engelenburg, Casper, et al.
Published: (2024)