Using Vision Language Models for Safety Hazard Identification in Construction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adil, Muhammad, Lee, Gaang, Gonzalez, Vicente A., Mei, Qipei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Integration of Object Detection and Small VLMs for Construction Safety Hazard Identification
von: Adil, Muhammad, et al.
Veröffentlicht: (2026)
von: Adil, Muhammad, et al.
Veröffentlicht: (2026)
TrajGATFormer: A Graph-Based Transformer Approach for Worker and Obstacle Trajectory Prediction in Off-site Construction Environments
von: Alduais, Mohammed, et al.
Veröffentlicht: (2025)
von: Alduais, Mohammed, et al.
Veröffentlicht: (2025)
ErgoChat: a Visual Query System for the Ergonomic Risk Assessment of Construction Workers
von: Fan, Chao, et al.
Veröffentlicht: (2024)
von: Fan, Chao, et al.
Veröffentlicht: (2024)
Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors?
von: Chen, Xuezheng, et al.
Veröffentlicht: (2025)
von: Chen, Xuezheng, et al.
Veröffentlicht: (2025)
HazardNet: A Small-Scale Vision Language Model for Real-Time Traffic Safety Detection at Edge Devices
von: Tami, Mohammad Abu, et al.
Veröffentlicht: (2025)
von: Tami, Mohammad Abu, et al.
Veröffentlicht: (2025)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
Multi-Modal Interpretability for Enhanced Localization in Vision-Language Models
von: Imran, Muhammad, et al.
Veröffentlicht: (2025)
von: Imran, Muhammad, et al.
Veröffentlicht: (2025)
AR-Facilitated Safety Inspection and Fall Hazard Detection on Construction Sites
von: Liu, Jiazhou, et al.
Veröffentlicht: (2024)
von: Liu, Jiazhou, et al.
Veröffentlicht: (2024)
Toward Autonomous Laboratory Safety Monitoring with Vision Language Models: Learning to See Hazards Through Scene Structure
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2026)
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2026)
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
INSIGHT: Enhancing Autonomous Driving Safety through Vision-Language Models on Context-Aware Hazard Detection and Edge Case Evaluation
von: Chen, Dianwei, et al.
Veröffentlicht: (2025)
von: Chen, Dianwei, et al.
Veröffentlicht: (2025)
HoliSafe: Holistic Safety Benchmarking and Modeling for Vision-Language Model
von: Lee, Youngwan, et al.
Veröffentlicht: (2025)
von: Lee, Youngwan, et al.
Veröffentlicht: (2025)
TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios
von: Yu, Qiucheng, et al.
Veröffentlicht: (2026)
von: Yu, Qiucheng, et al.
Veröffentlicht: (2026)
SIA: Enhancing Safety via Intent Awareness for Vision-Language Models
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
Lingua-SafetyBench: A Benchmark for Safety Evaluation of Multilingual Vision-Language Models
von: Shi, Enyi, et al.
Veröffentlicht: (2026)
von: Shi, Enyi, et al.
Veröffentlicht: (2026)
A Vision-Language Foundation Model for Leaf Disease Identification
von: Quoc, Khang Nguyen, et al.
Veröffentlicht: (2025)
von: Quoc, Khang Nguyen, et al.
Veröffentlicht: (2025)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
Safety Alignment for Vision Language Models
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
Beyond-Labels: Advancing Open-Vocabulary Segmentation With Vision-Language Models
von: Rahman, Muhammad Atta ur, et al.
Veröffentlicht: (2025)
von: Rahman, Muhammad Atta ur, et al.
Veröffentlicht: (2025)
JailBound: Jailbreaking Internal Safety Boundaries of Vision-Language Models
von: Song, Jiaxin, et al.
Veröffentlicht: (2025)
von: Song, Jiaxin, et al.
Veröffentlicht: (2025)
When Large Vision-Language Models Meet Person Re-Identification
von: Wang, Qizao, et al.
Veröffentlicht: (2024)
von: Wang, Qizao, et al.
Veröffentlicht: (2024)
VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language Models
von: Mei, Hefei, et al.
Veröffentlicht: (2025)
von: Mei, Hefei, et al.
Veröffentlicht: (2025)
Learning to Prompt with Text Only Supervision for Vision-Language Models
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding
von: Shi, Liang, et al.
Veröffentlicht: (2024)
von: Shi, Liang, et al.
Veröffentlicht: (2024)
CLIP-HandID: Vision-Language Model for Hand-Based Person Identification
von: Baisa, Nathanael L., et al.
Veröffentlicht: (2025)
von: Baisa, Nathanael L., et al.
Veröffentlicht: (2025)
CILP-FGDI: Exploiting Vision-Language Model for Generalizable Person Re-Identification
von: Zhao, Huazhong, et al.
Veröffentlicht: (2025)
von: Zhao, Huazhong, et al.
Veröffentlicht: (2025)
An Application-Agnostic Automatic Target Recognition System Using Vision Language Models
von: Palladino, Anthony, et al.
Veröffentlicht: (2024)
von: Palladino, Anthony, et al.
Veröffentlicht: (2024)
Fast Certification of Vision-Language Models Using Incremental Randomized Smoothing
von: Nirala, A K, et al.
Veröffentlicht: (2023)
von: Nirala, A K, et al.
Veröffentlicht: (2023)
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
von: George, Franky, et al.
Veröffentlicht: (2026)
von: George, Franky, et al.
Veröffentlicht: (2026)
Zero-shot Hazard Identification in Autonomous Driving: A Case Study on the COOOL Benchmark
von: Picek, Lukas, et al.
Veröffentlicht: (2024)
von: Picek, Lukas, et al.
Veröffentlicht: (2024)
Color Names in Vision-Language Models
von: Gomez-Villa, Alexandra, et al.
Veröffentlicht: (2025)
von: Gomez-Villa, Alexandra, et al.
Veröffentlicht: (2025)
Towards Calibrating Prompt Tuning of Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2026)
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2026)
Improving Large Vision and Language Models by Learning from a Panel of Peers
von: Hernandez, Jefferson, et al.
Veröffentlicht: (2025)
von: Hernandez, Jefferson, et al.
Veröffentlicht: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2024)
von: Ali, Eman, et al.
Veröffentlicht: (2024)
GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2024)
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2024)
Active Prompt Learning in Vision Language Models
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
DynaSplat: Dynamic-Static Gaussian Splatting with Hierarchical Motion Decomposition for Scene Reconstruction
von: Deng, Junli, et al.
Veröffentlicht: (2025)
von: Deng, Junli, et al.
Veröffentlicht: (2025)
GeoVLM: Improving Automated Vehicle Geolocalisation Using Vision-Language Matching
von: Dagda, Barkin, et al.
Veröffentlicht: (2025)
von: Dagda, Barkin, et al.
Veröffentlicht: (2025)
When Language Model Guides Vision: Grounding DINO for Cattle Muzzle Detection
von: Dulal, Rabin, et al.
Veröffentlicht: (2025)
von: Dulal, Rabin, et al.
Veröffentlicht: (2025)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Integration of Object Detection and Small VLMs for Construction Safety Hazard Identification
von: Adil, Muhammad, et al.
Veröffentlicht: (2026) -
TrajGATFormer: A Graph-Based Transformer Approach for Worker and Obstacle Trajectory Prediction in Off-site Construction Environments
von: Alduais, Mohammed, et al.
Veröffentlicht: (2025) -
ErgoChat: a Visual Query System for the Ergonomic Risk Assessment of Construction Workers
von: Fan, Chao, et al.
Veröffentlicht: (2024) -
Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors?
von: Chen, Xuezheng, et al.
Veröffentlicht: (2025) -
HazardNet: A Small-Scale Vision Language Model for Real-Time Traffic Safety Detection at Edge Devices
von: Tami, Mohammad Abu, et al.
Veröffentlicht: (2025)