Integration of Object Detection and Small VLMs for Construction Safety Hazard Identification
Fuente:
arXiv
Saved in:
| Main Authors: | Adil, Muhammad, Ahmed, Mehmood, Aqib, Muhammad, Gonzalez, Vicente A., Lee, Gaang, Mei, Qipei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using Vision Language Models for Safety Hazard Identification in Construction
by: Adil, Muhammad, et al.
Published: (2025)
by: Adil, Muhammad, et al.
Published: (2025)
TrajGATFormer: A Graph-Based Transformer Approach for Worker and Obstacle Trajectory Prediction in Off-site Construction Environments
by: Alduais, Mohammed, et al.
Published: (2025)
by: Alduais, Mohammed, et al.
Published: (2025)
Small Object Detection with YOLO: A Performance Analysis Across Model Versions and Hardware
by: Tariq, Muhammad Fasih, et al.
Published: (2025)
by: Tariq, Muhammad Fasih, et al.
Published: (2025)
Hybrid State-Space and GRU-based Graph Tokenization Mamba for Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2025)
by: Ahmad, Muhammad, et al.
Published: (2025)
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026)
by: George, Franky, et al.
Published: (2026)
DiffFormer: a Differential Spatial-Spectral Transformer for Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2024)
by: Ahmad, Muhammad, et al.
Published: (2024)
AR-Facilitated Safety Inspection and Fall Hazard Detection on Construction Sites
by: Liu, Jiazhou, et al.
Published: (2024)
by: Liu, Jiazhou, et al.
Published: (2024)
ErgoChat: a Visual Query System for the Ergonomic Risk Assessment of Construction Workers
by: Fan, Chao, et al.
Published: (2024)
by: Fan, Chao, et al.
Published: (2024)
Transformer-Driven Active Transfer Learning for Cross-Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2024)
by: Ahmad, Muhammad, et al.
Published: (2024)
Permutation-Aware Action Segmentation via Unsupervised Frame-to-Segment Alignment
by: Tran, Quoc-Huy, et al.
Published: (2023)
by: Tran, Quoc-Huy, et al.
Published: (2023)
Underwater Object Detection Enhancement via Channel Stabilization
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
From Words to Wavelengths: VLMs for Few-Shot Multispectral Object Detection
by: Nkegoum, Manuel, et al.
Published: (2025)
by: Nkegoum, Manuel, et al.
Published: (2025)
Robust Object Detection with Pseudo Labels from VLMs using Per-Object Co-teaching
by: Bhaskar, Uday, et al.
Published: (2025)
by: Bhaskar, Uday, et al.
Published: (2025)
HazardNet: A Small-Scale Vision Language Model for Real-Time Traffic Safety Detection at Edge Devices
by: Tami, Mohammad Abu, et al.
Published: (2025)
by: Tami, Mohammad Abu, et al.
Published: (2025)
CLGRPO: Reasoning Ability Enhancement for Small VLMs
by: Wang, Fanyi, et al.
Published: (2025)
by: Wang, Fanyi, et al.
Published: (2025)
LFA-Net: A Lightweight Network with LiteFusion Attention for Retinal Vessel Segmentation
by: Mehmood, Mehwish, et al.
Published: (2025)
by: Mehmood, Mehwish, et al.
Published: (2025)
Enhancing Wireless Device Identification through RF Fingerprinting: Leveraging Transient Energy Spectrum Analysis
by: Ahmed, Nisar, et al.
Published: (2025)
by: Ahmed, Nisar, et al.
Published: (2025)
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors
by: Jin, Sheng, et al.
Published: (2024)
by: Jin, Sheng, et al.
Published: (2024)
From Filters to VLMs: Benchmarking Defogging Methods through Object Detection and Segmentation Performance
by: Aryashad, Ardalan, et al.
Published: (2025)
by: Aryashad, Ardalan, et al.
Published: (2025)
General Hazard Detection
by: Ng, Stephanie, et al.
Published: (2026)
by: Ng, Stephanie, et al.
Published: (2026)
Empowering Small VLMs to Think with Dynamic Memorization and Exploration
by: Liu, Jiazhen, et al.
Published: (2025)
by: Liu, Jiazhen, et al.
Published: (2025)
What is YOLOv8: An In-Depth Exploration of the Internal Features of the Next-Generation Object Detector
by: Yaseen, Muhammad
Published: (2024)
by: Yaseen, Muhammad
Published: (2024)
What is YOLOv9: An In-Depth Exploration of the Internal Features of the Next-Generation Object Detector
by: Yaseen, Muhammad
Published: (2024)
by: Yaseen, Muhammad
Published: (2024)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
by: Akan, Adil Kaan
Published: (2025)
by: Akan, Adil Kaan
Published: (2025)
AdaptPrompt: Parameter-Efficient Adaptation of VLMs for Generalizable Deepfake Detection
by: Jiang, Yichen, et al.
Published: (2025)
by: Jiang, Yichen, et al.
Published: (2025)
Object Depth and Size Estimation using Stereo-vision and Integration with SLAM
by: Hamad, Layth, et al.
Published: (2024)
by: Hamad, Layth, et al.
Published: (2024)
Small Object Detection for Birds with Swin Transformer
by: Huo, Da, et al.
Published: (2025)
by: Huo, Da, et al.
Published: (2025)
ShrinkBox: Backdoor Attack on Object Detection to Disrupt Collision Avoidance in Machine Learning-based Advanced Driver Assistance Systems
by: Shahzad, Muhammad Zaeem, et al.
Published: (2025)
by: Shahzad, Muhammad Zaeem, et al.
Published: (2025)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
by: Singha, Mainak, et al.
Published: (2026)
by: Singha, Mainak, et al.
Published: (2026)
Semi-Supervised Object Detection: A Survey on Progress from CNN to Transformer
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
Spatial and Spatial-Spectral Morphological Mamba for Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2024)
by: Ahmad, Muhammad, et al.
Published: (2024)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
by: Shriram, Shashank, et al.
Published: (2025)
by: Shriram, Shashank, et al.
Published: (2025)
Teaching VLMs to Localize Specific Objects from In-context Examples
by: Doveh, Sivan, et al.
Published: (2024)
by: Doveh, Sivan, et al.
Published: (2024)
PIN: Positional Insert Unlocks Object Localisation Abilities in VLMs
by: Dorkenwald, Michael, et al.
Published: (2024)
by: Dorkenwald, Michael, et al.
Published: (2024)
VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
by: Lu, Zhijing, et al.
Published: (2026)
by: Lu, Zhijing, et al.
Published: (2026)
Semi-supervised Open-World Object Detection
by: Mullappilly, Sahal Shaji, et al.
Published: (2024)
by: Mullappilly, Sahal Shaji, et al.
Published: (2024)
Maritime Small Object Detection from UAVs using Deep Learning with Altitude-Aware Dynamic Tiling
by: Ahmed, Sakib, et al.
Published: (2025)
by: Ahmed, Sakib, et al.
Published: (2025)
Compositional Video Synthesis by Temporal Object-Centric Learning
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
Similar Items
-
Using Vision Language Models for Safety Hazard Identification in Construction
by: Adil, Muhammad, et al.
Published: (2025) -
TrajGATFormer: A Graph-Based Transformer Approach for Worker and Obstacle Trajectory Prediction in Off-site Construction Environments
by: Alduais, Mohammed, et al.
Published: (2025) -
Small Object Detection with YOLO: A Performance Analysis Across Model Versions and Hardware
by: Tariq, Muhammad Fasih, et al.
Published: (2025) -
Hybrid State-Space and GRU-based Graph Tokenization Mamba for Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2025) -
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026)