Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Kyem, Blessing Agyei, Asamoah, Joshua Kofi, Dontoh, Anthony, Aboah, Armstrong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PaveSync: A Unified and Comprehensive Dataset for Pavement Distress Analysis and Classification
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation
by: Kyem, Blessing Agyei, et al.
Published: (2024)
by: Kyem, Blessing Agyei, et al.
Published: (2024)
Context-CrackNet: A Context-Aware Framework for Precise Segmentation of Tiny Cracks in Pavement images
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
A Road-Conditioned Traffic Movie Prediction Network with Spatiotemporal and Structure-Consistent Learning
by: Asamoah, Joshua Kofi, et al.
Published: (2026)
by: Asamoah, Joshua Kofi, et al.
Published: (2026)
Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems
by: Kyem, Blessing Agyei, et al.
Published: (2026)
by: Kyem, Blessing Agyei, et al.
Published: (2026)
Advancing Pavement Distress Detection in Developing Countries: A Novel Deep Learning Approach with Locally-Collected Datasets
by: Kyem, Blessing Agyei, et al.
Published: (2024)
by: Kyem, Blessing Agyei, et al.
Published: (2024)
Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2026)
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2026)
Self-Supervised Multi-Scale Transformer with Attention-Guided Fusion for Efficient Crack Detection
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
An Improved ResNet50 Model for Predicting Pavement Condition Index (PCI) Directly from Pavement Images
by: Danyo, Andrews, et al.
Published: (2025)
by: Danyo, Andrews, et al.
Published: (2025)
Demographics-Informed Neural Network for Multi-Modal Spatiotemporal forecasting of Urban Growth and Travel Patterns Using Satellite Imagery
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2025)
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2025)
Task-Specific Dual-Model Framework for Comprehensive Traffic Safety Video Description and Analysis
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
Integrating Travel Behavior Forecasting and Generative Modeling for Predicting Future Urban Mobility and Spatial Transformations
by: Denteh, Eugene, et al.
Published: (2025)
by: Denteh, Eugene, et al.
Published: (2025)
Prompt-Guided Spatial Understanding with RGB-D Transformers for Fine-Grained Object Relation Reasoning
by: Muturi, Tanner, et al.
Published: (2025)
by: Muturi, Tanner, et al.
Published: (2025)
A Contextual Analysis of Driver-Facing and Dual-View Video Inputs for Distraction Detection in Naturalistic Driving Environments
by: Dontoh, Anthony, et al.
Published: (2025)
by: Dontoh, Anthony, et al.
Published: (2025)
A Unified Detection Pipeline for Robust Object Detection in Fisheye-Based Traffic Surveillance
by: Owor, Neema Jakisa, et al.
Published: (2025)
by: Owor, Neema Jakisa, et al.
Published: (2025)
A Review Paper of the Effects of Distinct Modalities and ML Techniques to Distracted Driving Detection
by: Dontoh, Anthony., et al.
Published: (2025)
by: Dontoh, Anthony., et al.
Published: (2025)
Visual Dominance and Emerging Multimodal Approaches in Distracted Driving Detection: A Review of Machine Learning Techniques
by: Dontoh, Anthony, et al.
Published: (2025)
by: Dontoh, Anthony, et al.
Published: (2025)
PaveSAM Segment Anything for Pavement Distress
by: Owor, Neema Jakisa, et al.
Published: (2024)
by: Owor, Neema Jakisa, et al.
Published: (2024)
DART: A Vision-Language Foundation Model for Comprehensive Rope Condition Monitoring
by: Rani, Anju, et al.
Published: (2026)
by: Rani, Anju, et al.
Published: (2026)
Enhancing Traffic Safety with Parallel Dense Video Captioning for End-to-End Event Analysis
by: Shoman, Maged, et al.
Published: (2024)
by: Shoman, Maged, et al.
Published: (2024)
Automated Pavement Cracks Detection and Classification Using Deep Learning
by: Nafaa, Selvia, et al.
Published: (2024)
by: Nafaa, Selvia, et al.
Published: (2024)
Revisiting Vision Language Foundations for No-Reference Image Quality Assessment
by: Yadav, Ankit, et al.
Published: (2025)
by: Yadav, Ankit, et al.
Published: (2025)
Low-Light Image Enhancement Framework for Improved Object Detection in Fisheye Lens Datasets
by: Tran, Dai Quoc, et al.
Published: (2024)
by: Tran, Dai Quoc, et al.
Published: (2024)
PEDESTRIAN: An Egocentric Vision Dataset for Obstacle Detection on Pavements
by: Thoma, Marios, et al.
Published: (2025)
by: Thoma, Marios, et al.
Published: (2025)
Journey into Automation: Image-Derived Pavement Texture Extraction and Evaluation
by: Lu, Bingjie, et al.
Published: (2025)
by: Lu, Bingjie, et al.
Published: (2025)
Real-Time Helmet Violation Detection in AI City Challenge 2023 with Genetic Algorithm-Enhanced YOLOv5
by: Soltanikazemi, Elham, et al.
Published: (2023)
by: Soltanikazemi, Elham, et al.
Published: (2023)
Comparative Analysis of Advanced AI-based Object Detection Models for Pavement Marking Quality Assessment during Daytime
by: Antariksa, Gian, et al.
Published: (2025)
by: Antariksa, Gian, et al.
Published: (2025)
Deep Learning for Pavement Condition Evaluation Using Satellite Imagery
by: Lebaku, Prathyush Kumar Reddy, et al.
Published: (2025)
by: Lebaku, Prathyush Kumar Reddy, et al.
Published: (2025)
ELEV-VISION-SAM: Integrated Vision Language and Foundation Model for Automated Estimation of Building Lowest Floor Elevation
by: Ho, Yu-Hsuan, et al.
Published: (2024)
by: Ho, Yu-Hsuan, et al.
Published: (2024)
3D Object Detection and High-Resolution Traffic Parameters Extraction Using Low-Resolution LiDAR Data
by: Zhang, Linlin, et al.
Published: (2024)
by: Zhang, Linlin, et al.
Published: (2024)
PainFormer: a Vision Foundation Model for Automatic Pain Assessment
by: Gkikas, Stefanos, et al.
Published: (2025)
by: Gkikas, Stefanos, et al.
Published: (2025)
RoadFusion: Latent Diffusion Model for Pavement Defect Detection
by: Aqeel, Muhammad, et al.
Published: (2025)
by: Aqeel, Muhammad, et al.
Published: (2025)
Automated Wildfire Damage Assessment from Multi view Ground level Imagery Via Vision Language Models
by: Esparza, Miguel, et al.
Published: (2025)
by: Esparza, Miguel, et al.
Published: (2025)
Evaluating Attribute Comprehension in Large Vision-Language Models
by: Zhang, Haiwen, et al.
Published: (2024)
by: Zhang, Haiwen, et al.
Published: (2024)
Generalizable Disaster Damage Assessment via Change Detection with Vision Foundation Model
by: Ahn, Kyeongjin, et al.
Published: (2024)
by: Ahn, Kyeongjin, et al.
Published: (2024)
Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving
by: Theodoridis, Nikos, et al.
Published: (2026)
by: Theodoridis, Nikos, et al.
Published: (2026)
Modeling Multi-Granularity Context Information Flow for Pavement Crack Detection
by: Pang, Junbiao, et al.
Published: (2024)
by: Pang, Junbiao, et al.
Published: (2024)
Towards Vision-Language Geo-Foundation Model: A Survey
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
EVLF-FM: Explainable Vision Language Foundation Model for Medicine
by: Bai, Yang, et al.
Published: (2025)
by: Bai, Yang, et al.
Published: (2025)
A Vision-Language Foundation Model for Leaf Disease Identification
by: Quoc, Khang Nguyen, et al.
Published: (2025)
by: Quoc, Khang Nguyen, et al.
Published: (2025)
Similar Items
-
PaveSync: A Unified and Comprehensive Dataset for Pavement Distress Analysis and Classification
by: Kyem, Blessing Agyei, et al.
Published: (2025) -
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation
by: Kyem, Blessing Agyei, et al.
Published: (2024) -
Context-CrackNet: A Context-Aware Framework for Precise Segmentation of Tiny Cracks in Pavement images
by: Kyem, Blessing Agyei, et al.
Published: (2025) -
A Road-Conditioned Traffic Movie Prediction Network with Spatiotemporal and Structure-Consistent Learning
by: Asamoah, Joshua Kofi, et al.
Published: (2026) -
Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems
by: Kyem, Blessing Agyei, et al.
Published: (2026)