Self-Supervised Multi-Scale Transformer with Attention-Guided Fusion for Efficient Crack Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Kyem, Blessing Agyei, Asamoah, Joshua Kofi, Denteh, Eugene, Danyo, Andrews, Aboah, Armstrong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Demographics-Informed Neural Network for Multi-Modal Spatiotemporal forecasting of Urban Growth and Travel Patterns Using Satellite Imagery
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2025)
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2025)
Integrating Travel Behavior Forecasting and Generative Modeling for Predicting Future Urban Mobility and Spatial Transformations
by: Denteh, Eugene, et al.
Published: (2025)
by: Denteh, Eugene, et al.
Published: (2025)
Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2026)
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2026)
PaveSync: A Unified and Comprehensive Dataset for Pavement Distress Analysis and Classification
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
Context-CrackNet: A Context-Aware Framework for Precise Segmentation of Tiny Cracks in Pavement images
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation
by: Kyem, Blessing Agyei, et al.
Published: (2024)
by: Kyem, Blessing Agyei, et al.
Published: (2024)
Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems
by: Kyem, Blessing Agyei, et al.
Published: (2026)
by: Kyem, Blessing Agyei, et al.
Published: (2026)
A Road-Conditioned Traffic Movie Prediction Network with Spatiotemporal and Structure-Consistent Learning
by: Asamoah, Joshua Kofi, et al.
Published: (2026)
by: Asamoah, Joshua Kofi, et al.
Published: (2026)
Advancing Pavement Distress Detection in Developing Countries: A Novel Deep Learning Approach with Locally-Collected Datasets
by: Kyem, Blessing Agyei, et al.
Published: (2024)
by: Kyem, Blessing Agyei, et al.
Published: (2024)
Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment
by: Kyem, Blessing Agyei, et al.
Published: (2026)
by: Kyem, Blessing Agyei, et al.
Published: (2026)
Task-Specific Dual-Model Framework for Comprehensive Traffic Safety Video Description and Analysis
by: Kyem, Blessing Agyei, et al.
Published: (2025)
by: Kyem, Blessing Agyei, et al.
Published: (2025)
Prompt-Guided Spatial Understanding with RGB-D Transformers for Fine-Grained Object Relation Reasoning
by: Muturi, Tanner, et al.
Published: (2025)
by: Muturi, Tanner, et al.
Published: (2025)
A Unified Detection Pipeline for Robust Object Detection in Fisheye-Based Traffic Surveillance
by: Owor, Neema Jakisa, et al.
Published: (2025)
by: Owor, Neema Jakisa, et al.
Published: (2025)
An Improved ResNet50 Model for Predicting Pavement Condition Index (PCI) Directly from Pavement Images
by: Danyo, Andrews, et al.
Published: (2025)
by: Danyo, Andrews, et al.
Published: (2025)
Visual Dominance and Emerging Multimodal Approaches in Distracted Driving Detection: A Review of Machine Learning Techniques
by: Dontoh, Anthony, et al.
Published: (2025)
by: Dontoh, Anthony, et al.
Published: (2025)
A Contextual Analysis of Driver-Facing and Dual-View Video Inputs for Distraction Detection in Naturalistic Driving Environments
by: Dontoh, Anthony, et al.
Published: (2025)
by: Dontoh, Anthony, et al.
Published: (2025)
A Review Paper of the Effects of Distinct Modalities and ML Techniques to Distracted Driving Detection
by: Dontoh, Anthony., et al.
Published: (2025)
by: Dontoh, Anthony., et al.
Published: (2025)
Attention-Guided Multi-Scale Local Reconstruction for Point Clouds via Masked Autoencoder Self-Supervised Learning
by: Cao, Xin, et al.
Published: (2025)
by: Cao, Xin, et al.
Published: (2025)
3D Object Detection and High-Resolution Traffic Parameters Extraction Using Low-Resolution LiDAR Data
by: Zhang, Linlin, et al.
Published: (2024)
by: Zhang, Linlin, et al.
Published: (2024)
Enhancing Traffic Safety with Parallel Dense Video Captioning for End-to-End Event Analysis
by: Shoman, Maged, et al.
Published: (2024)
by: Shoman, Maged, et al.
Published: (2024)
Low-Light Image Enhancement Framework for Improved Object Detection in Fisheye Lens Datasets
by: Tran, Dai Quoc, et al.
Published: (2024)
by: Tran, Dai Quoc, et al.
Published: (2024)
ScaleFusionNet: Transformer-Guided Multi-Scale Feature Fusion for Skin Lesion Segmentation
by: Qamar, Saqib, et al.
Published: (2025)
by: Qamar, Saqib, et al.
Published: (2025)
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
by: Lee, Sanghyeok, et al.
Published: (2024)
by: Lee, Sanghyeok, et al.
Published: (2024)
Adaptive Transformer Attention and Multi-Scale Fusion for Spine 3D Segmentation
by: Xiang, Yanlin, et al.
Published: (2025)
by: Xiang, Yanlin, et al.
Published: (2025)
EECD-Net: Energy-Efficient Crack Detection with Spiking Neural Networks and Gated Attention
by: Zhang, Shuo
Published: (2025)
by: Zhang, Shuo
Published: (2025)
Real-Time Helmet Violation Detection in AI City Challenge 2023 with Genetic Algorithm-Enhanced YOLOv5
by: Soltanikazemi, Elham, et al.
Published: (2023)
by: Soltanikazemi, Elham, et al.
Published: (2023)
LASFNet: A Lightweight Attention-Guided Self-Modulation Feature Fusion Network for Multimodal Object Detection
by: Hao, Lei, et al.
Published: (2025)
by: Hao, Lei, et al.
Published: (2025)
PaveSAM Segment Anything for Pavement Distress
by: Owor, Neema Jakisa, et al.
Published: (2024)
by: Owor, Neema Jakisa, et al.
Published: (2024)
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection
by: Ghadiya, Ayush, et al.
Published: (2024)
by: Ghadiya, Ayush, et al.
Published: (2024)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
by: Leem, Saebom, et al.
Published: (2024)
by: Leem, Saebom, et al.
Published: (2024)
Enhancing Road Crack Detection Accuracy with BsS-YOLO: Optimizing Feature Fusion and Attention Mechanisms
by: Tang, Jiaze, et al.
Published: (2024)
by: Tang, Jiaze, et al.
Published: (2024)
Cross-Layer Feature Self-Attention Module for Multi-Scale Object Detection
by: Xie, Dingzhou, et al.
Published: (2025)
by: Xie, Dingzhou, et al.
Published: (2025)
LungX: A Hybrid EfficientNet-Vision Transformer Architecture with Multi-Scale Attention for Accurate Pneumonia Detection
by: Yerzhanuly, Mansur
Published: (2025)
by: Yerzhanuly, Mansur
Published: (2025)
A Deformable Attention-Based Detection Transformer with Cross-Scale Feature Fusion for Industrial Coil Spring Inspection
by: Rossi, Matteo, et al.
Published: (2026)
by: Rossi, Matteo, et al.
Published: (2026)
Event-based Facial Keypoint Alignment via Cross-Modal Fusion Attention and Self-Supervised Multi-Event Representation Learning
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
Multimodal Attention-Enhanced Feature Fusion-based Weekly Supervised Anomaly Violence Detection
by: Kaneko, Yuta, et al.
Published: (2024)
by: Kaneko, Yuta, et al.
Published: (2024)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
by: Li, Shenglan, et al.
Published: (2025)
by: Li, Shenglan, et al.
Published: (2025)
ExFusion: Efficient Transformer Training via Multi-Experts Fusion
by: Ruan, Jiacheng, et al.
Published: (2026)
by: Ruan, Jiacheng, et al.
Published: (2026)
WP-CrackNet: A Collaborative Adversarial Learning Framework for End-to-End Weakly-Supervised Road Crack Detection
by: Ma, Nachuan, et al.
Published: (2025)
by: Ma, Nachuan, et al.
Published: (2025)
CASA: Cross-Attention over Self-Attention for Efficient Vision-Language Fusion
by: Böhle, Moritz, et al.
Published: (2025)
by: Böhle, Moritz, et al.
Published: (2025)
Similar Items
-
Demographics-Informed Neural Network for Multi-Modal Spatiotemporal forecasting of Urban Growth and Travel Patterns Using Satellite Imagery
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2025) -
Integrating Travel Behavior Forecasting and Generative Modeling for Predicting Future Urban Mobility and Spatial Transformations
by: Denteh, Eugene, et al.
Published: (2025) -
Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition
by: Denteh, Eugene Kofi Okrah, et al.
Published: (2026) -
PaveSync: A Unified and Comprehensive Dataset for Pavement Distress Analysis and Classification
by: Kyem, Blessing Agyei, et al.
Published: (2025) -
Context-CrackNet: A Context-Aware Framework for Precise Segmentation of Tiny Cracks in Pavement images
by: Kyem, Blessing Agyei, et al.
Published: (2025)