Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Alhadidi, Taqwa, Jaber, Ahmed, Jaradat, Shadi, Ashqar, Huthaifa I, Elhenawy, Mohammed |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Pavement Crack Classification with Bidirectional Cascaded Neural Networks
by: Alhadidi, Taqwa I., et al.
Published: (2025)
by: Alhadidi, Taqwa I., et al.
Published: (2025)
Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing
by: Ashqar, Huthaifa I., et al.
Published: (2024)
by: Ashqar, Huthaifa I., et al.
Published: (2024)
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
by: Elhenawy, Mohammed, et al.
Published: (2025)
by: Elhenawy, Mohammed, et al.
Published: (2025)
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
by: Ashqar, Huthaifa I., et al.
Published: (2024)
by: Ashqar, Huthaifa I., et al.
Published: (2024)
Exploring Traffic Crash Narratives in Jordan Using Text Mining Analytics
by: Jaradat, Shadi, et al.
Published: (2024)
by: Jaradat, Shadi, et al.
Published: (2024)
Advancing Roadway Sign Detection with YOLO Models and Transfer Learning
by: Nafaa, Selvia, et al.
Published: (2024)
by: Nafaa, Selvia, et al.
Published: (2024)
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
by: Elhenawy, Mohammed, et al.
Published: (2025)
by: Elhenawy, Mohammed, et al.
Published: (2025)
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems
by: Elhenawy, Mohammed, et al.
Published: (2024)
by: Elhenawy, Mohammed, et al.
Published: (2024)
Visual Reasoning and Multi-Agent Approach in Multimodal Large Language Models (MLLMs): Solving TSP and mTSP Combinatorial Challenges
by: Elhenawy, Mohammed, et al.
Published: (2024)
by: Elhenawy, Mohammed, et al.
Published: (2024)
Visual Reasoning at Urban Intersections: FineTuning GPT-4o for Traffic Conflict Detection
by: Masri, Sari, et al.
Published: (2025)
by: Masri, Sari, et al.
Published: (2025)
Automated Pavement Cracks Detection and Classification Using Deep Learning
by: Nafaa, Selvia, et al.
Published: (2024)
by: Nafaa, Selvia, et al.
Published: (2024)
Using Multimodal Large Language Models for Automated Detection of Traffic Safety Critical Events
by: Tami, Mohammad Abu, et al.
Published: (2024)
by: Tami, Mohammad Abu, et al.
Published: (2024)
HazardNet: A Small-Scale Vision Language Model for Real-Time Traffic Safety Detection at Edge Devices
by: Tami, Mohammad Abu, et al.
Published: (2025)
by: Tami, Mohammad Abu, et al.
Published: (2025)
Multimodal Large Language Models for Enhanced Traffic Safety: A Comprehensive Review and Future Trends
by: Tami, Mohammad Abu, et al.
Published: (2025)
by: Tami, Mohammad Abu, et al.
Published: (2025)
Large Language Models (LLMs) as Traffic Control Systems at Urban Intersections: A New Paradigm
by: Masri, Sari, et al.
Published: (2024)
by: Masri, Sari, et al.
Published: (2024)
Automated Question Generation for Science Tests in Arabic Language Using NLP Techniques
by: Tami, Mohammad, et al.
Published: (2024)
by: Tami, Mohammad, et al.
Published: (2024)
Leveraging Large Language Models (LLMs) for Traffic Management at Urban Intersections: The Case of Mixed Traffic Scenarios
by: Masri, Sari, et al.
Published: (2024)
by: Masri, Sari, et al.
Published: (2024)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
by: Sammoudi, Mohammad, et al.
Published: (2024)
by: Sammoudi, Mohammad, et al.
Published: (2024)
Transformer Models in Education: Summarizing Science Textbooks with AraBART, MT5, AraT5, and mBART
by: Masri, Sari, et al.
Published: (2024)
by: Masri, Sari, et al.
Published: (2024)
3D Roadway Scene Object Detection with LIDARs in Snowfall Conditions
by: Farhani, Ghazal, et al.
Published: (2025)
by: Farhani, Ghazal, et al.
Published: (2025)
DeepHistoViT: An Interpretable Vision Transformer Framework for Histopathological Cancer Classification
by: Mosalpuri, Ravi, et al.
Published: (2026)
by: Mosalpuri, Ravi, et al.
Published: (2026)
Spatial Transform Decoupling for Oriented Object Detection
by: Yu, Hongtian, et al.
Published: (2023)
by: Yu, Hongtian, et al.
Published: (2023)
Efficient Oriented Object Detection with Enhanced Small Object Recognition in Aerial Images
by: Shi, Zhifei, et al.
Published: (2024)
by: Shi, Zhifei, et al.
Published: (2024)
Ride-sharing Determinants: Spatial and Spatio-temporal Bayesian Analysis for Chicago Service in 2022
by: Elkhouly, Mohamed, et al.
Published: (2024)
by: Elkhouly, Mohamed, et al.
Published: (2024)
Flexible ViG: Learning the Self-Saliency for Flexible Object Recognition
by: Zuo, Lin, et al.
Published: (2024)
by: Zuo, Lin, et al.
Published: (2024)
ViTOC: Vision Transformer and Object-aware Captioner
by: Huang, Feiyang
Published: (2024)
by: Huang, Feiyang
Published: (2024)
RQFormer: Rotated Query Transformer for End-to-End Oriented Object Detection
by: Zhao, Jiaqi, et al.
Published: (2023)
by: Zhao, Jiaqi, et al.
Published: (2023)
Real-Time Oriented Object Detection Transformer in Remote Sensing Images
by: Ding, Zeyu, et al.
Published: (2026)
by: Ding, Zeyu, et al.
Published: (2026)
DenSe-AdViT: A novel Vision Transformer for Dense SAR Object Detection
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
Towards Equitable AI: Detecting Bias in Using Large Language Models for Marketing
by: Yilmaz, Berk, et al.
Published: (2025)
by: Yilmaz, Berk, et al.
Published: (2025)
CarFormer: Self-Driving with Learned Object-Centric Representations
by: Hamdan, Shadi, et al.
Published: (2024)
by: Hamdan, Shadi, et al.
Published: (2024)
Unsupervised Fault Detection using SAM with a Moving Window Approach
by: Maged, Ahmed, et al.
Published: (2024)
by: Maged, Ahmed, et al.
Published: (2024)
i-WiViG: Interpretable Window Vision GNN
by: Obadic, Ivica, et al.
Published: (2025)
by: Obadic, Ivica, et al.
Published: (2025)
Hands-on Evaluation of Visual Transformers for Object Recognition and Detection
by: Vlachogiannis, Dimitrios N., et al.
Published: (2025)
by: Vlachogiannis, Dimitrios N., et al.
Published: (2025)
HCLSM: Hierarchical Causal Latent State Machines for Object-Centric World Modeling
by: Jaber, Jaber, et al.
Published: (2026)
by: Jaber, Jaber, et al.
Published: (2026)
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer
by: Elharrouss, Omar, et al.
Published: (2025)
by: Elharrouss, Omar, et al.
Published: (2025)
Enhancing Mathematics Learning for Hard-of-Hearing Students Through Real-Time Palestinian Sign Language Recognition: A New Dataset
by: Khandaqji, Fidaa, et al.
Published: (2025)
by: Khandaqji, Fidaa, et al.
Published: (2025)
LoViT: Long Video Transformer for Surgical Phase Recognition
by: Liu, Yang, et al.
Published: (2023)
by: Liu, Yang, et al.
Published: (2023)
ARS-DETR: Aspect Ratio-Sensitive Detection Transformer for Aerial Oriented Object Detection
by: Zeng, Ying, et al.
Published: (2023)
by: Zeng, Ying, et al.
Published: (2023)
Cross-level Attention with Overlapped Windows for Camouflaged Object Detection
by: Li, Jiepan, et al.
Published: (2023)
by: Li, Jiepan, et al.
Published: (2023)
Similar Items
-
Enhancing Pavement Crack Classification with Bidirectional Cascaded Neural Networks
by: Alhadidi, Taqwa I., et al.
Published: (2025) -
Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing
by: Ashqar, Huthaifa I., et al.
Published: (2024) -
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
by: Elhenawy, Mohammed, et al.
Published: (2025) -
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
by: Ashqar, Huthaifa I., et al.
Published: (2024) -
Exploring Traffic Crash Narratives in Jordan Using Text Mining Analytics
by: Jaradat, Shadi, et al.
Published: (2024)