Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition

Alhadidi, Taqwa; Jaber, Ahmed; Jaradat, Shadi; Ashqar, Huthaifa; Elhenawy, Mohammed

Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition

dc.contributor.author	Alhadidi, Taqwa
dc.contributor.author	Jaber, Ahmed
dc.contributor.author	Jaradat, Shadi
dc.contributor.author	Ashqar, Huthaifa
dc.contributor.author	Elhenawy, Mohammed
dc.date.accessioned	2024-10-28T14:31:06Z
dc.date.available	2024-10-28T14:31:06Z
dc.date.issued	2024-06-15
dc.description.abstract	Object detection is a critical component of transportation systems, particularly for applications such as autonomous driving, traffic monitoring, and infrastructure maintenance. Traditional object detection methods often struggle with limited data and variability in object appearance. The Oriented Window Learning Vision Transformer (OWL-ViT) offers a novel approach by adapting window orientations to the geometry and existence of objects, making it highly suitable for detecting diverse roadway assets. This study leverages OWL-ViT within a one-shot learning framework to recognize transportation infrastructure components, such as traffic signs, poles, pavement, and cracks. This study presents a novel method for roadway asset detection using OWL-ViT. We conducted a series of experiments to evaluate the performance of the model in terms of detection consistency, semantic flexibility, visual context adaptability, resolution robustness, and impact of non-max suppression. The results demonstrate the high efficiency and reliability of the OWL-ViT across various scenarios, underscoring its potential to enhance the safety and efficiency of intelligent transportation systems.
dc.description.uri	https://arxiv.org/abs/2406.10712v1
dc.format.extent	16 pages
dc.genre	journal articles
dc.genre	preprints
dc.identifier	doi:10.13016/m23sru-vnsy
dc.identifier.uri	https://doi.org/10.48550/arXiv.2406.10712
dc.identifier.uri	http://hdl.handle.net/11603/36797
dc.language.iso	en
dc.relation.isAvailableAt	The University of Maryland, Baltimore County (UMBC)
dc.relation.ispartof	UMBC Faculty Collection
dc.relation.ispartof	UMBC Data Science
dc.rights	Attribution 4.0 International
dc.rights.uri	https://creativecommons.org/licenses/by/4.0/
dc.title	Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition
dc.type	Text
dcterms.creator	https://orcid.org/0000-0002-6835-8338

Files

Original bundle

Now showing 1 - 1 of 1

Name:: 2406.10712v1 (1).pdf
Size:: 6.52 MB
Format:: Adobe Portable Document Format

Download

Collections

UMBC Faculty Collection
UMBC Data Science