Skip to main navigation Skip to search Skip to main content

A Cross-Domain Threat Screening and Localization Framework Using Vision Transformers and Self-supervised Learning

  • Ammara Nasim*
  • , Muhammad Usman Akram
  • , Asad Mansoor Khan
  • , Muhammad Belal Afsar Khan
  • , Taimur Hassan
  • *Corresponding author for this work
  • National University of Sciences and Technology Pakistan

Research output: Chapter in Book/Published conference outputConference publication

Abstract

Due to the ever-changing global security landscape, various countries are constantly updating their safety protocols at designated airports and train stations. Consequently, there is an expanding inventory of prohibited items that are not permitted to be carried in luggage. This renders training of prior models on large datasets but lesser threat classes, ineffective. The scarcity of extensive, accurately labeled X-ray images with new threat items hinders the feasibility of training data-intensive, fully supervised deep models, primarily due to the high cost and time associated with manual labeling. In this paper, a self-supervised threat detection approach is discussed that is combined with vision transformer(VIT)-based pre-screening. The approach is adaptable to new classes and can easily detect new threat classes in the inference stage using very few annotated support images. The VIT based classification module showcases promising results with an accuracy of 98% and F1-score of 99%. The self-supervised threat detection module surpasses other self-supervised and weakly supervised frameworks with Mean Average Precision (mAP) of 0.89 and an Intersection over Union (IoU) value of 0.70 on GDXRAY dataset. The framework performs exceptionally well on SIXRAY dataset with an IOU score of 0.67 and mAP value of 0.865 despite being trained on GDXRAY dataset.

Original languageEnglish
Title of host publication2024 14th International Conference on Pattern Recognition Systems, ICPRS 2024
PublisherIEEE
Number of pages7
ISBN (Electronic)9798350375657
ISBN (Print)9798350375664
DOIs
Publication statusPublished - 23 Sept 2024
Event14th International Conference on Pattern Recognition Systems, ICPRS 2024 - London, United Kingdom
Duration: 15 Jul 202418 Jul 2024

Publication series

Name2024 14th International Conference on Pattern Recognition Systems, ICPRS 2024

Conference

Conference14th International Conference on Pattern Recognition Systems, ICPRS 2024
Country/TerritoryUnited Kingdom
CityLondon
Period15/07/2418/07/24

Keywords

  • Cross-domain
  • Self-supervised learning
  • threat localization
  • Vision transformer

Fingerprint

Dive into the research topics of 'A Cross-Domain Threat Screening and Localization Framework Using Vision Transformers and Self-supervised Learning'. Together they form a unique fingerprint.

Cite this