A Moving Object Tracking Technique Using Few Frames with Feature Map Extraction and Feature Fusion

N/ACitations
Citations of this article
4Readers
Mendeley users who have this article in their library.

Abstract

Moving object tracking techniques using machine and deep learning require large datasets for neural model training. New strategies need to be invented that utilize smaller data training sizes to realize the impact of large-sized datasets. However, current research does not balance the training data size and neural parameters, which creates the problem of inadequacy of the information provided by the low visual data content for parameter optimization. To enhance the performance of moving object tracking that appears in only a few frames, this research proposes a deep learning model using an abundant encoder–decoder (a high-resolution transformer (HRT) encoder–decoder). An HRT encoder–decoder employs feature map extraction that focuses on high resolution feature maps that are more representative of the moving object. In addition, we employ the proposed HRT encoder–decoder for feature map extraction and fusion to reimburse the few frames that have the visual information. Our extensive experiments on the Pascal DOC19 and MS-DS17 datasets have implied that the HRT encoder–decoder abundant model outperforms those of previous studies involving few frames that include moving objects.

Cite

CITATION STYLE

APA

Alarfaj, A. A., & Mahmoud, H. A. H. (2022). A Moving Object Tracking Technique Using Few Frames with Feature Map Extraction and Feature Fusion. ISPRS International Journal of Geo-Information, 11(7). https://doi.org/10.3390/ijgi11070379

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free