Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Vision-based Human Detection by Fine-Tuned SSD Models

Author 1: Tang Jin Cheng Author 2: Ahmad Fakhri Ab. Nasir Author 3: Anwar P. P. Abdul Majeed Author 4: Mohd Azraai Mohd Razman Author 5: Thai Li Lim
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 13, No. 11 · Published 2022

DOI: https://doi.org/10.14569/IJACSA.2022.0131143

Abstract

Human-robot interaction (HRI) and human-robot collaboration (HRC) has become more popular as the industries are taking initiative to idealize the era of automation and digitalization. Introduction of robots are often considered as a risk due to the fact that robots do not own the intelligent as human does. However, the literature that uses deep learning technologies as the base to improve HRI safety are limited, not to mention transfer learning approach. Hence, this study intended to empirically examine the efficacy of transfer learning approach in human detection task by fine-tuning the SSD models. A custom image dataset is developed by using the surveillance system in TT Vision Holdings Berhad and annotated accordingly. Thereafter, the dataset is partitioned into the train, validation, and test set by a ratio of 70:20:10. The learning behaviour of the models was monitored throughout the fine-tuning process via total loss graph. The result reveals that the SSD fine-tuned model with MobileNetV1 achieved 87.20% test AP, which is 6.1% higher than the SSD fine-tuned model with MobileNetV2. As a trade-off, the SSD fine-tuned model with MobileNetV1 attained 46.2 ms inference time on RTX 3070, which is 9.6 ms slower as compared to SSD fine-tuned model with MobileNetV2. Taking test AP as the key metric, SSD fine-tuned model with MobileNetV1 is considered as the best fine-tuned model in this study. In conclusion, it has shown that the transfer learning approach within the deep learning domain can help to protect human from the risk by detecting human at the first place.

Keywords

How to Cite this Article

Cheng, T. J., Nasir, A. F. A., Majeed, A. P. P. A., Razman, M. A. M., & Lim, T. L. (2022). Vision-based Human Detection by Fine-Tuned SSD Models. International Journal of Advanced Computer Science and Applications, 13(11). https://doi.org/10.14569/IJACSA.2022.0131143

Cheng, Tang Jin, et al.. "Vision-based Human Detection by Fine-Tuned SSD Models." International Journal of Advanced Computer Science and Applications, vol. 13, no. 11, 2022, https://doi.org/10.14569/IJACSA.2022.0131143.

@article{Cheng2022,
  title     = {Vision-based Human Detection by Fine-Tuned SSD Models},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {13},
  number    = {11},
  year      = {2022},
  publisher = {The Science and Information Organization},
  author    = {Tang Jin Cheng and Ahmad Fakhri Ab. Nasir and Anwar P. P. Abdul Majeed and Mohd Azraai Mohd Razman and Thai Li Lim},
  doi       = {10.14569/IJACSA.2022.0131143},
  url       = {https://doi.org/10.14569/IJACSA.2022.0131143}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.