A one-shot face detection and recognition using deep learning method for access control system

Tsai, Tsung-Han; Tsai, Chi-En; Chi, Po-Ting

doi:10.1007/s11760-022-02366-1

A one-shot face detection and recognition using deep learning method for access control system

Original Paper
Published: 30 October 2022

Volume 17, pages 1571–1579, (2023)
Cite this article

Signal, Image and Video Processing Aims and scope Submit manuscript

Tsung-Han Tsai¹,
Chi-En Tsai¹ &
Po-Ting Chi¹

410 Accesses
Explore all metrics

Abstract

In this paper, we propose a face detection and recognition system using deep learning method. It can be used as an access control system that performs face detection and recognition in real-time processing. Our goal is to achieve a one-shot recognition instead of traditional two-step methods. We use SSD as the main model for face detection and VGG-Face as the main model for face recognition. We perform the deep learning method through the collection of datasets. Moreover, we use some techniques, such as data augmentation, preprocessing of the image, and post-processing of the image to train the robust face detection and recognition subsystems. We use continuous frames as input to avoid false-positive cases and make the system output without wrong results. A real demonstration system is constructed to determine the identification of the laboratory members. We use 1280 × 960 resolution video for experimental testing and achieve about 30 fps speed under GPU acceleration.

This is a preview of subscription content, log in via an institution to check access.

Access this article

Log in via an institution

Price excludes VAT (USA)
Tax calculation will be finalised during checkout.

Instant access to the full article PDF.

Institutional subscriptions

MTCNN and FACENET Based Access Control System for Face Detection and Recognition

Article 01 January 2021

Deep Learning Feature Extraction Architectures for Real-Time Face Detection

Article 28 August 2023

Improved Face Recognition from Weighed Face Representations for Deepcam

Data availability

Please contact the author for data requests.

References

Qu, X., Wei, T., Peng C., Du, P.: “A fast face recognition system based on deep learning,” IEEE Conf. in proceedings of the 2018 11th international symposium on computational intelligence and design (ISCID), Hangzhou, China, 8–9 December 2018; pp. 289–292
Ranjan, R., et al.: A fast and accurate system for face detection, identification, and verification. IEEE Trans. Biometr. Behav. Ident. Sci. 1(2), 82–96 (2019)
Article Google Scholar
Rai, A., Karnani R., et al.: “An end-to-end real-time face identification and attendance system using convolutional neural networks,” 2019 IEEE 16th India Council International Conference (INDICON), 2019
Liu, W., Anguelov, D., Erhan D., et al.: SSD: Single shot multibox detector. European conference on computer vision. Springer, Amsterdam, The Netherlands, October, (2016) pp. 21–37
Parkhi, O.M., Vedaldi A., Zisserman, A.: Deep face recognition. proceedings of the british machine vision conference, Swansea, UK, September, pp. 41.1–41.12 (2015)
Dalal, N., Triggs, B.: Histograms of oriented gradients for human detection. IEEE Conf. Computer Vision and Pattern Recognition, San Diego, CA, USA, June (2005)
Lowe, D.G., Object recognition from local scale-invariant features. Proceedings of the international conference on computer vision, Kerkyra, Corfu, Greece, September 20–25, pp.1150–1157 (1999)
Yang, M., Kriegman, D., Ahuja, N.: Detecting faces in images: a survey. IEEE Trans. Pattern Anal. Mach. Intell. 24, 34–58 (2002)
Article Google Scholar
Zhang, Z., Zhang, C.: A survey of recent advances in face detection. Microsoft research technical report (2010)
Girshick, R., Donahue, J., Darrell, T., Malik, J., Rich feature hierarchies for accurate object detection and semantic segmentation. IEEE Conference on Computer Vision and Pattern Recognition, Columbus, OH, USA, Nov (2013)
Girshick, R.: Fast R-CNN. International conference on computer vision, Santiago, Chile, Apr (2015)
He, K., Zhang, X., Ren, S., Sun, J.: Spatial pyramid pooling in deep convolutional networks for visual recognition. IEEE Trans. Pattern Anal. Mach. Intell. 37(9), 1904–1916 (2015)
Article Google Scholar
Ren, S., He, K., Girshick, R., Sun, J.: Faster R-CNN: towards real-time object detection with region proposal networks. IEEE Trans. Pattern Anal. Mach. Intell. 39, 1137–1149 (2017)
Article Google Scholar
Shrivastava, A., Gupta, A., Girshick, R.: Training region based object detectors with online hard example mining. In Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pages 761–769 (2016)
Zhu, Y., Zhao, C., Wang, J., Zhao, X., Wu, Y., Lu, H.: Couplenet: Coupling global structure with local parts for object detection. In Proc. IEEE Int. Conf. Comput. Vis., pages 4126–4134 (2017)
Dai, J., Qi, H., Xiong, Y., Li, Y., Zhang, G., Hu, H., Wei, Y.: Deformable convolutional networks. In Proc. IEEE Int. Conf. Comput. Vis., pages 764–773 (2017)
Singh, B., Davis, L.S.: An analysis of scale invariance in object detection–snip. In: Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pages 3578–3587 (2018)
Shrivastava, A., Sukthankar, R., Malik, J., Gupta, A.: Beyond skip connections: top-down modulation for object detection. arXiv:1612.06851 (2016)
Tychsen-Smith, L., Petersson, L.: Improving object localization with fitness NMS and bounded iou loss. In Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pages 6877–6885 (2018)
Lin, T.-Y., Dollar, P., Girshick, R., He, K., Hariharan, B., Belongie, S.:, Feature pyramid networks for object detection. In: Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pp 2117–2125 (2017)
Cai, Z., Vasconcelos, N.: Cascade r-cnn: delving into high quality object detection. In Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pp 6154–6162 (2018)
He, K., Gkioxari, G., Dollár, P., Girshick, R.: Mask r-cnn. In Proc. IEEE Int. Conf. Comput. Vis., pp 2961–2969 (2017)
Redmon, J., Divvala, S., Girshick, R., Farhadi, A.: You only look once: unified, real-time object detection. IEEE Conf. Computer Vision and Pattern Recognition, Las Vegas, NV, USA, Jun (2016)
Redmon, J., Farhadi, A.: Yolov3: an incremental improvement. arXiv:1804.02767 (2018)
Bochkovskiy, A., Wang, C.-Y., Mark Liao, H.-Y.: YOLOv4: optimal speed and accuracy of object detection. arXiv.2004.10934 (2020)
Qi, D., Tan, W., Yao Q., Liu, J.: YOLO5Face: why reinventing a face detector. arXiv:2105.12931v1 (2021)
Zhang, S., Zhu, X., Lei, Z., Shi, H., Wang, X., Li, S.Z., (2017) FD: single shot scale-invariant face detector. ICCV
Najibi, M., Samangouei, P., Chellappa R., Davis, L.S.: SSH: single stage headless face detector. ICCV (2017)
Li, J., Wang, Y., Wang, C., Tai, Y., Qian, J., Yang, J., Wang, C., Li J., Huang, F.: DSFD: Dual Shot Face Detector. CVPR (2019)
Turk, M., Pentland, A.: Eigenfaces for recognition. J. Cogn. Neurosci. 3(1), 71–86 (1991)
Article Google Scholar
Turk, M., Pentland, A.: Face recognition using eigen faces. Proceeding of IEEE conference on computer vision and pattern recognition, Maui, HI, USA, June, pp. 586–591 (1991)
Belhumeur, P.N., Hespanha, J.P., Kriegman, D.J., Eigenfaces vs. fisherfaces: recognition using class specific linear projection. IEEE transaction pattern analysis and machine intelligence, 19, 7, pp 711–720 (1997)
Viola, P., Jones, M.: Robust real-time face detection”. Int. J. Comput. Vis. 57(2), 137–154 (2004)
Article Google Scholar
Zhuang, Z., Cheng, Y., Sun, Q., Wang, Y.: Robust face detection and analysis”. J. Electron. 3, 193–201 (2000)
Google Scholar
Taigman, Y., Yang, M., Ranzato, M., Wolf, L.: Deepface: closing the gap to human-level performance in face verification. IEEE conference on computer vision and pattern recognition, pp. 1701–1708 (2014)
Huang, G.B., Mattar, M., Berg, T., Learned-Miller, E.: Labeled faces in the wild: a database for studying face recognition in unconstrained environments (2008)
Sun, Y., Liang, D., Wang, X., Tang, X.: Deepid3: face recognition with very deep neural networks. arXiv:1502.00873 (2015)
Schroff, F., Kalenichenko, D., Philbin, J.: Facenet: a unified embedding for face recognition and clustering. IEEE conference on computer vision and pattern recognition, pp. 815–823 (2015)
Liu, W., Wen, Y., Yu, Z., Yang, M.: Large-margin softmax loss for convolutional neural networks. ICML 2, 7 (2016)
Google Scholar
Liu, W., Wen, Y., Yu, Z., Li, M., Raj, B., Song, L.: Sphereface: deep hypersphere embedding for face recognition, IEEE conference on computer vision and pattern recognition, pp. 212–220 (2017)
Wang, F., Cheng, J., Liu, W., Liu, H.: Additive margin softmax for face verification. IEEE Signal Process. Lett. 25(7), 926–930 (2018)
Article Google Scholar
Deng, J., Guo, J., Xue, N., Zafeiriou, S.: Arcface: Additive angular margin loss for deep face recognition. IEEE Conference on computer vision and pattern recognition, pp. 4690–4699 (2019)
Masi, I., Wu, Y., Hassner, T., Natarajan, P.: Deep face recognition: a survey. 31st SIBGRAPI conference on graphics, patterns and images (SIBGRAPI), IEEE, pp. 471–478 (2018)
Wang, Z., Wang, G., Huang, B., Xiong, Z., Hong, Q., Wu, H., Yi, P., Jiang, K., Wang, N., Pei, Y.: Masked face recognition dataset and application. arXiv:2003.09093v2 (2020)
Huang, B., Wang, Z., Wang, G., Jiang, K., Zeng, K., Han, Z., Tian, X., Yang, Y.: When face recognition meets occlusion: a new benchmark.” ICASSP (2021)
Yang, S., Luo, P., Loy, CC., Tang, X.: WIDER FACE: a face detection benchmark. CVPR (2016)
Huang, G.B., Ramesh, M., Berg, T., Learned-Miller, E.: Labeled faces in the wild: a database for studying face recognition in unconstrained environments. Univ. Massachusetts Amherst Tech. Rep. 1, 07–49 (2007)
Google Scholar
Bodla, N., Singh, B., Chellappa, R., Davis, L.S.: Soft-nms - improving object detection with one line of code. Proceedings of the IEEE international conference on computer vision, Venice, Italy, October, pp 5562–5570 (2017)
Zhong, Z., Zheng, L., Kang, G., Li, S., Yang, Y.: Random erasing data augmentation. arXiv:1708.04896 (2017)
DeVries T., Taylor, G.W.: Improved regularization of convolutional neural networks with cutout. arXiv:1708.04552. (2017)
Tamara, L., Berg, Alexander C. Berg, Edwards, J., Forsyth, D.A.: Who’s in the Picture? IEEE international conference on multimedia & expo workshops (ICMEW) (2016)

Download references

Funding

Not applicable.

Author information

Authors and Affiliations

Department of Electrical Engineering, National Central University, No.300, Jung-Da Rd., Zhongli, 320, Taiwan, ROC
Tsung-Han Tsai, Chi-En Tsai & Po-Ting Chi

Authors

Tsung-Han Tsai
View author publications
You can also search for this author in PubMed Google Scholar
Chi-En Tsai
View author publications
You can also search for this author in PubMed Google Scholar
Po-Ting Chi
View author publications
You can also search for this author in PubMed Google Scholar

Contributions

CET conceived and designed the study. PTC performed the experiments. THT reviewed and edited the manuscript. All authors read and approved the manuscript.

Corresponding author

Correspondence to Tsung-Han Tsai.

Ethics declarations

Conflict of interest

The authors declare that they have no conflict of interest.

Additional information

Publisher's Note

Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.

Rights and permissions

Springer Nature or its licensor (e.g. a society or other partner) holds exclusive rights to this article under a publishing agreement with the author(s) or other rightsholder(s); author self-archiving of the accepted manuscript version of this article is solely governed by the terms of such publishing agreement and applicable law.

Reprints and permissions

About this article

Cite this article

Tsai, TH., Tsai, CE. & Chi, PT. A one-shot face detection and recognition using deep learning method for access control system. SIViP 17, 1571–1579 (2023). https://doi.org/10.1007/s11760-022-02366-1

Download citation

Received: 14 July 2022
Revised: 14 July 2022
Accepted: 18 September 2022
Published: 30 October 2022
Issue Date: June 2023
DOI: https://doi.org/10.1007/s11760-022-02366-1

Keywords

Access this article

Log in via an institution

Price excludes VAT (USA)
Tax calculation will be finalised during checkout.

Instant access to the full article PDF.

Institutional subscriptions

A one-shot face detection and recognition using deep learning method for access control system

Abstract

Access this article

Similar content being viewed by others

MTCNN and FACENET Based Access Control System for Face Detection and Recognition

Deep Learning Feature Extraction Architectures for Real-Time Face Detection

Improved Face Recognition from Weighed Face Representations for Deepcam

Data availability

References

Funding

Author information

Authors and Affiliations

Contributions

Corresponding author

Ethics declarations

Conflict of interest

Additional information

Publisher's Note

Rights and permissions

About this article

Cite this article

Keywords

Navigation

A one-shot face detection and recognition using deep learning method for access control system

Abstract

Access this article

Similar content being viewed by others

MTCNN and FACENET Based Access Control System for Face Detection and Recognition

Deep Learning Feature Extraction Architectures for Real-Time Face Detection

Improved Face Recognition from Weighed Face Representations for Deepcam

Data availability

References

Funding

Author information

Authors and Affiliations

Contributions

Corresponding author

Ethics declarations

Conflict of interest

Additional information

Publisher's Note

Rights and permissions

About this article

Cite this article

Share this article

Keywords

Search

Navigation