Human Activity Recognition in Videos Using Deep Learning

Kumar, Mohit; Rana, Adarsh; Ankita; Yadav, Arun Kumar; Yadav, Divakar

doi:10.1007/978-3-031-27609-5_23

Mohit Kumar⁹,
Adarsh Rana⁹,
Ankita⁹,
Arun Kumar Yadav⁹ &
…
Divakar Yadav⁹

Part of the book series: Communications in Computer and Information Science ((CCIS,volume 1788))

Included in the following conference series:

International Conference on Soft Computing and its Engineering Applications

411 Accesses

Abstract

Human Activity Recognition (HAR) is a challenging classification task. In the past, it traditionally involved the identification of the movement and activities of a person based on sensor inputs, apply signal processing to receive features and fit the features into a machine learning model. In recent times, deep learning methods have shown good results in automatic Human Activity Recognition. In this paper, we propose a pre-trained CNN (Inception-v3) and LSTM based methodology for Human Activity Recognition. The proposed methodology is evaluated on the publicly available UCF-101 dataset. The results show that the proposed methodology outperforms recent state-of-art methods in terms of accuracy (79.21%) and top-5 accuracy (92.92%) on the HAR task.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Log in via an institution

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 79.99; Price excludes VAT (USA)

Softcover Book: USD 99.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

References

Ahmad, Z., Illanko, K., Khan, N., Androutsos, D.: Human action recognition using convolutional neural network and depth sensor data. In: Proceedings of the 2019 International Conference on Information Technology and Computer Communications, pp. 1–5 (2019)
Google Scholar
Avilés-Cruz, C., Ferreyra-Ramírez, A., Zúñiga-López, A., Villegas-Cortéz, J.: Coarse-fine convolutional deep-learning strategy for human activity recognition. Sensors 19(7), 1556 (2019)
Article Google Scholar
Banjarey, K., Sahu, S.P., Dewangan, D.K.: A survey on human activity recognition using sensors and deep learning methods. In: 2021 5th International Conference on Computing Methodologies and Communication (ICCMC), pp. 1610–1617. IEEE (2021)
Google Scholar
Choutas, V., Weinzaepfel, P., Revaud, J., Schmid, C.: Potion: Pose motion representation for action recognition. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 7024–7033 (2018)
Google Scholar
Das, Srijan, Thonnat, Monique, Sakhalkar, Kaustubh, Koperski, Michal, Bremond, Francois, Francesca, Gianpiero: A new hybrid architecture for human activity recognition from RGB-D videos. In: Kompatsiaris, Ioannis, Huet, Benoit, Mezaris, Vasileios, Gurrin, Cathal, Cheng, Wen-Huang., Vrochidis, Stefanos (eds.) MMM 2019. LNCS, vol. 11296, pp. 493–505. Springer, Cham (2019). https://doi.org/10.1007/978-3-030-05716-9_40
Chapter Google Scholar
El-Ghaish, H., Hussien, M.E., Shoukry, A., Onai, R.: Human action recognition based on integrating body pose, part shape, and motion. IEEE Access 6, 49040–49055 (2018)
Article Google Scholar
Geng, C., Song, J.: Human action recognition based on convolutional neural networks with a convolutional auto-encoder. In: 2015 5th International Conference on Computer Sciences and Automation Engineering (ICCSAE 2015), pp. 933–938. Atlantis Press (2016)
Google Scholar
Karpathy, A., Toderici, G., Shetty, S., Leung, T., Sukthankar, R., Fei-Fei, L.: Large-scale video classification with convolutional neural networks. In: Proceedings of the IEEE conference on Computer Vision and Pattern Recognition, pp. 1725–1732 (2014)
Google Scholar
Khan, S., et al.: Human action recognition: a paradigm of best deep learning features selection and serial based extended fusion. Sensors 21(23), 7941 (2021)
Article Google Scholar
Khattar, L., Kapoor, C., Aggarwal, G.: Analysis of human activity recognition using deep learning. In: 2021 11th International Conference on Cloud Computing, Data Science & Engineering (Confluence), pp. 100–104. IEEE (2021)
Google Scholar
Kong, Y., Fu, Y.: Human action recognition and prediction: A survey. arXiv preprint arXiv:1806.11230 (2018)
Kopuklu, O., Kose, N., Gunduz, A., Rigoll, G.: Resource efficient 3d convolutional neural networks. In: Proceedings of the IEEE/CVF International Conference on Computer Vision Workshops (2019)
Google Scholar
Mazari, A., Sahbi, H.: Mlgcn: Multi-laplacian graph convolutional networks for human action recognition. In: The British Machine Vision Conference (BMVC) (2019)
Google Scholar
Moussa, M.M., Hamayed, E., Fayek, M.B., El Nemr, H.A.: An enhanced method for human action recognition. J. Adv. Res. 6(2), 163–169 (2015)
Article Google Scholar
Orozco, C.I., Xamena, E., Buemi, M.E., Berlles, J.J.: Human action recognition in videos using a robust cnn lstm approach. Ciencia y Tecnologí 23–36 (2020)
Google Scholar
Özyer, T., Ak, D.S., Alhajj, R.: Human action recognition approaches with video datasets-a survey. Knowledge-Based Systems 222, 106995 (2021)
Article Google Scholar
Pan, T., Song, Y., Yang, T., Jiang, W., Liu, W.: Videomoco: Contrastive video representation learning with temporally adversarial examples. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 11205–11214 (2021)
Google Scholar
Pareek, P., Thakkar, A.: A survey on video-based human action recognition: recent updates, datasets, challenges, and applications. Artif. Intell. Rev. 54(3), 2259–2322 (2021)
Article Google Scholar
Pienaar, S.W., Malekian, R.: Human activity recognition using lstm-rnn deep neural network architecture. In 2019 IEEE 2nd Wireless Africa Conference (WAC), pp. 1–5. IEEE (2019)
Google Scholar
Roshan Singh, Alok Kumar Singh Kushwaha, Rajeev Srivastava, et al. Recent trends in human activity recognition-a comparative study. Cognitive Systems Research, 2022
Google Scholar
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah. Ucf101: A dataset of 101 human actions classes from videos in the wild. arXiv preprint arXiv:1212.0402, 2012
Sultani, W., Shah, M.: Human action recognition in drone videos using a few aerial training examples. Comput. Vis. Image Underst. 206, 103186 (2021)
Article Google Scholar
Ankit Vijayvargiya, Nidhi Kumari, Palak Gupta, and Rajesh Kumar. Implementation of machine learning algorithms for human activity recognition. In 2021 3rd International Conference on Signal Processing and Communication (ICPSC), pages 440–444. IEEE, 2021
Google Scholar
Michalis Vrigkas, Christophoros Nikou, and Ioannis A Kakadiaris. A review of human activity recognition methods. Frontiers in Robotics and AI, 2:28, 2015
Google Scholar
Wang, L., Yangyang, X., Cheng, J., Xia, H., Yin, J., Jiaji, W.: Human action recognition by learning spatio-temporal features with deep neural networks. IEEE access 6, 17913–17922 (2018)
Article Google Scholar
Xia, K., Huang, J., Wang, H.: Lstm-cnn architecture for human activity recognition. IEEE Access 8, 56855–56866 (2020)
Article Google Scholar
Joe Yue-Hei Ng, Matthew Hausknecht, Sudheendra Vijayanarasimhan, Oriol Vinyals, Rajat Monga, and George Toderici. Beyond short snippets: Deep networks for video classification. In Proceedings of the IEEE conference on computer vision and pattern recognition, pages 4694–4702, 2015
Google Scholar
Yi Zhu, Yang Long, Yu Guan, Shawn Newsam, and Ling Shao. Towards universal representation for unseen action recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition, pages 9436–9445, 2018
Google Scholar

Download references

Author information

Authors and Affiliations

Department of Computer Science and Engineering, NIT Hamirpur, Hamirpur, 177005, HP, India
Mohit Kumar, Adarsh Rana, Ankita, Arun Kumar Yadav & Divakar Yadav

Authors

Mohit Kumar
View author publications
You can also search for this author in PubMed Google Scholar
Adarsh Rana
View author publications
You can also search for this author in PubMed Google Scholar
Ankita
View author publications
You can also search for this author in PubMed Google Scholar
Arun Kumar Yadav
View author publications
You can also search for this author in PubMed Google Scholar
Divakar Yadav
View author publications
You can also search for this author in PubMed Google Scholar

Corresponding author

Correspondence to Mohit Kumar .

Editor information

Editors and Affiliations

Charotar University of Science and Technology, Changa, Anand, Gujarat, India
Kanubhai K. Patel
University of South Dakota, Vermillion, SD, USA
K. C. Santosh
Charotar University of Science and Technology, Changa, Anand, India
Atul Patel
Indian Statistical Institute, Kolkata, India
Ashish Ghosh

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Kumar, M., Rana, A., Ankita, Yadav, A.K., Yadav, D. (2023). Human Activity Recognition in Videos Using Deep Learning. In: Patel, K.K., Santosh, K.C., Patel, A., Ghosh, A. (eds) Soft Computing and Its Engineering Applications. icSoftComp 2022. Communications in Computer and Information Science, vol 1788. Springer, Cham. https://doi.org/10.1007/978-3-031-27609-5_23

Download citation

DOI: https://doi.org/10.1007/978-3-031-27609-5_23
Published: 08 March 2023
Publisher Name: Springer, Cham
Print ISBN: 978-3-031-27608-8
Online ISBN: 978-3-031-27609-5
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics

Human Activity Recognition in Videos Using Deep Learning