CNN-Based Action Recognition and Pose Estimation for Classifying Animal Behavior from Videos: A Survey

01/15/2023
by   Michael Perez, et al.
0

Classifying the behavior of humans or animals from videos is important in biomedical fields for understanding brain function and response to stimuli. Action recognition, classifying activities performed by one or more subjects in a trimmed video, forms the basis of many of these techniques. Deep learning models for human action recognition have progressed significantly over the last decade. Recently, there is an increased interest in research that incorporates deep learning-based action recognition for animal behavior classification. However, human action recognition methods are more developed. This survey presents an overview of human action recognition and pose estimation methods that are based on convolutional neural network (CNN) architectures and have been adapted for animal behavior classification in neuroscience. Pose estimation, estimating joint positions from an image frame, is included because it is often applied before classifying animal behavior. First, we provide foundational information on algorithms that learn spatiotemporal features through 2D, two-stream, and 3D CNNs. We explore motivating factors that determine optimizers, loss functions and training procedures, and compare their performance on benchmark datasets. Next, we review animal behavior frameworks that use or build upon these methods, organized by the level of supervision they require. Our discussion is uniquely focused on the technical evolution of the underlying CNN models and their architectural adaptations (which we illustrate), rather than their usability in a neuroscience lab. We conclude by discussing open research problems, and possible research directions. Our survey is designed to be a resource for researchers developing fully unsupervised animal behavior classification systems of which there are only a few examples in the literature.

READ FULL TEXT

page 2

page 10

page 11

page 13

page 15

page 17

page 23

page 24

research
06/13/2022

A Training Method For VideoPose3D With Ideology of Action Recognition

Action recognition and pose estimation from videos are closely related t...
research
12/18/2022

2D Pose Estimation based Child Action Recognition

We present a graph convolutional network with 2D pose estimation for the...
research
07/31/2018

Understanding human-human interactions: a survey

Many videos depict people, and it is their interactions that inform us o...
research
07/23/2022

Intelligent 3D Network Protocol for Multimedia Data Classification using Deep Learning

In videos, the human's actions are of three-dimensional (3D) signals. Th...
research
04/02/2021

UAV-Human: A Large Benchmark for Human Behavior Understanding with Unmanned Aerial Vehicles

Human behavior understanding with unmanned aerial vehicles (UAVs) is of ...
research
12/23/2015

Convolutional Architecture Exploration for Action Recognition and Image Classification

Convolutional Architecture for Fast Feature Encoding (CAFFE) [11] is a s...

Please sign up or login with your details

Forgot password? Click here to reset