Video Action Understanding

Matthew S. Hutchinson,Vijay N. Gadepally

doi:10.1109/access.2021.3115476

Matthew S. Hutchinson, Vijay N. Gadepally

Open Access

https://doi.org/10.1109/access.2021.3115476

Copy DOI

Abstract

Many believe that the successes of deep learning on image understanding problems can be replicated in the realm of video understanding. However, due to the scale and temporal nature of video, the span of video understanding problems and the set of proposed deep learning solutions is arguably wider and more diverse than those of their 2D image siblings. Finding, identifying, and predicting actions are a few of the most salient tasks in this emerging and rapidly evolving field. With a pedagogical emphasis, this tutorial introduces and systematizes fundamental topics, basic concepts, and notable examples in supervised video action understanding. Specifically, we clarify a taxonomy of action problems, catalog and highlight video datasets, describe common video data preparation methods, present the building blocks of state-of-the art deep learning model architectures, and formalize domain-specific metrics to baseline proposed solutions. This tutorial is intended to be accessible to a general computer science audience and assumes a conceptual understanding of supervised learning.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: IEEE Access	Publication Date: Jan 1, 2021
Citations: 18	License type: CC BY-NC-ND 4.0

R Discovery Prime

R Discovery Prime

Video Action Understanding

Abstract

Talk to us

Similar Papers

More From: IEEE Access

Lead the way for us

Similar Papers

Ses Tanıma için Derin Öğrenme Mimarileri Üzerine Derleme
Yeşim Dokuz ... Zekeriya Tüfekci̇
European Journal of Science and Technology | VOL. -
Yeşim Dokuz, et. al.Yeşim Dokuz ... Zekeriya Tüfekci̇
30 Apr 2020
European Journal of Science and Technology | VOL. -

Generative Reasoning Integrated Label Noise Robust Deep Image Representation Learning.
Gencer Sumbul ... Begum Demir
IEEE Transactions on Image Processing | VOL. PP
Gencer Sumbul, et. al.Gencer Sumbul ... Begum Demir
01 Jan 2023
IEEE Transactions on Image Processing | VOL. PP

Deep self-taught learning for facial beauty prediction
Junying Gan ... Lichen Li
Neurocomputing | VOL. 144
Junying Gan, et. al.Junying Gan ... Lichen Li
05 Jun 2014
Neurocomputing | VOL. 144

A systematic review of deep transfer learning for machinery fault diagnosis
Chuan Li ... Edgar Estupinan
Neurocomputing | VOL. 407
Chuan Li, et. al.Chuan Li ... Edgar Estupinan
12 May 2020
Neurocomputing | VOL. 407

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Video Action Understanding

Abstract

Talk to us

Similar Papers

More From: IEEE Access