DIY Human Action Data Set Generation

by   Mehran Khodabandeh, et al.

The recent successes in applying deep learning techniques to solve standard computer vision problems has aspired researchers to propose new computer vision problems in different domains. As previously established in the field, training data itself plays a significant role in the machine learning process, especially deep learning approaches which are data hungry. In order to solve each new problem and get a decent performance, a large amount of data needs to be captured which may in many cases pose logistical difficulties. Therefore, the ability to generate de novo data or expand an existing data set, however small, in order to satisfy data requirement of current networks may be invaluable. Herein, we introduce a novel way to partition an action video clip into action, subject and context. Each part is manipulated separately and reassembled with our proposed video generation technique. Furthermore, our novel human skeleton trajectory generation along with our proposed video generation technique, enables us to generate unlimited action recognition training data. These techniques enables us to generate video action clips from an small set without costly and time-consuming data acquisition. Lastly, we prove through extensive set of experiments on two small human action recognition data sets, that this new data generation technique can improve the performance of current action recognition neural nets.


page 3

page 4

page 7


Video-based Human Action Recognition using Deep Learning: A Review

Human action recognition is an important application domain in computer ...

Procedural Generation of Videos to Train Deep Action Recognition Networks

Deep learning for human action recognition in videos is making significa...

Deep Neural Networks in Video Human Action Recognition: A Review

Currently, video behavior recognition is one of the most foundational ta...

Roweisposes, Including Eigenposes, Supervised Eigenposes, and Fisherposes, for 3D Action Recognition

Human action recognition is one of the important fields of computer visi...

Bringing Online Egocentric Action Recognition into the wild

To enable a safe and effective human-robot cooperation, it is crucial to...

MS-ASL: A Large-Scale Data Set and Benchmark for Understanding American Sign Language

Computer Vision has been improved significantly in the past few decades....

Iterative Projection and Matching: Finding Structure-preserving Representatives and Its Application to Computer Vision

The goal of data selection is to capture the most structural information...

Please sign up or login with your details

Forgot password? Click here to reset