First Person Action Recognition Using Deep Learned Descriptors

资源分类

2019-12-23 |

63 |

Abstract

We focus on the problem of wearer’s action recognition in fifirst person a.k.a. egocentric videos. This problem is more challenging than third person activity recognition due to unavailability of wearer’s pose and sharp movements in the videos caused by the natural head motion of the wearer. Carefully crafted features based on hands and objects cues for the problem have been shown to be successful for limited targeted datasets. We propose convolutional neural networks (CNNs) for end to end learning and classifification of wearer’s actions. The proposed network makes use of egocentric cues by capturing hand pose, head motion and saliency map. It is compact. It can also be trained from relatively small number of labeled egocentric videos that are available. We show that the proposed network can generalize and give state of the art performance on various disparate egocentric action datasets.

上一篇：Simultaneous Optical Flow and Intensity Estimation from an Event Camera

下一篇：Learning to Match Aerial Images with Deep Attentive Architectures

用户评价

全部评价

还没有评论，说两句吧！

热门资源

A Mathematical Mo...

Direct democracy, where each voter casts one vo...
Learning to Predi...

Much of model-based reinforcement learning invo...
The Variational S...

Unlike traditional images which do not offer in...
Hierarchical Task...

We extend hierarchical task network planning wi...
Shape-based Autom...

We present an algorithm for automatic detection...

智能在线

400-630-6780
聆听.建议反馈

E-mail: support@tusaishared.com