LgNet: A Local-global Network for Action Recognition and Beyond

编辑：映维 | 分类：XR | 2022年11月10日

Note: We don't have the ability to review paper

PubDate: July 2022

Teams: Beihang University

Writers: Jiaqi Zhou; Zehua Fu; Qiuyu Huang; Qingjie Liu; Yunhong Wang

PDF: LgNet: A Local-global Network for Action Recognition and Beyond

LgNet: A Local-global Network for Action Recognition and Beyond

Abstract

This work addresses the task of action recognition in video sequences. In real world applications, this task is quite challenging due to the complex background of video content, the similarities between different types of actions, the dependence on a large amount of annotated data, and so on. Most of the existing methods fail to distinguish similar actions with the same static appearance and motion pattern. We attempt to address this issue from the perspective of a local-global view, considering videos as combinations of a set of action units (local semantic information) and their relations along temporal dimension (global relation information). To achieve this end, we propose a novel Local-global Networks (LgNet) to enhance recognition of similar action. Besides, we propose an end-to-end training method to decrease the reliance on annotated data. It combines self-supervised learning and supervised learning, which not only enables the model to learn video representations from a large number unannotated data but also avoids subsequent finetuning. The proposed training method can be flexibly equipped to a wide array of vision tasks. Experiments on several benchmark datasets show that our proposed model and training method achieve state-of-the-art performance.

本文链接：https://paper.nweon.com/13434

LgNet: A Local-global Network for Action Recognition and Beyond

您可能还喜欢...

最新AR/VR行业分享

最新AR/VR专利

最新AR/VR行业招聘

LgNet: A Local-global Network for Action Recognition and Beyond

您可能还喜欢...

Room2Room: Enabling Life-Size Telepresence in a Projected Augmented Reality Environment

Implementation and Evaluation of a 50 kHz, 28μs Motion-to-Pose Latency Head Tracking Instrument

Measuring User Experience of Mobile Augmented Reality Systems Through Non-instrumental Quality Attributes

最新AR/VR行业分享

最新AR/VR专利

最新AR/VR行业招聘