跳到主要导航 跳到搜索 跳到主要内容

Mutual Context Network for Jointly Estimating Egocentric Gaze and Action

  • Yifei Huang
  • , Minjie Cai*
  • , Zhenqiang Li
  • , Feng Lu
  • , Yoichi Sato
  • *此作品的通讯作者
  • The University of Tokyo
  • Hunan University

科研成果: 期刊稿件文章同行评审

摘要

In this work, we address two coupled tasks of gaze prediction and action recognition in egocentric videos by exploring their mutual context: the information from gaze prediction facilitates action recognition and vice versa. Our assumption is that during the procedure of performing a manipulation task, on the one hand, what a person is doing determines where the person is looking at. On the other hand, the gaze location reveals gaze regions which contain important and information about the undergoing action and also the non-gaze regions that include complimentary clues for differentiating some fine-grained actions. We propose a novel mutual context network (MCN) that jointly learns action-dependent gaze prediction and gaze-guided action recognition in an end-to-end manner. Experiments on multiple egocentric video datasets demonstrate that our MCN achieves state-of-the-art performance of both gaze prediction and action recognition. The experiments also show that action-dependent gaze patterns could be learned with our method.

源语言英语
文章编号9139335
页(从-至)7795-7806
页数12
期刊IEEE Transactions on Image Processing
29
DOI
出版状态已出版 - 2020

指纹

探究 'Mutual Context Network for Jointly Estimating Egocentric Gaze and Action' 的科研主题。它们共同构成独一无二的指纹。

引用此