跳到主要导航 跳到搜索 跳到主要内容

Matching-CNN meets KNN: Quasi-parametric human parsing

  • Si Liu
  • , Xiaodan Liang
  • , Luoqi Liu
  • , Xiaohui Shen
  • , Jianchao Yang
  • , Changsheng Xu
  • , Liang Lin
  • , Xiaochun Cao
  • , Shuicheng Yan
  • CAS - Institute of Information Engineering
  • National University of Singapore
  • Sun Yat-Sen University
  • Adobe Systems Incorporated
  • Chinese Academy of Sciences

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Both parametric and non-parametric approaches have demonstrated encouraging performances in the human parsing task, namely segmenting a human image into several semantic regions (e.g., hat, bag, left arm, face). In this work, we aim to develop a new solution with the advantages of both methodologies, namely supervision from annotated data and the flexibility to use newly annotated (possibly uncommon) images, and present a quasi-parametric human parsing model. Under the classic K Nearest Neighbor (KNN)-based nonparametric framework, the parametric Matching Convolutional Neural Network (M-CNN) is proposed to predict the matching confidence and displacements of the best matched region in the testing image for a particular semantic region in one KNN image. Given a testing image, we first retrieve its KNN images from the annotated/manually-parsed human image corpus. Then each semantic region in each KNN image is matched with confidence to the testing image using M-CNN, and the matched regions from all KNN images are further fused, followed by a superpixel smoothing procedure to obtain the ultimate human parsing result. The M-CNN differs from the classic CNN [12] in that the tailored cross image matching filters are introduced to characterize the matching between the testing image and the semantic region of a KNN image. The cross image matching filters are defined at different convolutional layers, each aiming to capture a particular range of displacements. Comprehensive evaluations over a large dataset with 7,700 annotated human images well demonstrate the significant performance gain from the quasi-parametric model over the state-of-the-arts [29, 30], for the human parsing task.

源语言英语
主期刊名IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2015
出版商IEEE Computer Society
1419-1427
页数9
ISBN(电子版)9781467369640
DOI
出版状态已出版 - 14 10月 2015
已对外发布
活动IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2015 - Boston, 美国
期限: 7 6月 201512 6月 2015

出版系列

姓名Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
07-12-June-2015
ISSN(印刷版)1063-6919

会议

会议IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2015
国家/地区美国
Boston
时期7/06/1512/06/15

学术指纹

探究 'Matching-CNN meets KNN: Quasi-parametric human parsing' 的科研主题。它们共同构成独一无二的学术指纹。

引用此