跳到主要导航 跳到搜索 跳到主要内容

Learning From Architectural Redundancy: Enhanced Deep Supervision in Deep Multipath Encoder-Decoder Networks

  • Ying Luo
  • , Jinhu Lu*
  • , Xiaolong Jiang
  • , Baochang Zhang
  • *此作品的通讯作者
  • Beihang University
  • Alibaba Group Holding Ltd.

科研成果: 期刊稿件文章同行评审

摘要

Deep encoder-decoders are the model of choice for pixel-level estimation due to their redundant deep architectures. Yet they still suffer from the vanishing supervision information issue that affects convergence because of their overly deep architectures. In this work, we propose and theoretically derive an enhanced deep supervision (EDS) method which improves on conventional deep supervision (DS) by incorporating variance minimization into the optimization. A new structure variance loss is introduced to build a bridge between deep encoder-decoders and variance minimization, and provides a new way to minimize the variance by forcing different intermediate decoding outputs (paths) to reach an agreement. We also design a focal weighting strategy to effectively combine multiple losses in a scale-balanced way, so that the supervision information is sufficiently enforced throughout the encoder-decoders. To evaluate the proposed method on the pixel-level estimation task, a novel multipath residual encoder is proposed and extensive experiments are conducted on four challenging density estimation and crowd counting benchmarks. The experimental results demonstrate the superiority of our EDS over other paradigms, and improved estimation performance is reported using our deeply supervised encoder-decoder.

源语言英语
页(从-至)4271-4284
页数14
期刊IEEE Transactions on Neural Networks and Learning Systems
33
9
DOI
出版状态已出版 - 1 9月 2022

学术指纹

探究 'Learning From Architectural Redundancy: Enhanced Deep Supervision in Deep Multipath Encoder-Decoder Networks' 的科研主题。它们共同构成独一无二的学术指纹。

引用此