跳到主要导航 跳到搜索 跳到主要内容

Bandit Interpretability of Deep Models via Confidence Selection

  • Xiaoyue Duan
  • , Hong Li*
  • , Panpan Wang
  • , Tiancheng Wang
  • , Boyu Liu
  • , Baochang Zhang
  • *此作品的通讯作者
  • Beihang University
  • Nanchang Institute of Technology
  • Zhongguancun Laboratory

科研成果: 期刊稿件文章同行评审

摘要

Interpretability of black-box deep models is yet challenging because existing model-agnostic methods mainly locally explain the behavior of the classifier by learning a linear proxy around the instance being predicted. The explanation can be faithful locally, but may not be accurate globally. In this paper, we for the first time formulate the interpretation of classifiers as a bandit problem and introduce a Bandit Interpretation method via Confidence Selection (BICS). We statistically impose disturbances on different arms (image regions) and examine non-linear changes of the model's output to fairly select important regions via Upper Confidence Bounds (UCB). Unlike previous model-agnostic methods that directly occlude super-pixels, our method softly applies perturbations at a pixel level and thus can fully explore more regions with multiple granularities, leading to a more precise and robust interpretation. Quantitative and qualitative experimental results demonstrate that our approach provides reasonable and precise explanations for various image recognition tasks on different models.

源语言英语
期刊论文编号126250
期刊Neurocomputing
544
DOI
出版状态已出版 - 1 8月 2023

学术指纹

探究 'Bandit Interpretability of Deep Models via Confidence Selection' 的科研主题。它们共同构成独一无二的学术指纹。

引用此