DA-DGCEx:确保分发软件自动编码器损失的深方向反事实解释的有效性 (DA-DGCEx: Ensuring Validity of Deep Guided Counterfactual Explanations With Distribution-Aware Autoencoder Loss)

Deep Learning has become a very valuable tool in different fields, and no one doubts the learning capacity of these models. Nevertheless, since Deep Learning models are often seen as black boxes due to their lack of interpretability, there is a general mistrust in their decision-making process. To find a balance between effectiveness and interpretability, Explainable Artificial Intelligence (XAI) is gaining popularity in recent years, and some of the methods within this area are used to generate counterfactual explanations. The process of generating these explanations generally consists of solving an optimization problem for each input to be explained, which is unfeasible when real-time feedback is needed. To speed up this process, some methods have made use of autoencoders to generate instant counterfactual explanations. Recently, a method called Deep Guided Counterfactual Explanations (DGCEx) has been proposed, which trains an autoencoder attached a the classification model, in order to generate straightforward counterfactual explanations. However, this method does not ensure that the generated counterfactual instances are close to the data manifold, so unrealistic counterfactual instances may be generated. To overcome this issue, this paper presents Distribution Aware Deep Guided Counterfactual Explanations (DA-DGCEx), which adds a term to the DGCEx cost function that penalizes out of distribution counterfactual instances.

翻译：深层学习已成为不同领域一个非常宝贵的工具,没有人怀疑这些模型的学习能力。然而,深层学习模型由于缺乏可解释性,往往被视为黑盒,因此在决策过程中普遍存在着不信任。为了在有效性和可解释性之间找到平衡,近年来,可以解释的人工智能(XAI)越来越受欢迎,而且该领域的一些方法被用来产生直接反事实解释。这些解释的过程一般包括解决每个要解释的输入的优化问题,而当需要实时反馈时,这是不可行的。为了加快这一进程,有些方法已经利用自动编码器来产生即时反事实解释。最近,提出了一种名为“深导反事实解释(DGCExExExExex)”的方法,该方法培养了一个附有分类模型的自动编码器,以便产生直接反事实解释。然而,这种方法并不能确保生成的反事实实例接近数据多重,因此可能产生不切实际的反事实实例。为了克服这一问题,本文件展示了“深导反事实解释”(ExD)的传播成本功能,从而增加了“深导反事实解释(D)”一词。

相关内容

自编码器

关注 140

自动编码器是一种人工神经网络，用于以无监督的方式学习有效的数据编码。自动编码器的目的是通过训练网络忽略信号“噪声”来学习一组数据的表示（编码），通常用于降维。与简化方面一起，学习了重构方面，在此，自动编码器尝试从简化编码中生成尽可能接近其原始输入的表示形式，从而得到其名称。基本模型存在几种变体，其目的是迫使学习的输入表示形式具有有用的属性。自动编码器可有效地解决许多应用问题，从面部识别到获取单词的语义。

【图与几何深度学习】Graph and geometric deep learning，49页ppt

专知会员服务

65+阅读 · 2021年4月24日

最新《生成式对抗网络》简介，25页ppt

专知会员服务

175+阅读 · 2020年6月28日

因果图，Causal Graphs，52页ppt

专知会员服务

252+阅读 · 2020年4月19日

【CVPR2020-浙江大学-阿里巴巴】深层知识迁移的深层归因图，DEPARA: Deep Attribution Graph for Deep Knowledge Transferability

专知会员服务

29+阅读 · 2020年4月17日