GLoltalX -- -- 从地方到全球对 " 黑盒 " AI 模式的 " 从地方到全球解释 " (GLocalX -- From Local to Global Explanations of Black Box AI Models)

Artificial Intelligence (AI) has come to prominence as one of the major components of our society, with applications in most aspects of our lives. In this field, complex and highly nonlinear machine learning models such as ensemble models, deep neural networks, and Support Vector Machines have consistently shown remarkable accuracy in solving complex tasks. Although accurate, AI models often are "black boxes" which we are not able to understand. Relying on these models has a multifaceted impact and raises significant concerns about their transparency. Applications in sensitive and critical domains are a strong motivational factor in trying to understand the behavior of black boxes. We propose to address this issue by providing an interpretable layer on top of black box models by aggregating "local" explanations. We present GLocalX, a "local-first" model agnostic explanation method. Starting from local explanations expressed in form of local decision rules, GLocalX iteratively generalizes them into global explanations by hierarchically aggregating them. Our goal is to learn accurate yet simple interpretable models to emulate the given black box, and, if possible, replace it entirely. We validate GLocalX in a set of experiments in standard and constrained settings with limited or no access to either data or local explanations. Experiments show that GLocalX is able to accurately emulate several models with simple and small models, reaching state-of-the-art performance against natively global solutions. Our findings show how it is often possible to achieve a high level of both accuracy and comprehensibility of classification models, even in complex domains with high-dimensional data, without necessarily trading one property for the other. This is a key requirement for a trustworthy AI, necessary for adoption in high-stakes decision making applications.

翻译：人工智能(AI)已成为我们社会的主要组成部分之一,成为我们生活的大多数方面的应用。在这个领域,复杂且高度非线性机器学习模型,如连字符模型、深神经网络和矢量机支持等,在解决复杂任务时始终表现出惊人的准确性。虽然准确性,但AI模型往往是我们无法理解的“黑盒 ” 。依靠这些模型会产生多方面的影响,并引起对其透明度的极大关注。在敏感和关键领域的应用是试图理解黑盒行为的一个非常复杂的激励因素。我们提议通过集成“本地”解释,在黑盒模型顶部提供一个可解释的准确性格。我们介绍GLotelX,“当地一级”模型在解决复杂任务时一贯性解释。从以当地决策规则的形式表达的地方解释开始,GlocalX反复概括这些模型成为全球解释的缩略图。我们的目标是学习准确而简单易解的模型,以便效仿黑盒的行为,如果可能的话,则完全取代它。我们经常在黑盒上提供一个可解释性的域域域域域域域,我们用一个可解释性模型来进行一个可解释的精确性实验,在标准和高度模型中进行。我们的标准和高度的模型, 显示一些数据显示,在标准和高度的模型中, 显示高度的实验性能显示高度的状态。我们的标准和高度的实验性能显示高度的状态。我们的标准和高度的模型。

相关内容

MoDELS

关注 44

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

一文读懂可解释机器学习简史，让你的模型再也不是「Black Box」

专知会员服务

39+阅读 · 2020年10月30日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日