The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics
标准可解释模型:一种基于拉格朗日力学的可解释机器学习通用理论,用于演绎设计可解释方法
机构 * IBM Research (CH)(IBM研究院(瑞士)) ; University of Oxford (UK)(牛津大学(英国)) ; University of Cambridge (UK)(剑桥大学(英国)) ; KU Leuven (BE)(鲁汶大学(比利时)) ; Institute of Physics of the Czech Academy of Sciences (CZ)(捷克科学院物理研究所(捷克))
AI总结 提出标准可解释模型(SIM),基于拉格朗日力学从前提演绎出可解释性对称性和约束,通过最小化拉格朗日函数得到最优可解释模型,解决现有方法局限性并指导新方法设计。