Knowledge-Based Systems · 2021 · 41 citations · 44 references
Artificial IntelligenceEngineeringMachine LearningNeural Networks (Machine Learning)Ai FoundationIntelligent SystemsInterpretable Case-based ReasoningLanguage ProcessingSocial SciencesRepresentation LearningData ScienceInterpretabilityTraditional Multilayer PerceptronRobot LearningOptimal FeatureLarge Ai ModelFeature LearningMachine Learning ModelComputer ScienceNeural Networks (Computational Neuroscience)Deep LearningDeep Neural NetworksExplanation-based LearningModel InterpretabilityTwin SystemsExplainable Artificial IntelligenceExplainable AiIntelligent Systems Engineering
In this paper, the twin-systems approach is reviewed, implemented, and competitively tested as a post-hoc explanation-by-example solution to the eXplainable Artificial Intelligence (XAI) problem. In twin-systems, an opaque artificial neural network (ANN) is explained by “twinning” it with a more interpretable case-based reasoning (CBR) system, by mapping the feature weights from the former to the latter. Extensive comparative tests are performed, over four experiments, to determine the optimal feature-weighting method for such twin-systems. Twin-systems for traditional multilayer perceptron (MLP) networks (MLP–CBR twins), convolutional neural networks (CNNs; CNN–CBR twins), and transformers for NLP (BERT–CBR twins) are examined. In addition, Feature Activation Maps (FAMs) are explored to enhance explainability by providing an additional layer of explanatory insight. The wider implications of this research on XAI is discussed, and a code library is provided to ease replicability.
44
Learning Deep Features for Discriminative Localization
Bolei Zhou, Aditya Khosla, Àgata Lapedriza et al. · 2016 · 10.6K citations
Convolutional Neural Network, Engineering, Machine Learning +16
A Unified Approach to Interpreting Model Predictions
Scott Lundberg, Su‐In Lee · arXiv (Cornell University) · 2017 · 7.6K citations · Full text