Statistics Surveys · 2022 · 784 citations · 277 references
Artificial IntelligenceReasoningInterpretable Machine LearningEngineeringMachine LearningData ScienceMachine Learning ModelAutomated ReasoningPredictive AnalyticsInterpretable MlModel InterpretabilityInterpretabilityComputer ScienceSymbolic Machine LearningExplainable AiRepresentation Learning
Interpretability in machine learning is essential for high‑stakes decisions and troubleshooting, and this survey serves as a starting point for researchers in the field. The paper aims to establish fundamental principles for interpretable ML, clarify misconceptions, and outline ten technical challenge areas. The authors present a framework of ten key challenge areas—ranging from sparse logical models to interpretable reinforcement learning—and discuss each in the context of principled interpretability.
Interpretability in machine learning (ML) is crucial for high stakes decisions and troubleshooting. In this work, we provide fundamental principles for interpretable ML, and dispel common misunderstandings that dilute the importance of this crucial topic. We also identify 10 technical challenge areas in interpretable machine learning and provide history and background on each problem. Some of these problems are classically important, and some are recent problems that have arisen in the last few years. These problems are: (1) Optimizing sparse logical models such as decision trees; (2) Optimization of scoring systems; (3) Placing constraints into generalized additive models to encourage sparsity and better interpretability; (4) Modern case-based reasoning, including neural networks and matching for causal inference; (5) Complete supervised disentanglement of neural networks; (6) Complete or even partial unsupervised disentanglement of neural networks; (7) Dimensionality reduction for data visualization; (8) Machine learning models that can incorporate physics and other generative or causal constraints; (9) Characterization of the “Rashomon set” of good models; and (10) Interpretable reinforcement learning. This survey is suitable as a starting point for statisticians and computer scientists interested in working in interpretable machine learning.
277
Scikit-learn: Machine Learning in Python
Fabián Pedregosa, Gaël Varoquaux, Alexandre Gramfort et al. · arXiv (Cornell University) · 2012 · 63.3K citations · Full text
Laurens van der Maaten, Geoffrey E. Hinton · Journal of Machine Learning Research · 2008 · 35.7K citations
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver et al. · Nature · 2015 · 28.8K citations
Artificial Intelligence, Engineering, Deep Reinforcement Learning +3
Greedy function approximation: A gradient boosting machine.
Jerome H. Friedman · The Annals of Statistics · 2001 · 27.3K citations · Full text