2018 · 193 citations · 18 references
Artificial IntelligenceEngineeringMachine LearningVerificationSoftware AnalysisFormal VerificationData ScienceAeqitas ApproachFairness (Computer Systems)Fair Data PrincipleFairness TestingDecision MakingMachine-learning ModelsAlgorithmic BiasFair Resource AllocationComputer ScienceProgram AnalysisAutomated ReasoningSoftware TestingAlgorithmic Fairness
Fairness is a critical trait in decision making. As machine-learning models are increasingly being used in sensitive application domains (e.g. education and employment) for decision making, it is crucial that the decisions computed by such models are free of unintended bias. But how can we automatically validate the fairness of arbitrary machine-learning models? For a given machine-learning model and a set of sensitive input parameters, our Aeqitas approach automatically discovers discriminatory inputs that highlight fairness violation. At the core of Aeqitas are three novel strategies to employ probabilistic search over the input space with the objective of uncovering fairness violation. Our Aeqitas approach leverages inherent robustness property in common machine-learning models to design and implement scalable test generation methodologies. An appealing feature of our generated test inputs is that they can be systematically added to the training set of the underlying model and improve its fairness. To this end, we design a fully automated module that guarantees to improve the fairness of the model. We implemented Aeqitas and we have evaluated it on six stateof- the-art classifiers. Our subjects also include a classifier that was designed with fairness in mind. We show that Aeqitas effectively generates inputs to uncover fairness violation in all the subject classifiers and systematically improves the fairness of respective models using the generated test inputs. In our evaluation, Aeqitas generates up to 70% discriminatory inputs (w.r.t. the total number of inputs generated) and leverages these inputs to improve the fairness up to 94%.
18
UCI Machine Learning Repository
Arthur Asuncion · Medical Entomology and Zoology · 2007 · 24.3K citations
Practical Black-Box Attacks against Machine Learning
Nicolas Papernot, Patrick McDaniel, Ian Goodfellow et al. · 2017 · 3.4K citations
Artificial Intelligence, Deep Neural Networks, Engineering +13
Cynthia Dwork, Moritz Hardt, Toniann Pitassi et al. · 2012 · 3.3K citations
The Unreasonable Effectiveness of Data
Alon Halevy, Peter Norvig, Fernando Pereira · IEEE Intelligent Systems · 2009 · 1.8K citations
Artificial Intelligence, Structured Prediction, Engineering +27
Certifying and Removing Disparate Impact
Michael Feldman, Sorelle A. Friedler, John Moeller et al. · 2015 · 1.7K citations
Protected Class, Discrimination, Law +16