Department of Mathematics
 Search | Help | Login | printable version

Math @ Duke





.......................

.......................


Publications [#373531] of Cynthia D. Rudin

Papers Published

  1. Chen, Z; Tan, S; Chajewska, U; Rudin, C; Caruana, R, Missing Values and Imputation in Healthcare Data: Can Interpretable Machine Learning Help?, Proceedings of Machine Learning Research, vol. 209 (January, 2023), pp. 86-99
    (last updated on 2024/11/20)

    Abstract:
    Missing values are a fundamental problem in data science. Many datasets have missing values that must be properly handled because the way missing values are treated can have large impact on the resulting machine learning model. In medical applications, the consequences may affect healthcare decisions. There are many methods in the literature for dealing with missing values, including state-of-the-art methods which often depend on black-box models for imputation. In this work, we show how recent advances in interpretable machine learning provide a new perspective for understanding and tackling the missing value problem. We propose methods based on high-accuracy glass-box Explainable Boosting Machines (EBMs) that can help users (1) gain new insights on missingness mechanisms and better understand the causes of missingness, and (2) detect – or even alleviate – potential risks introduced by imputation algorithms. Experiments on real-world medical datasets illustrate the effectiveness of the proposed methods.

 

dept@math.duke.edu
ph: 919.660.2800
fax: 919.660.2821

Mathematics Department
Duke University, Box 90320
Durham, NC 27708-0320