Should College Dropout Prediction Models Include Protected Attributes?
Should College Dropout Prediction Models Include Protected Attributes?
复制标题
大学辍学预测模型是否应该包含受保护的属性?
DOI:
--
复制
发表时间:
2021
期刊:
影响因子:
--
通讯作者:
René F. Kizilcec
中科院分区:
文献类型:
--
作者:
Renzhe Yu;Hansol Lee;René F. Kizilcec
Early identification of college dropouts can provide tremendous value for improving student success and institutional effectiveness, and predictive analytics are increasingly used for this purpose. However, ethical concerns have emerged about whether including protected attributes in these prediction models discriminates against underrepresented student groups and exacerbates existing inequities. We examine this issue in the context of a large U.S. research university with both residential and fully online degree-seeking students. Based on comprehensive institutional records for the entire student population across multiple years (N = 93,457), we build machine learning models to predict student dropout after one academic year of study and compare the overall performance and fairness of model predictions with or without four protected attributes (gender, URM, first-generation student, and high financial need). We find that including protected attributes does not impact the overall prediction performance and it only marginally improves the algorithmic fairness of predictions. These findings suggest that including protected attributes is preferable. We offer guidance on how to evaluate the impact of including protected attributes in a local context, where institutional stakeholders seek to leverage predictive analytics to support student success.