课题基金 / 基金详情

CAREER: Ethical Machine Learning in Health: Robustness in Data, Learning and Deployment

CAREER: Ethical Machine Learning in Health: Robustness in Data, Learning and Deployment
职业:健康领域的道德机器学习:数据、学习和部署的稳健性
批准号:
2339381
负责人:
Marzyeh Ghassemi
金额:
$60.0万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2024
资助国家:
美国
项目状态:
未结题
起止时间:
2024-07-01 至 2029-06-30

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
由于护理管理的日益复杂和大量数据的可用,健康是机器学习(ML)的一个具有巨大潜力的领域。最近的研究表明,医疗保健中的模型缺乏健壮性,并不是在所有患者和环境中都表现得同样好。最近在一般模型稳健性方面的工作未能转化为健康设置,部分原因是它们没有考虑到将在其中使用模型的患者、条件和背景的多样性。该项目将创造新的方法来提高模型的稳健性,并使研究人员能够针对更符合道德的部署。这项研究将确定数据使用和模型培训的改进,通过关注卫生数据的细微差别和复杂性,确定卫生领域可操作模型的优先顺序。归根结底,这些进展还将有助于贷款、教育和法律制度等其他高风险领域的机器学习,这些领域依赖于常规收集的数据来产生洞察力。除了这些进步带来的直接和长期的社会影响外,这项工作还将有助于为新的本科暑期课程奠定基础,重点是将更多、更多样化的学生带入健康领域的机器学习。患者安全的重要性加上糟糕的模型稳健性限制了ML在医疗保健中的实际应用,伦理部署需要开发方法和指标以确保最先进的模型是健壮的。该项目的目标是通过三种方法开发稳健的健康模型:确保表示和下游模型经受住不正确的数据关联,实现公平和稳健的模型学习,以及在测试期间增强对离群值数据的事后稳健性。首先,针对数据错误和变化的代表性稳健性,它将通过深度度量模型中的对比自我监督,在患者亚群和护理变化之间建立弹性模型。其次,在模型学习方面,它将改进稳定训练的算法,通过结合临床预测任务的私人和公共数据来平衡公平性和稳健性。第三,它将针对测试时间方法进行离群值检测,并将预先训练的模型扩展到涵盖少数族裔亚群。该项目将产生解决数据、学习和测试中的稳健性的方法,作为在道德上部署健康模型的关键步骤。该奖项反映了NSF的法定使命,并通过使用基金会的智力优势和更广泛的影响审查标准进行评估,被认为值得支持。
英文摘要
Health is an area of immense potential for machine learning (ML), due to the increasing complexity of care management and large volume of data becoming available. Recent work has shown that models in healthcare lack robustness, and do not perform equally well across all patients and settings. Recent work in general model robustness have failed to translate to health settings in part because they do not consider the diversity of patients, conditions, and contexts that models will be used in. This project will create new ways to improve model robustness, and empower researchers to target more ethical deployments. This research will identify improvements for data use and model training that prioritize actionable models in health, by focusing on the nuance and complexity of health data. Ultimately these advances will also contribute to machine learning in other high-stakes areas such as lending, education and legal systems, that rely on routinely collected data to generate insights. Beyond the direct and long-term societal impact of these advances, this work will help lay the foundation for a new undergraduate-focused summer course focusing on bringing a larger, and more diverse, pipeline of students into machine learning in health. The importance of patient safety combined with poor model robustness limits the practical utility of ML in healthcare, and ethical deployment requires developing methods and metrics to ensure state-of-the-art models are robust. This project targets three ways to develop robust health models: ensuring representations and downstream models withstand incorrect data associations, achieving fair and robust model learning, and enhancing post-hoc robustness to outlier data during testing. First, targeting representational robustness to data error and change, it will build resilient models across patient subpopulations and variations in care through contrastive self-supervision in deep metric models. Second, in model learning, it will improve algorithms for stable training, balancing fairness/robustness trade-offs by combining private and public data for clinical prediction tasks. Third, it will target test-time methods for outlier detection and extending pre-trained models to cover minority subgroups. The project will result in methods that address robustness in data, learning, and testing, as crucial steps toward ethically deploying health models.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金