课题基金 / 基金详情

DATA MINING AND MODEL BUILDING IN MEDICAL INFORMATICS

DATA MINING AND MODEL BUILDING IN MEDICAL INFORMATICS
医疗信息学中的数据挖掘和模型构建
批准号:
6391275
负责人:
BRUCE G. BUCHANAN
金额:
$21.55万
依托单位国家:
美国
项目类别:
财政年份:
1999
资助国家:
美国
项目状态:
已结题
起止时间:
1999-05-01 至 2003-04-30

项目摘要

项目成果

BRUCE G. BUCHANAN的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Our long-term goal is to assist biomedical scientists by extracting and codifying new knowledge from large biomedical databases routinely by computer. As large collections of data become more readily accessibly, the opportunities for discovering new information increase. We propose here to work toward this goal by extending our prior research on machine learning in two important directions: (1) codification of disparate pieces of knowledge into a coherent model (model building), and (2) discovery of new information in medical databases (data mining). Machine learning programs find classification rules (or decision trees or networks) that separate members of a target class from other individuals. They have emphasized predictive accuracy, with some attention to tradeoffs between accuracy and cost of errors or between accuracy and simplicity. We propose a framework in which these, and other, tradeoffs are explicit and the criteria by which tradeoffs are made are available for modification. We also include semantic considerations among the criteria to control the internal coherence of models. "Data mining" is a recently-coined term for using computers to explore large databases, with a goal of discovering new relationships but usually with no specific target defined at the outset. In addition to accuracy, simplicity, coherence, and cost, a program that purports to discover new relationships must be able to assess novelty. We propose to measure the extent to which proposed relationships are novel by comparing them against existing knowledge in the domain of discourse, and to look for unusual rules (and other relations) that would be very interesting if true. The computer program we are primarily building on, RL, is a knowledge- based learning program that learns classification rules from a collection of data. RL has been demonstrated to be flexible enough to allow guidance from prior knowledge, and powerful enough to learn publishable information for scientists working in several different domains. Both parts of the research will requires extending the RL system in new ways detailed in the research plan, which are consistent with the overall design philosophy of the present system. We will primarily work with data already collected on pneumonia patients with with which we have considerable. We will test the generality of the criteria used to evaluate models and discoveries with a Baynesian Net learning. We will test the generality of the generality of the criteria used to evaluate models and discoveries with Bayesian Net learning system, K2.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
DATA MINING AND MODEL BUILDING IN MEDICAL INFORMATICS
ARTIFICIAL INTELLIGENCE METHODS FOR CRYSTALIZATION
DATA MINING AND MODEL BUILDING IN MEDICAL INFORMATICS
ARTIFICIAL INTELLIGENCE METHODS FOR CRYSTALIZATION
海外基金