Understanding Diabetes Heterogeneity via Mining Multimodality Interconnected Data
Understanding Diabetes Heterogeneity via Mining Multimodality Interconnected Data
批准号:
10644701
负责人:
Ji (Carl) Yang
金额:
$17.28万
依托单位:
依托单位国家:
美国
项目类别:
财政年份:
2023
资助国家:
美国
项目状态:
未结题
起止时间:
2023-05-01 至 2028-04-30
关键词:
AmputationBehavioralBiological MarkersBiotechnologyBlindnessCaringCharacteristicsChronic DiseaseClinicalClinical DataComplexComputer softwareComputersDataData ScienceData SetDetectionDevelopmentDiabetes MellitusDiseaseDisease ProgressionEarly DiagnosisEarly InterventionEconomic BurdenEconomicsElectronic Health RecordElectronicsEnvironmental Risk FactorEpidemicGenomicsGoalsGrainGraphHealthHealthcareHealthcare SystemsHeart DiseasesHeterogeneityHumanIndividualInformation NetworksInstitutionKidney DiseasesLife ExperienceMachine LearningMeasuresMethodsMiningModalityModelingModernizationMolecularNational Institute of Diabetes and Digestive and Kidney DiseasesNeural Network SimulationNon-Insulin-Dependent Diabetes MellitusOutcomePathologicPatientsPatternPhenotypePopulationPrediabetes syndromePrevalenceProductionRecordsResearchResearch PersonnelScientistSocial InteractionSourceStructureSurveysSymptomsSystemTrainingTranslatingTranslational ResearchUnited StatesUnited States National Institutes of Healthadvanced analyticsanalytical toolclinical phenotypeclinical practiceclinically actionablecohortcombinatorialcomorbiditycomplex datacomputer sciencecostcost effective treatmentdata structuredesigndiabetes managementdiabeticdiabetic patientdigitaleffective therapygraph neural networkhealth datahigh riskimprovedindividual patientindividual variationinsightmachine learning modelmultimodal datamultimodalitymultiple omicsnovelopen sourceoutcome predictionpatient responsepersonalized medicineprecision medicineprototypesecondary analysissensorsocialsocial factorstherapy designtime intervaltreatment effecttreatment strategy
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Understanding Diabetes Heterogeneity via Mining Multimodality Interconnected Data
Abstract. Diabetes is a prevalent and highly heterogeneous disease that incurs tremendous human, economic, and so-
cial costs globally. Prediabetes and early-stage type 2 diabetes often do not have single strong indicates or symptoms,
posing great challenges for early detection and intervention. Moreover, once into a later-stage, diabetic patients are at
high risk of developing various health problems such as heart disease, vision loss and kidney disease, which further com-
plicates effective healthcare and may eventually lead to consequences from blindness to amputations to limited social
interactions due to mobility. Unfortunately, current subtyping of diabetes has failed to decouple such heterogeneity.
Recent remarkable advances in biotechnology have led to a significant production of high throughput patient data such
as electronic health records (EHRs), multi-omics, and structured surveys, providing tremendous promises to powerful
quantitative approaches towards the understanding of diabetes heterogeneity. However, existing machine learning (ML)
models ignore the higher-order interconnections among various disease variables, thus failing to differentiate complex
fine-grained subtypes and extract subtle corresponding phenotypes– regarding specific combinations of disease vari-
ables. Moreover, most existing studies focus on single sources of data such as clinical, molecular or behavioral, failing to
discover integrative biomarkers towards even more effective disease detection, analysis and treatment. Often case, these
methods also rely heavily on dataset-specific feature preprocessing and fail to transfer from one cohort to another.
As a computer scientist aiming at bridging data science and diabetes management, I have developed a well-structured
training pan in this proposed K25 project, and my primary goal is to develop a high-impact and practical ML system that
can be used to perform precise detection, in-depth analysis and cost-effective treatment of diabetes. To fully decouple
the heterogeneity of diabetes from complex patient data, I propose (1) a hyper-hetero-graph data structure (H2G) to
facilitate the comprehensive representation of patients and deep identification of diabetic characteristics regarding the
interconnections among various disease variables and (2) a specialized graph neural network model (H2GNN) capable
of modeling H2G along with a temporal component to capture the full trajectories of disease progression and a self-
clustering component to identify novel subtypes of diabetes. Leveraging the national All of Us dataset from NIH with
EHRs, genomics and surveys of 329K+ patients (42K+ diabetic), I propose to (1) apply the model to clinical data (EHRs)
towards precise early diabetes detection, (2) incorporate molecular data (genomics) towards in-depth diabetes patho-
logical analysis, and (3) further incorporate behavioral data (surveys) towards personalized diabetes treatment design.
Finally, we will iteratively evaluate our system and its discovered subtypes based on both All of Us and our independent
local dataset NELL curated from Emory Healthcare System with multimodality data of 295K patients (39K diabetic).
This project is consistent with NIDDK's commitment to translational research on chronic diseases, and aligns well with
ADA's recent initiative towards diabetes precision medicine. Results of this project will also inform the development of
powerful quantitative approaches for a broader spectrum of diseases.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
国内基金
海外基金
Behavioral Insights on Cooperation in Social Dilemmas
-
批准号:--
-
项目类别:外国优秀青年学者研究基金项目
-
资助金额:--
-
批准年份:2024
-
负责人:LIEN,Jaimie Wei-Hung
-
依托单位: