课题基金 / 基金详情

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
开发用于医学成像应用的人工智能技术需要在大型和多样化的数据集上训练模型。目前,包括放射学和病理学图像在内的大型数据存储库的聚合受到对患者隐私的担忧的限制。为了成功地共享医学图像,机构必须能够快速准确地批量去识别大量图像。这个过程目前是手动的,而且很耗时。
英文摘要
Developing artificial intelligence technology for medical imaging applications requires training models on large and diverse datasets. Currently, aggregation of large data repositories, including radiology and pathology images, is limited by concerns around patient privacy. In order to successfully share medical images, an institution must be able to quickly and accurately de-identify large numbers of images in batches. This process is currently manual and time-consuming. We propose a pipeline to remove PHI from both radiology DICOM images and pathology whole slide images by leveraging machine learning, natural language processing, and compartmentalized workflow techniques to significantly reduce the human intervention needed to anonymize medical images. In addition to examining header data in the images, we will use optical character recognition and computer vision algorithms to detect text in any location or orientation in the image, then automatically record and subsequently purge these regions. These techniques will be configured to work on a variety of image types (CT, MRI, radiograph, etc) and cover multiple OEM vendors for both radiology and pathology images. This phase I statement of work will construct the software tools, methods, and datasets necessary to facilitate a phase II where the complex algorithms needed for autonomous deidentification will be developed. This phase II processing will be referred to throughout this document as the workflow.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金