AI in the Loop: functionalizing fold performance disagreement to monitor automated medical image segmentation workflows.
AI in the Loop: functionalizing fold performance disagreement to monitor automated medical image segmentation workflows.
复制标题
DOI:
10.3389/fradi.2023.1223294
复制
发表时间:
2023
期刊:
影响因子:
--
通讯作者:
Kline, Timothy L.
中科院分区:
文献类型:
--
作者:
Gottlich, Harrison C.;Korfiatis, Panagiotis;Gregory, Adriana V.;Kline, Timothy L.
关键词:
Methods that automatically flag poor performing predictions are drastically needed to safely implement machine learning workflows into clinical practice as well as to identify difficult cases during model training. Disagreement between the fivefold cross-validation sub-models was quantified using dice scores between folds and summarized as a surrogate for model confidence. The summarized Interfold Dices were compared with thresholds informed by human interobserver values to determine whether final ensemble model performance should be manually reviewed. The method on all tasks efficiently flagged poor segmented images without consulting a reference standard. Using the median Interfold Dice for comparison, substantial dice score improvements after excluding flagged images was noted for the in-domain CT (0.85 ± 0.20 to 0.91 ± 0.08, 8/50 images flagged) and MR (0.76 ± 0.27 to 0.85 ± 0.09, 8/50 images flagged). Most impressively, there were dramatic dice score improvements in the simulated out-of-distribution task where the model was trained on a radical nephrectomy dataset with different contrast phases predicting a partial nephrectomy all cortico-medullary phase dataset (0.67 ± 0.36 to 0.89 ± 0.10, 122/300 images flagged). Comparing interfold sub-model disagreement against human interobserver values is an effective and efficient way to assess automated predictions when a reference standard is not available. This functionality provides a necessary safeguard to patient care important to safely implement automated medical image segmentation workflows.
登录
查看更多内容
影响因子:
15.8
作者:
Vayena E;Blasimme A;Cohen IG
通讯作者:
Cohen IG
影响因子:
10.9
作者:
Bilic, Patrick;Christ, Patrick;Li, Hongwei Bran;Vorontsov, Eugene;Ben-Cohen, Avi;Kaissis, Georgios;Szeskin, Adi;Jacobs, Colin;Mamani, Gabriel Efrain Humpire;Chartrand, Gabriel;Lohoefer, Fabian;Holch, Julian Walter;Sommer, Wieland;Hofmann, Felix;Hostettler, Alexandre;Lev-Cohain, Naama;Drozdzal, Michal;Amitai, Michal Marianne;Vivanti, Refael;Sosna, Jacob;Ezhov, Ivan;Sekuboyina, Anjany;Navarro, Fernando;Kofler, Florian;Paetzold, Johannes C.;Shit, Suprosanna;Hu, Xiaobin;Lipkova, Jana;Rempfler, Markus;Piraud, Marie;Kirschke, Jan;Wiestler, Benedikt;Zhang, Zhiheng;Huelsemeyer, Christian;Beetz, Marcel;Ettlinger, Florian;Antonelli, Michela;Bae, Woong;Bellver, Miriam;Bi, Lei;Chen, Hao;Chlebus, Grzegorz;Dam, Erik B.;Dou, Qi;Fu, Chi-Wing;Georgescu, Bogdan;Giro-I-Nieto, Xavier;Gruen, Felix;Han, Xu;Heng, Pheng-Ann;Hesser, Jurgen;Moltz, Jan Hendrik;Igel, Christian;Isensee, Fabian;Jaeger, Paul;Jia, Fucang;Kaluva, Krishna Chaitanya;Khened, Mahendra;Kim, Ildoo;Kim, Jae-Hun;Kim, Sungwoong;Kohl, Simon;Konopczynski, Tomasz;Kori, Avinash;Krishnamurthi, Ganapathy;Li, Fan;Li, Hongchao;Li, Junbo;Li, Xiaomeng;Lowengrub, John;Ma, Jun;Maier-Hein, Klaus;Maninis, Kevis-Kokitsi;Meine, Hans;Merhof, Dorit;Pai, Akshay;Perslev, Mathias;Petersen, Jens;Pont-Tuset, Jordi;Qi, Jin;Qi, Xiaojuan;Rippel, Oliver;Roth, Karsten;Sarasua, Ignacio;Schenk, Andrea;Shen, Zengming;Torres, Jordi;Wachinger, Christian;Wang, Chunliang;Weninger, Leon;Wu, Jianrong;Xu, Daguang;Yang, Xiaoping;Yu, Simon Chun-Ho;Yuan, Yading;Yue, Miao;Zhang, Liping;Cardoso, Jorge;Bakas, Spyridon;Braren, Rickmer;Heinemann, Volker;Pal, Christopher;Tang, An;Kadoury, Samuel;Soler, Luc;van Ginneken, Bram;Greenspan, Hayit;Joskowicz, Leo;Menze, Bjoern
通讯作者:
Menze, Bjoern
影响因子:
13.6
作者:
Denic, Aleksandar;Elsherbiny, Hisham;Rule, Andrew D.
通讯作者:
Rule, Andrew D.
影响因子:
7.4
作者:
Shaw, James;Rudzicz, Frank;Goldfarb, Avi
通讯作者:
Goldfarb, Avi
影响因子:
2.4
作者:
Mueller, Sabine;Farag, Iva;Graf, Norbert
通讯作者:
Graf, Norbert