ALGES: Active Learning with Gradient Embeddings for Semantic Segmentation of Laparoscopic Surgical Images
ALGES: Active Learning with Gradient Embeddings for Semantic Segmentation of Laparoscopic Surgical Images
复制标题
DOI:
--
复制
发表时间:
2022
期刊:
影响因子:
--
通讯作者:
Josiah Aklilu;Serena Yeung
中科院分区:
文献类型:
--
作者:
Josiah Aklilu;Serena Yeung
Annotating medical images for the purposes of training computer vision models is an ex-tremely laborious task that takes time and resources away from expert clinicians. Active learning (AL) is a machine learning paradigm that mitigates this problem by deliberately proposing data points that should be labeled in order to maximize model performance. We propose a novel AL algorithm for segmentation, ALGES, that utilizes gradient embeddings to effectively select laparoscopic images to be labeled by some external oracle while reducing annotation effort. Given any unlabeled image, our algorithm treats predicted segmenta-tions as truth and computes gradients with respect to the model parameters of the last layer in a segmentation network. The norms of these per-pixel gradient vectors correspond to the magnitude of the induced change in model parameters and contain rich information about the model’s predictive uncertainty. Our algorithm then computes gradients embeddings in two ways, and we employ a center-finding algorithm with these embeddings to procure representative and diverse batches in each round of AL. An advantage of our approach is extensibility to any model architecture and differentiable loss scheme for semantic segmentation. We apply our approach to a public data set of laparoscopic cholecystectomy images and show that it outperforms current AL algorithms in selecting the most informative data points for improving the segmentation model. Our code is available at https://github.com/josaklil-ai/surg-active-learning . encapsulate uncertainty information while providing a way to ensure diversity in acquired batches during AL. Deep learning models are trained via gradient updates, and data points that bring about large gradient updates are data points that the model should learn. Thus, uncertainty for an image can be captured by computing gradients of the weights for segmentation models, and representing these gradients as an image-level embedding. We have also shown that at the semantic level, considering pixels of a particular semantic class that induce large updates to the model parameters gives us insight into model uncertainty for particular regions of an image. We substantiate this intuition with empirical studies, but we recognize that this frame of reference can provide meaningful insight in understanding uncertainty for deep learning models.