Reusing Deep Learning Models: Challenges and Directions in Software Engineering
Reusing Deep Learning Models: Challenges and Directions in Software Engineering
复制标题
DOI:
10.1109/jva60410.2023.00015
复制
发表时间:
2023-07
期刊:
影响因子:
--
通讯作者:
James C. Davis;Purvish Jajal;Wenxin Jiang;Taylor R. Schorlemmer;Nicholas Synovic;G. Thiruvathukal
中科院分区:
文献类型:
--
作者:
James C. Davis;Purvish Jajal;Wenxin Jiang;Taylor R. Schorlemmer;Nicholas Synovic;G. Thiruvathukal
Deep neural networks (DNNs) achieve state-of-the-art performance in many areas, including computer vision, system configuration, and question-answering. However, DNNs are expensive to develop, both in intellectual effort (e.g., devising new architectures) and computational costs (e.g., training). Re-using DNNs is a promising direction to amortize costs within a company and across the computing industry. As with any new technology, however, there are many challenges in re-using DNNs. These challenges include both missing technical capabilities and missing engineering practices. This vision paper describes challenges in current approaches to DNN re-use. We summarize studies of re-use failures across the spectrum of re-use techniques, including conceptual (e.g., re-using based on a research paper), adaptation (e.g., re-using by building on an existing implementation), and deployment (e.g., direct re-use on a new device). We outline possible advances that would improve each kind of re-use.