Supporting Program Comprehension through Fast Query response in Large-Scale Systems

Supporting Program Comprehension through Fast Query response in Large-Scale Systems
复制标题

通过大型系统中的快速查询响应支持程序理解

DOI:
10.1145/3387904.338926
复制
发表时间:
2020
期刊:
2020 IEEE/ACM 28th International Conference on Program Comprehension (ICPC
影响因子:
--
通讯作者:
Cleland-Huang, Jane
Cleland-Huang, Jane
中科院分区:
--
文献类型:
--
作者:
Lin, Jinfeng;Liu, Yalin;Cleland-Huang, Jane

文献摘要

参考文献

相似文献

软件可追溯性为包括程序理解在内的各种工程活动提供支持;然而,大型工业项目的完成可能充满挑战且艰巨。研究人员提出了自动追踪技术来创建、维护和利用追踪链接。存储库挖掘和深度学习等计算密集型技术已显示出提供准确跟踪链接的能力。由于实际的性能挑战,在工业规模上实现可信、自动跟踪技术的目标尚未成功实现。本文评估了在大型工业项目中部署有效、计算成本昂贵的可追溯性算法的高性能解决方案,并利用生成的跟踪链接来回答程序理解查询。我们比较评估了支持工业规模跟踪解决方案的四种不同平台,这些平台能够处理具有数百万工件的软件项目。我们证明,使用大数据框架构建的跟踪解决方案可以很好地扩展大型项目,并且我们的 Spark 实现优于关系数据库、图形数据库 (GraphDB) 和普通 Java 实现。这些发现与早期的结果相矛盾,早期的结果表明应该采用 GraphDB 解决方案来解决大规模跟踪问题。
Software traceability provides support for various engineering activities including Program Comprehension; however, it can be challenging and arduous to complete in large industrial projects. Researchers have proposed automated traceability techniques to create, maintain and leverage trace links. Computationally intensive techniques, such as repository mining and deep learning, have showed the capability to deliver accurate trace links. The objective of achieving trusted, automated tracing techniques at industrial scale has not yet been successfully accomplished due to practical performance challenges. This paper evaluates high-performance solutions for deploying effective, computationally expensive trace-ability algorithms in large scale industrial projects and leverages generated trace links to answer Program Comprehension Queries. We comparatively evaluate four different platforms for supporting industrial-scale tracing solutions, capable of tackling software projects with millions of artifacts. We demonstrate that tracing solutions built using big data frameworks scale well for large projects and that our Spark implementation outperforms relational database, graph database (GraphDB), and plain Java implementations. These findings contradict earlier results which suggested that GraphDB solutions should be adopted for large-scale tracing problems.
DOI: --
发表时间: 2014
期刊: Journal of management science
影响因子: --
作者:
อนิรุธ สืบสิงห์
通讯作者: อนิรุธ สืบสิงห์
DOI: 10.1145/2245276.2231943
发表时间: 2012-03
期刊: --
影响因子: --
作者:
Yonghee Shin;J. Cleland-Huang
通讯作者: Yonghee Shin;J. Cleland-Huang
DOI: 10.1145/1806799.1806828
发表时间: 2010-05
期刊: 2010 ACM/IEEE 32nd International Conference on Software Engineering
影响因子: --
作者:
Thomas Fritz;G. Murphy
通讯作者: Thomas Fritz;G. Murphy
MUMPS:马萨诸塞州综合医院公用事业多道程序系统
DOI: --
发表时间: 1976
期刊:
影响因子: --
作者:
Norman F. Hirst
通讯作者: Norman F. Hirst
DOI: 10.1109/re.2007.17
发表时间: 2007
期刊: 15th IEEE International Requirements Engineering Conference (RE 2007)
影响因子: --
作者:
Alex Dekhtyar;J. Hayes;S. Sundaram;E. A. Holbrook;O. Dekhtyar
通讯作者: O. Dekhtyar