Recovering traceability links between code and documentation
Recovering traceability links between code and documentation
复制标题
DOI:
10.1109/tse.2002.1041053
复制
发表时间:
2002-10-01
影响因子:
7.4
通讯作者:
Merlo, E
中科院分区:
文献类型:
--
作者:
Antoniol, G;Canfora, G;Merlo, E
Software system documentation is almost always expressed informally in natural language and free text. Examples include requirement specifications, design documents, manual pages, system development journals, error logs, and related maintenance reports. We propose a method based on information retrieval to recover traceability links between source code and free text documents. A premise of our work is that programmers use meaningful names for program items, such as functions, variables, types, classes, and methods. We believe that the application-domain knowledge that programmers process when writing the code is often captured by the mnemonics for identifiers; therefore, the analysis of these mnemonics can help to associate high-level concepts with program concepts and vice-versa. We apply both a probabilistic and a vector space information retrieval model in two case studies to trace C++ source code onto manual pages and Java code to functional requirements. We compare the results of applying the two models, discuss the benefits and limitations, and describe directions for improvements.