Bias-variance tradeoffs in program analysis
Bias-variance tradeoffs in program analysis
复制标题
DOI:
10.1145/2535838.2535853
复制
发表时间:
2014-01
期刊:
影响因子:
--
通讯作者:
Rahul Sharma;A. Nori;A. Aiken
中科院分区:
文献类型:
--
作者:
Rahul Sharma;A. Nori;A. Aiken
It is often the case that increasing the precision of a program analysis leads to worse results. It is our thesis that this phenomenon is the result of fundamental limits on the ability to use precise abstract domains as the basis for inferring strong invariants of programs. We show that bias-variance tradeoffs, an idea from learning theory, can be used to explain why more precise abstractions do not necessarily lead to better results and also provides practical techniques for coping with such limitations. Learning theory captures precision using a combinatorial quantity called the VC dimension. We compute the VC dimension for different abstractions and report on its usefulness as a precision metric for program analyses. We evaluate cross validation, a technique for addressing bias-variance tradeoffs, on an industrial strength program verification tool called YOGI. The tool produced using cross validation has significantly better running time, finds new defects, and has fewer time-outs than the current production version. Finally, we make some recommendations for tackling bias-variance tradeoffs in program analysis.