Adaptive Reward Computation in Reinforcement Learning-Based Continuous Integration Testing

Adaptive Reward Computation in Reinforcement Learning-Based Continuous Integration Testing
复制标题

基于强化学习的持续集成测试中的自适应奖励计算

DOI:
10.1109/access.2021.3063232
复制
发表时间:
2021
期刊:
影响因子:
3.9
通讯作者:
Zhao Ruilian
Zhao Ruilian
中科院分区:
计算机科学3区
文献类型:
--
作者:
Yang Yang;Pan Chaoyue;Li Zheng;Zhao Ruilian

文献摘要

参考文献

相似文献

强化学习(RL)已应用于在连续整合(CI)测试中优先考虑测试案例,其中奖励起着至关重要的作用。已经证明,基于历史信息的奖励功能可以提高有效性
Reinforcement learning (RL) has been applied to prioritizing test cases in Continuous Integration (CI) testing, where the reward plays a crucial role. It has been demonstrated that historical information-based reward function can improve the effectiveness of the test case prioritization (TCP). However, the inherent character of frequent iterations in CI can produce a considerable accumulation of historical information, which may decrease TCP efficiency and result in slow feedback. In this paper, the partial historical information is considered in the reward computation, where sliding window techniques are adopted to capture the possible efficient information. Firstly, the fixed-size sliding window is introduced to set a fixed length of recent historical information for each CI test. Then dynamic sliding window techniques are proposed, where the window size is continuously adaptive to each CI testing. Two methods are proposed, the test suite-based dynamic sliding window and the individual test case-based dynamic sliding window. The empirical studies are conducted on fourteen industrial-level programs, and the results reveal that under limited time, the sliding window-based reward function can effectively improve the TCP effect, where the NAPFD (Normalized Average Percentage of Faults Detected) and Recall of the dynamic sliding windows are better than that of the fixed-size sliding window. In particular, the individual test case-based dynamic sliding window approach can rank 74.18% failed test cases in the top 50% of the sorting sequence, with 1.35% improvement of NAPFD and 6.66 positions increased in TTF (Test to Fail).
DOI: 10.1109/tnn.1998.712192
发表时间: 1998
期刊: IEEE Trans. Neural Networks
影响因子: --
作者:
R. S. Sutton;A. Barto
通讯作者: R. S. Sutton;A. Barto
DOI: 10.3233/jifs-181998
发表时间: 2019
期刊: J. Intell. Fuzzy Syst.
影响因子: --
作者:
Hosney Jahan;Ziliang Feng;S. Mahmud;Penglin Dong
通讯作者: Hosney Jahan;Ziliang Feng;S. Mahmud;Penglin Dong
DOI: 10.1145/3361242.3361258
发表时间: 2019-10
期刊: Proceedings of the 11th Asia-Pacific Symposium on Internetware
影响因子: --
作者:
Zhaolin Wu;Yang Yang-Yang;Zheng Li;Ruilian Zhao
通讯作者: Zhaolin Wu;Yang Yang-Yang;Zheng Li;Ruilian Zhao
DOI: --
发表时间: 2014
期刊: --
影响因子: --
作者:
Dan Dewey
通讯作者: Dan Dewey
DOI: 10.1109/access.2017.2685629
发表时间: 2017-03
期刊: IEEE Access
影响因子: 3.9
作者:
Mojtaba Shahin;Muhammad Ali Babar;Liming Zhu
通讯作者: Mojtaba Shahin;Muhammad Ali Babar;Liming Zhu