TMBL kernels for CUDA GPUs compile faster using PTX: computational intelligence on consumer games and graphics hardware
TMBL kernels for CUDA GPUs compile faster using PTX: computational intelligence on consumer games and graphics hardware
复制标题
CUDA GPU 的 TMBL 内核使用 PTX 编译速度更快:消费游戏和图形硬件上的计算智能
DOI:
--
复制
发表时间:
2011
期刊:
影响因子:
--
通讯作者:
G. D. Magoulas
中科院分区:
文献类型:
--
作者:
Tony E. Lewis;G. D. Magoulas
Many of the most effective attempts to harness the power of the Graphics Processing Unit (GPU) to accelerate Genetic Programming (GP) have dynamically compiled code for individuals as they are to be evaluated. This approach executes very quickly on the GPU but is slow to compile, hence only vast data-sets fully reap its rewards. To reduce compilation time, we generate and compile code in the lower-level language PTX. We investigate this in the context of implementing Tweaking Mutation Behaviour Learning (TMBL) on the GPU. We find that for programs of 300 instructions, using PTX reduces the compile time 5.861 times and even increases the evaluation speed by 23.029%.