AVX-512 extension to OpenQCD 1.6

AVX-512 extension to OpenQCD 1.6
复制标题

AVX-512 扩展至 OpenQCD 1.6

DOI:
--
复制
发表时间:
2018
期刊:
影响因子:
--
通讯作者:
J. Rantaharju
J. Rantaharju
中科院分区:
--
文献类型:
--
作者:
E. Bennett;M. Dawson;M. Mesiti;J. Rantaharju

文献摘要

被引文献

相似文献

我们发布了openQCD-1.6的扩展,其中包含使用Intel intrinsic的AVX-512矢量指令。最近的英特尔处理器支持在512位宽向量上进行操作的扩展指令集,从而增加了浮点运算和寄存器内存的容量。新功能的最佳使用需要将数据和浮点运算重新组织到这些更宽的矢量单位中。我们报告了AVX-512 OpenQCD扩展在使用Intel Knights Landing和Xeon Scalable (Skylake) cpu的集群上的实现和性能。在具有物理相关参数的完整HMC轨迹中,我们观察到性能提高了5%至10%。
We publish an extension of openQCD-1.6 with AVX-512 vector instructions using Intel intrinsics. Recent Intel processors support extended instruction sets with operations on 512-bit wide vectors, increasing both the capacity for floating point operations and register memory. Optimal use of the new capabilities requires reorganising data and floating point operations into these wider vector units. We report on the implementation and performance of the AVX-512 OpenQCD extension on clusters using Intel Knights Landing and Xeon Scalable (Skylake) CPUs. In complete HMC trajectories with physically relevant parameters we observe a performance increase of 5% to 10%.