A new petabyte-scale data derivation framework for ATLAS

A new petabyte-scale data derivation framework for ATLAS
复制标题

ATLAS 的新 PB 级数据导出框架

DOI:
--
复制
发表时间:
2015
期刊:
影响因子:
--
通讯作者:
G. Stewart
G. Stewart
中科院分区:
--
文献类型:
--
作者:
J. Catmore;Justin Cranshaw;T. Gillam;E. Gramstad;P. Laycock;Nurcan Ozturk;G. Stewart

文献摘要

被引文献

相似文献

在LHC的长周期运行期间,ATLAS协作组根据运行1期间获得的经验对其分析模型进行了大修。该模型的一个重要组成部分是一个“衍生框架”,它采用ATLAS重建的PB级AOD输出,并产生样本,通常是TB级的大小,针对特定的分析。该框架结合了核心重建软件的所有功能,同时产生简单配置的输出。事件选择是通过一种简洁的特定于领域的语言来指定的,包括对逻辑操作的支持。输出内容可以高度优化,以最大限度地减少磁盘需求,同时保持相同的C++界面。该框架包括后期物理分析工具的接口,确保最终输出符合工具要求。最后,该框架允许为相同的输入产生多个输出,从而提供了优化计算资源配置的可能性。
During the Long Shutdown of the LHC, the ATLAS collaboration overhauled its analysis model based on experience gained during Run 1. A significant component of the model is a “Derivation Framework” that takes the petabyte-scale AOD output from ATLAS reconstruction and produces samples, typically terabytes in size, targeted at specific analyses. The framework incorporates all of the functionality of the core reconstruction software, while producing outputs that are simply configured. Event selections are specified via a concise domain-specific language, including support for logical operations. The output content can be highly optimised to minimise disk requirements, while maintaining the same C++ interface. The framework includes an interface to the late-stage physics analysis tools, ensuring that the final outputs are consistent with tool requirements. Finally, the framework allows several outputs to be produced for the same input, providing the possibility to optimise configurations to computing resources.