Managing Larger Data on a GitHub Repository

Managing Larger Data on a GitHub Repository
复制标题

管理 GitHub 存储库上的更大数据

DOI:
--
复制
发表时间:
2018
影响因子:
--
通讯作者:
C. Boettiger
C. Boettiger
中科院分区:
--
文献类型:
--
作者:
C. Boettiger

文献摘要

被引文献

相似文献

GitHub已经成为学术研究中保存和共享软件驱动分析的核心组件(Ram,2013)。随着科学家采用这种工作流程,很快就出现了以相同方式管理与分析相关的数据的愿望。虽然小数据可以很容易地与源代码和分析脚本一起提交到GitHub存储库,但大于50 MB的文件则不能。现有的解决方案引入了显着的复杂性,并打破了共享的易用性(Boettiger,2018 a)。
GitHub has become a central component for preserving and sharing software-driven analysis in academic research (Ram, 2013). As scientists adopt this workflow, a desire to manage data associated with the analysis in the same manner soon emerges. While small data can easily be committed to GitHub repositories along-side source code and analysis scripts, files larger than 50 MB cannot. Existing work-arounds introduce significant complexity and break the ease of sharing (Boettiger, 2018a).