A Collaborative Approach to Research Data Management in a Web Archive Context

A Collaborative Approach to Research Data Management in a Web Archive Context
复制标题

网络档案环境中研究数据管理的协作方法

DOI:
--
复制
发表时间:
2018
期刊:
影响因子:
--
通讯作者:
J. Kamps
J. Kamps
中科院分区:
--
文献类型:
--
作者:
Hugo C. Huurdeman;J. Kamps

文献摘要

被引文献

相似文献

在我们这个时代,研究人员创建和收集他们的数据集,不可避免地从单纯的“小”数据变成越来越“大”的数据。机构提供越来越多的服务来管理,存储和访问这些重要的研究数据。然而,尽管存储库提供了宝贵的服务,研究人员进行的无数操作和广义存储库提供的存储服务之间仍然存在不匹配:通常,只有研究过程的最终产品被存储。这可能意味着在研究过程中丢失了有价值的操作和数据转换,实际上也使未来的研究人员更难解释这些数据。为了解决这个问题,并使研究人员和机构更紧密地联系在一起,我们提出了一种更具协作性的方法,可能涉及研究人员的整个工作流程,从数据创建到使用和潜在的重用。这个命题是通过web存档的具体用例来实现的。地球仪上的组织和个人将Web内容存档,并将其组合成庞大的Web存档。这些网络档案可能被用作各种环境中的研究数据集,从计算机科学到人文科学。然而,由于一些限制,包括研究的次优接入服务,它们迄今几乎没有用于研究。本章讨论了在网络档案的具体案例中吸取的经验和教训,以及它们对整个研究数据管理的影响。
In our times, researchers create and gather their datasets, irrevocably changing from solely ‘small’ to increasingly ‘big’ data. Institutions provide more and more services to manage, store and access this essential research data. However, despite the invaluable services provided by repositories, there is still a mismatch between the myriad of operations carried out by researchers, and the storage services offered by generalized repositories: often, only the end products of the research process are stored. This may mean that valuable operations and transformations of data during the research process are lost, in effect also making it harder for future researchers to interpret this data. To address this issue, and to bring researchers and institutions closer together, we propose a more collaborative approach, potentially involving researchers in their entire workflow, from data creation to use and potential reuse. This proposition is contextualized via the concrete use case of the web archive. Organizations and individuals across the globe archive web content, assembled into vast web archives. These web archives could potentially be used as research datasets in various settings, ranging from computer science to the humanities. However, they have scarcely been used for research thus far, due to a number of limitations, including suboptimal access services for research. This chapter discusses the experiences and lessons learned in the concrete case of the web archive, and their implications for research data management at large.