Providing Spatial Data for Secondary Analysis

Providing Spatial Data for Secondary Analysis
复制标题

提供空间数据进行二次分析

DOI:
--
复制
发表时间:
2009
期刊:
影响因子:
--
通讯作者:
J. Mcnally
J. Mcnally
中科院分区:
--
文献类型:
--
作者:
M. Gutmann;K. Witkowski;C. Colyer;J. Mcnally

文献摘要

被引文献

相似文献

摘要空间上明确的数据为所有参与者提供长期保存和二次分析的数据--数据生产者、数据存档者和数据使用者--带来了一系列的机遇和挑战。我们报告的机会和挑战,为每三个球员,然后转向当前的思考如何最好地准备,存档,传播和利用社会科学数据,具有空间明确的标识摘要。贯穿全文的核心问题是被调查者身份泄露的风险。如果我们知道他们住在哪里,他们在哪里工作,或者他们在哪里拥有财产,就有可能找出他们是谁。参与收集、存档和使用数据的人员需要了解披露的风险,并熟悉最佳做法,以避免对受访者有害的披露。关键词档案;保密性;数据;披露;位置本文是关于生产,存档和共享社会科学数据,其中有空间上明确的信息嵌入其中,同时避免披露个人的私人信息的风险,谁同意分享自己的信息,在调查研究的情况下,或谁是宇宙的一部分,个人包括在行政记录系统或数据库中的挑战。它以数据档案管理员的视角为出发点,但它试图保持对数据生产者、数据用户、调查受访者和数据存储库管理者的竞争利益的清晰理解,更不用说提供收集、清理、记录和传播数据所需资源的组织了。像其他关心保护调查受访者的机密性的人一样,我们敏锐地意识到,今天公开的信息财富增加了有人违反大多数社会科学数据收集时所做的保密承诺的风险。空间上明确的数据,因为它们根据定义与特定位置相关联,可能是某人的家或另一个容易识别的地方,有可能加剧这种风险。我们的目标是描述许多问题,确定一些可以保护机密性的方法,然后得出关于当前最佳实践的结论。
Abstract Spatially explicit data pose a series of opportunities and challenges for all the actors involved inproviding data for long-term preservation and secondary analysis -- the data producer, the dataarchive, and the data user. We report on opportunities and challenges for each of the three players,and then turn to a summary of current thinking about how best to prepare, archive, disseminate, andmake use of social science data that have spatially explicit identification. The core issue that runsthrough the paper is the risk of the disclosure of the identity of respondents. If we know where theylive, where they work, or where they own property, it is possible to find out who they are. Thoseinvolved in collecting, archiving, and using data need to be aware of the risks of disclosure andbecome familiar with best practices to avoid disclosures that will be harmful to respondents. Keywords archives; confidentiality; data; disclosure; locationThis paper is about the challenges involved in producing, archiving, and sharing social sciencedata that have spatially explicit information embedded within them, all while avoiding the riskof disclosing private information about the individuals who have consented to shareinformation about themselves, in the case of survey research, or who are part of the universeof individuals included in an administrative record system or database. It takes as its startingpoint the perspective of the data archivist, but it tries to maintain a clear understanding of thecompeting interests of the data producer, the data user, the survey respondent, and the managerof the data repository, not to mention whatever organization has provided the resources requiredto collect, clean, document, and disseminate the data. Like others who are concerned aboutprotecting the confidentiality of survey respondents, we are acutely aware that the wealth ofinformation publicly available today increases the risk that someone will breach the promiseof confidentiality that is made when most social science data are collected. Spatially explicitdata, because they are by definition linked to a specific location that might be someone’s homeor another easily identifiable place, have the potential to aggravate that risk. Our goal here isto describe many of the issues, identify some of the ways that confidentiality can be protected,and then draw conclusions about current best practices.